Float Sum combine1: Maximum use of data abstraction: 21.22 cycles/element Float Sum combine2: Take vec_length() out of loop: 15.35 cycles/element Float Sum combine3: Array reference to vector data: 3.38 cycles/element Float Sum combine4: Array reference, accumulate in temporary: 3.33 cycles/element Float Sum combine4p: Pointer reference, accumulate in temporary: 3.33 cycles/element Float Sum Array code, unrolled by 2: 3.00 cycles/element Float Sum combine5p: Pointer code, unrolled by 3, for loop: 3.00 cycles/element Function Array code, unrolled by 3, while loop, Should be 18, Got 17 Float Sum Array code, unrolled by 3, while loop: 3.12 cycles/element Float Sum Array code, unrolled by 4: 3.08 cycles/element Float Sum Array code, unrolled by 8: 3.10 cycles/element Float Sum Array code, unrolled by 16: 3.00 cycles/element Float Sum Pointer code, unrolled by 2: 3.00 cycles/element Float Sum Pointer code, unrolled by 3: 3.00 cycles/element Float Sum Pointer code, unrolled by 4: 3.08 cycles/element Float Sum Pointer code, unrolled by 8: 3.06 cycles/element Float Sum Pointer code, unrolled by 16: 3.02 cycles/element Float Sum combine6: Array code, unrolled by 2, Superscalar x2: 1.81 cycles/element Float Sum Array code, unrolled by 4, Superscalar x2: 1.64 cycles/element Float Sum Array code, unrolled by 8, Superscalar x2: 1.58 cycles/element Float Sum Array code, unrolled by 3, Superscalar x3: 1.67 cycles/element Float Sum Array code, unrolled by 4, Superscalar x4: 1.50 cycles/element Float Sum Array code, unrolled by 8, Superscalar x4: 1.44 cycles/element Float Sum Array code, unrolled by 6, Superscalar x6: 1.34 cycles/element Float Sum Array code, unrolled by 8, Superscalar x8: 1.65 cycles/element Float Sum Array code, unrolled by 10, Superscalar x10: 1.59 cycles/element Float Sum Array code, unrolled by 12, Superscalar x6: 1.47 cycles/element Float Sum Array code, unrolled by 12, Superscalar x12: 1.48 cycles/element Float Sum Pointer code, unrolled by 8, Superscalar x2: 1.58 cycles/element Float Sum Pointer code, unrolled by 8, Superscalar x4: 1.37 cycles/element Float Sum Pointer code, unrolled by 8, Superscalar x8: 1.40 cycles/element Float Sum Pointer code, unrolled by 9, Superscalar x3: 1.34 cycles/element Float Sum Array code, Unroll x2, Superscalar x2, noninterleaved: 1.80 cycles/element Float Sum Array code, unrolled by 2, different associativity: 1.83 cycles/element Float Sum Array code, unrolled by 3, Different Associativity: 1.58 cycles/element Float Sum Array code, unrolled by 4, Different Associativity: 1.50 cycles/element Float Sum Array code, unrolled by 6, Different Associativity: 1.42 cycles/element Float Sum Array code, unrolled by 8, Different Associativity: 1.32 cycles/element Float Sum SSE code, 1*VSIZE-way parallelism: 0.89 cycles/element Float Sum SSE code, 2*VSIZE-way parallelism: 0.72 cycles/element Float Sum SSE code, 4*VSIZE-way parallelism: 0.74 cycles/element Float Sum SSE code, 8*VSIZE-way parallelism: 0.58 cycles/element Float Sum SSE code, 12*VSIZE-way parallelism: 0.70 cycles/element Float Sum SSE code, 2*VSIZE-way parallelism, reassociate: 0.84 cycles/element Float Sum SSE code, 4*VSIZE-way parallelism, reassociate: 0.70 cycles/element Float Sum SSE code, 8*VSIZE-way parallelism, reassociate: 0.68 cycles/element