Loops
MultiBsplineRef.hpp: 68 - 163.37 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 849 | 8.68 | 8.20 | 6.80 | 100 | 100 | 314.3 | 1037 | 33.07 | 33.91 | 27.06 | 100 | 100 | 304.18 | 966 | 33.20 | 32.96 | 27.23 | 100 | 100 | 312.47 | 809 | 8.80 | 8.60 | 6.83 | 100 | 100 | 300.47 | 1037 | 33.26 | 32.66 | 27.23 | 100 | 100 | 315.91 | 966 | 33.38 | 32.96 | 27.23 | 100 | 100 | 312.18 |
| 851 | 8.65 | 8.24 | 6.84 | 100 | 100 | 313.95 | 807 | 8.69 | 8.56 | 6.80 | 100 | 100 | 300.89 | ||||||||||||||||||||||||||||
| 848 | 8.82 | 8.31 | 6.89 | 100 | 100 | 308.47 | 806 | 8.64 | 8.66 | 6.87 | 100 | 100 | 296.4 | ||||||||||||||||||||||||||||
| 850 | 8.54 | 8.21 | 6.81 | 100 | 100 | 314.98 | 808 | 8.70 | 8.54 | 6.78 | 100 | 100 | 302.69 | ||||||||||||||||||||||||||||
| Sum on 4 analyzed binary loops (exec - 849, exec - 851, exec - 848, exec - 850) | Sum on 1 analyzed binary loop (exec - 1037) | Sum on 1 analyzed binary loop (exec - 966) | Sum on 4 analyzed binary loops (exec - 809, exec - 807, exec - 806, exec - 808) | Sum on 1 analyzed binary loop (exec - 1037) | Sum on 1 analyzed binary loop (exec - 966) | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
MultiBsplineRef.hpp: 242 - 67.09 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 865 | 13.66 | 13.21 | 10.96 | 100 | 100 | 1083.5 | 1048 | 13.94 | 13.90 | 11.10 | 100 | 100 | 1029.39 | 1030 | 14.61 | 13.91 | 11.49 | 15.22 | 14.4 | 1019.77 | 823 | 13.77 | 13.81 | 10.97 | 100 | 100 | 1035.05 | 1048 | 13.81 | 13.30 | 11.09 | 100 | 100 | 1076.42 | 1030 | 14.48 | 13.91 | 11.49 | 15.22 | 14.4 | 1015.8 |
| Sum on 1 analyzed binary loop (exec - 865) | Sum on 1 analyzed binary loop (exec - 1048) | Sum on 1 analyzed binary loop (exec - 1030) | Sum on 1 analyzed binary loop (exec - 823) | Sum on 1 analyzed binary loop (exec - 1048) | Sum on 1 analyzed binary loop (exec - 1030) | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
| Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | ||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | ||||||||||||||||||||||||||||||||
| More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | ||||||||||||||||||||||||||||||||
| More than 20% of the loads are accessing the stack | 1 | More than 20% of the loads are accessing the stack | 0 | More than 20% of the loads are accessing the stack | More than 20% of the loads are accessing the stack | 1 | More than 20% of the loads are accessing the stack | 0 | More than 20% of the loads are accessing the stack | ||||||||||||||||||||||||||||||||
| Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | ||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | Presence of constant non-unit stride data access | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | Presence of constant non-unit stride data access | ||||||||||||||||||||||||||||||||||
SoaDistanceTableAAOMPTarget.h: 440 - 9.06 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1791 | 2.50 | 1.81 | 1.51 | 54.55 | 15.91 | 0 | 1976 | 2.27 | 1.86 | 1.49 | 0 | 12.5 | 0 | 2682 | 2.28 | 1.86 | 1.54 | 27.27 | 15.91 | 0 | 1689 | 2.23 | 1.90 | 1.51 | 54.55 | 15.91 | 0 | 1976 | 2.31 | 1.80 | 1.50 | 0 | 12.5 | 0 | 2682 | 2.35 | 1.85 | 1.53 | 27.27 | 15.91 | 0 |
| Sum on 1 analyzed binary loop (exec - 1791) | Sum on 1 analyzed binary loop (exec - 1976) | Sum on 1 analyzed binary loop (exec - 2682) | Sum on 1 analyzed binary loop (exec - 1689) | Sum on 1 analyzed binary loop (exec - 1976) | Sum on 1 analyzed binary loop (exec - 2682) | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
| Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | ||||||||||||||||||||||||||||||||||||
| Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | ||||||||||||||||||||||||||||||
| Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | ||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | ||||||||||||||||||||||||||||||
| Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | 0 | Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | 0 | ||||||||||||||||||||||||||||||
| More than 20% of the loads are accessing the stack | 0 | More than 20% of the loads are accessing the stack | 1 | More than 20% of the loads are accessing the stack | 0 | More than 20% of the loads are accessing the stack | 0 | More than 20% of the loads are accessing the stack | 1 | More than 20% of the loads are accessing the stack | 0 | ||||||||||||||||||||||||||||||
| Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | ||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | ||||||||||||||||||||||||||||||
| Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | 0 | Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | 0 | ||||||||||||||||||||||||||||||
BsplineFunctor.h: 236 - 7.00 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 326 | 1.44 | 1.30 | 1.08 | 75 | 76.56 | 29.7 | 463 | 1.75 | 1.60 | 1.28 | 94.12 | 94.49 | 23.73 | 248 | 0.05 | 0.02 | 0.02 | 0 | 10 | 11.56 | 258 | 1.52 | 1.41 | 1.12 | 75 | 76.56 | 25.97 | 463 | 1.57 | 1.41 | 1.17 | 94.12 | 94.49 | 28.32 | 248 | 0.06 | 0.02 | 0.02 | 0 | 10 | 11.12 |
| 228 | 0.06 | 0.03 | 0.02 | 75 | 76.56 | 12.41 | 338 | 1.60 | 1.39 | 1.15 | 0 | 10 | 19.38 | 338 | 1.57 | 1.39 | 1.15 | 0 | 10 | 19.16 | |||||||||||||||||||||
| Sum on 1 analyzed binary loop (exec - 326) | Sum on 1 analyzed binary loop (exec - 463) | Sum on 1 analyzed binary loop (exec - 338) | Sum on 1 analyzed binary loop (exec - 258) | Sum on 1 analyzed binary loop (exec - 463) | Sum on 1 analyzed binary loop (exec - 338) | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
| Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | ||||||||||||||||||||||||||||||||||||
| Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | ||||||||||||||||||||||||||||||
| Control Flow Issues | Control Flow Issues | Control Flow Issues | Control Flow Issues | Control Flow Issues | Control Flow Issues | ||||||||||||||||||||||||||||||||||||
| Presence of more than 4 paths | Presence of more than 4 paths | Presence of more than 4 paths | 1 | Presence of more than 4 paths | Presence of more than 4 paths | Presence of more than 4 paths | 1 | ||||||||||||||||||||||||||||||||||
| Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | ||||||||||||||||||||||||||||||||||||
| Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | ||||||||||||||||||||||||||||||||
| More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | ||||||||||||||||||||||||||||||||
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | ||||||||||||||||||||||||||||||||
| Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | ||||||||||||||||||||||||||||||||||||
| Presence of more than 4 paths | 0 | Presence of more than 4 paths | 0 | Presence of more than 4 paths | 1 | Presence of more than 4 paths | 0 | Presence of more than 4 paths | 0 | Presence of more than 4 paths | 1 | ||||||||||||||||||||||||||||||
| Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | 0 | Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | 0 | ||||||||||||||||||||||||||||||
| Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | ||||||||||||||||||||||||||||||||||||
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | ||||||||||||||||||||||||||||||||
| Use of masked instructions | 1 | Use of masked instructions | 1 | Use of masked instructions | Use of masked instructions | 1 | Use of masked instructions | 1 | Use of masked instructions | ||||||||||||||||||||||||||||||||
inner_product.hpp: 155 - 4.31 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1060 | 0.07 | 0.02 | 0.02 | 100 | 100 | 3895.43 | 1143 | 0.59 | 0.47 | 0.38 | 100 | 81.82 | 227.42 | 1143 | 0.56 | 0.48 | 0.40 | 100 | 100 | 222.88 | 1010 | 0.06 | 0.02 | 0.02 | 100 | 100 | 3664.88 | 1143 | 0.56 | 0.45 | 0.38 | 100 | 81.82 | 238.72 | 1143 | 0.58 | 0.48 | 0.40 | 100 | 100 | 223.21 |
| 1049 | 0.14 | 0.08 | 0.07 | 100 | 100 | 252.47 | 1145 | 0.07 | 0.03 | 0.03 | 100 | 81.82 | 3004.81 | 1152 | 0.13 | 0.09 | 0.08 | 100 | 100 | 235.73 | 1020 | 0.59 | 0.53 | 0.42 | 100 | 100 | 205.45 | 1145 | 0.08 | 0.03 | 0.03 | 100 | 81.82 | 3243.43 | 1152 | 0.13 | 0.09 | 0.07 | 100 | 100 | 240.25 |
| 1056 | 0.56 | 0.45 | 0.37 | 100 | 100 | 234.64 | 1139 | 0.06 | 0.03 | 0.02 | 100 | 100 | 3677.83 | 1006 | 0.56 | 0.47 | 0.37 | 100 | 100 | 229.88 | 1139 | 0.05 | 0.02 | 0.02 | 100 | 100 | 3914.36 | ||||||||||||||
| 1070 | 0.59 | 0.50 | 0.42 | 100 | 100 | 214.14 | 1135 | 0.61 | 0.45 | 0.37 | 100 | 100 | 234.58 | 1000 | 0.13 | 0.09 | 0.07 | 100 | 100 | 227.41 | 1135 | 0.58 | 0.46 | 0.38 | 100 | 100 | 234.12 | ||||||||||||||
| Sum on 1 analyzed binary loop (exec - 1070) | Sum on 1 analyzed binary loop (exec - 1143) | Sum on 2 analyzed binary loops (exec - 1143, exec - 1135) | Sum on 1 analyzed binary loop (exec - 1020) | Sum on 1 analyzed binary loop (exec - 1143) | Sum on 2 analyzed binary loops (exec - 1143, exec - 1135) | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
| Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | ||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | 1 | ||||||||||||||||||||||||||||||
| More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | ||||||||||||||||||||||||||||||
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | ||||||||||||||||||||||||||||||
| Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | ||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | Presence of constant non-unit stride data access | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | Presence of constant non-unit stride data access | Presence of constant non-unit stride data access | 1 | ||||||||||||||||||||||||||||||||||
| Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | ||||||||||||||||||||||||||||||||||||
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | ||||||||||||||||||||||||||||||
| Use of masked instructions | 0 | Use of masked instructions | 1 | Use of masked instructions | 0 | Use of masked instructions | 0 | Use of masked instructions | 1 | Use of masked instructions | 0 | ||||||||||||||||||||||||||||||
TwoBodyJastrowRef.h: 342 - 2.71 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 311 | 0.26 | 0.18 | 0.15 | 100 | 100 | 494.17 | 450 | 0.27 | 0.19 | 0.15 | 100 | 100 | 473.16 | 596 | 0.69 | 0.55 | 0.46 | 100 | 100 | 584.53 | 306 | 0.28 | 0.19 | 0.15 | 100 | 100 | 476.93 | 450 | 0.28 | 0.18 | 0.15 | 100 | 100 | 499.34 | 596 | 0.69 | 0.54 | 0.45 | 100 | 100 | 594.69 |
| 315 | 0.28 | 0.18 | 0.15 | 100 | 100 | 488.89 | 451 | 0.27 | 0.19 | 0.15 | 100 | 100 | 479.37 | 304 | 0.26 | 0.19 | 0.15 | 100 | 100 | 467.88 | 451 | 0.26 | 0.18 | 0.15 | 100 | 100 | 493.23 | ||||||||||||||
| 313 | 0.25 | 0.18 | 0.15 | 100 | 100 | 492.57 | 449 | 0.27 | 0.19 | 0.15 | 100 | 100 | 459.33 | 302 | 0.28 | 0.20 | 0.16 | 100 | 100 | 449.54 | 449 | 0.28 | 0.18 | 0.15 | 100 | 100 | 493.1 | ||||||||||||||
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | Sum on 1 analyzed binary loop (exec - 596) | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | Sum on 1 analyzed binary loop (exec - 596) | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
| Data Access Issues | Data Access Issues | ||||||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | ||||||||||||||||||||||||||||||||||||||
| More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | ||||||||||||||||||||||||||||||||||||||
| Vectorization Roadblocks | Vectorization Roadblocks | ||||||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | ||||||||||||||||||||||||||||||||||||||
einspline_spo_ref.hpp: 223 - 1.98 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 839 | 0.50 | 0.40 | 0.33 | 0 | 11.93 | 1.49 | 1039 | 0.70 | 0.51 | 0.41 | 20 | 13.13 | 0.48 | 1035 | 0.42 | 0.30 | 0.25 | 11.11 | 13.89 | 1.15 | 798 | 0.53 | 0.41 | 0.33 | 0 | 11.93 | 1.34 | 1039 | 0.70 | 0.49 | 0.41 | 20 | 13.13 | 0.5 | 1035 | 0.42 | 0.31 | 0.25 | 11.11 | 13.89 | 1.24 |
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | Sum on 1 analyzed binary loop (exec - 1039) | Sum on 1 analyzed binary loop (exec - 1035) | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | Sum on 1 analyzed binary loop (exec - 1039) | Sum on 1 analyzed binary loop (exec - 1035) | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
| Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | Loop Computation Issues | ||||||||||||||||||||||||||||||||||||||
| Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | ||||||||||||||||||||||||||||||||||||
| Data Access Issues | Data Access Issues | Data Access Issues | Data Access Issues | ||||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | ||||||||||||||||||||||||||||||||||
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | ||||||||||||||||||||||||||||||||||
| Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | Vectorization Roadblocks | ||||||||||||||||||||||||||||||||||||||
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | ||||||||||||||||||||||||||||||||||
| Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | Inefficient Vectorization | ||||||||||||||||||||||||||||||||||||||
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | ||||||||||||||||||||||||||||||||||
BsplineFunctor.h: 291 - 1.31 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 262 | 0.34 | 0.25 | 0.21 | 80 | 71.25 | 16.4 | 423 | 0.34 | 0.25 | 0.20 | 86.96 | 88.04 | 22.81 | 289 | 0.17 | 0.11 | 0.09 | 0 | 9.38 | 17.97 | 253 | 0.36 | 0.25 | 0.20 | 80 | 71.25 | 15.61 | 423 | 0.33 | 0.24 | 0.20 | 86.96 | 88.04 | 23.91 | 289 | 0.17 | 0.11 | 0.09 | 0 | 9.38 | 17.27 |
| 591 | 0.24 | 0.17 | 0.14 | 0 | 9.38 | 6.5 | 591 | 0.28 | 0.18 | 0.15 | 0 | 9.38 | 5.52 | ||||||||||||||||||||||||||||
| 615 | 0.05 | 0.02 | 0.02 | 0 | 9.38 | 5.47 | 615 | 0.05 | 0.02 | 0.02 | 0 | 9.38 | 4.75 | ||||||||||||||||||||||||||||
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | Sum on 1 analyzed binary loop (exec - 423) | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | Sum on 1 analyzed binary loop (exec - 423) | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
| Loop Computation Issues | Loop Computation Issues | ||||||||||||||||||||||||||||||||||||||||
| Presence of a large number of scalar integer instructions | 1 | Presence of a large number of scalar integer instructions | 1 | ||||||||||||||||||||||||||||||||||||||
| Data Access Issues | Data Access Issues | ||||||||||||||||||||||||||||||||||||||||
| Presence of indirect access | 1 | Presence of indirect access | 1 | ||||||||||||||||||||||||||||||||||||||
| More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 1 | ||||||||||||||||||||||||||||||||||||||
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | ||||||||||||||||||||||||||||||||||||||
| Vectorization Roadblocks | Vectorization Roadblocks | ||||||||||||||||||||||||||||||||||||||||
| Presence of indirect access | 1 | Presence of indirect access | 1 | ||||||||||||||||||||||||||||||||||||||
| Inefficient Vectorization | Inefficient Vectorization | ||||||||||||||||||||||||||||||||||||||||
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | ||||||||||||||||||||||||||||||||||||||
| Use of masked instructions | 1 | Use of masked instructions | 1 | ||||||||||||||||||||||||||||||||||||||
inner_product.hpp: 211 - 0.89 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1032 | 0.18 | 0.17 | 0.14 | 0 | 12.5 | 0 | 1122 | 0.15 | 0.11 | 0.08 | 85.71 | 76.79 | 0 | 1162 | 0.22 | 0.19 | 0.16 | 33.33 | 16.67 | 0 | 982 | 0.19 | 0.19 | 0.15 | 0 | 12.5 | 0 | 1122 | 0.14 | 0.10 | 0.08 | 85.71 | 76.79 | 0 | 1162 | 0.22 | 0.19 | 0.16 | 33.33 | 16.67 | 0 |
| 1121 | 0.12 | 0.07 | 0.06 | 85.71 | 76.79 | 0 | 1121 | 0.12 | 0.07 | 0.06 | 85.71 | 76.79 | 0 | ||||||||||||||||||||||||||||
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
TwoBodyJastrowRef.h: 324 - 0.76 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 318 | 0.22 | 0.15 | 0.13 | 100 | 100 | 1014.34 | 453 | 0.21 | 0.15 | 0.12 | 100 | 100 | 1027.94 | 598 | 0.27 | 0.16 | 0.13 | 100 | 100 | 992.62 | 309 | 0.24 | 0.17 | 0.13 | 100 | 100 | 942.91 | 453 | 0.22 | 0.15 | 0.12 | 100 | 100 | 1074.73 | 598 | 0.22 | 0.15 | 0.13 | 100 | 100 | 1024.22 |
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
TwoBodyJastrowRef.h: 155 - 0.53 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 295 | 0.09 | 0.04 | 0.03 | 100 | 100 | 1562.55 | 437 | 0.07 | 0.03 | 0.03 | 100 | 100 | 1898.64 | 285 | 0.24 | 0.11 | 0.09 | 100 | 100 | 1913.49 | 288 | 0.07 | 0.04 | 0.03 | 100 | 100 | 1882.1 | 437 | 0.09 | 0.04 | 0.03 | 100 | 100 | 1992.6 | 285 | 0.17 | 0.11 | 0.09 | 100 | 100 | 1899.1 |
| 297 | 0.08 | 0.04 | 0.03 | 100 | 100 | 1956.51 | 440 | 0.08 | 0.03 | 0.03 | 100 | 100 | 1985.91 | 286 | 0.08 | 0.04 | 0.03 | 100 | 100 | 1569.1 | 440 | 0.07 | 0.03 | 0.03 | 100 | 100 | 2120.26 | ||||||||||||||
| 296 | 0.07 | 0.03 | 0.03 | 100 | 100 | 1998.67 | 444 | 0.09 | 0.03 | 0.03 | 100 | 100 | 1737.22 | 287 | 0.08 | 0.04 | 0.03 | 100 | 100 | 1831.74 | 444 | 0.08 | 0.03 | 0.03 | 100 | 100 | 1766.08 | ||||||||||||||
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||
MultiBsplineRef.hpp: 276 - 0.48 %
| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_1 | Run gcc_2 | ||||||||||||||||||||||||||||||||||||
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| Loop Source Regions |
| ||||||||||||||||||||||||||||||
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 863 | 0.09 | 0.05 | 0.04 | 100 | 100 | 4220.22 | 1041 | 0.09 | 0.05 | 0.04 | 100 | 100 | 4010.17 | 1033 | 0.28 | 0.20 | 0.16 | 0 | 12.5 | 991.72 | 821 | 0.09 | 0.05 | 0.04 | 100 | 100 | 4013.79 | 1041 | 0.08 | 0.05 | 0.04 | 100 | 100 | 4287.04 | 1033 | 0.31 | 0.20 | 0.16 | 0 | 12.5 | 981.63 |
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | ||||||||||||||||||||||||||||||||||||
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | ||||||||||||||||||||||||||||||

