| Run orig_default | Run icx_default | Run gcc_default | Run aocc_6 | Run icx_5 | Run gcc_5 |
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-90
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 96-106
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Layout.hpp: 191-191
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/internal/Iterators.hpp: 449-449
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-89
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 96-106
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-106
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-90
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 96-106
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-90
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 96-106
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 366-366
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-106
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9338/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/TypedViewBase.hpp: 211-211
|
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 984 | 16.24 | 10.79 | 25.50 | 100 | 25 | 27.73 | 813 | 15.49 | 10.29 | 18.50 | 100 | 25 | 29.14 | 1977 | 13.84 | 12.62 | 24.91 | 100 | 25 | 23.86 | 1356 | 1.41 | 0.95 | 1.71 | 16.13 | 14.52 | 124.3 | 814 | 14.76 | 9.84 | 19.17 | 0 | 12.5 | 30.56 | 2023 | 13.49 | 11.81 | 25.56 | 0 | 12.5 | 25.37 |
| 660 | 4.12 | 3.37 | 7.97 | 0 | 12.5 | 88.81 | 957 | 11.93 | 8.16 | 14.66 | 100 | 25 | 36.7 | 1866 | 15.70 | 14.06 | 27.74 | 100 | 25 | 21.43 | 679 | 11.89 | 8.11 | 14.57 | 0 | 12.5 | 36.98 | 1089 | 0.41 | 0.27 | 0.52 | 0 | 12.5 | 67.07 | 1928 | 14.95 | 12.81 | 27.73 | 0 | 12.5 | 23.39 |
| 1089 | 0.10 | 0.07 | 0.17 | 0 | 12.5 | 254.69 | 1393 | 1.40 | 0.97 | 1.74 | 25 | 15.63 | 122.59 | 2406 | 0.76 | 0.69 | 1.35 | 6.25 | 13.28 | 86.42 | 1135 | 0.43 | 0.27 | 0.49 | 100 | 100 | 65.58 | 948 | 10.70 | 7.35 | 14.32 | 0 | 12.5 | 40.84 | 2120 | 0.50 | 0.42 | 0.90 | 0 | 12.5 | 43.36 |
| 1290 | 1.19 | 0.83 | 1.96 | 14.71 | 14.34 | 142.84 | 1105 | 0.43 | 0.28 | 0.50 | 100 | 25 | 64.73 | 2089 | 0.51 | 0.45 | 0.88 | 62.5 | 20.31 | 40.19 | 1014 | 16.19 | 10.72 | 19.27 | 100 | 100 | 27.92 | 1355 | 1.40 | 0.96 | 1.88 | 0 | 12.5 | 123.79 | 2443 | 0.63 | 0.53 | 1.14 | 7.41 | 13.43 | 112.07 |
| | 2413 | 0.74 | 0.67 | 1.33 | 6.25 | 13.28 | 88.61 | | | 2437 | 0.68 | 0.57 | 1.23 | 7.27 | 13.41 | 108.5 |
| | | | | |
| Sum on 4 analyzed binary loops (exec - 984, exec - 660, exec - 1089, exec - 1290) | Sum on 4 analyzed binary loops (exec - 813, exec - 957, exec - 1393, exec - 1105) | Sum on 5 analyzed binary loops (exec - 1977, exec - 1866, exec - 2406, exec - 2089, exec - 2413) | Sum on 4 analyzed binary loops (exec - 1356, exec - 679, exec - 1135, exec - 1014) | Sum on 4 analyzed binary loops (exec - 814, exec - 1089, exec - 948, exec - 1355) | Sum on 5 analyzed binary loops (exec - 2023, exec - 1928, exec - 2120, exec - 2443, exec - 2437) |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Loop Computation Issues | | Loop Computation Issues | | Loop Computation Issues | | Loop Computation Issues | | Loop Computation Issues | | Loop Computation Issues | |
| Presence of expensive FP instructions | 1 | Presence of expensive FP instructions | 0 | Presence of expensive FP instructions | 1 | Presence of expensive FP instructions | | Presence of expensive FP instructions | 1 | Presence of expensive FP instructions | 1 |
| Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 1 | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 1 | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 1 | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 0 | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 0 |
| Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 |
| Presence of indirect access | 1 | Presence of indirect access | 0 | Presence of indirect access | | Presence of indirect access | 0 | Presence of indirect access | 1 | Presence of indirect access | 0 |
| More than 10% of the vector loads instructions are unaligned | 0 | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 0 | More than 10% of the vector loads instructions are unaligned | 0 |
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 0 | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | 0 | Presence of special instructions executing on a single port | 0 | Presence of special instructions executing on a single port | 0 |
| More than 20% of the loads are accessing the stack | 1 | More than 20% of the loads are accessing the stack | 0 | More than 20% of the loads are accessing the stack | | More than 20% of the loads are accessing the stack | 0 | More than 20% of the loads are accessing the stack | 1 | More than 20% of the loads are accessing the stack | 0 |
| Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 |
| Presence of indirect access | 1 | Presence of indirect access | | Presence of indirect access | | Presence of indirect access | | Presence of indirect access | 1 | Presence of indirect access | 0 |
| Inefficient Vectorization | | Inefficient Vectorization | | Inefficient Vectorization | | Inefficient Vectorization | | Inefficient Vectorization | | Inefficient Vectorization | |
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | |