| Run orig_default | Run aocc_default | Run gcc_default | Run icx_10 | Run aocc_6 | Run gcc_9 |
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/internal/Iterators.hpp: 449-449
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Layout.hpp: 191-191
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-89
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 96-106
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-90
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 96-106
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-106
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-90
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 96-106
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-90
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 96-106
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
| Loop Source Regions | - /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/Population.cpp: 58-58
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/pattern/kernel/For.hpp: 142-142
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 369-369
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/tpl/raja/include/RAJA/util/Operators.hpp: 366-366
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LTimes.cpp: 62-62
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/LPlusTimes.cpp: 57-57
- /beegfs/hackathon/users/eoseret/qaas_runs_test/178-668-9279/intel/Kripke/build/Kripke/src/Kripke/Kernel/SweepSubdomain.cpp: 88-106
|
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 813 | 4.46 | 3.44 | 9.12 | 100 | 25 | 88.41 | 984 | 4.70 | 3.59 | 10.30 | 100 | 25 | 84.16 | 1977 | 3.94 | 5.56 | 15.87 | 100 | 25 | 53.98 | 819 | 4.40 | 3.16 | 8.23 | 0 | 12.5 | 90.77 | 1149 | 0.09 | 0.06 | 0.17 | 100 | 50 | 364.07 | 2186 | 0.25 | 0.28 | 0.77 | 100 | 25 | 60.28 |
| 957 | 3.42 | 3.14 | 8.33 | 100 | 25 | 94.39 | 660 | 3.15 | 3.06 | 8.79 | 0 | 12.5 | 98.21 | 1866 | 1.69 | 2.61 | 7.46 | 100 | 25 | 116.56 | 1026 | 2.60 | 2.34 | 6.10 | 0 | 12.5 | 84.83 | 1360 | 0.63 | 0.54 | 1.45 | 16.13 | 14.52 | 213.31 | 2073 | 4.61 | 5.53 | 15.51 | 100 | 25 | 54.42 |
| 1105 | 0.14 | 0.11 | 0.30 | 100 | 25 | 159.19 | 1089 | 0.10 | 0.08 | 0.22 | 0 | 12.5 | 211.44 | 2406 | 0.34 | 0.41 | 1.17 | 6.25 | 13.28 | 151.29 | 1530 | 0.65 | 0.58 | 1.52 | 0 | 12.5 | 198.69 | 689 | 3.06 | 2.98 | 7.99 | 0 | 12.5 | 100.73 | 1962 | 1.74 | 2.45 | 6.88 | 100 | 25 | 121.14 |
| 1393 | 0.69 | 0.61 | 1.61 | 25 | 15.63 | 191.16 | 1290 | 0.60 | 0.54 | 1.56 | 14.71 | 14.34 | 212.39 | 2089 | 0.16 | 0.18 | 0.52 | 62.5 | 20.31 | 66.49 | 1201 | 0.15 | 0.10 | 0.27 | 0 | 12.5 | 176.3 | 1029 | 5.53 | 3.99 | 10.70 | 100 | 50 | 72.76 | 2509 | 0.38 | 0.43 | 1.22 | 7.41 | 13.43 | 145.3 |
| | 2413 | 0.31 | 0.39 | 1.10 | 6.25 | 13.28 | 150.85 | 1025 | 1.38 | 1.04 | 2.71 | 0 | 12.5 | 91.33 | | 2503 | 0.35 | 0.43 | 1.20 | 7.27 | 13.41 | 158.19 |
| | | 818 | 0.36 | 0.18 | 0.48 | 0 | 12.5 | 88.16 | | |
| | | | | |
| Sum on 4 analyzed binary loops (exec - 813, exec - 957, exec - 1105, exec - 1393) | Sum on 4 analyzed binary loops (exec - 984, exec - 660, exec - 1089, exec - 1290) | Sum on 5 analyzed binary loops (exec - 1977, exec - 1866, exec - 2406, exec - 2089, exec - 2413) | Sum on 6 analyzed binary loops (exec - 819, exec - 1026, exec - 1530, exec - 1201, exec - 1025, exec - 818) | Sum on 4 analyzed binary loops (exec - 1149, exec - 1360, exec - 689, exec - 1029) | Sum on 5 analyzed binary loops (exec - 2186, exec - 2073, exec - 1962, exec - 2509, exec - 2503) |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Loop Computation Issues | | Loop Computation Issues | | Loop Computation Issues | | Loop Computation Issues | | Loop Computation Issues | | Loop Computation Issues | |
| Presence of expensive FP instructions | 1 | Presence of expensive FP instructions | 1 | Presence of expensive FP instructions | 1 | Presence of expensive FP instructions | | Presence of expensive FP instructions | | Presence of expensive FP instructions | 1 |
| Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 1 | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 1 | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 1 | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 0 |
| Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 0 | Presence of constant non-unit stride data access | 1 |
| Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | | Presence of indirect access | | Presence of indirect access | 0 | Presence of indirect access | 0 |
| More than 10% of the vector loads instructions are unaligned | 0 | More than 10% of the vector loads instructions are unaligned | 0 | More than 10% of the vector loads instructions are unaligned | | More than 10% of the vector loads instructions are unaligned | | More than 10% of the vector loads instructions are unaligned | 1 | More than 10% of the vector loads instructions are unaligned | 0 |
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | 0 | Presence of special instructions executing on a single port | 0 |
| More than 20% of the loads are accessing the stack | 1 | More than 20% of the loads are accessing the stack | 1 | More than 20% of the loads are accessing the stack | | More than 20% of the loads are accessing the stack | | More than 20% of the loads are accessing the stack | 0 | More than 20% of the loads are accessing the stack | 0 |
| Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 |
| Presence of indirect access | 1 | Presence of indirect access | 1 | Presence of indirect access | | Presence of indirect access | | Presence of indirect access | | Presence of indirect access | 0 |
| Inefficient Vectorization | | Inefficient Vectorization | | Inefficient Vectorization | | Inefficient Vectorization | | Inefficient Vectorization | | Inefficient Vectorization | |
| Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | 1 | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | | Presence of special instructions executing on a single port | |