| Run orig_default | Run gcc_default | Run armclang_3 | Run gcc_3 |
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 108-113
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 108-113
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 108-113
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 108-113
|
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 19 | 64.76 | 61.43 | 39.50 | 0 | 50 | 66.14 | 27 | 64.28 | 61.42 | 39.13 | 75 | 87.5 | 66.12 | 32 | 64.76 | 61.62 | 39.46 | 100 | 100 | 65.83 | 28 | 64.00 | 61.02 | 38.98 | 100 | 100 | 66.63 |
| | | |
| Sum on 1 analyzed binary loop (exec - 19) | Sum on 1 analyzed binary loop (exec - 27) | Sum on 1 analyzed binary loop (exec - 32) | Sum on 1 analyzed binary loop (exec - 28) |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | |
| Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 |
| Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | |
| Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 |
| Run orig_default | Run gcc_default | Run armclang_3 | Run gcc_3 |
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 86-90
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 86-90
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 86-90
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 86-90
|
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 16 | 59.39 | 58.03 | 37.32 | 4.35 | 52.17 | 174.94 | 22 | 59.69 | 58.93 | 37.54 | 88.46 | 94.23 | 183.77 | 27 | 60.69 | 59.48 | 38.09 | 94.59 | 97.3 | 170.36 | 23 | 59.36 | 58.67 | 37.48 | 100 | 100 | 184.68 |
| | | |
| Sum on 1 analyzed binary loop (exec - 16) | Sum on 1 analyzed binary loop (exec - 22) | Sum on 1 analyzed binary loop (exec - 27) | Sum on 1 analyzed binary loop (exec - 23) |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 |
| Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 |
| Run orig_default | Run gcc_default | Run armclang_3 | Run gcc_3 |
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 128-131
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 128-131
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 128-131
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 128-131
|
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 23 | 26.96 | 26.29 | 16.91 | 100 | 100 | 51.51 | 31 | 26.88 | 26.19 | 16.68 | 100 | 100 | 51.79 | 37 | 26.50 | 25.74 | 16.48 | 100 | 100 | 52.68 | 32 | 26.88 | 26.36 | 16.84 | 100 | 100 | 51.45 |
| | | |
| Sum on 1 analyzed binary loop (exec - 23) | Sum on 1 analyzed binary loop (exec - 31) | Sum on 1 analyzed binary loop (exec - 37) | Sum on 1 analyzed binary loop (exec - 32) |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | |
| Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 |
| Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | |
| Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 |
| Run orig_default | Run gcc_default | Run armclang_3 | Run gcc_3 |
| Loop Source Regions | | Loop Source Regions | | Loop Source Regions | | Loop Source Regions | |
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 0 | 0.04 | 0.01 | 0.01 | 0 | 0 | 6.84 | 4 | 0.03 | 0.01 | 0.00 | 0 | 0 | 70.34 | 0 | 0.07 | 0.01 | 0.00 | 0 | 0 | 6.23 | 5 | 0.03 | 0.00 | 0.00 | 0 | 0 | 58.61 |
| 132 | 0.03 | 0.00 | 0.00 | 0 | 0 | 639.96 | 8 | 0.02 | 0.00 | 0.00 | 0 | 0 | 0 | 192 | 0.04 | 0.02 | 0.01 | 0 | 0 | 90.04 | 161 | 0.03 | 0.00 | 0.00 | 0 | 0 | 8.78 |
| 49 | 0.02 | 0.00 | 0.00 | 0 | 0 | 72 | 15 | 0.03 | 0.00 | 0.00 | 0 | 0 | 4.94 | 184 | 0.02 | 0.00 | 0.00 | 0 | 0 | 0 | 153 | 0.03 | 0.01 | 0.01 | 0 | 0 | 99.08 |
| 76 | 0.03 | 0.01 | 0.00 | 0 | 0 | 0 | 19 | 0.03 | 0.00 | 0.00 | 0 | 0 | 0 | 205 | 0.02 | 0.00 | 0.00 | 0 | 0 | 953.96 | 167 | 0.03 | 0.00 | 0.00 | 0 | 0 | 815.21 |
| 11 | 0.03 | 0.00 | 0.00 | 0 | 0 | 77.79 | 140 | 0.02 | 0.00 | 0.00 | 0 | 0 | 0 | 94 | 0.07 | 0.00 | 0.00 | 0 | 0 | 198.85 | 8 | 0.02 | 0.00 | 0.00 | 0 | 0 | 0 |
| 131 | 0.02 | 0.00 | 0.00 | 0 | 0 | 15 | 155 | 0.02 | 0.00 | 0.00 | 0 | 0 | 947.55 | 203 | 0.03 | 0.00 | 0.00 | 0 | 0 | 2.88 | 20 | 0.04 | 0.00 | 0.00 | 0 | 0 | 0.31 |
| 79 | 0.05 | 0.00 | 0.00 | 0 | 0 | 0 | 95 | 0.04 | 0.01 | 0.00 | 0 | 0 | 0 | 18 | 0.04 | 0.01 | 0.00 | 0 | 0 | 64.43 | 98 | 0.04 | 0.00 | 0.00 | 0 | 0 | 0 |
| 8 | 0.02 | 0.00 | 0.00 | 0 | 0 | 0 | 143 | 0.04 | 0.01 | 0.01 | 0 | 0 | 95.35 | 112 | 0.04 | 0.01 | 0.00 | 0 | 0 | 0 | 16 | 0.03 | 0.01 | 0.00 | 0 | 0 | 5.86 |
| 63 | 0.07 | 0.00 | 0.00 | 0 | 0 | 214.14 | 97 | 0.03 | 0.00 | 0.00 | 0 | 0 | 0 | 116 | 0.03 | 0.00 | 0.00 | 0 | 0 | 0 | 83 | 0.06 | 0.00 | 0.00 | 0 | 0 | 247.03 |
| 125 | 0.04 | 0.01 | 0.01 | 0 | 0 | 100.35 | 150 | 0.02 | 0.00 | 0.00 | 0 | 0 | 5.96 | | 96 | 0.03 | 0.00 | 0.00 | 0 | 0 | 0 |
| 77 | 0.07 | 0.00 | 0.00 | 0 | 0 | 194.11 | | |
| | | |
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Run orig_default | Run gcc_default | Run armclang_3 | Run gcc_3 |
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 62-68
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 62-68
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 62-68
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 62-68
|
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 13 | 0.04 | 0.03 | 0.02 | 3.7 | 51.85 | 89.05 | 18 | 0.04 | 0.02 | 0.01 | 0 | 50 | 95.24 | 23 | 0.04 | 0.02 | 0.01 | 97.5 | 98.75 | 97.1 | 19 | 0.04 | 0.02 | 0.01 | 0 | 50 | 104.38 |
| | | |
| Sum on 1 analyzed binary loop (exec - 13) | Sum on 1 analyzed binary loop (exec - 18) | Sum on 1 analyzed binary loop (exec - 23) | Sum on 1 analyzed binary loop (exec - 19) |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Data Access Issues | | Data Access Issues | | Data Access Issues | | Data Access Issues | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 |
| Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 |
| Run orig_default | Run gcc_default | Run armclang_3 | Run gcc_3 |
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/local_halos.cpp: 13-15
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/local_halos.cpp: 13-15
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/local_halos.cpp: 13-15
| Loop Source Regions | |
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 70 | 0.07 | 0.03 | 0.02 | 0 | 50 | 0 | 86 | 0.06 | 0.03 | 0.02 | 0 | 50 | 0 | 105 | 0.09 | 0.03 | 0.02 | 0 | 50 | 0 | |
| | | |
| Sum on 1 analyzed binary loop (exec - 70) | Sum on 1 analyzed binary loop (exec - 86) | Sum on 1 analyzed binary loop (exec - 105) | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Data Access Issues | | Data Access Issues | | Data Access Issues | | | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | | |
| Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | | |
| Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | 1 | | |
| Run orig_default | Run gcc_default | Run armclang_3 | Run gcc_3 |
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/local_halos.cpp: 28-30
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/local_halos.cpp: 28-30
| Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/local_halos.cpp: 28-30
| Loop Source Regions | |
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 72 | 0.05 | 0.02 | 0.01 | 0 | 50 | 0 | 90 | 0.05 | 0.02 | 0.01 | 0 | 50 | 0 | 108 | 0.09 | 0.02 | 0.01 | 0 | 50 | 0 | |
| | | |
| Sum on 1 analyzed binary loop (exec - 72) | Sum on 1 analyzed binary loop (exec - 90) | Sum on 1 analyzed binary loop (exec - 108) | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| Data Access Issues | | Data Access Issues | | Data Access Issues | | | |
| Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | | |
| Vectorization Roadblocks | | Vectorization Roadblocks | | Vectorization Roadblocks | | | |
| Presence of constant non-unit stride data access | | Presence of constant non-unit stride data access | 1 | Presence of constant non-unit stride data access | | | |
| Run orig_default | Run gcc_default | Run armclang_3 | Run gcc_3 |
| Loop Source Regions | | Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 105-105
| Loop Source Regions | | Loop Source Regions | - /home/eoseret/qaas/qaas_runs/178-550-2839/intel/TeaLeaf/build/TeaLeaf/src/omp/cg.cpp: 105-105
|
| ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s | ASM Loop ID | Max Time Over Threads (s) | Time w.r.t. Wall Time (s) | Cov (%) | Vect. Ratio (%) | Vector Length Use (%) | GFLOP/s |
| 24 | 0.09 | 0.02 | 0.01 | 0 | 50 | 0.08 | | 25 | 0.06 | 0.02 | 0.01 | 0 | 50 | 0.3 |
| | | |
| No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | Sum on 1 analyzed binary loop (exec - 24) | No Loops Overview analysis found for any assembly loop. More loops can be analyzed using option --summary-loop-count. | Sum on 1 analyzed binary loop (exec - 25) |
| Analysis | Count | Analysis | Count | Analysis | Count | Analysis | Count |
| | Loop Computation Issues | | | | Loop Computation Issues | |
| | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 1 | | | Less than 10% of the FP ADD/SUB/MUL arithmetic operations are performed using FMA | 1 |