Help is available by moving the cursor above any  symbol or by checking MAQAO website.
 symbol or by checking MAQAO website.
  - r0: run_1_thread
- r1: run_2_threads
- r2: run_4_threads
- r3: run_8_threads
- r4: run_16_threads
- r5: run_32_threads
- r6: run_64_threads
- r7: run_96_threads
| Metric | r0 | r1 | r2 | r3 | r4 | r5 | r6 | r7 | 
|---|
| Total Time (s) | 107.35 | 54.13 | 27.49 | 14.09 | 7.12 | 3.69 | 1.92 | 2.04 | 
| Max (Thread Active Time) (s) | 99.82 | 49.96 | 25.32 | 13.00 | 6.56 | 3.30 | 1.47 | 0.80 | 
| Average Active Time (s) | 99.82 | 49.95 | 25.03 | 12.63 | 6.34 | 3.20 | 1.35 | 0.69 | 
| Activity Ratio (%) | 93.0 | 92.3 | 91.1 | 89.7 | 89.2 | 87.0 | 70.8 | 34.1 | 
| Average number of active threads | 0.930 | 1.845 | 3.642 | 7.172 | 14.260 | 27.792 | 45.092 | 32.499 | 
| Affinity Stability (%) | 99.7 | 99.7 | 99.5 | 99.4 | 99.2 | 98.8 | 97.8 | 96.4 | 
| GFLOPS | 4.314 | 8.562 | 16.920 | 33.388 | 66.576 | 129.536 | 212.494 | 154.053 | 
| Time in analyzed loops (%) | 100.0 | 100.0 | 100.0 | 100.0 | 100.0 | 99.9 | 99.6 | 99.1 | 
| Time in analyzed innermost loops (%) | 99.1 | 99.1 | 99.0 | 99.1 | 99.2 | 99.0 | 98.7 | 98.2 | 
| Time in user code (%) | 100 | 100 | 100.0 | 100.0 | 100.0 | 99.9 | 99.6 | 99.1 | 
| Compilation Options Score (%) | 62.5 | 62.5 | 62.5 | 62.5 | 62.5 | 62.5 | 62.5 | 62.5 | 
| Array Access Efficiency (%) | 99.6 | 99.5 | 99.5 | 99.5 | 99.6 | 99.5 | 99.5 | 99.5 | 
|  | 
| Potential Speedups |  | 
| Perfect Flow Complexity | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 
| Perfect OpenMP/MPI/Pthread/TBB | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.01 | 1.01 | 
| Perfect OpenMP/MPI/Pthread/TBB + Perfect Load Distribution | 1.00 | 1.00 | 1.01 | 1.03 | 1.04 | 1.03 | 1.09 | 1.17 | 
| Scalability - Gap | 1.00 | 1.01 | 1.02 | 1.05 | 1.06 | 1.10 | 1.14 | 1.82 | 
| No Scalar Integer | Potential Speedup | 1.37 | 1.37 | 1.37 | 1.37 | 1.37 | 1.36 | 1.33 | 1.33 | 
| Nb Loops to get 80% | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 
| FP Vectorised | Potential Speedup | 1.37 | 1.37 | 1.37 | 1.37 | 1.37 | 1.36 | 1.33 | 1.33 | 
| Nb Loops to get 80% | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 
| Fully Vectorised | Potential Speedup | 1.31 | 1.31 | 1.31 | 1.31 | 1.31 | 1.30 | 1.28 | 1.28 | 
| Nb Loops to get 80% | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 
| Only FP Arithmetic | Potential Speedup | 1.37 | 1.37 | 1.37 | 1.37 | 1.37 | 1.36 | 1.33 | 1.33 | 
| Nb Loops to get 80% | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 
| Source Object | Issue | 
|---|
| ▼kmeans-gcc-Ofast– |  | 
| ▼main.cpp– |  | 
| ○ | -O3 or -Ofast is missing. | 
| ○ | -funroll-loops is missing. | 
 
 
| Source Object | Issue | 
|---|
| ▼kmeans-gcc-Ofast– |  | 
| ▼main.cpp– |  | 
| ○ | -O3 or -Ofast is missing. | 
| ○ | -funroll-loops is missing. | 
 
 
| Source Object | Issue | 
|---|
| ▼kmeans-gcc-Ofast– |  | 
| ▼main.cpp– |  | 
| ○ | -O3 or -Ofast is missing. | 
| ○ | -funroll-loops is missing. | 
 
 
| Source Object | Issue | 
|---|
| ▼kmeans-gcc-Ofast– |  | 
| ▼main.cpp– |  | 
| ○ | -O3 or -Ofast is missing. | 
| ○ | -funroll-loops is missing. | 
 
 
| Source Object | Issue | 
|---|
| ▼kmeans-gcc-Ofast– |  | 
| ▼main.cpp– |  | 
| ○ | -O3 or -Ofast is missing. | 
| ○ | -funroll-loops is missing. | 
 
 
| Source Object | Issue | 
|---|
| ▼kmeans-gcc-Ofast– |  | 
| ▼main.cpp– |  | 
| ○ | -O3 or -Ofast is missing. | 
| ○ | -funroll-loops is missing. | 
 
 
| Source Object | Issue | 
|---|
| ▼kmeans-gcc-Ofast– |  | 
| ▼main.cpp– |  | 
| ○ | -O3 or -Ofast is missing. | 
| ○ | -funroll-loops is missing. | 
 
 
| Source Object | Issue | 
|---|
| ▼kmeans-gcc-Ofast– |  | 
| ▼main.cpp– |  | 
| ○ | -O3 or -Ofast is missing. | 
| ○ | -funroll-loops is missing. | 
 
 
 
|  | r0 | r1 | r2 | r3 | r4 | r5 | r6 | r7 | 
|---|
| Experiment Name | K-Means scalability gcc-Ofast 100000000 | K-Means scalability gcc-Ofast 100000000 | K-Means scalability gcc-Ofast 100000000 | K-Means scalability gcc-Ofast 100000000 | K-Means scalability gcc-Ofast 100000000 | K-Means scalability gcc-Ofast 100000000 | K-Means scalability gcc-Ofast 100000000 | K-Means scalability gcc-Ofast 100000000 | 
|---|
| Application | ./kmeans/kmeans-gcc-Ofast | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Timestamp | 2025-07-17 11:01:26 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Experiment Type | Sequential | OpenMP; | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | 
|---|
| Machine | ip-172-31-47-249.ec2.internal | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Architecture | aarch64 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Micro Architecture | ARM_NEOVERSE_V2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Model Name |  |  |  |  |  |  |  |  | 
|---|
| Cache Size |  |  |  |  |  |  |  |  | 
|---|
| Number of Cores |  |  |  |  |  |  |  |  | 
|---|
| Maximal Frequency | 0 GHz | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| OS Version | Linux 6.1.109-118.189.amzn2023.aarch64 #1 SMP Tue Sep 10 08:58:40 UTC 2024 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Architecture used during static analysis | aarch64 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Micro Architecture used during static analysis | ARM_NEOVERSE_V2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Compilation Options | kmeans-gcc-Ofast: GNU C++14 14.2.0 -mlittle-endian -mabi=lp64 -mcpu=neoverse-v2+crc+sve2-aes+sve2-sha3+nossbs -g -Ofast -std=c++14 -fno-omit-frame-pointer -fopenmp GNU C17 14.2.0 -mlittle-endian -mabi=lp64 -g -g -g -O2 -O2 -O2 -fbuilding-libgcc -fno-stack-protector -fPIC | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Number of processes observed | 1 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Number of threads observed | 1 | 2 | 4 | 8 | 16 | 32 | 64 | 96 | 
|---|
| Frequency Driver | NA | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Frequency Governor | NA | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Huge Pages | madvise | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Hyperthreading | off | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Number of sockets | 1 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Number of cores per socket | 96 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| MAQAO version | 2025.1.1 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| MAQAO build | 2302fb4b01f3b07cccb215042f4e5c7e9fcc3718::20250717-122749 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|
| Comments | AWS Graviton 4 (Neoverse V2) CPU, 1-96 threads runs | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | 
|---|