Help is available by moving the cursor above any
symbol or by checking MAQAO website.
- r0: 2x1
- r1: 2x2
- r2: 2x4
- r3: 2x8
- r4: 2x16
- r5: 2x18
- r6: 2x24
- r7: 2x32
- r8: 2x36
Metric | r0 | r1 | r2 | r3 | r4 | r5 | r6 | r7 | r8 |
---|
Total Time (s) | 42.58 | 45.05 | 45.18 | 50.79 | 67.20 | 72.67 | 91.01 | 116.40 | 129.70 |
Max (Thread Active Time) (s) | 40.88 | 43.40 | 43.48 | 48.85 | 64.61 | 69.88 | 87.81 | 112.16 | 124.49 |
Average Active Time (s) | 40.88 | 42.10 | 42.92 | 48.48 | 64.14 | 69.37 | 86.65 | 110.95 | 123.48 |
Activity Ratio (%) | 96.3 | 94.2 | 96.0 | 96.4 | 96.2 | 96.2 | 95.8 | 95.8 | 95.6 |
Average number of active threads | 1.920 | 3.738 | 7.600 | 15.272 | 30.543 | 34.363 | 45.703 | 61.000 | 68.548 |
Affinity Stability (%) | 98.4 | 98.6 | 98.6 | 98.7 | 98.4 | 98.3 | 98.3 | 98.4 | 98.2 |
GFLOPS | 39.030 | 73.595 | 147.279 | 262.135 | 396.249 | 412.213 | 438.968 | 457.606 | 461.971 |
Time in analyzed loops (%) | 65.8 | 65.1 | 64.1 | 64.1 | 63.0 | 62.6 | 61.9 | 61.2 | 60.8 |
Time in analyzed innermost loops (%) | 65.6 | 64.9 | 63.9 | 63.9 | 62.8 | 62.5 | 61.8 | 61.0 | 60.6 |
Time in user code (%) | 65.2 | 64.5 | 63.5 | 63.6 | 62.1 | 61.7 | 60.8 | 60.0 | 59.6 |
Compilation Options Score (%) | 100 | 100 | 100 | 100 | 100 | 100 | 100 | 100 | 100 |
Array Access Efficiency (%) | 95.0 | 95.2 | 94.9 | 94.7 | 93.9 | 93.6 | 93.2 | 92.9 | 93.0 |
|
Potential Speedups |
Perfect Flow Complexity | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 |
Perfect OpenMP + MPI + Pthread | 1.00 | 1.00 | 1.00 | 1.01 | 1.00 | 1.01 | 1.00 | 1.01 | 1.01 |
Perfect OpenMP + MPI + Pthread + Perfect Load Distribution | 1.00 | 1.05 | 1.04 | 1.03 | 1.03 | 1.03 | 1.03 | 1.02 | 1.02 |
Scalability - Gap | 1.00 | 1.06 | 1.06 | 1.19 | 1.58 | 1.71 | 2.14 | 2.73 | 3.05 |
No Scalar Integer | Potential Speedup | 1.04 | 1.03 | 1.04 | 1.04 | 1.04 | 1.04 | 1.05 | 1.05 | 1.05 |
Nb Loops to get 80% | 3 | 3 | 3 | 3 | 3 | 3 | 2 | 3 | 3 |
FP Vectorised | Potential Speedup | 1.12 | 1.12 | 1.12 | 1.11 | 1.09 | 1.09 | 1.08 | 1.07 | 1.07 |
Nb Loops to get 80% | 1 | 1 | 1 | 1 | 2 | 2 | 2 | 2 | 2 |
Fully Vectorised | Potential Speedup | 1.61 | 1.58 | 1.59 | 1.59 | 1.57 | 1.56 | 1.55 | 1.53 | 1.52 |
Nb Loops to get 80% | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 |
Only FP Arithmetic | Potential Speedup | 1.11 | 1.11 | 1.11 | 1.12 | 1.13 | 1.13 | 1.13 | 1.13 | 1.13 |
Nb Loops to get 80% | 5 | 5 | 5 | 5 | 5 | 5 | 5 | 5 | 5 |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
Source Object | Issue |
▼exec– | |
▼WaveFunction.cpp– | |
○ | |
▼stl_map.h– | |
○ | |
▼TwoBodyJastrowRef.h– | |
○ | |
▼NewTimer.cpp– | |
○ | |
▼SoaDistanceTableAAOMPTarget.h– | |
○ | |
▼DiracMatrix.h– | |
○ | |
▼ParticleSet.cpp– | |
○ | |
▼BsplineAllocator.hpp– | |
○ | |
▼OhmmsVector.h– | |
○ | |
▼ParticleBConds3DSoa.h– | |
○ | |
▼SPOSet.h– | |
○ | |
▼einspline_spo_ref.hpp– | |
○ | |
▼DelayedUpdate.h– | |
○ | |
▼OneBodyJastrowRef.h– | |
○ | |
▼DiracDeterminantRef.cpp– | |
○ | |
▼TimerManager.cpp– | |
○ | |
▼BsplineFunctor.h– | |
○ | |
▼SoaDistanceTableABOMPTarget.h– | |
○ | |
▼NonLocalPP.hpp– | |
○ | |
| r0 | r1 | r2 | r3 | r4 | r5 | r6 | r7 | r8 |
Application | /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/run/binaries/icx_2/exec | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Timestamp | 2025-04-08 20:05:40 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Experiment Type | MPI; | MPI; OpenMP; | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 |
Machine | itp09.benchmarkcenter.megware.com | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Architecture | x86_64 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Micro Architecture | ICELAKE_SP | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Model Name | Intel(R) Xeon(R) Platinum 8360Y CPU @ 2.40GHz | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Cache Size | 55296 KB | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Number of Cores | 36 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Maximal Frequency | 3.5 GHz | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
OS Version | Linux 5.14.0-503.16.1.el9_5.x86_64 #1 SMP PREEMPT_DYNAMIC Fri Dec 13 01:47:05 EST 2024 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Architecture used during static analysis | x86_64 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Micro Architecture used during static analysis | ICELAKE_SP | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Compilation Options |
exec: clang based Intel(R) oneAPI DPC++/C++ Compiler 2024.0.0 (2024.0.0.20231017) /cluster/intel/oneapi/2024.0.0/compiler/2024.0/bin/compiler/clang --driver-mode=g++ --intel -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/icx_2/src -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Particle -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Utilities -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Platforms -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Platforms/Host -I /cluster/intel/oneapi/2024.0.0/mpi/2021.11/include -D ADD_ -D H5_USE_16_API -D HAVE_CONFIG_H -D HAVE_MKL -D MPICH_SKIP_MPICXX -D OMPI_SKIP_MPICXX -D OPENMP_NO_COMPLEX -D _MPICC_H -D restrict=__restrict__ -isystem /cluster/intel/oneapi/2024.0.0/mkl/2024.0/include -O3 -O3 -x ICELAKE-SERVER -mprefer-vector-width=512 -g -fno-omit-frame-pointer -fcf-protection=none -nopie -grecord-command-line -fiopenmp -fstrict-aliasing -O3 -D NDEBUG -std=c++17 -MD -MT src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o -MF src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o.d -o src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o -c /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/QMCWaveFunctions/SPOSet_builder.cpp -fveclib=SVML -fheinous-gnu-extensions --driver-mode=g++ --intel -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/icx_2/src -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Particle -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Utilities -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Platforms -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Platforms/Host -I /cluster/intel/oneapi/2024.0.0/mpi/2021.11/include -D ADD_ -D H5_USE_16_API -D HAVE_CONFIG_H -D HAVE_MKL -D MPICH_SKIP_MPICXX -D OMPI_SKIP_MPICXX -D OPENMP_NO_COMPLEX -D _MPICC_H -D restrict=__restrict__ -isystem /cluster/intel/oneapi/2024.0.0/mkl/2024.0/include -O3 -O3 -x ICELAKE-SERVER -mprefer-vector-width=512 -g -fno-omit-frame-pointer -fcf-protection=none -nopie -grecord-command-line -fiopenmp -fstrict-aliasing -O3 -D NDEBUG -std=c++17 -MD -MT src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o -MF src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o.d -o src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o -c /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/QMCWaveFunctions/SPOSet_builder.cpp -fveclib=SVML -fheinous-gnu-extensions | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Number of processes observed | 2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Number of threads observed | 2 | 4 | 8 | 16 | 32 | 36 | 48 | 64 | 72 |
Frequency Driver | intel_pstate | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Frequency Governor | performance | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Huge Pages | always | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Hyperthreading | on | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Number of sockets | 2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Number of cores per socket | 36 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
MAQAO version | 2.21.1 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
MAQAO build | 8271f65b618decdd516f3bd4a943e5566ffabed6::20250211-191351 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
Comments | OV scalability run using icx_2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |