Help is available by moving the cursor above any
symbol or by checking MAQAO website.
- r0: 2x1
- r1: 2x2
- r2: 2x4
- r3: 2x8
- r4: 2x16
- r5: 2x18
- r6: 2x24
- r7: 2x32
- r8: 2x36
| Metric | r0 | r1 | r2 | r3 | r4 | r5 | r6 | r7 | r8 |
|---|
| Total Time (s) | 42.58 | 45.05 | 45.18 | 50.79 | 67.20 | 72.67 | 91.01 | 116.40 | 129.70 |
| Max (Thread Active Time) (s) | 40.88 | 43.40 | 43.48 | 48.85 | 64.61 | 69.88 | 87.81 | 112.16 | 124.49 |
| Average Active Time (s) | 40.88 | 42.10 | 42.92 | 48.48 | 64.14 | 69.37 | 86.65 | 110.95 | 123.48 |
| Activity Ratio (%) | 96.3 | 94.2 | 96.0 | 96.4 | 96.2 | 96.2 | 95.8 | 95.8 | 95.6 |
| Average number of active threads | 1.920 | 3.738 | 7.600 | 15.272 | 30.543 | 34.363 | 45.703 | 61.000 | 68.548 |
| Affinity Stability (%) | 98.4 | 98.6 | 98.6 | 98.7 | 98.4 | 98.3 | 98.3 | 98.4 | 98.2 |
| GFLOPS | 39.030 | 73.595 | 147.279 | 262.135 | 396.249 | 412.213 | 438.968 | 457.606 | 461.971 |
| Time in analyzed loops (%) | 65.8 | 65.1 | 64.1 | 64.1 | 63.0 | 62.6 | 61.9 | 61.2 | 60.8 |
| Time in analyzed innermost loops (%) | 65.6 | 64.9 | 63.9 | 63.9 | 62.8 | 62.5 | 61.8 | 61.0 | 60.6 |
| Time in user code (%) | 65.2 | 64.5 | 63.5 | 63.6 | 62.1 | 61.7 | 60.8 | 60.0 | 59.6 |
| Compilation Options Score (%) | 100 | 100 | 100 | 100 | 100 | 100 | 100 | 100 | 100 |
| Array Access Efficiency (%) | 95.0 | 95.2 | 94.9 | 94.7 | 93.9 | 93.6 | 93.2 | 92.9 | 93.0 |
|
| Potential Speedups |
| Perfect Flow Complexity | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 |
| Perfect OpenMP + MPI + Pthread | 1.00 | 1.00 | 1.00 | 1.01 | 1.00 | 1.01 | 1.00 | 1.01 | 1.01 |
| Perfect OpenMP + MPI + Pthread + Perfect Load Distribution | 1.00 | 1.05 | 1.04 | 1.03 | 1.03 | 1.03 | 1.03 | 1.02 | 1.02 |
| Scalability - Gap | 1.00 | 1.06 | 1.06 | 1.19 | 1.58 | 1.71 | 2.14 | 2.73 | 3.05 |
| No Scalar Integer | Potential Speedup | 1.04 | 1.03 | 1.04 | 1.04 | 1.04 | 1.04 | 1.05 | 1.05 | 1.05 |
| Nb Loops to get 80% | 3 | 3 | 3 | 3 | 3 | 3 | 2 | 3 | 3 |
| FP Vectorised | Potential Speedup | 1.12 | 1.12 | 1.12 | 1.11 | 1.09 | 1.09 | 1.08 | 1.07 | 1.07 |
| Nb Loops to get 80% | 1 | 1 | 1 | 1 | 2 | 2 | 2 | 2 | 2 |
| Fully Vectorised | Potential Speedup | 1.61 | 1.58 | 1.59 | 1.59 | 1.57 | 1.56 | 1.55 | 1.53 | 1.52 |
| Nb Loops to get 80% | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 |
| Only FP Arithmetic | Potential Speedup | 1.11 | 1.11 | 1.11 | 1.12 | 1.13 | 1.13 | 1.13 | 1.13 | 1.13 |
| Nb Loops to get 80% | 5 | 5 | 5 | 5 | 5 | 5 | 5 | 5 | 5 |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| Source Object | Issue |
| ▼exec– | |
| ▼WaveFunction.cpp– | |
| ○ | |
| ▼stl_map.h– | |
| ○ | |
| ▼TwoBodyJastrowRef.h– | |
| ○ | |
| ▼NewTimer.cpp– | |
| ○ | |
| ▼SoaDistanceTableAAOMPTarget.h– | |
| ○ | |
| ▼DiracMatrix.h– | |
| ○ | |
| ▼ParticleSet.cpp– | |
| ○ | |
| ▼BsplineAllocator.hpp– | |
| ○ | |
| ▼OhmmsVector.h– | |
| ○ | |
| ▼ParticleBConds3DSoa.h– | |
| ○ | |
| ▼SPOSet.h– | |
| ○ | |
| ▼einspline_spo_ref.hpp– | |
| ○ | |
| ▼DelayedUpdate.h– | |
| ○ | |
| ▼OneBodyJastrowRef.h– | |
| ○ | |
| ▼DiracDeterminantRef.cpp– | |
| ○ | |
| ▼TimerManager.cpp– | |
| ○ | |
| ▼BsplineFunctor.h– | |
| ○ | |
| ▼SoaDistanceTableABOMPTarget.h– | |
| ○ | |
| ▼NonLocalPP.hpp– | |
| ○ | |
| r0 | r1 | r2 | r3 | r4 | r5 | r6 | r7 | r8 |
| Application | /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/run/binaries/icx_2/exec | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Timestamp | 2025-04-08 20:05:40 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Experiment Type | MPI; | MPI; OpenMP; | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 |
| Machine | itp09.benchmarkcenter.megware.com | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Architecture | x86_64 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Micro Architecture | ICELAKE_SP | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Model Name | Intel(R) Xeon(R) Platinum 8360Y CPU @ 2.40GHz | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Cache Size | 55296 KB | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Number of Cores | 36 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Maximal Frequency | 3.5 GHz | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| OS Version | Linux 5.14.0-503.16.1.el9_5.x86_64 #1 SMP PREEMPT_DYNAMIC Fri Dec 13 01:47:05 EST 2024 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Architecture used during static analysis | x86_64 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Micro Architecture used during static analysis | ICELAKE_SP | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Compilation Options |
exec: clang based Intel(R) oneAPI DPC++/C++ Compiler 2024.0.0 (2024.0.0.20231017) /cluster/intel/oneapi/2024.0.0/compiler/2024.0/bin/compiler/clang --driver-mode=g++ --intel -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/icx_2/src -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Particle -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Utilities -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Platforms -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Platforms/Host -I /cluster/intel/oneapi/2024.0.0/mpi/2021.11/include -D ADD_ -D H5_USE_16_API -D HAVE_CONFIG_H -D HAVE_MKL -D MPICH_SKIP_MPICXX -D OMPI_SKIP_MPICXX -D OPENMP_NO_COMPLEX -D _MPICC_H -D restrict=__restrict__ -isystem /cluster/intel/oneapi/2024.0.0/mkl/2024.0/include -O3 -O3 -x ICELAKE-SERVER -mprefer-vector-width=512 -g -fno-omit-frame-pointer -fcf-protection=none -nopie -grecord-command-line -fiopenmp -fstrict-aliasing -O3 -D NDEBUG -std=c++17 -MD -MT src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o -MF src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o.d -o src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o -c /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/QMCWaveFunctions/SPOSet_builder.cpp -fveclib=SVML -fheinous-gnu-extensions --driver-mode=g++ --intel -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/icx_2/src -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Particle -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Utilities -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Platforms -I /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/Platforms/Host -I /cluster/intel/oneapi/2024.0.0/mpi/2021.11/include -D ADD_ -D H5_USE_16_API -D HAVE_CONFIG_H -D HAVE_MKL -D MPICH_SKIP_MPICXX -D OMPI_SKIP_MPICXX -D OPENMP_NO_COMPLEX -D _MPICC_H -D restrict=__restrict__ -isystem /cluster/intel/oneapi/2024.0.0/mkl/2024.0/include -O3 -O3 -x ICELAKE-SERVER -mprefer-vector-width=512 -g -fno-omit-frame-pointer -fcf-protection=none -nopie -grecord-command-line -fiopenmp -fstrict-aliasing -O3 -D NDEBUG -std=c++17 -MD -MT src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o -MF src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o.d -o src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/SPOSet_builder.cpp.o -c /beegfs/hackathon/users/eoseret/qaas_runs_CPU_8360Y/174-411-9252/intel/miniqmc/build/miniqmc/src/QMCWaveFunctions/SPOSet_builder.cpp -fveclib=SVML -fheinous-gnu-extensions | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Number of processes observed | 2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Number of threads observed | 2 | 4 | 8 | 16 | 32 | 36 | 48 | 64 | 72 |
| Frequency Driver | intel_pstate | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Frequency Governor | performance | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Huge Pages | always | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Hyperthreading | on | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Number of sockets | 2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Number of cores per socket | 36 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| MAQAO version | 2.21.1 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| MAQAO build | 8271f65b618decdd516f3bd4a943e5566ffabed6::20250211-191351 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Comments | OV scalability run using icx_2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |