options

miniqmc - MAQAO 2.17.4

Help is available by moving the cursor above any symbol or by checking MAQAO website.

Global Metrics

Total Time (s)33.31
Profiled Time (s)32.71
Time in analyzed loops (%)87.9
Time in analyzed innermost loops (%)87.5
Time in user code (%)88.1
Compilation Options Score (%)100
Perfect Flow Complexity1.00
Iterations Count1.00
Array Access Efficiency (%)92.2
Perfect OpenMP + MPI + Pthread1.01
Perfect OpenMP + MPI + Pthread + Perfect Load Distribution1.03
No Scalar IntegerPotential Speedup1.12
Nb Loops to get 80%2
FP VectorisedPotential Speedup1.37
Nb Loops to get 80%2
Fully VectorisedPotential Speedup3.52
Nb Loops to get 80%3
Data In L1 CachePotential Speedup2.46
Nb Loops to get 80%3
FP Arithmetic OnlyPotential Speedup1.17
Nb Loops to get 80%3

CQA Potential Speedups Summary

Loop Based Profile⏎

Innermost Loop Based Profile⏎

Application Categorization⏎

Compilation Options⏎

Source ObjectIssue
▼miniqmc–
○OneBodyJastrow.h
○iostream
○NonLocalPP.hpp
○einspline_spo_omp.cpp
○NewTimer.cpp
○DiracDeterminant.cpp
○ParticleSet.cpp
○SoaDistanceTableABOMPTarget.h
○WaveFunction.cpp
○stl_vector.h
○SoaDistanceTableAAOMPTarget.h
○vector.tcc
○miniqmc.cpp
○TwoBodyJastrow.h

Loop Path Count Profile⏎

Loop Iteration Count Profile⏎

Cumulated Speedup If No Scalar Integer⏎

Cumulated Speedup If FP Vectorized⏎

Cumulated Speedup If Fully Vectorized⏎

Cumulated Speedup If Data In L1⏎

Cumulated Speedup If FP Arithmetic Only⏎

Experiment Summary

Application./miniqmc
TimestampNA Universal TimestampNA
Number of processes observed1 Number of threads observed16
Experiment TypeOpenMP;
Machineskylake
Model NameIntel(R) Xeon(R) Platinum 8170 CPU @ 2.10GHz
Architecturex86_64 Micro ArchitectureSKYLAKE
Cache Size36608 KB Number of Cores26
OS VersionLinux 6.2.12-arch1-1 #1 SMP PREEMPT_DYNAMIC Thu, 20 Apr 2023 16:11:55 +0000
Architecture used during static analysisx86_64 Micro Architecture used during static analysisSKYLAKE
Frequency Driverintel_cpufreq Frequency Governorschedutil
Huge Pagesalways Hyperthreadingoff
Number of sockets2 Number of cores per socket26
Compilation Options
miniqmc: Intel(R) C++ Intel(R) 64 Compiler Classic for applications running on Intel(R) 64, Version 2021.8.0 Build 20221119_000000 -I/home/eoseret/miniqmc/src -I/home/eoseret/miniqmc/build_icc/src -I/home/eoseret/miniqmc/src/Particle -I/home/eoseret/miniqmc/src/Utilities -I/home/eoseret/miniqmc/src/Platforms -I/home/eoseret/miniqmc/src/Platforms/Host -DADD_ -DH5_USE_16_API -DHAVE_CONFIG_H -DHAVE_MKL -DOPENMP_NO_COMPLEX -isystem /opt/intel/oneapi/mkl/2023.0.0/include -fno-omit-frame-pointer -qopt-zmm-usage=high -qopenmp -Wno-deprecated -restrict -unroll -ip -qopt-prefetch -ftz -xHost -O2 -g -DNDEBUG -std=c++17 -MD -MT src/QMCWaveFunctions/CMakeFiles/qmcwfs.dir/einspline_spo_omp.cpp.o -MF CMakeFiles/qmcwfs.dir/einspline_spo_omp.cpp.o.d -o CMakeFiles/qmcwfs.dir/einspline_spo_omp.cpp.o -c
Commentsminiqmc compiled with icpc -qopt-zmm-usage=high, run on Skylake-SP using 16 threads

Configuration Summary

Dataset
Run Command<executable> -g "2 2 2"
Number Processes1
Number Nodes1
Filter{type = number ; value = 10 ; }
Profile StartNot Used
Maximal Path Number4
×