Amazon AWS Graviton3E vs. Graviton 2/3 benchmarks

Benchmarks by Michael Larabel for a future article on Phoronix.com.

HTML result view exported from: https://openbenchmarking.org/result/2405291-NE-2308110NE13&grr&sor.

Amazon AWS Graviton3E vs. Graviton 2/3 benchmarksProcessorMotherboardChipsetMemoryDiskNetworkGraphicsAudioMonitorOSKernelCompilerFile-SystemSystem LayerVulkanDisplay ServerDisplay DriverOpenCLScreen Resolutionm7g.16xlarge Graviton3c6g.16xlarge Graviton2c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3egeo-07ARMv8 Neoverse-V1 (64 Cores)Amazon EC2 m7g.16xlarge (1.0 BIOS)Amazon Device 0200256GB215GB Amazon Elastic Block StoreAmazon ElasticUbuntu 22.045.19.0-1025-aws (aarch64)GCC 11.3.0ext4amazonARMv8 Neoverse-N1 (64 Cores)Amazon EC2 c6g.16xlarge (1.0 BIOS)128GBARMv8 Neoverse-V1 (64 Cores)Amazon EC2 c7g.16xlarge (1.0 BIOS)Amazon EC2 c7gn.16xlarge (1.0 BIOS)AMD EPYC 7R13 (32 Cores / 64 Threads)Amazon EC2 c6a.16xlarge (1.0 BIOS)Intel 440FX 82441FX PMC322GB Amazon Elastic Block Store5.19.0-1025-aws (x86_64)1.3.238GCC 11.4.02 x Intel Xeon Silver 4208 @ 3.20GHz (16 Cores / 32 Threads)Dell Precision 7920 Rack 0DY2X0 (2.21.2 BIOS)Intel Sky Lake-E DMI3 Registers64GB2000GB TOSHIBA DT01ACA2Matrox G200eW3 15GBNVIDIA TU104 HD AudioDELL 17FP4 x Intel I350Debian 115.10.0-28-amd64 (x86_64)X ServerNVIDIAOpenCL 3.0 CUDA 12.2.1381.3.242GCC 10.2.1 20210110 + Clang 11.0.1-2 + CUDA 11.21280x1024OpenBenchmarking.orgKernel Details- m7g.16xlarge Graviton3: Transparent Huge Pages: madvise- c6g.16xlarge Graviton2: Transparent Huge Pages: madvise- c7g.16xlarge Graviton3: Transparent Huge Pages: madvise- c7gn.16xlarge Graviton3E: Transparent Huge Pages: madvise- c6a.16xlarge AMD Zen 3: Transparent Huge Pages: madvise- egeo-07: Transparent Huge Pages: alwaysCompiler Details- m7g.16xlarge Graviton3: --build=aarch64-linux-gnu --disable-libquadmath --disable-libquadmath-support --disable-werror --enable-bootstrap --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-fix-cortex-a53-843419 --enable-gnu-unique-object --enable-languages=c,ada,c++,go,d,fortran,objc,obj-c++,m2 --enable-libphobos-checking=release --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-link-serialization=2 --enable-multiarch --enable-nls --enable-objc-gc=auto --enable-plugin --enable-shared --enable-threads=posix --host=aarch64-linux-gnu --program-prefix=aarch64-linux-gnu- --target=aarch64-linux-gnu --with-build-config=bootstrap-lto-lean --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-target-system-zlib=auto -v - c6g.16xlarge Graviton2: --build=aarch64-linux-gnu --disable-libquadmath --disable-libquadmath-support --disable-werror --enable-bootstrap --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-fix-cortex-a53-843419 --enable-gnu-unique-object --enable-languages=c,ada,c++,go,d,fortran,objc,obj-c++,m2 --enable-libphobos-checking=release --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-link-serialization=2 --enable-multiarch --enable-nls --enable-objc-gc=auto --enable-plugin --enable-shared --enable-threads=posix --host=aarch64-linux-gnu --program-prefix=aarch64-linux-gnu- --target=aarch64-linux-gnu --with-build-config=bootstrap-lto-lean --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-target-system-zlib=auto -v - c7g.16xlarge Graviton3: --build=aarch64-linux-gnu --disable-libquadmath --disable-libquadmath-support --disable-werror --enable-bootstrap --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-fix-cortex-a53-843419 --enable-gnu-unique-object --enable-languages=c,ada,c++,go,d,fortran,objc,obj-c++,m2 --enable-libphobos-checking=release --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-link-serialization=2 --enable-multiarch --enable-nls --enable-objc-gc=auto --enable-plugin --enable-shared --enable-threads=posix --host=aarch64-linux-gnu --program-prefix=aarch64-linux-gnu- --target=aarch64-linux-gnu --with-build-config=bootstrap-lto-lean --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-target-system-zlib=auto -v - c7gn.16xlarge Graviton3E: --build=aarch64-linux-gnu --disable-libquadmath --disable-libquadmath-support --disable-werror --enable-bootstrap --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-fix-cortex-a53-843419 --enable-gnu-unique-object --enable-languages=c,ada,c++,go,d,fortran,objc,obj-c++,m2 --enable-libphobos-checking=release --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-link-serialization=2 --enable-multiarch --enable-nls --enable-objc-gc=auto --enable-plugin --enable-shared --enable-threads=posix --host=aarch64-linux-gnu --program-prefix=aarch64-linux-gnu- --target=aarch64-linux-gnu --with-build-config=bootstrap-lto-lean --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-target-system-zlib=auto -v - c6a.16xlarge AMD Zen 3: --build=x86_64-linux-gnu --disable-vtable-verify --disable-werror --enable-bootstrap --enable-cet --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-gnu-unique-object --enable-languages=c,ada,c++,go,brig,d,fortran,objc,obj-c++,m2 --enable-libphobos-checking=release --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-link-serialization=2 --enable-multiarch --enable-multilib --enable-nls --enable-objc-gc=auto --enable-offload-targets=nvptx-none=/build/gcc-11-XeT9lY/gcc-11-11.4.0/debian/tmp-nvptx/usr,amdgcn-amdhsa=/build/gcc-11-XeT9lY/gcc-11-11.4.0/debian/tmp-gcn/usr --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --program-prefix=x86_64-linux-gnu- --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-build-config=bootstrap-lto-lean --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-multilib-list=m32,m64,mx32 --with-target-system-zlib=auto --with-tune=generic --without-cuda-driver -v - egeo-07: --build=x86_64-linux-gnu --disable-vtable-verify --disable-werror --enable-bootstrap --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-gnu-unique-object --enable-languages=c,ada,c++,go,brig,d,fortran,objc,obj-c++,m2 --enable-libphobos-checking=release --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-link-mutex --enable-multiarch --enable-multilib --enable-nls --enable-objc-gc=auto --enable-offload-targets=nvptx-none=/build/gcc-10-Km9U7s/gcc-10-10.2.1/debian/tmp-nvptx/usr,amdgcn-amdhsa=/build/gcc-10-Km9U7s/gcc-10-10.2.1/debian/tmp-gcn/usr,hsa --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --program-prefix=x86_64-linux-gnu- --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-build-config=bootstrap-lto-lean --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-multilib-list=m32,m64,mx32 --with-target-system-zlib=auto --with-tune=generic --without-cuda-driver -v Python Details- m7g.16xlarge Graviton3: Python 3.10.6- c6g.16xlarge Graviton2: Python 3.10.6- c7g.16xlarge Graviton3: Python 3.10.6- c7gn.16xlarge Graviton3E: Python 3.10.6- c6a.16xlarge AMD Zen 3: Python 3.10.12- egeo-07: Python 2.7.18 + Python 3.9.2Security Details- m7g.16xlarge Graviton3: itlb_multihit: Not affected + l1tf: Not affected + mds: Not affected + meltdown: Not affected + mmio_stale_data: Not affected + retbleed: Not affected + spec_store_bypass: Mitigation of SSB disabled via prctl + spectre_v1: Mitigation of __user pointer sanitization + spectre_v2: Mitigation of CSV2 BHB + srbds: Not affected + tsx_async_abort: Not affected- c6g.16xlarge Graviton2: itlb_multihit: Not affected + l1tf: Not affected + mds: Not affected + meltdown: Not affected + mmio_stale_data: Not affected + retbleed: Not affected + spec_store_bypass: Mitigation of SSB disabled via prctl + spectre_v1: Mitigation of __user pointer sanitization + spectre_v2: Mitigation of CSV2 BHB + srbds: Not affected + tsx_async_abort: Not affected- c7g.16xlarge Graviton3: itlb_multihit: Not affected + l1tf: Not affected + mds: Not affected + meltdown: Not affected + mmio_stale_data: Not affected + retbleed: Not affected + spec_store_bypass: Mitigation of SSB disabled via prctl + spectre_v1: Mitigation of __user pointer sanitization + spectre_v2: Mitigation of CSV2 BHB + srbds: Not affected + tsx_async_abort: Not affected- c7gn.16xlarge Graviton3E: itlb_multihit: Not affected + l1tf: Not affected + mds: Not affected + meltdown: Not affected + mmio_stale_data: Not affected + retbleed: Not affected + spec_store_bypass: Mitigation of SSB disabled via prctl + spectre_v1: Mitigation of __user pointer sanitization + spectre_v2: Mitigation of CSV2 BHB + srbds: Not affected + tsx_async_abort: Not affected- c6a.16xlarge AMD Zen 3: itlb_multihit: Not affected + l1tf: Not affected + mds: Not affected + meltdown: Not affected + mmio_stale_data: Not affected + retbleed: Not affected + spec_store_bypass: Mitigation of SSB disabled via prctl + spectre_v1: Mitigation of usercopy/swapgs barriers and __user pointer sanitization + spectre_v2: Mitigation of Retpolines IBPB: conditional IBRS_FW STIBP: conditional RSB filling PBRSB-eIBRS: Not affected + srbds: Not affected + tsx_async_abort: Not affected - egeo-07: gather_data_sampling: Mitigation of Microcode + itlb_multihit: KVM: Mitigation of VMX disabled + l1tf: Not affected + mds: Not affected + meltdown: Not affected + mmio_stale_data: Mitigation of Clear buffers; SMT vulnerable + retbleed: Mitigation of Enhanced IBRS + spec_rstack_overflow: Not affected + spec_store_bypass: Mitigation of SSB disabled via prctl and seccomp + spectre_v1: Mitigation of usercopy/swapgs barriers and __user pointer sanitization + spectre_v2: Mitigation of Enhanced IBRS IBPB: conditional RSB filling PBRSB-eIBRS: SW sequence + srbds: Not affected + tsx_async_abort: Mitigation of TSX disabled Processor Details- c6a.16xlarge AMD Zen 3: CPU Microcode: 0xa0011cf- egeo-07: Scaling Governor: intel_pstate powersave (EPP: balance_performance) - CPU Microcode: 0x5003605

Amazon AWS Graviton3E vs. Graviton 2/3 benchmarksnwchem: C240 Buckyballgraph500: 26graph500: 26graph500: 26graph500: 26lammps: 20k Atomsbuild-gem5: Time To Compilebrl-cad: VGR Performance Metricstockfish: Total Timelczero: BLASbuild-nodejs: Time To Compilelczero: Eigenqmcpack: FeCO6_b3lyp_gmsqmcpack: FeCO6_b3lyp_gmsbuild-godot: Time To Compilemocassin: Dust 2D tau100.0openssl: SHA256openssl: AES-128-GCMopenssl: ChaCha20openssl: ChaCha20-Poly1305openssl: AES-256-GCMopenssl: SHA512qmcpack: Li2_STO_aestress-ng: CPU Cachelaghos: Sedov Blast Wave, ube_922_hex.meshnekrs: TurboPipe Periodicmt-dgemm: Sustained Floating-Point Ratenpb: EP.Dgpaw: Carbon Nanotubenekrs: Kershawnpb: SP.Cnginx: 1000nginx: 500stress-ng: Wide Vector Mathlaghos: Triple Point Problemrodinia: OpenMP LavaMDheffte: c2c - FFTW - double - 512gromacs: MPI CPU - water_GMX50_barenpb: LU.Copenssl: RSA4096openssl: RSA4096coremark: CoreMark Size 666 - Iterations Per Secondrodinia: OpenMP Streamclusterstress-ng: NUMAqmcpack: simple-H2Osrsran: PUSCH Processor Benchmark, Throughput Totalstress-ng: Vector Floating Pointsrsran: Downlink Processor Benchmarksrsran: PUSCH Processor Benchmark, Throughput Threadheffte: r2c - FFTW - double - 512heffte: c2c - FFTW - float - 512kripke: compress-7zip: Decompression Ratingcompress-7zip: Compression Ratingincompact3d: input.i3d 193 Cells Per Directionliquid-dsp: 64 - 256 - 512liquid-dsp: 32 - 256 - 512stress-ng: Fused Multiply-Addstress-ng: Vector Shuffleliquid-dsp: 64 - 256 - 32liquid-dsp: 64 - 256 - 57stress-ng: Matrix 3D Mathstress-ng: Matrix Mathstress-ng: Memory Copyingliquid-dsp: 32 - 256 - 32liquid-dsp: 32 - 256 - 57stress-ng: Vector Mathremhos: Sample Remap Examplepennant: sedovbigamg: heffte: r2c - FFTW - float - 512mocassin: Gas HII40lulesh: pennant: leblancbigincompact3d: input.i3d 129 Cells Per Directionheffte: c2c - FFTW - double - 256npb: CG.Crodinia: OpenMP CFD Solverheffte: c2c - FFTW - float - 256heffte: r2c - FFTW - double - 256npb: MG.Cheffte: r2c - FFTW - float - 256lammps: Rhodopsin Proteinheffte: c2c - FFTW - double - 128heffte: r2c - FFTW - float - 128heffte: c2c - FFTW - float - 128heffte: r2c - FFTW - double - 128m7g.16xlarge Graviton3c6g.16xlarge Graviton2c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3egeo-071940.24197540002994970001227790000119432000036.927180.2477837771121197111301237.7831398211.60205.72154.37882.669542125155803320331719001032267845177428746099028333311363032125448870112.613892396.34410.55397630000024.3623533738.9861.831315068000017244.85255616.04255768.441542834.94232.0143.78846.25044.22328341.68713859.510181.91601880.34226411.6633759.1028.0415413.876102.55318.595.884.473988.048233900040028554031682513.94541801627533338139666763762252.7654143.402270500000144240000010403.93368750.6720484.241136066667721493333217235.5914.0409.2064901646761667162.95613.57528296.3786.7205373.0987103840.892321988.994.37581.444278.504950126.29164.87337.55857.1503306.540186.356138.0142976.928468900020935000087438900086043200025.171225.30553302086609284947287.814891302.19297.94218.276145.37442472798847158436163857672925412034671763680712919959315714393925490165.121921785.20322.37222019000020.4179522216.2692.76017603366679711.70158676.40148964.69997272.65180.8062.22424.26582.76718741.90214040.92624.31260642.17702413.7352112.6645.2253938.742850.82197.263.844.929742.828422012023323420224070225.88256581349266676748633337732190.5435614.5115314000009782000005752.17284713.6311324.79765466667489270000147886.1420.74016.48050103558633381.941220.75817557.48512.176835.6372073520.627913103.626.05141.981640.110425671.2992.399625.95032.7468209.496135.35881.44981962.74157580002938260001206990000117771000036.862181.7797890661173164761333238.5431382211.32204.77156.68782.822542165612633320643498431032755169977431884221328337379573732145914147112.643844101.98408.01397898333324.1406053664.5462.083326185333317219.95255552.05255145.521535336.57230.6843.96346.37064.20028375.71713945.910181.41605948.67464511.6253523.5827.9905356.876178.46319.795.784.745188.184235444273328563331105613.83266931627666678141200063818458.6154472.072271966667144236666710813.59368671.3920478.671136133333721386667217446.1214.1209.4222701765277667163.27613.65928708.6566.9613453.1444799940.828321911.024.44281.009677.768549742.30162.01037.41255.1055301.418184.026133.51419144117620002961640001207760000117564000036.838182.4717447431170271211392238.6361444188.28204.25155.95182.974541542185934111304699431141181194237996946548735115246542032126059040113.203860335.38423.11414144000024.0785293657.6756.440330282333317163.11256585.83253518.511530043.52236.2244.04446.53004.82028369.11713754.810183.31611801.55926510.6903525.1727.9995431.276911.74323.297.485.006088.455135423406728567731200913.76067261627566678139400063723431.5554695.042266833333144266666710882.02369258.8920475.961136000000721380000217567.1014.0829.3409531765966333163.55913.52528736.2266.8399983.1148982840.970822155.364.42981.167178.165849860.68162.36137.48255.1038300.396184.110133.4223440.420455000015768800041777700041057100020.342192.118485038969056091316230.4231152184.10187.32147.737194.435458575347771514492693171383893787539252299937313845788945015291283297123.951447265.35275.9243375366679.3880503061.4289.818430881000034025.35163178.67165847.751380146.63227.4064.17923.52123.96595221.40548396.58392.41466587.0365808.396552.6826.8676479.196529.51691.3215.942.439444.317623708765023578723097030.314528846007666727480333330920910.9222255.84218486666717108000004571.96147576.418080.4311939666671444266667221776.1522.10416.5305083699930082.758412.66916708.2589.9175657.0197528820.871920210.009.34243.590741.586845946.81102.65219.56348.9432158.85898.702686.37301098485755800690706002105720002082880007.539462.59910268426092344641.867608.86552.72391.737281.2493466925203611313202535621349462726507041907446191285603878566000408.451525522.4670.772.0942061329.20257.46411726.5570348.7773654.19399851.8561.58212.54510.66531.11636556.34200342.73041.8362563.28837023.2470.8977.5231129.521347.32342.388.518.990219.6715109107233590557584084.664220212956000012825000010084119.217162.566412433334778966671841.5855962.003209.8363348333342959666738515.8968.95689.0049044445613334.367426.9205676.776243.4588021.23927379.323799661.9118.63016.997117.117019477.2332.42877.1769.6229658.201226.736723.0160OpenBenchmarking.org

NWChem

Input: C240 Buckyball

OpenBenchmarking.orgSeconds, Fewer Is BetterNWChem 7.0.2Input: C240 Buckyballc7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-072K4K6K8K10K1914.01940.21962.72976.93440.410984.0-m64-ldl -lutil -m641. (F9X) gfortran options: -lnwctask -lccsd -lmcscf -lselci -lmp2 -lmoints -lstepper -ldriver -loptim -lnwdft -lgradients -lcphf -lesp -lddscf -ldangchang -lguess -lhessian -lvib -lnwcutil -lrimp2 -lproperty -lsolvation -lnwints -lprepar -lnwmd -lnwpw -lofpw -lpaw -lpspw -lband -lnwpwlib -lcafe -lspace -lanalyze -lqhop -lpfft -ldplot -ldrdy -lvscf -lqmmm -lqmd -letrans -ltce -lbq -lmm -lcons -lperfm -ldntmc -lccca -ldimqm -lga -larmci -lpeigs -l64to32 -lopenblas -lpthread -lrt -llapack -lnwcblas -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz -lcomex -ffast-math -std=legacy -fdefault-integer-8 -finline-functions -O2

Graph500

Scale: 26

OpenBenchmarking.orgsssp max_TEPS, More Is BetterGraph500 3.0Scale: 26m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0790M180M270M360M450M41975400041575800041176200028468900020455000085755800-pthread1. (CC) gcc options: -fcommon -O3 -lpthread -lm -lmpi

Graph500

Scale: 26

OpenBenchmarking.orgsssp median_TEPS, More Is BetterGraph500 3.0Scale: 26m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0760M120M180M240M300M29949700029616400029382600020935000015768800069070600-pthread1. (CC) gcc options: -fcommon -O3 -lpthread -lm -lmpi

Graph500

Scale: 26

OpenBenchmarking.orgbfs max_TEPS, More Is BetterGraph500 3.0Scale: 26m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07300M600M900M1200M1500M1227790000120776000012069900008743890004177770002105720001. (CC) gcc options: -fcommon -O3 -lpthread -lm -lmpi

Graph500

Scale: 26

OpenBenchmarking.orgbfs median_TEPS, More Is BetterGraph500 3.0Scale: 26m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07300M600M900M1200M1500M1194320000117771000011756400008604320004105710002082880001. (CC) gcc options: -fcommon -O3 -lpthread -lm -lmpi

LAMMPS Molecular Dynamics Simulator

Model: 20k Atoms

OpenBenchmarking.orgns/day, More Is BetterLAMMPS Molecular Dynamics Simulator 23Jun2022Model: 20k Atomsm7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07816243240SE +/- 0.034, N = 3SE +/- 0.025, N = 3SE +/- 0.018, N = 3SE +/- 0.009, N = 3SE +/- 0.066, N = 3SE +/- 0.006, N = 336.92736.86236.83825.17120.3427.539-lm-pthread -lm1. (CXX) g++ options: -O3 -ldl

Timed Gem5 Compilation

Time To Compile

OpenBenchmarking.orgSeconds, Fewer Is BetterTimed Gem5 Compilation 21.2Time To Compilem7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-07100200300400500SE +/- 0.13, N = 3SE +/- 0.26, N = 3SE +/- 0.38, N = 3SE +/- 0.26, N = 3SE +/- 0.35, N = 3SE +/- 9.98, N = 9180.25181.78182.47192.12225.31462.60

BRL-CAD

VGR Performance Metric

OpenBenchmarking.orgVGR Performance Metric, More Is BetterBRL-CAD 7.34VGR Performance Metricc7g.16xlarge Graviton3m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07200K400K600K800K1000K789066783777744743533020485038102684-m641. (CXX) g++ options: -std=c++14 -pipe -fvisibility=hidden -fno-strict-aliasing -fno-common -fexceptions -ftemplate-depth-128 -ggdb3 -O3 -fipa-pta -fstrength-reduce -finline-functions -flto -ltcl8.6 -lregex_brl -lz_brl -lnetpbm -ldl -lm -ltk8.6

Stockfish

Total Time

OpenBenchmarking.orgNodes Per Second, More Is BetterStockfish 15Total Timec7g.16xlarge Graviton3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0730M60M90M120M150MSE +/- 2998209.87, N = 12SE +/- 1531345.46, N = 15SE +/- 2854071.93, N = 15SE +/- 1430593.84, N = 15SE +/- 2597495.37, N = 15SE +/- 349749.32, N = 12117316476117027121112119711969056098660928426092344-m64 -msse -msse3 -mpopcnt -mavx2 -msse4.1 -mssse3 -msse2 -mbmi2-m64 -msse -msse3 -mpopcnt -mavx2 -mavx512f -mavx512bw -mavx512vnni -mavx512dq -mavx512vl -msse4.1 -mssse3 -msse2 -mbmi21. (CXX) g++ options: -lgcov -lpthread -fno-exceptions -std=c++17 -fno-peel-loops -fno-tracer -pedantic -O3 -flto -flto=jobserver

LeelaChessZero

Backend: BLAS

OpenBenchmarking.orgNodes Per Second, More Is BetterLeelaChessZero 0.28Backend: BLASc7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3m7g.16xlarge Graviton3c6g.16xlarge Graviton230060090012001500SE +/- 7.22, N = 3SE +/- 3.53, N = 3SE +/- 13.29, N = 5SE +/- 4.67, N = 3SE +/- 11.79, N = 313921333131613019471. (CXX) g++ options: -flto -pthread

Timed Node.js Compilation

Time To Compile

OpenBenchmarking.orgSeconds, Fewer Is BetterTimed Node.js Compilation 19.8.1Time To Compilec6a.16xlarge AMD Zen 3m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2egeo-07140280420560700SE +/- 0.40, N = 3SE +/- 0.33, N = 3SE +/- 0.20, N = 3SE +/- 0.32, N = 3SE +/- 0.16, N = 3SE +/- 1.11, N = 3230.42237.78238.54238.64287.81641.87

LeelaChessZero

Backend: Eigen

OpenBenchmarking.orgNodes Per Second, More Is BetterLeelaChessZero 0.28Backend: Eigenc7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton230060090012001500SE +/- 14.88, N = 3SE +/- 8.74, N = 3SE +/- 15.65, N = 3SE +/- 7.37, N = 3SE +/- 4.73, N = 314441398138211528911. (CXX) g++ options: -flto -pthread

QMCPACK

Input: FeCO6_b3lyp_gms

OpenBenchmarking.orgTotal Execution Time - Seconds, Fewer Is BetterQMCPACK 3.16Input: FeCO6_b3lyp_gmsc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-07130260390520650SE +/- 1.03, N = 3SE +/- 0.29, N = 3SE +/- 0.19, N = 3SE +/- 0.22, N = 3SE +/- 0.37, N = 3SE +/- 0.11, N = 3184.10188.28211.32211.60302.19608.86-march=native-mcpu=native-mcpu=native-mcpu=native-mcpu=native-march=native -pthread1. (CXX) g++ options: -fopenmp -foffload=disable -finline-limit=1000 -fstrict-aliasing -funroll-all-loops -ffast-math -O3 -lm -ldl

QMCPACK

Input: FeCO6_b3lyp_gms

OpenBenchmarking.orgTotal Execution Time - Seconds, Fewer Is BetterQMCPACK 3.16Input: FeCO6_b3lyp_gmsc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-07120240360480600SE +/- 2.30, N = 3SE +/- 0.21, N = 3SE +/- 0.82, N = 3SE +/- 0.45, N = 3SE +/- 1.75, N = 3SE +/- 7.86, N = 3187.32204.25204.77205.72297.94552.72-march=native-mcpu=native-mcpu=native-mcpu=native-mcpu=native-march=native -pthread1. (CXX) g++ options: -fopenmp -foffload=disable -finline-limit=1000 -fstrict-aliasing -funroll-all-loops -ffast-math -O3 -lm -ldl

Timed Godot Game Engine Compilation

Time To Compile

OpenBenchmarking.orgSeconds, Fewer Is BetterTimed Godot Game Engine Compilation 4.0Time To Compilec6a.16xlarge AMD Zen 3m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-0780160240320400SE +/- 0.12, N = 3SE +/- 0.32, N = 3SE +/- 0.45, N = 3SE +/- 0.63, N = 3SE +/- 0.30, N = 3SE +/- 0.36, N = 3147.74154.38155.95156.69218.28391.74

Monte Carlo Simulations of Ionised Nebulae

Input: Dust 2D tau100.0

OpenBenchmarking.orgSeconds, Fewer Is BetterMonte Carlo Simulations of Ionised Nebulae 2.02.73.3Input: Dust 2D tau100.0m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0760120180240300SE +/- 0.01, N = 3SE +/- 0.00, N = 3SE +/- 0.07, N = 3SE +/- 0.86, N = 3SE +/- 1.84, N = 7SE +/- 0.37, N = 382.6782.8282.97145.37194.44281.25-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -cpp -Jsource/ -ffree-line-length-0 -lm -std=legacy -O2 -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lz

OpenSSL

Algorithm: SHA256

OpenBenchmarking.orgbyte/s, More Is BetterOpenSSL 3.1Algorithm: SHA256c7g.16xlarge Graviton3m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0712000M24000M36000M48000M60000MSE +/- 16491036.11, N = 3SE +/- 18610524.10, N = 3SE +/- 19542665.92, N = 3SE +/- 26770675.21, N = 3SE +/- 245440310.03, N = 3SE +/- 404619.57, N = 354216561263542125155805415421859345857534777424727988473466925203-m641. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

OpenSSL

Algorithm: AES-128-GCM

OpenBenchmarking.orgbyte/s, More Is BetterOpenSSL 3.1Algorithm: AES-128-GCMc7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0790000M180000M270000M360000M450000MSE +/- 11273100.69, N = 3SE +/- 12264074.61, N = 3SE +/- 81289574.27, N = 3SE +/- 9833681.11, N = 3SE +/- 4227452.23, N = 3SE +/- 11737066.92, N = 341113046994333206434984333203317190015843616385715144926931761131320253-m641. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

OpenSSL

Algorithm: ChaCha20

OpenBenchmarking.orgbyte/s, More Is BetterOpenSSL 3.1Algorithm: ChaCha20c6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-0730000M60000M90000M120000M150000MSE +/- 36376378.52, N = 3SE +/- 771581.87, N = 3SE +/- 1725060.95, N = 3SE +/- 1293723.80, N = 3SE +/- 35952887.59, N = 3SE +/- 13595278.49, N = 31383893787531141181194231032755169971032267845176729254120356213494627-m64-m641. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

OpenSSL

Algorithm: ChaCha20-Poly1305

OpenBenchmarking.orgbyte/s, More Is BetterOpenSSL 3.1Algorithm: ChaCha20-Poly1305c6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-0720000M40000M60000M80000M100000MSE +/- 232372675.93, N = 3SE +/- 1769561.47, N = 3SE +/- 1218886.42, N = 3SE +/- 1340503.89, N = 3SE +/- 1132293.08, N = 3SE +/- 1523000.86, N = 3925229993737996946548774318842213742874609904671763680726507041907-m64-m641. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

OpenSSL

Algorithm: AES-256-GCM

OpenBenchmarking.orgbyte/s, More Is BetterOpenSSL 3.1Algorithm: AES-256-GCMc7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0780000M160000M240000M320000M400000MSE +/- 24279491.44, N = 3SE +/- 33807617.40, N = 3SE +/- 6411836.47, N = 3SE +/- 41584947.90, N = 3SE +/- 2312792.64, N = 3SE +/- 2585526.42, N = 335115246542028337379573728333311363013845788945012919959315744619128560-m641. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

OpenSSL

Algorithm: SHA512

OpenBenchmarking.orgbyte/s, More Is BetterOpenSSL 3.1Algorithm: SHA512c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-077000M14000M21000M28000M35000MSE +/- 4573992.60, N = 3SE +/- 16155877.53, N = 3SE +/- 17714077.14, N = 3SE +/- 207279.55, N = 3SE +/- 9173912.49, N = 3SE +/- 1513929.31, N = 332145914147321260590403212544887015291283297143939254903878566000-m641. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

QMCPACK

Input: Li2_STO_ae

OpenBenchmarking.orgTotal Execution Time - Seconds, Fewer Is BetterQMCPACK 3.16Input: Li2_STO_aem7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0790180270360450SE +/- 0.08, N = 3SE +/- 0.12, N = 3SE +/- 0.31, N = 3SE +/- 0.13, N = 3SE +/- 1.13, N = 3SE +/- 4.45, N = 3112.61112.64113.20123.95165.12408.45-mcpu=native-mcpu=native-mcpu=native-march=native-mcpu=native-march=native -pthread1. (CXX) g++ options: -fopenmp -foffload=disable -finline-limit=1000 -fstrict-aliasing -funroll-all-loops -ffast-math -O3 -lm -ldl

Stress-NG

Test: CPU Cache

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: CPU Cachem7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-07c6a.16xlarge AMD Zen 3800K1600K2400K3200K4000KSE +/- 57217.78, N = 15SE +/- 40698.46, N = 15SE +/- 59376.56, N = 15SE +/- 21905.72, N = 15SE +/- 22640.51, N = 15SE +/- 30785.49, N = 123892396.343860335.383844101.981921785.201525522.461447265.35-laio -lbsd -lEGL -lGLESv2 -lmd1. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

Laghos

Test: Sedov Blast Wave, ube_922_hex.mesh

OpenBenchmarking.orgMajor Kernels Total Rate, More Is BetterLaghos 3.1Test: Sedov Blast Wave, ube_922_hex.meshc7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0790180270360450SE +/- 0.79, N = 3SE +/- 0.42, N = 3SE +/- 0.89, N = 3SE +/- 0.89, N = 3SE +/- 0.48, N = 3SE +/- 0.27, N = 3423.11410.55408.01322.37275.9270.77-pthread1. (CXX) g++ options: -O3 -std=c++11 -lmfem -lHYPRE -lmetis -lrt -lmpi_cxx -lmpi

nekRS

Input: TurboPipe Periodic

OpenBenchmarking.orgflops/rank, More Is BetternekRS 23.0Input: TurboPipe Periodicc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2900M1800M2700M3600M4500MSE +/- 12801180.07, N = 3SE +/- 1394740.12, N = 3SE +/- 169148.19, N = 3SE +/- 1199180.28, N = 3SE +/- 144222.05, N = 3433753666741414400003978983333397630000022201900001. (CXX) g++ options: -fopenmp -O2 -march=native -mtune=native -ftree-vectorize -rdynamic -lmpi_cxx -lmpi

ACES DGEMM

Sustained Floating-Point Rate

OpenBenchmarking.orgGFLOP/s, More Is BetterACES DGEMM 1.0Sustained Floating-Point Ratem7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07612182430SE +/- 0.171001, N = 13SE +/- 0.285590, N = 4SE +/- 0.297525, N = 4SE +/- 0.154503, N = 3SE +/- 0.038051, N = 3SE +/- 0.035680, N = 1524.36235324.14060524.07852920.4179529.3880502.0942061. (CC) gcc options: -O3 -march=native -fopenmp

NAS Parallel Benchmarks

Test / Class: EP.D

OpenBenchmarking.orgTotal Mop/s, More Is BetterNAS Parallel Benchmarks 3.4Test / Class: EP.Dm7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-078001600240032004000SE +/- 1.69, N = 3SE +/- 34.07, N = 15SE +/- 32.06, N = 15SE +/- 4.77, N = 3SE +/- 2.22, N = 3SE +/- 0.38, N = 33738.983664.543657.673061.422216.261329.20-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -O3 -march=native -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz 2. c6a.16xlarge AMD Zen 3: Open MPI 4.1.23. egeo-07: Open MPI 4.1.0

GPAW

Input: Carbon Nanotube

OpenBenchmarking.orgSeconds, Fewer Is BetterGPAW 23.6Input: Carbon Nanotubec7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0760120180240300SE +/- 0.04, N = 3SE +/- 0.03, N = 3SE +/- 0.04, N = 3SE +/- 0.13, N = 3SE +/- 0.02, N = 3SE +/- 0.19, N = 356.4461.8362.0889.8292.76257.46-pthread1. (CC) gcc options: -shared -fwrapv -O2 -lxc -lblas -lmpi

nekRS

Input: Kershaw

OpenBenchmarking.orgflops/rank, More Is BetternekRS 23.0Input: Kershawc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2900M1800M2700M3600M4500MSE +/- 22342148.51, N = 3SE +/- 5414395.42, N = 3SE +/- 2490845.46, N = 3SE +/- 1575066.14, N = 3SE +/- 737119.02, N = 3430881000033028233333261853333315068000017603366671. (CXX) g++ options: -fopenmp -O2 -march=native -mtune=native -ftree-vectorize -rdynamic -lmpi_cxx -lmpi

NAS Parallel Benchmarks

Test / Class: SP.C

OpenBenchmarking.orgTotal Mop/s, More Is BetterNAS Parallel Benchmarks 3.4Test / Class: SP.Cc6a.16xlarge AMD Zen 3m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Eegeo-07c6g.16xlarge Graviton27K14K21K28K35KSE +/- 20.85, N = 3SE +/- 10.19, N = 3SE +/- 7.21, N = 3SE +/- 31.31, N = 3SE +/- 15.52, N = 3SE +/- 1.54, N = 334025.3517244.8517219.9517163.1111726.559711.70-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -O3 -march=native -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz 2. c6a.16xlarge AMD Zen 3: Open MPI 4.1.23. egeo-07: Open MPI 4.1.0

nginx

Connections: 1000

OpenBenchmarking.orgRequests Per Second, More Is Betternginx 1.23.2Connections: 1000c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0750K100K150K200K250KSE +/- 402.16, N = 3SE +/- 137.20, N = 3SE +/- 55.97, N = 3SE +/- 136.82, N = 3SE +/- 185.79, N = 3SE +/- 141.39, N = 3256585.83255616.04255552.05163178.67158676.4070348.771. (CC) gcc options: -lluajit-5.1 -lm -lssl -lcrypto -lpthread -ldl -std=c99 -O2

nginx

Connections: 500

OpenBenchmarking.orgRequests Per Second, More Is Betternginx 1.23.2Connections: 500m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0750K100K150K200K250KSE +/- 323.56, N = 3SE +/- 243.69, N = 3SE +/- 317.05, N = 3SE +/- 60.38, N = 3SE +/- 90.87, N = 3SE +/- 47.73, N = 3255768.44255145.52253518.51165847.75148964.6973654.191. (CC) gcc options: -lluajit-5.1 -lm -lssl -lcrypto -lpthread -ldl -std=c99 -O2

Stress-NG

Test: Wide Vector Math

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: Wide Vector Mathm7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-07300K600K900K1200K1500KSE +/- 16116.93, N = 15SE +/- 16521.46, N = 15SE +/- 16444.95, N = 15SE +/- 2507.18, N = 3SE +/- 505.84, N = 3SE +/- 641.34, N = 31542834.941535336.571530043.521380146.63997272.65399851.85-laio -lbsd -lEGL -lGLESv2 -lmd1. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

Laghos

Test: Triple Point Problem

OpenBenchmarking.orgMajor Kernels Total Rate, More Is BetterLaghos 3.1Test: Triple Point Problemc7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0750100150200250SE +/- 0.27, N = 3SE +/- 0.28, N = 3SE +/- 0.16, N = 3SE +/- 1.06, N = 3SE +/- 0.48, N = 3SE +/- 0.71, N = 4236.22232.01230.68227.40180.8061.58-pthread1. (CXX) g++ options: -O3 -std=c++11 -lmfem -lHYPRE -lmetis -lrt -lmpi_cxx -lmpi

Rodinia

Test: OpenMP LavaMD

OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenMP LavaMDm7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0750100150200250SE +/- 0.15, N = 3SE +/- 0.11, N = 3SE +/- 0.15, N = 3SE +/- 0.04, N = 3SE +/- 0.53, N = 3SE +/- 0.02, N = 343.7943.9644.0462.2264.18212.55-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-m64 -lm -lcuda -lcudart -lcudadevrt -lcudart_static -lrt -lpthread -ldl1. (CXX) g++ options:

HeFFTe - Highly Efficient FFT for Exascale

Test: c2c - Backend: FFTW - Precision: double - X Y Z: 512

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: c2c - Backend: FFTW - Precision: double - X Y Z: 512c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-071122334455SE +/- 0.01, N = 3SE +/- 0.01, N = 3SE +/- 0.01, N = 3SE +/- 0.01, N = 3SE +/- 0.05, N = 3SE +/- 0.02, N = 346.5346.3746.2524.2723.5210.67-pthread1. (CXX) g++ options: -O3

GROMACS

Implementation: MPI CPU - Input: water_GMX50_bare

OpenBenchmarking.orgNs Per Day, More Is BetterGROMACS 2023Implementation: MPI CPU - Input: water_GMX50_barec7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-071.08452.1693.25354.3385.4225SE +/- 0.003, N = 3SE +/- 0.003, N = 3SE +/- 0.004, N = 3SE +/- 0.013, N = 3SE +/- 0.002, N = 3SE +/- 0.002, N = 34.8204.2234.2003.9652.7671.116-lm1. (CXX) g++ options: -O3

NAS Parallel Benchmarks

Test / Class: LU.C

OpenBenchmarking.orgTotal Mop/s, More Is BetterNAS Parallel Benchmarks 3.4Test / Class: LU.Cc6a.16xlarge AMD Zen 3egeo-07c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c6g.16xlarge Graviton220K40K60K80K100KSE +/- 90.22, N = 3SE +/- 21.23, N = 3SE +/- 36.09, N = 3SE +/- 43.73, N = 3SE +/- 48.62, N = 3SE +/- 26.12, N = 395221.4036556.3428375.7128369.1128341.6818741.90-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -O3 -march=native -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz 2. c6a.16xlarge AMD Zen 3: Open MPI 4.1.23. egeo-07: Open MPI 4.1.0

OpenSSL

Algorithm: RSA4096

OpenBenchmarking.orgverify/s, More Is BetterOpenSSL 3.1Algorithm: RSA4096c7g.16xlarge Graviton3m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-07150K300K450K600K750KSE +/- 12.03, N = 3SE +/- 21.82, N = 3SE +/- 198.10, N = 3SE +/- 34.73, N = 3SE +/- 88.30, N = 3SE +/- 170.27, N = 3713945.9713859.5713754.8548396.5214040.9200342.7-m64-m641. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

OpenSSL

Algorithm: RSA4096

OpenBenchmarking.orgsign/s, More Is BetterOpenSSL 3.1Algorithm: RSA4096c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3egeo-07c6g.16xlarge Graviton22K4K6K8K10KSE +/- 0.84, N = 3SE +/- 1.27, N = 3SE +/- 1.54, N = 3SE +/- 3.06, N = 3SE +/- 5.58, N = 3SE +/- 1.71, N = 310183.310181.910181.48392.43041.82624.3-m64-m641. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

Coremark

CoreMark Size 666 - Iterations Per Second

OpenBenchmarking.orgIterations/Sec, More Is BetterCoremark 1.0CoreMark Size 666 - Iterations Per Secondc7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-07300K600K900K1200K1500KSE +/- 14869.41, N = 7SE +/- 13274.76, N = 15SE +/- 11449.37, N = 15SE +/- 6710.50, N = 3SE +/- 153.60, N = 3SE +/- 4039.35, N = 31611801.561605948.671601880.341466587.041260642.18362563.291. (CC) gcc options: -O2 -lrt" -lrt

Rodinia

Test: OpenMP Streamcluster

OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenMP Streamclusterc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-07612182430SE +/- 0.101, N = 15SE +/- 0.233, N = 12SE +/- 0.099, N = 8SE +/- 0.138, N = 3SE +/- 0.211, N = 15SE +/- 0.397, N = 158.39610.69011.62511.66313.73523.247-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-m64 -lm -lcuda -lcudart -lcudadevrt -lcudart_static -lrt -lpthread -ldl1. (CXX) g++ options:

Stress-NG

Test: NUMA

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: NUMAm7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-078001600240032004000SE +/- 5.17, N = 3SE +/- 7.31, N = 3SE +/- 3.39, N = 3SE +/- 1.53, N = 3SE +/- 9.75, N = 15SE +/- 0.00, N = 33759.103525.173523.582112.66552.680.891. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

QMCPACK

Input: simple-H2O

OpenBenchmarking.orgTotal Execution Time - Seconds, Fewer Is BetterQMCPACK 3.16Input: simple-H2Oc6a.16xlarge AMD Zen 3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-0720406080100SE +/- 0.08, N = 3SE +/- 0.02, N = 3SE +/- 0.04, N = 3SE +/- 0.03, N = 3SE +/- 0.24, N = 3SE +/- 0.86, N = 526.8727.9928.0028.0445.2377.52-march=native-mcpu=native-mcpu=native-mcpu=native-mcpu=native-march=native -pthread1. (CXX) g++ options: -fopenmp -foffload=disable -finline-limit=1000 -fstrict-aliasing -funroll-all-loops -ffast-math -O3 -lm -ldl

srsRAN Project

Test: PUSCH Processor Benchmark, Throughput Total

OpenBenchmarking.orgMbps, More Is BettersrsRAN Project 23.5Test: PUSCH Processor Benchmark, Throughput Totalc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-0714002800420056007000SE +/- 21.76, N = 3SE +/- 3.32, N = 3SE +/- 4.08, N = 3SE +/- 1.80, N = 3SE +/- 2.53, N = 3SE +/- 6.96, N = 36479.15431.25413.85356.83938.71129.5-march=native -mfma-march=native -mfma -lpthread1. (CXX) g++ options: -O3 -fno-trapping-math -fno-math-errno -lgtest

Stress-NG

Test: Vector Floating Point

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: Vector Floating Pointc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-0720K40K60K80K100KSE +/- 864.23, N = 13SE +/- 1.74, N = 3SE +/- 71.97, N = 3SE +/- 190.19, N = 3SE +/- 31.31, N = 3SE +/- 160.53, N = 396529.5176911.7476178.4676102.5542850.8221347.32-laio -lbsd -lEGL -lGLESv2 -lmd1. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

srsRAN Project

Test: Downlink Processor Benchmark

OpenBenchmarking.orgMbps, More Is BettersrsRAN Project 23.5Test: Downlink Processor Benchmarkc6a.16xlarge AMD Zen 3egeo-07c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2150300450600750SE +/- 1.26, N = 3SE +/- 4.00, N = 4SE +/- 0.06, N = 3SE +/- 0.95, N = 3SE +/- 0.91, N = 3SE +/- 0.25, N = 3691.3342.3323.2319.7318.5197.2-march=native -mfma-march=native -mfma -lpthread1. (CXX) g++ options: -O3 -fno-trapping-math -fno-math-errno -lgtest

srsRAN Project

Test: PUSCH Processor Benchmark, Throughput Thread

OpenBenchmarking.orgMbps, More Is BettersrsRAN Project 23.5Test: PUSCH Processor Benchmark, Throughput Threadc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3egeo-07c6g.16xlarge Graviton250100150200250SE +/- 0.55, N = 3SE +/- 0.06, N = 3SE +/- 0.03, N = 3SE +/- 0.03, N = 3SE +/- 0.68, N = 10SE +/- 0.03, N = 3215.997.495.895.788.563.8-march=native -mfma-march=native -mfma -lpthread1. (CXX) g++ options: -O3 -fno-trapping-math -fno-math-errno -lgtest

HeFFTe - Highly Efficient FFT for Exascale

Test: r2c - Backend: FFTW - Precision: double - X Y Z: 512

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: r2c - Backend: FFTW - Precision: double - X Y Z: 512c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0720406080100SE +/- 0.02, N = 3SE +/- 0.01, N = 3SE +/- 0.02, N = 3SE +/- 0.03, N = 3SE +/- 0.05, N = 3SE +/- 0.01, N = 385.0184.7584.4744.9342.4418.99-pthread1. (CXX) g++ options: -O3

HeFFTe - Highly Efficient FFT for Exascale

Test: c2c - Backend: FFTW - Precision: float - X Y Z: 512

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: c2c - Backend: FFTW - Precision: float - X Y Z: 512c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0720406080100SE +/- 0.02, N = 3SE +/- 0.05, N = 3SE +/- 0.02, N = 3SE +/- 0.14, N = 3SE +/- 0.01, N = 3SE +/- 0.02, N = 388.4688.1888.0544.3242.8319.67-pthread1. (CXX) g++ options: -O3

Kripke

OpenBenchmarking.orgThroughput FoM, More Is BetterKripke 1.2.6c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0780M160M240M320M400MSE +/- 525406.56, N = 3SE +/- 445212.18, N = 3SE +/- 619419.33, N = 3SE +/- 2932840.19, N = 4SE +/- 102787.75, N = 3SE +/- 523405.33, N = 3354442733354234067339000400237087650220120233109107233-pthread1. (CXX) g++ options: -O3 -fopenmp -ldl

7-Zip Compression

Test: Decompression Rating

OpenBenchmarking.orgMIPS, More Is Better7-Zip Compression 22.01Test: Decompression Ratingc7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0760K120K180K240K300KSE +/- 54.90, N = 3SE +/- 146.43, N = 3SE +/- 93.51, N = 3SE +/- 1190.65, N = 3SE +/- 15.43, N = 3SE +/- 265.06, N = 3285677285633285540235787234202590551. (CXX) g++ options: -lpthread -ldl -O2 -fPIC

7-Zip Compression

Test: Compression Rating

OpenBenchmarking.orgMIPS, More Is Better7-Zip Compression 22.01Test: Compression Ratingm7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0770K140K210K280K350KSE +/- 154.72, N = 3SE +/- 308.14, N = 3SE +/- 72.90, N = 3SE +/- 209.44, N = 3SE +/- 670.46, N = 3SE +/- 414.62, N = 3316825312009311056240702230970758401. (CXX) g++ options: -lpthread -ldl -O2 -fPIC

Xcompact3d Incompact3d

Input: input.i3d 193 Cells Per Direction

OpenBenchmarking.orgSeconds, Fewer Is BetterXcompact3d Incompact3d 2021-03-11Input: input.i3d 193 Cells Per Directionc7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0720406080100SE +/- 0.05, N = 3SE +/- 0.09, N = 3SE +/- 0.02, N = 3SE +/- 0.03, N = 3SE +/- 0.28, N = 3SE +/- 0.03, N = 313.7613.8313.9525.8830.3184.66-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -cpp -O2 -funroll-loops -floop-optimize -fcray-pointer -fbacktrace -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz

Liquid-DSP

Threads: 64 - Buffer Length: 256 - Filter Length: 512

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 1.6Threads: 64 - Buffer Length: 256 - Filter Length: 512c6a.16xlarge AMD Zen 3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-07100M200M300M400M500MSE +/- 392527.42, N = 3SE +/- 3333.33, N = 3SE +/- 8819.17, N = 3SE +/- 6666.67, N = 3SE +/- 3333.33, N = 3SE +/- 92915.73, N = 34600766671627666671627566671627533331349266671295600001. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

Liquid-DSP

Threads: 32 - Buffer Length: 256 - Filter Length: 512

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 1.6Threads: 32 - Buffer Length: 256 - Filter Length: 512c6a.16xlarge AMD Zen 3egeo-07c7g.16xlarge Graviton3m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton260M120M180M240M300MSE +/- 193419.52, N = 3SE +/- 120554.28, N = 3SE +/- 1000.00, N = 3SE +/- 1855.92, N = 3SE +/- 577.35, N = 3SE +/- 333.33, N = 3274803333128250000814120008139666781394000674863331. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

Stress-NG

Test: Fused Multiply-Add

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: Fused Multiply-Addc7g.16xlarge Graviton3m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0714M28M42M56M70MSE +/- 4431.60, N = 3SE +/- 4870.19, N = 3SE +/- 10061.51, N = 3SE +/- 3687.67, N = 3SE +/- 32747.05, N = 3SE +/- 16948.60, N = 363818458.6163762252.7663723431.5537732190.5430920910.9210084119.211. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

Stress-NG

Test: Vector Shuffle

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: Vector Shufflec7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0712K24K36K48K60KSE +/- 294.96, N = 3SE +/- 139.03, N = 3SE +/- 21.44, N = 3SE +/- 74.80, N = 3SE +/- 0.50, N = 3SE +/- 0.40, N = 354695.0454472.0754143.4035614.5122255.847162.561. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

Liquid-DSP

Threads: 64 - Buffer Length: 256 - Filter Length: 32

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 1.6Threads: 64 - Buffer Length: 256 - Filter Length: 32c7g.16xlarge Graviton3m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-07500M1000M1500M2000M2500MSE +/- 284800.12, N = 3SE +/- 435889.89, N = 3SE +/- 2915666.50, N = 3SE +/- 218581.28, N = 3SE +/- 251661.15, N = 3SE +/- 707515.21, N = 3227196666722705000002266833333218486666715314000006412433331. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

Liquid-DSP

Threads: 64 - Buffer Length: 256 - Filter Length: 57

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 1.6Threads: 64 - Buffer Length: 256 - Filter Length: 57c6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-07400M800M1200M1600M2000MSE +/- 1014889.16, N = 3SE +/- 88191.71, N = 3SE +/- 152752.52, N = 3SE +/- 284800.12, N = 3SE +/- 11547.01, N = 3SE +/- 851162.60, N = 317108000001442666667144240000014423666679782000004778966671. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

Stress-NG

Test: Matrix 3D Math

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: Matrix 3D Mathc7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-072K4K6K8K10KSE +/- 19.16, N = 3SE +/- 9.35, N = 3SE +/- 6.38, N = 3SE +/- 1.40, N = 3SE +/- 1.96, N = 3SE +/- 9.17, N = 310882.0210813.5910403.935752.174571.961841.58-laio -lbsd -lEGL -lGLESv2 -lmd1. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

Stress-NG

Test: Matrix Math

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: Matrix Mathc7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0780K160K240K320K400KSE +/- 28.60, N = 3SE +/- 53.44, N = 3SE +/- 38.76, N = 3SE +/- 8.13, N = 3SE +/- 167.77, N = 3SE +/- 6.88, N = 3369258.89368750.67368671.39284713.63147576.4155962.001. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

Stress-NG

Test: Memory Copying

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: Memory Copyingm7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-074K8K12K16K20KSE +/- 3.80, N = 3SE +/- 4.65, N = 3SE +/- 1.36, N = 3SE +/- 1.12, N = 3SE +/- 0.46, N = 3SE +/- 1.93, N = 320484.2420478.6720475.9611324.798080.433209.83-laio -lbsd -lEGL -lGLESv2 -lmd1. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

Liquid-DSP

Threads: 32 - Buffer Length: 256 - Filter Length: 32

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 1.6Threads: 32 - Buffer Length: 256 - Filter Length: 32c6a.16xlarge AMD Zen 3c7g.16xlarge Graviton3m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2egeo-07300M600M900M1200M1500MSE +/- 578311.72, N = 3SE +/- 33333.33, N = 3SE +/- 233333.33, N = 3SE +/- 57735.03, N = 3SE +/- 456520.66, N = 3SE +/- 571382.34, N = 311939666671136133333113606666711360000007654666676334833331. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

Liquid-DSP

Threads: 32 - Buffer Length: 256 - Filter Length: 57

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 1.6Threads: 32 - Buffer Length: 256 - Filter Length: 57c6a.16xlarge AMD Zen 3m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2egeo-07300M600M900M1200M1500MSE +/- 9533333.33, N = 3SE +/- 3333.33, N = 3SE +/- 168358.08, N = 3SE +/- 150111.07, N = 3SE +/- 23094.01, N = 3SE +/- 846666.67, N = 314442666677214933337213866677213800004892700004295966671. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

Stress-NG

Test: Vector Math

OpenBenchmarking.orgBogo Ops/s, More Is BetterStress-NG 0.15.10Test: Vector Mathc6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-0750K100K150K200K250KSE +/- 100.78, N = 3SE +/- 27.00, N = 3SE +/- 20.95, N = 3SE +/- 47.94, N = 3SE +/- 37.96, N = 3SE +/- 9.74, N = 3221776.15217567.10217446.12217235.59147886.1438515.891. (CXX) g++ options: -lm -lapparmor -latomic -lc -lcrypt -ldl -ljpeg -lpthread -lrt -lsctp -lz

Remhos

Test: Sample Remap Example

OpenBenchmarking.orgSeconds, Fewer Is BetterRemhos 1.0Test: Sample Remap Examplem7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-071530456075SE +/- 0.04, N = 3SE +/- 0.02, N = 3SE +/- 0.04, N = 3SE +/- 0.08, N = 3SE +/- 0.11, N = 3SE +/- 0.44, N = 314.0414.0814.1220.7422.1068.96-pthread1. (CXX) g++ options: -O3 -std=c++11 -lmfem -lHYPRE -lmetis -lrt -lmpi_cxx -lmpi

Pennant

Test: sedovbig

OpenBenchmarking.orgHydro Cycle Time - Seconds, Fewer Is BetterPennant 1.0.1Test: sedovbigm7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0720406080100SE +/- 0.011347, N = 3SE +/- 0.003721, N = 3SE +/- 0.011497, N = 3SE +/- 0.018218, N = 3SE +/- 0.036687, N = 3SE +/- 0.050055, N = 39.2064909.3409539.42227016.48050016.53050089.004900-pthread1. (CXX) g++ options: -fopenmp -lmpi_cxx -lmpi

Algebraic Multi-Grid Benchmark

OpenBenchmarking.orgFigure Of Merit, More Is BetterAlgebraic Multi-Grid Benchmark 1.2c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07400M800M1200M1600M2000MSE +/- 488508.39, N = 3SE +/- 192645.90, N = 3SE +/- 103191.30, N = 3SE +/- 140169.34, N = 3SE +/- 1055539.30, N = 3SE +/- 394420.25, N = 31765966333176527766716467616671035586333836999300444456133-pthread1. (CC) gcc options: -lparcsr_ls -lparcsr_mv -lseq_mv -lIJ_mv -lkrylov -lHYPRE_utilities -lm -fopenmp -lmpi

HeFFTe - Highly Efficient FFT for Exascale

Test: r2c - Backend: FFTW - Precision: float - X Y Z: 512

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: r2c - Backend: FFTW - Precision: float - X Y Z: 512c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-074080120160200SE +/- 0.03, N = 3SE +/- 0.05, N = 3SE +/- 0.13, N = 3SE +/- 0.08, N = 3SE +/- 0.03, N = 3SE +/- 0.05, N = 3163.56163.28162.9682.7681.9434.37-pthread1. (CXX) g++ options: -O3

Monte Carlo Simulations of Ionised Nebulae

Input: Gas HII40

OpenBenchmarking.orgSeconds, Fewer Is BetterMonte Carlo Simulations of Ionised Nebulae 2.02.73.3Input: Gas HII40c6a.16xlarge AMD Zen 3c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6g.16xlarge Graviton2egeo-07612182430SE +/- 0.02, N = 3SE +/- 0.03, N = 3SE +/- 0.05, N = 3SE +/- 0.03, N = 3SE +/- 0.17, N = 3SE +/- 0.04, N = 312.6713.5313.5813.6620.7626.92-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -cpp -Jsource/ -ffree-line-length-0 -lm -std=legacy -O2 -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lz

LULESH

OpenBenchmarking.orgz/s, More Is BetterLULESH 2.0.3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3m7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-076K12K18K24K30KSE +/- 12.73, N = 3SE +/- 11.81, N = 3SE +/- 27.09, N = 3SE +/- 38.55, N = 3SE +/- 90.11, N = 3SE +/- 5.42, N = 328736.2328708.6628296.3817557.4916708.265676.78-pthread1. (CXX) g++ options: -O3 -fopenmp -lm -lmpi_cxx -lmpi

Pennant

Test: leblancbig

OpenBenchmarking.orgHydro Cycle Time - Seconds, Fewer Is BetterPennant 1.0.1Test: leblancbigm7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-071020304050SE +/- 0.000869, N = 3SE +/- 0.000467, N = 3SE +/- 0.005468, N = 3SE +/- 0.013289, N = 3SE +/- 0.018924, N = 3SE +/- 0.025073, N = 36.7205376.8399986.9613459.91756512.17683043.458800-pthread1. (CXX) g++ options: -fopenmp -lmpi_cxx -lmpi

Xcompact3d Incompact3d

Input: input.i3d 129 Cells Per Direction

OpenBenchmarking.orgSeconds, Fewer Is BetterXcompact3d Incompact3d 2021-03-11Input: input.i3d 129 Cells Per Directionm7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07510152025SE +/- 0.02702838, N = 3SE +/- 0.01738352, N = 3SE +/- 0.03233273, N = 3SE +/- 0.02560507, N = 3SE +/- 0.08686597, N = 15SE +/- 0.12655798, N = 33.098710383.114898283.144479995.637207357.0197528821.23927370-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -cpp -O2 -funroll-loops -floop-optimize -fcray-pointer -fbacktrace -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz

HeFFTe - Highly Efficient FFT for Exascale

Test: c2c - Backend: FFTW - Precision: double - X Y Z: 256

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: c2c - Backend: FFTW - Precision: double - X Y Z: 256c7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-07918273645SE +/- 0.02971, N = 3SE +/- 0.01031, N = 3SE +/- 0.02659, N = 3SE +/- 0.17467, N = 3SE +/- 0.01033, N = 3SE +/- 0.01615, N = 340.9708040.8923040.8283020.8719020.627909.32379-pthread1. (CXX) g++ options: -O3

NAS Parallel Benchmarks

Test / Class: CG.C

OpenBenchmarking.orgTotal Mop/s, More Is BetterNAS Parallel Benchmarks 3.4Test / Class: CG.Cc7gn.16xlarge Graviton3Em7g.16xlarge Graviton3c7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-075K10K15K20K25KSE +/- 125.21, N = 3SE +/- 130.18, N = 3SE +/- 283.23, N = 3SE +/- 14.83, N = 3SE +/- 31.56, N = 3SE +/- 12.05, N = 322155.3621988.9921911.0220210.0013103.629661.91-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -O3 -march=native -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz 2. c6a.16xlarge AMD Zen 3: Open MPI 4.1.23. egeo-07: Open MPI 4.1.0

Rodinia

Test: OpenMP CFD Solver

OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenMP CFD Solverm7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07510152025SE +/- 0.011, N = 3SE +/- 0.027, N = 3SE +/- 0.021, N = 3SE +/- 0.016, N = 3SE +/- 0.002, N = 3SE +/- 0.208, N = 44.3754.4294.4426.0519.34218.630-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-O2 -lOpenCL-m64 -lm -lcuda -lcudart -lcudadevrt -lcudart_static -lrt -lpthread -ldl1. (CXX) g++ options:

HeFFTe - Highly Efficient FFT for Exascale

Test: c2c - Backend: FFTW - Precision: float - X Y Z: 256

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: c2c - Backend: FFTW - Precision: float - X Y Z: 256m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0720406080100SE +/- 0.01, N = 3SE +/- 0.09, N = 3SE +/- 0.07, N = 3SE +/- 0.42, N = 6SE +/- 0.05, N = 3SE +/- 0.10, N = 381.4481.1781.0143.5941.9817.00-pthread1. (CXX) g++ options: -O3

HeFFTe - Highly Efficient FFT for Exascale

Test: r2c - Backend: FFTW - Precision: double - X Y Z: 256

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: r2c - Backend: FFTW - Precision: double - X Y Z: 256m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0720406080100SE +/- 0.02, N = 3SE +/- 0.03, N = 3SE +/- 0.31, N = 3SE +/- 0.17, N = 3SE +/- 0.01, N = 3SE +/- 0.10, N = 378.5078.1777.7741.5940.1117.12-pthread1. (CXX) g++ options: -O3

NAS Parallel Benchmarks

Test / Class: MG.C

OpenBenchmarking.orgTotal Mop/s, More Is BetterNAS Parallel Benchmarks 3.4Test / Class: MG.Cm7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-0711K22K33K44K55KSE +/- 24.30, N = 3SE +/- 14.65, N = 3SE +/- 32.94, N = 3SE +/- 167.32, N = 3SE +/- 7.02, N = 3SE +/- 35.78, N = 350126.2949860.6849742.3045946.8125671.2919477.23-pthread -ldl -lutil -lrt1. (F9X) gfortran options: -O3 -march=native -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz 2. c6a.16xlarge AMD Zen 3: Open MPI 4.1.23. egeo-07: Open MPI 4.1.0

HeFFTe - Highly Efficient FFT for Exascale

Test: r2c - Backend: FFTW - Precision: float - X Y Z: 256

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: r2c - Backend: FFTW - Precision: float - X Y Z: 256m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-074080120160200SE +/- 0.27, N = 3SE +/- 0.04, N = 3SE +/- 0.11, N = 3SE +/- 1.28, N = 3SE +/- 0.19, N = 3SE +/- 0.04, N = 3164.87162.36162.01102.6592.4032.43-pthread1. (CXX) g++ options: -O3

LAMMPS Molecular Dynamics Simulator

Model: Rhodopsin Protein

OpenBenchmarking.orgns/day, More Is BetterLAMMPS Molecular Dynamics Simulator 23Jun2022Model: Rhodopsin Proteinm7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-07918273645SE +/- 0.057, N = 3SE +/- 0.026, N = 3SE +/- 0.033, N = 3SE +/- 0.083, N = 3SE +/- 0.257, N = 12SE +/- 0.024, N = 337.55837.48237.41225.95019.5637.176-lm-pthread -lm1. (CXX) g++ options: -O3 -ldl

HeFFTe - Highly Efficient FFT for Exascale

Test: c2c - Backend: FFTW - Precision: double - X Y Z: 128

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: c2c - Backend: FFTW - Precision: double - X Y Z: 128m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-071326395265SE +/- 0.28294, N = 3SE +/- 0.14885, N = 3SE +/- 0.32202, N = 3SE +/- 0.84547, N = 15SE +/- 0.08221, N = 3SE +/- 0.03519, N = 357.1503055.1055055.1038048.9432032.746809.622961. (CXX) g++ options: -O3

HeFFTe - Highly Efficient FFT for Exascale

Test: r2c - Backend: FFTW - Precision: float - X Y Z: 128

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: r2c - Backend: FFTW - Precision: float - X Y Z: 128m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-0770140210280350SE +/- 0.83, N = 3SE +/- 0.56, N = 3SE +/- 1.62, N = 3SE +/- 0.64, N = 3SE +/- 1.94, N = 3SE +/- 0.85, N = 15306.54301.42300.40209.50158.8658.20-pthread1. (CXX) g++ options: -O3

HeFFTe - Highly Efficient FFT for Exascale

Test: c2c - Backend: FFTW - Precision: float - X Y Z: 128

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: c2c - Backend: FFTW - Precision: float - X Y Z: 128m7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec7g.16xlarge Graviton3c6g.16xlarge Graviton2c6a.16xlarge AMD Zen 3egeo-074080120160200SE +/- 0.27, N = 3SE +/- 0.20, N = 3SE +/- 0.47, N = 3SE +/- 0.35, N = 3SE +/- 1.25, N = 14SE +/- 0.04, N = 3186.36184.11184.03135.3698.7026.74-pthread1. (CXX) g++ options: -O3

HeFFTe - Highly Efficient FFT for Exascale

Test: r2c - Backend: FFTW - Precision: double - X Y Z: 128

OpenBenchmarking.orgGFLOP/s, More Is BetterHeFFTe - Highly Efficient FFT for Exascale 2.3Test: r2c - Backend: FFTW - Precision: double - X Y Z: 128m7g.16xlarge Graviton3c7g.16xlarge Graviton3c7gn.16xlarge Graviton3Ec6a.16xlarge AMD Zen 3c6g.16xlarge Graviton2egeo-07306090120150SE +/- 0.12, N = 3SE +/- 0.47, N = 3SE +/- 0.04, N = 3SE +/- 1.46, N = 12SE +/- 0.61, N = 3SE +/- 0.06, N = 3138.01133.51133.4286.3781.4523.02-pthread1. (CXX) g++ options: -O3


Phoronix Test Suite v10.8.4