Neoverse-V1 Compiler Tests

amazon testing on Ubuntu 22.04 via the Phoronix Test Suite.

HTML result view exported from: https://openbenchmarking.org/result/2206076-NE-NEOVERSEV38.

Neoverse-V1 Compiler TestsProcessorMotherboardChipsetMemoryDiskNetworkOSKernelCompilerFile-SystemSystem Layerarmv8.4-a+svearmv8.4-aARMv8 Neoverse-V1 (32 Cores)Amazon EC2 c7g.8xlarge (1.0 BIOS)Amazon Device 020062GB301GB Amazon Elastic Block StoreAmazon ElasticUbuntu 22.045.15.0-1004-aws (aarch64)GCC 12.0.0 20220117ext4amazonOpenBenchmarking.orgKernel Details- Transparent Huge Pages: madviseEnvironment Details- armv8.4-a+sve: CXXFLAGS="-O3 -march=armv8.4-a+sve" CFLAGS="-O3 -march=armv8.4-a+sve" - armv8.4-a: CXXFLAGS="-O3 -march=armv8.4-a" CFLAGS="-O3 -march=armv8.4-a" Compiler Details- --build=aarch64-linux-gnu --disable-libquadmath --disable-libquadmath-support --disable-nls --disable-werror --enable-checking=yes,extra,rtl --enable-clocale=gnu --enable-fix-cortex-a53-843419 --enable-gnu-unique-object --enable-languages=c,ada,c++,go,d,fortran,objc,obj-c++ --enable-libphobos-checking=release --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-objc-gc=auto --enable-plugin --enable-shared --host=aarch64-linux-gnu --program-prefix= --target=aarch64-linux-gnu --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-target-system-zlib=auto -v Python Details- Python 3.10.4Security Details- itlb_multihit: Not affected + l1tf: Not affected + mds: Not affected + meltdown: Not affected + spec_store_bypass: Mitigation of SSB disabled via prctl + spectre_v1: Mitigation of __user pointer sanitization + spectre_v2: Mitigation of CSV2 BHB + srbds: Not affected + tsx_async_abort: Not affected

Neoverse-V1 Compiler Testscryptopp: Unkeyed Algorithmsetcpak: Multi-Threaded - ETC2lczero: BLASlczero: Eigenclomp: Static OMP Speedupmrbayes: Primate Phylogeny Analysisqe: AUSURF112lammps: Rhodopsin Proteinwebp: Defaultwebp: Quality 100webp: Quality 100, Losslesswebp: Quality 100, Highest Compressiongmpbench: Total Timexmrig: Monero - 1Mxmrig: Wownero - 1Mcompress-zstd: 3 - Compression Speedcompress-zstd: 19 - Compression Speedcompress-zstd: 19 - Decompression Speedcompress-zstd: 3, Long Mode - Compression Speedcompress-zstd: 3, Long Mode - Decompression Speedcompress-zstd: 19, Long Mode - Compression Speedcompress-zstd: 19, Long Mode - Decompression Speedjpegxl: PNG - 7jpegxl: PNG - 8jpegxl: JPEG - 7jpegxl: JPEG - 8nettle: aes256nettle: chachanettle: sha512nettle: poly1305-aesluajit: Compositeluajit: Monte Carloluajit: Fast Fourier Transformluajit: Sparse Matrix Multiplyluajit: Dense LU Matrix Factorizationluajit: Jacobi Successive Over-Relaxationbotan: KASUMIbotan: KASUMI - Decryptbotan: AES-256botan: AES-256 - Decryptbotan: Twofishbotan: Twofish - Decryptbotan: Blowfishbotan: Blowfish - Decryptbotan: CAST-256botan: CAST-256 - Decryptbotan: ChaCha20Poly1305botan: ChaCha20Poly1305 - Decryptgraphics-magick: Swirlgraphics-magick: Rotategraphics-magick: Sharpengraphics-magick: Enhancedgraphics-magick: Resizinggraphics-magick: Noise-Gaussiangraphics-magick: HWB Color Spaceaom-av1: Speed 8 Realtime - Bosphorus 4Kaom-av1: Speed 9 Realtime - Bosphorus 4Kaom-av1: Speed 10 Realtime - Bosphorus 4Kaom-av1: Speed 8 Realtime - Bosphorus 1080paom-av1: Speed 9 Realtime - Bosphorus 1080paom-av1: Speed 10 Realtime - Bosphorus 1080px264: Bosphorus 4Kx264: Bosphorus 1080pmt-dgemm: Sustained Floating-Point Ratecoremark: CoreMark Size 666 - Iterations Per Secondhimeno: Poisson Pressure Solverstockfish: Total Timestargate: 44100 - 512stargate: 96000 - 512stargate: 44100 - 1024stargate: 480000 - 512stargate: 96000 - 1024stargate: 480000 - 1024c-ray: Total Time - 4K, 16 Rays Per Pixelpovray: Trace Timeprimesieve: 1e12 Prime Number Generationsmallpt: Global Illumination Renderer; 128 Samplesaobench: 2048 x 2048 - Total Timeencode-flac: WAV To FLACencode-mp3: WAV To MP3encode-opus: WAV To Opus Encodeespeak: Text-To-Speech Synthesisngspice: C2670ngspice: C7552rnnoise: synthmark: VoiceMark_100openjpeg: NASA Curiosity Panorama M34openssl: SHA256openssl: RSA4096openssl: RSA4096liquid-dsp: 8 - 256 - 57liquid-dsp: 16 - 256 - 57liquid-dsp: 32 - 256 - 57gromacs: MPI CPU - water_GMX50_bareastcenc: Mediumastcenc: Thoroughastcenc: Exhaustivebasis: ETC1Sbasis: UASTC Level 0basis: UASTC Level 2basis: UASTC Level 3sqlite-speedtest: Timed Time - Size 1,000draco: Liondraco: Church Facaderedis: GETredis: SETcaffe: AlexNet - CPU - 200caffe: GoogleNet - CPU - 200tnn: CPU - DenseNettnn: CPU - MobileNet v2tnn: CPU - SqueezeNet v2tnn: CPU - SqueezeNet v1.1sysbench: CPUonnx: GPT-2 - CPU - Standardonnx: bertsquad-12 - CPU - Standardonnx: fcn-resnet101-11 - CPU - Standardonnx: ArcFace ResNet-100 - CPU - Standardonnx: super-resolution-10 - CPU - Standardencode-wavpack: WAV To WavPackgnupg: 2.7GB Sample File Encryptionkripke: armv8.4-a+svearmv8.4-a449.0242082006.7041297133329.6234.255545.5621.1492.4333.56123.4038.6404155.68669.811877.87027.274.03094.81242.73824.840.33263.78.360.6779.5827.344447.04733.59504.33820.511309.03343.85615.711162.333521.13902.1662.0062.2615442.6495474.321248.887258.148280.570289.032108.754108.623390.311383.94512726112977182414515106762.1983.8265.62123.95156.71193.6848.51169.5813.442233762066.5011635508.282435558233406.2135004.4748486.5381636.1244554.8077976.45018219.29920.2638.5333.89633.49438.5157.44014.40229.982106.910111.64417.385670.21055196274281768805088.1356407.81677333333354233336686366672.2754.80929.013135.192624.8136.89513.97622.60080.355530978432513289.21861924.13439311251252346.322280.24376.301205.79996666.761231777273935541120.51558.138192709233459.8701632005.1681281131129.8237.542545.9721.3302.4303.56223.8528.6304152.38645.411811.26937.872.93083.41241.33820.840.03250.98.320.6773.2126.304435.91740.25498.83871.901282.59343.27661.551151.573355.53901.0262.01762.2775494.3135477.571239.703246.155278.874288.505108.786108.599389.375382.5141225577677732233949497862.1383.8861.88120.13152.46190.2748.43168.9212.813927789646.9248005561.559287574856806.0729164.4140556.3700356.0053864.7298486.32268419.29619.8478.4383.89533.48138.3118.05418.32036.587102.558103.93317.622670.97657205276039435705090.5356359.61763633333527000007052333332.2774.88339.143535.364724.8216.91013.98722.63080.487535479352523377.921865840.13436341238072730.40260.77871.130257.70496726.401236477373938541320.48857.198204143167OpenBenchmarking.org

Crypto++

Test: Unkeyed Algorithms

OpenBenchmarking.orgMiB/second, More Is BetterCrypto++ 8.2Test: Unkeyed Algorithmsarmv8.4-a+svearmv8.4-a100200300400500SE +/- 0.25, N = 3SE +/- 0.25, N = 3449.02459.87-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -fPIC -pthread -pipe

Etcpak

Benchmark: Multi-Threaded - Configuration: ETC2

OpenBenchmarking.orgMpx/s, More Is BetterEtcpak 1.0Benchmark: Multi-Threaded - Configuration: ETC2armv8.4-a+svearmv8.4-a400800120016002000SE +/- 0.47, N = 3SE +/- 1.80, N = 32006.702005.171. (CXX) g++ options: -O3 -mcpu=native -std=c++11 -lpthread

LeelaChessZero

Backend: BLAS

OpenBenchmarking.orgNodes Per Second, More Is BetterLeelaChessZero 0.28Backend: BLASarmv8.4-a+svearmv8.4-a30060090012001500SE +/- 6.96, N = 3SE +/- 13.99, N = 512971281-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -flto -O3 -pthread

LeelaChessZero

Backend: Eigen

OpenBenchmarking.orgNodes Per Second, More Is BetterLeelaChessZero 0.28Backend: Eigenarmv8.4-a+svearmv8.4-a30060090012001500SE +/- 14.64, N = 5SE +/- 13.65, N = 313331311-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -flto -O3 -pthread

CLOMP

Static OMP Speedup

OpenBenchmarking.orgSpeedup, More Is BetterCLOMP 1.2Static OMP Speeduparmv8.4-a+svearmv8.4-a714212835SE +/- 0.18, N = 3SE +/- 0.34, N = 329.629.8-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fopenmp -O3 -lm

Timed MrBayes Analysis

Primate Phylogeny Analysis

OpenBenchmarking.orgSeconds, Fewer Is BetterTimed MrBayes Analysis 3.2.7Primate Phylogeny Analysisarmv8.4-a+svearmv8.4-a50100150200250SE +/- 0.18, N = 3SE +/- 0.07, N = 3234.26237.54-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -std=c99 -pedantic -lm

Quantum ESPRESSO

Input: AUSURF112

OpenBenchmarking.orgSeconds, Fewer Is BetterQuantum ESPRESSO 7.0Input: AUSURF112armv8.4-a+svearmv8.4-a120240360480600SE +/- 0.32, N = 3SE +/- 0.38, N = 3545.56545.971. (F9X) gfortran options: -pthread -fopenmp -ldevXlib -lopenblas -lFoX_dom -lFoX_sax -lFoX_wxml -lFoX_common -lFoX_utils -lFoX_fsys -lfftw3_omp -lfftw3 -lmpi_usempif08 -lmpi_mpifh -lmpi -lopen-rte -lopen-pal -lhwloc -levent_core -levent_pthreads -lm -lz

LAMMPS Molecular Dynamics Simulator

Model: Rhodopsin Protein

OpenBenchmarking.orgns/day, More Is BetterLAMMPS Molecular Dynamics Simulator 29Oct2020Model: Rhodopsin Proteinarmv8.4-a+svearmv8.4-a510152025SE +/- 0.01, N = 3SE +/- 0.09, N = 321.1521.33-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -lm

WebP Image Encode

Encode Settings: Default

OpenBenchmarking.orgEncode Time - Seconds, Fewer Is BetterWebP Image Encode 1.1Encode Settings: Defaultarmv8.4-a+svearmv8.4-a0.54741.09481.64222.18962.737SE +/- 0.001, N = 3SE +/- 0.001, N = 32.4332.430-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fvisibility=hidden -O3 -lm -ljpeg -lpng16 -ltiff

WebP Image Encode

Encode Settings: Quality 100

OpenBenchmarking.orgEncode Time - Seconds, Fewer Is BetterWebP Image Encode 1.1Encode Settings: Quality 100armv8.4-a+svearmv8.4-a0.80151.6032.40453.2064.0075SE +/- 0.001, N = 3SE +/- 0.006, N = 33.5613.562-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fvisibility=hidden -O3 -lm -ljpeg -lpng16 -ltiff

WebP Image Encode

Encode Settings: Quality 100, Lossless

OpenBenchmarking.orgEncode Time - Seconds, Fewer Is BetterWebP Image Encode 1.1Encode Settings: Quality 100, Losslessarmv8.4-a+svearmv8.4-a612182430SE +/- 0.18, N = 3SE +/- 0.01, N = 323.4023.85-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fvisibility=hidden -O3 -lm -ljpeg -lpng16 -ltiff

WebP Image Encode

Encode Settings: Quality 100, Highest Compression

OpenBenchmarking.orgEncode Time - Seconds, Fewer Is BetterWebP Image Encode 1.1Encode Settings: Quality 100, Highest Compressionarmv8.4-a+svearmv8.4-a246810SE +/- 0.017, N = 3SE +/- 0.006, N = 38.6408.630-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fvisibility=hidden -O3 -lm -ljpeg -lpng16 -ltiff

GNU GMP GMPbench

Total Time

OpenBenchmarking.orgGMPbench Score, More Is BetterGNU GMP GMPbench 6.2.1Total Timearmv8.4-a+svearmv8.4-a90018002700360045004155.64152.3-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -lm

Xmrig

Variant: Monero - Hash Count: 1M

OpenBenchmarking.orgH/s, More Is BetterXmrig 6.12.1Variant: Monero - Hash Count: 1Marmv8.4-a+svearmv8.4-a2K4K6K8K10KSE +/- 6.18, N = 3SE +/- 9.56, N = 38669.88645.4-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -fexceptions -fno-rtti -Ofast -static-libgcc -static-libstdc++ -rdynamic -lssl -lcrypto -luv -lpthread -lrt -ldl -lhwloc

Xmrig

Variant: Wownero - Hash Count: 1M

OpenBenchmarking.orgH/s, More Is BetterXmrig 6.12.1Variant: Wownero - Hash Count: 1Marmv8.4-a+svearmv8.4-a3K6K9K12K15KSE +/- 26.88, N = 3SE +/- 27.59, N = 311877.811811.2-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -fexceptions -fno-rtti -Ofast -static-libgcc -static-libstdc++ -rdynamic -lssl -lcrypto -luv -lpthread -lrt -ldl -lhwloc

Zstd Compression

Compression Level: 3 - Compression Speed

OpenBenchmarking.orgMB/s, More Is BetterZstd Compression 1.5.0Compression Level: 3 - Compression Speedarmv8.4-a+svearmv8.4-a15003000450060007500SE +/- 15.42, N = 3SE +/- 14.45, N = 37027.26937.8-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lz -llzma

Zstd Compression

Compression Level: 19 - Compression Speed

OpenBenchmarking.orgMB/s, More Is BetterZstd Compression 1.5.0Compression Level: 19 - Compression Speedarmv8.4-a+svearmv8.4-a1632486480SE +/- 0.03, N = 3SE +/- 0.23, N = 374.072.9-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lz -llzma

Zstd Compression

Compression Level: 19 - Decompression Speed

OpenBenchmarking.orgMB/s, More Is BetterZstd Compression 1.5.0Compression Level: 19 - Decompression Speedarmv8.4-a+svearmv8.4-a7001400210028003500SE +/- 6.60, N = 3SE +/- 8.46, N = 33094.83083.4-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lz -llzma

Zstd Compression

Compression Level: 3, Long Mode - Compression Speed

OpenBenchmarking.orgMB/s, More Is BetterZstd Compression 1.5.0Compression Level: 3, Long Mode - Compression Speedarmv8.4-a+svearmv8.4-a30060090012001500SE +/- 4.88, N = 3SE +/- 5.37, N = 31242.71241.3-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lz -llzma

Zstd Compression

Compression Level: 3, Long Mode - Decompression Speed

OpenBenchmarking.orgMB/s, More Is BetterZstd Compression 1.5.0Compression Level: 3, Long Mode - Decompression Speedarmv8.4-a+svearmv8.4-a8001600240032004000SE +/- 1.28, N = 3SE +/- 3.95, N = 33824.83820.8-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lz -llzma

Zstd Compression

Compression Level: 19, Long Mode - Compression Speed

OpenBenchmarking.orgMB/s, More Is BetterZstd Compression 1.5.0Compression Level: 19, Long Mode - Compression Speedarmv8.4-a+svearmv8.4-a918273645SE +/- 0.00, N = 3SE +/- 0.03, N = 340.340.0-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lz -llzma

Zstd Compression

Compression Level: 19, Long Mode - Decompression Speed

OpenBenchmarking.orgMB/s, More Is BetterZstd Compression 1.5.0Compression Level: 19, Long Mode - Decompression Speedarmv8.4-a+svearmv8.4-a7001400210028003500SE +/- 0.59, N = 3SE +/- 7.62, N = 33263.73250.9-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lz -llzma

JPEG XL libjxl

Input: PNG - Encode Speed: 7

OpenBenchmarking.orgMP/s, More Is BetterJPEG XL libjxl 0.6.1Input: PNG - Encode Speed: 7armv8.4-a+svearmv8.4-a246810SE +/- 0.01, N = 3SE +/- 0.02, N = 38.368.32-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -funwind-tables -O2 -fPIE -pie

JPEG XL libjxl

Input: PNG - Encode Speed: 8

OpenBenchmarking.orgMP/s, More Is BetterJPEG XL libjxl 0.6.1Input: PNG - Encode Speed: 8armv8.4-a+svearmv8.4-a0.15080.30160.45240.60320.754SE +/- 0.00, N = 3SE +/- 0.00, N = 30.670.67-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -funwind-tables -O2 -fPIE -pie

JPEG XL libjxl

Input: JPEG - Encode Speed: 7

OpenBenchmarking.orgMP/s, More Is BetterJPEG XL libjxl 0.6.1Input: JPEG - Encode Speed: 7armv8.4-a+svearmv8.4-a20406080100SE +/- 0.12, N = 3SE +/- 0.13, N = 379.5873.21-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -funwind-tables -O2 -fPIE -pie

JPEG XL libjxl

Input: JPEG - Encode Speed: 8

OpenBenchmarking.orgMP/s, More Is BetterJPEG XL libjxl 0.6.1Input: JPEG - Encode Speed: 8armv8.4-a+svearmv8.4-a612182430SE +/- 0.03, N = 3SE +/- 0.01, N = 327.3426.30-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -funwind-tables -O2 -fPIE -pie

Nettle

Test: aes256

OpenBenchmarking.orgMbyte/s, More Is BetterNettle 3.8Test: aes256armv8.4-a+svearmv8.4-a10002000300040005000SE +/- 0.39, N = 3SE +/- 3.11, N = 34447.044435.91-march=armv8.4-a+sve - MIN: 3925.11 / MAX: 5627.84-march=armv8.4-a - MIN: 3927.32 / MAX: 5628.861. (CC) gcc options: -O3 -ggdb3 -lnettle -lgmp -lm -lcrypto

Nettle

Test: chacha

OpenBenchmarking.orgMbyte/s, More Is BetterNettle 3.8Test: chachaarmv8.4-a+svearmv8.4-a160320480640800SE +/- 0.55, N = 3SE +/- 0.61, N = 3733.59740.25-march=armv8.4-a+sve - MIN: 442.26 / MAX: 956.22-march=armv8.4-a - MIN: 454.21 / MAX: 956.531. (CC) gcc options: -O3 -ggdb3 -lnettle -lgmp -lm -lcrypto

Nettle

Test: sha512

OpenBenchmarking.orgMbyte/s, More Is BetterNettle 3.8Test: sha512armv8.4-a+svearmv8.4-a110220330440550SE +/- 0.07, N = 3SE +/- 0.04, N = 3504.33498.83-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -ggdb3 -lnettle -lgmp -lm -lcrypto

Nettle

Test: poly1305-aes

OpenBenchmarking.orgMbyte/s, More Is BetterNettle 3.8Test: poly1305-aesarmv8.4-a+svearmv8.4-a2004006008001000SE +/- 5.37, N = 3SE +/- 1.52, N = 3820.51871.90-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -ggdb3 -lnettle -lgmp -lm -lcrypto

LuaJIT

Test: Composite

OpenBenchmarking.orgMflops, More Is BetterLuaJIT 2.1-gitTest: Compositearmv8.4-a+svearmv8.4-a30060090012001500SE +/- 18.19, N = 3SE +/- 0.41, N = 31309.031282.59-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -lm -ldl -O2 -fomit-frame-pointer -O3 -U_FORTIFY_SOURCE -fno-stack-protector

LuaJIT

Test: Monte Carlo

OpenBenchmarking.orgMflops, More Is BetterLuaJIT 2.1-gitTest: Monte Carloarmv8.4-a+svearmv8.4-a70140210280350SE +/- 0.54, N = 3SE +/- 0.35, N = 3343.85343.27-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -lm -ldl -O2 -fomit-frame-pointer -O3 -U_FORTIFY_SOURCE -fno-stack-protector

LuaJIT

Test: Fast Fourier Transform

OpenBenchmarking.orgMflops, More Is BetterLuaJIT 2.1-gitTest: Fast Fourier Transformarmv8.4-a+svearmv8.4-a140280420560700SE +/- 10.69, N = 3SE +/- 0.39, N = 3615.71661.55-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -lm -ldl -O2 -fomit-frame-pointer -O3 -U_FORTIFY_SOURCE -fno-stack-protector

LuaJIT

Test: Sparse Matrix Multiply

OpenBenchmarking.orgMflops, More Is BetterLuaJIT 2.1-gitTest: Sparse Matrix Multiplyarmv8.4-a+svearmv8.4-a30060090012001500SE +/- 7.20, N = 3SE +/- 3.14, N = 31162.331151.57-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -lm -ldl -O2 -fomit-frame-pointer -O3 -U_FORTIFY_SOURCE -fno-stack-protector

LuaJIT

Test: Dense LU Matrix Factorization

OpenBenchmarking.orgMflops, More Is BetterLuaJIT 2.1-gitTest: Dense LU Matrix Factorizationarmv8.4-a+svearmv8.4-a8001600240032004000SE +/- 86.66, N = 3SE +/- 6.02, N = 33521.133355.53-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -lm -ldl -O2 -fomit-frame-pointer -O3 -U_FORTIFY_SOURCE -fno-stack-protector

LuaJIT

Test: Jacobi Successive Over-Relaxation

OpenBenchmarking.orgMflops, More Is BetterLuaJIT 2.1-gitTest: Jacobi Successive Over-Relaxationarmv8.4-a+svearmv8.4-a2004006008001000SE +/- 1.00, N = 3SE +/- 0.68, N = 3902.16901.02-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -lm -ldl -O2 -fomit-frame-pointer -O3 -U_FORTIFY_SOURCE -fno-stack-protector

Botan

Test: KASUMI

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: KASUMIarmv8.4-a+svearmv8.4-a1428425670SE +/- 0.00, N = 3SE +/- 0.00, N = 362.0062.021. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: KASUMI - Decrypt

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: KASUMI - Decryptarmv8.4-a+svearmv8.4-a1428425670SE +/- 0.00, N = 3SE +/- 0.00, N = 362.2662.281. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: AES-256

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: AES-256armv8.4-a+svearmv8.4-a12002400360048006000SE +/- 9.30, N = 3SE +/- 14.75, N = 35442.655494.311. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: AES-256 - Decrypt

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: AES-256 - Decryptarmv8.4-a+svearmv8.4-a12002400360048006000SE +/- 8.64, N = 3SE +/- 5.58, N = 35474.325477.571. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: Twofish

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: Twofisharmv8.4-a+svearmv8.4-a50100150200250SE +/- 0.26, N = 3SE +/- 0.12, N = 3248.89239.701. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: Twofish - Decrypt

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: Twofish - Decryptarmv8.4-a+svearmv8.4-a60120180240300SE +/- 0.11, N = 3SE +/- 0.20, N = 3258.15246.161. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: Blowfish

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: Blowfisharmv8.4-a+svearmv8.4-a60120180240300SE +/- 0.10, N = 3SE +/- 0.29, N = 3280.57278.871. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: Blowfish - Decrypt

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: Blowfish - Decryptarmv8.4-a+svearmv8.4-a60120180240300SE +/- 0.03, N = 3SE +/- 0.07, N = 3289.03288.511. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: CAST-256

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: CAST-256armv8.4-a+svearmv8.4-a20406080100SE +/- 0.01, N = 3SE +/- 0.02, N = 3108.75108.791. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: CAST-256 - Decrypt

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: CAST-256 - Decryptarmv8.4-a+svearmv8.4-a20406080100SE +/- 0.01, N = 3SE +/- 0.02, N = 3108.62108.601. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: ChaCha20Poly1305

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: ChaCha20Poly1305armv8.4-a+svearmv8.4-a80160240320400SE +/- 0.07, N = 3SE +/- 0.04, N = 3390.31389.381. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

Botan

Test: ChaCha20Poly1305 - Decrypt

OpenBenchmarking.orgMiB/s, More Is BetterBotan 2.17.3Test: ChaCha20Poly1305 - Decryptarmv8.4-a+svearmv8.4-a80160240320400SE +/- 0.13, N = 3SE +/- 0.02, N = 3383.95382.511. (CXX) g++ options: -fstack-protector -pthread -lbotan-2 -ldl -lrt

GraphicsMagick

Operation: Swirl

OpenBenchmarking.orgIterations Per Minute, More Is BetterGraphicsMagick 1.3.33Operation: Swirlarmv8.4-a+svearmv8.4-a30060090012001500SE +/- 1.33, N = 3SE +/- 0.33, N = 312721225-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fopenmp -O3 -ljbig -ltiff -lfreetype -ljpeg -lXext -lSM -lICE -lX11 -llzma -lz -lm -lpthread

GraphicsMagick

Operation: Rotate

OpenBenchmarking.orgIterations Per Minute, More Is BetterGraphicsMagick 1.3.33Operation: Rotatearmv8.4-a+svearmv8.4-a130260390520650SE +/- 1.20, N = 3SE +/- 0.00, N = 3611577-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fopenmp -O3 -ljbig -ltiff -lfreetype -ljpeg -lXext -lSM -lICE -lX11 -llzma -lz -lm -lpthread

GraphicsMagick

Operation: Sharpen

OpenBenchmarking.orgIterations Per Minute, More Is BetterGraphicsMagick 1.3.33Operation: Sharpenarmv8.4-a+svearmv8.4-a150300450600750SE +/- 0.00, N = 3SE +/- 0.33, N = 3297677-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fopenmp -O3 -ljbig -ltiff -lfreetype -ljpeg -lXext -lSM -lICE -lX11 -llzma -lz -lm -lpthread

GraphicsMagick

Operation: Enhanced

OpenBenchmarking.orgIterations Per Minute, More Is BetterGraphicsMagick 1.3.33Operation: Enhancedarmv8.4-a+svearmv8.4-a160320480640800SE +/- 0.33, N = 3SE +/- 0.33, N = 3718732-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fopenmp -O3 -ljbig -ltiff -lfreetype -ljpeg -lXext -lSM -lICE -lX11 -llzma -lz -lm -lpthread

GraphicsMagick

Operation: Resizing

OpenBenchmarking.orgIterations Per Minute, More Is BetterGraphicsMagick 1.3.33Operation: Resizingarmv8.4-a+svearmv8.4-a5001000150020002500SE +/- 1.86, N = 3SE +/- 22.36, N = 324142339-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fopenmp -O3 -ljbig -ltiff -lfreetype -ljpeg -lXext -lSM -lICE -lX11 -llzma -lz -lm -lpthread

GraphicsMagick

Operation: Noise-Gaussian

OpenBenchmarking.orgIterations Per Minute, More Is BetterGraphicsMagick 1.3.33Operation: Noise-Gaussianarmv8.4-a+svearmv8.4-a110220330440550SE +/- 0.58, N = 3SE +/- 0.67, N = 3515494-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fopenmp -O3 -ljbig -ltiff -lfreetype -ljpeg -lXext -lSM -lICE -lX11 -llzma -lz -lm -lpthread

GraphicsMagick

Operation: HWB Color Space

OpenBenchmarking.orgIterations Per Minute, More Is BetterGraphicsMagick 1.3.33Operation: HWB Color Spacearmv8.4-a+svearmv8.4-a2004006008001000SE +/- 0.33, N = 3SE +/- 1.00, N = 31067978-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -fopenmp -O3 -ljbig -ltiff -lfreetype -ljpeg -lXext -lSM -lICE -lX11 -llzma -lz -lm -lpthread

AOM AV1

Encoder Mode: Speed 8 Realtime - Input: Bosphorus 4K

OpenBenchmarking.orgFrames Per Second, More Is BetterAOM AV1 3.3Encoder Mode: Speed 8 Realtime - Input: Bosphorus 4Karmv8.4-a+svearmv8.4-a1428425670SE +/- 0.61, N = 3SE +/- 0.43, N = 362.1962.13-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -std=c++11 -U_FORTIFY_SOURCE -lm

AOM AV1

Encoder Mode: Speed 9 Realtime - Input: Bosphorus 4K

OpenBenchmarking.orgFrames Per Second, More Is BetterAOM AV1 3.3Encoder Mode: Speed 9 Realtime - Input: Bosphorus 4Karmv8.4-a+svearmv8.4-a20406080100SE +/- 1.37, N = 12SE +/- 1.33, N = 1583.8283.88-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -std=c++11 -U_FORTIFY_SOURCE -lm

AOM AV1

Encoder Mode: Speed 10 Realtime - Input: Bosphorus 4K

OpenBenchmarking.orgFrames Per Second, More Is BetterAOM AV1 3.3Encoder Mode: Speed 10 Realtime - Input: Bosphorus 4Karmv8.4-a+svearmv8.4-a1530456075SE +/- 0.22, N = 3SE +/- 0.47, N = 365.6261.88-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -std=c++11 -U_FORTIFY_SOURCE -lm

AOM AV1

Encoder Mode: Speed 8 Realtime - Input: Bosphorus 1080p

OpenBenchmarking.orgFrames Per Second, More Is BetterAOM AV1 3.3Encoder Mode: Speed 8 Realtime - Input: Bosphorus 1080parmv8.4-a+svearmv8.4-a306090120150SE +/- 0.09, N = 3SE +/- 0.03, N = 3123.95120.13-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -std=c++11 -U_FORTIFY_SOURCE -lm

AOM AV1

Encoder Mode: Speed 9 Realtime - Input: Bosphorus 1080p

OpenBenchmarking.orgFrames Per Second, More Is BetterAOM AV1 3.3Encoder Mode: Speed 9 Realtime - Input: Bosphorus 1080parmv8.4-a+svearmv8.4-a306090120150SE +/- 0.20, N = 3SE +/- 0.03, N = 3156.71152.46-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -std=c++11 -U_FORTIFY_SOURCE -lm

AOM AV1

Encoder Mode: Speed 10 Realtime - Input: Bosphorus 1080p

OpenBenchmarking.orgFrames Per Second, More Is BetterAOM AV1 3.3Encoder Mode: Speed 10 Realtime - Input: Bosphorus 1080parmv8.4-a+svearmv8.4-a4080120160200SE +/- 0.20, N = 3SE +/- 0.28, N = 3193.68190.27-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -std=c++11 -U_FORTIFY_SOURCE -lm

x264

Video Input: Bosphorus 4K

OpenBenchmarking.orgFrames Per Second, More Is Betterx264 2022-02-22Video Input: Bosphorus 4Karmv8.4-a+svearmv8.4-a1122334455SE +/- 0.01, N = 3SE +/- 0.08, N = 348.5148.431. (CC) gcc options: -ldl -lavformat -lavcodec -lavutil -lswscale -lm -lpthread -O3 -flto

x264

Video Input: Bosphorus 1080p

OpenBenchmarking.orgFrames Per Second, More Is Betterx264 2022-02-22Video Input: Bosphorus 1080parmv8.4-a+svearmv8.4-a4080120160200SE +/- 0.10, N = 3SE +/- 0.07, N = 3169.58168.921. (CC) gcc options: -ldl -lavformat -lavcodec -lavutil -lswscale -lm -lpthread -O3 -flto

ACES DGEMM

Sustained Floating-Point Rate

OpenBenchmarking.orgGFLOP/s, More Is BetterACES DGEMM 1.0Sustained Floating-Point Ratearmv8.4-a+svearmv8.4-a3691215SE +/- 0.05, N = 3SE +/- 0.04, N = 313.4412.81-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -march=native -fopenmp

Coremark

CoreMark Size 666 - Iterations Per Second

OpenBenchmarking.orgIterations/Sec, More Is BetterCoremark 1.0CoreMark Size 666 - Iterations Per Secondarmv8.4-a+svearmv8.4-a200K400K600K800K1000KSE +/- 416.47, N = 3SE +/- 169.09, N = 3762066.50789646.92-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O2 -O3 -lrt" -lrt

Himeno Benchmark

Poisson Pressure Solver

OpenBenchmarking.orgMFLOPS, More Is BetterHimeno Benchmark 3.0Poisson Pressure Solverarmv8.4-a+svearmv8.4-a12002400360048006000SE +/- 5.76, N = 3SE +/- 2.66, N = 35508.285561.56-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3

Stockfish

Total Time

OpenBenchmarking.orgNodes Per Second, More Is BetterStockfish 13Total Timearmv8.4-a+svearmv8.4-a12M24M36M48M60MSE +/- 645132.01, N = 3SE +/- 721518.45, N = 145582334057485680-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -lgcov -lpthread -O3 -fno-exceptions -std=c++17 -pedantic -flto -fprofile-use -fno-peel-loops -fno-tracer -flto=jobserver

Stargate Digital Audio Workstation

Sample Rate: 44100 - Buffer Size: 512

OpenBenchmarking.orgRender Ratio, More Is BetterStargate Digital Audio Workstation 21.10.9Sample Rate: 44100 - Buffer Size: 512armv8.4-a+svearmv8.4-a246810SE +/- 0.003428, N = 3SE +/- 0.002564, N = 36.2135006.0729161. (CXX) g++ options: -lpthread -lsndfile -lm -O3 -march=native -ffast-math -funroll-loops -fstrength-reduce -fstrict-aliasing -finline-functions

Stargate Digital Audio Workstation

Sample Rate: 96000 - Buffer Size: 512

OpenBenchmarking.orgRender Ratio, More Is BetterStargate Digital Audio Workstation 21.10.9Sample Rate: 96000 - Buffer Size: 512armv8.4-a+svearmv8.4-a1.00682.01363.02044.02725.034SE +/- 0.002191, N = 3SE +/- 0.003502, N = 34.4748484.4140551. (CXX) g++ options: -lpthread -lsndfile -lm -O3 -march=native -ffast-math -funroll-loops -fstrength-reduce -fstrict-aliasing -finline-functions

Stargate Digital Audio Workstation

Sample Rate: 44100 - Buffer Size: 1024

OpenBenchmarking.orgRender Ratio, More Is BetterStargate Digital Audio Workstation 21.10.9Sample Rate: 44100 - Buffer Size: 1024armv8.4-a+svearmv8.4-a246810SE +/- 0.002411, N = 3SE +/- 0.002515, N = 36.5381636.3700351. (CXX) g++ options: -lpthread -lsndfile -lm -O3 -march=native -ffast-math -funroll-loops -fstrength-reduce -fstrict-aliasing -finline-functions

Stargate Digital Audio Workstation

Sample Rate: 480000 - Buffer Size: 512

OpenBenchmarking.orgRender Ratio, More Is BetterStargate Digital Audio Workstation 21.10.9Sample Rate: 480000 - Buffer Size: 512armv8.4-a+svearmv8.4-a246810SE +/- 0.002030, N = 3SE +/- 0.001956, N = 36.1244556.0053861. (CXX) g++ options: -lpthread -lsndfile -lm -O3 -march=native -ffast-math -funroll-loops -fstrength-reduce -fstrict-aliasing -finline-functions

Stargate Digital Audio Workstation

Sample Rate: 96000 - Buffer Size: 1024

OpenBenchmarking.orgRender Ratio, More Is BetterStargate Digital Audio Workstation 21.10.9Sample Rate: 96000 - Buffer Size: 1024armv8.4-a+svearmv8.4-a1.08182.16363.24544.32725.409SE +/- 0.002826, N = 3SE +/- 0.000903, N = 34.8077974.7298481. (CXX) g++ options: -lpthread -lsndfile -lm -O3 -march=native -ffast-math -funroll-loops -fstrength-reduce -fstrict-aliasing -finline-functions

Stargate Digital Audio Workstation

Sample Rate: 480000 - Buffer Size: 1024

OpenBenchmarking.orgRender Ratio, More Is BetterStargate Digital Audio Workstation 21.10.9Sample Rate: 480000 - Buffer Size: 1024armv8.4-a+svearmv8.4-a246810SE +/- 0.002250, N = 3SE +/- 0.002118, N = 36.4501826.3226841. (CXX) g++ options: -lpthread -lsndfile -lm -O3 -march=native -ffast-math -funroll-loops -fstrength-reduce -fstrict-aliasing -finline-functions

C-Ray

Total Time - 4K, 16 Rays Per Pixel

OpenBenchmarking.orgSeconds, Fewer Is BetterC-Ray 1.1Total Time - 4K, 16 Rays Per Pixelarmv8.4-a+svearmv8.4-a510152025SE +/- 0.00, N = 3SE +/- 0.03, N = 319.3019.30-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -lm -lpthread -O3

POV-Ray

Trace Time

OpenBenchmarking.orgSeconds, Fewer Is BetterPOV-Ray 3.7.0.7Trace Timearmv8.4-a+svearmv8.4-a510152025SE +/- 0.04, N = 3SE +/- 0.03, N = 320.2619.85-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -pipe -O3 -ffast-math -R/usr/lib -lXpm -lSM -lICE -lX11 -lIlmImf -lIlmImf-2_5 -lImath-2_5 -lHalf-2_5 -lIex-2_5 -lIexMath-2_5 -lIlmThread-2_5 -lIlmThread -ltiff -ljpeg -lpng -lz -lrt -lm -lboost_thread -lboost_system

Primesieve

1e12 Prime Number Generation

OpenBenchmarking.orgSeconds, Fewer Is BetterPrimesieve 7.71e12 Prime Number Generationarmv8.4-a+svearmv8.4-a246810SE +/- 0.043, N = 3SE +/- 0.022, N = 38.5338.438-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3

Smallpt

Global Illumination Renderer; 128 Samples

OpenBenchmarking.orgSeconds, Fewer Is BetterSmallpt 1.0Global Illumination Renderer; 128 Samplesarmv8.4-a+svearmv8.4-a0.87661.75322.62983.50644.383SE +/- 0.003, N = 3SE +/- 0.002, N = 33.8963.895-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -fopenmp -O3

AOBench

Size: 2048 x 2048 - Total Time

OpenBenchmarking.orgSeconds, Fewer Is BetterAOBenchSize: 2048 x 2048 - Total Timearmv8.4-a+svearmv8.4-a816243240SE +/- 0.01, N = 3SE +/- 0.00, N = 333.4933.48-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -lm -O3

FLAC Audio Encoding

WAV To FLAC

OpenBenchmarking.orgSeconds, Fewer Is BetterFLAC Audio Encoding 1.3.3WAV To FLACarmv8.4-a+svearmv8.4-a918273645SE +/- 0.01, N = 5SE +/- 0.02, N = 538.5238.31-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -fvisibility=hidden -logg -lm

LAME MP3 Encoding

WAV To MP3

OpenBenchmarking.orgSeconds, Fewer Is BetterLAME MP3 Encoding 3.100WAV To MP3armv8.4-a+svearmv8.4-a246810SE +/- 0.002, N = 3SE +/- 0.004, N = 37.4408.054-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -ffast-math -funroll-loops -fschedule-insns2 -fbranch-count-reg -fforce-addr -pipe -lm

Opus Codec Encoding

WAV To Opus Encode

OpenBenchmarking.orgSeconds, Fewer Is BetterOpus Codec Encoding 1.3.1WAV To Opus Encodearmv8.4-a+svearmv8.4-a510152025SE +/- 0.02, N = 5SE +/- 0.00, N = 514.4018.32-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -fvisibility=hidden -logg -lm

eSpeak-NG Speech Engine

Text-To-Speech Synthesis

OpenBenchmarking.orgSeconds, Fewer Is BettereSpeak-NG Speech Engine 20200907Text-To-Speech Synthesisarmv8.4-a+svearmv8.4-a816243240SE +/- 0.30, N = 20SE +/- 0.31, N = 1629.9836.59-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -std=c99

Ngspice

Circuit: C2670

OpenBenchmarking.orgSeconds, Fewer Is BetterNgspice 34Circuit: C2670armv8.4-a+svearmv8.4-a20406080100SE +/- 0.18, N = 3SE +/- 0.06, N = 3106.91102.56-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -fopenmp -lm -lstdc++ -lfftw3 -lXaw -lXmu -lXt -lXext -lX11 -lXft -lfontconfig -lXrender -lfreetype -lSM -lICE

Ngspice

Circuit: C7552

OpenBenchmarking.orgSeconds, Fewer Is BetterNgspice 34Circuit: C7552armv8.4-a+svearmv8.4-a20406080100SE +/- 1.04, N = 3SE +/- 0.51, N = 3111.64103.93-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -fopenmp -lm -lstdc++ -lfftw3 -lXaw -lXmu -lXt -lXext -lX11 -lXft -lfontconfig -lXrender -lfreetype -lSM -lICE

RNNoise

OpenBenchmarking.orgSeconds, Fewer Is BetterRNNoise 2020-06-28armv8.4-a+svearmv8.4-a48121620SE +/- 0.05, N = 3SE +/- 0.05, N = 317.3917.62-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pedantic -fvisibility=hidden

Google SynthMark

Test: VoiceMark_100

OpenBenchmarking.orgVoices, More Is BetterGoogle SynthMark 20201109Test: VoiceMark_100armv8.4-a+svearmv8.4-a140280420560700SE +/- 0.37, N = 3SE +/- 0.49, N = 3670.21670.981. (CXX) g++ options: -lm -lpthread -std=c++11 -Ofast

OpenJPEG

Encode: NASA Curiosity Panorama M34

OpenBenchmarking.orgms, Fewer Is BetterOpenJPEG 2.4Encode: NASA Curiosity Panorama M34armv8.4-a+svearmv8.4-a12K24K36K48K60KSE +/- 89.48, N = 3SE +/- 19.06, N = 35519657205-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -rdynamic

OpenSSL

Algorithm: SHA256

OpenBenchmarking.orgbyte/s, More Is BetterOpenSSL 3.0Algorithm: SHA256armv8.4-a+svearmv8.4-a6000M12000M18000M24000M30000MSE +/- 32102278.66, N = 3SE +/- 25639974.71, N = 32742817688027603943570-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

OpenSSL

Algorithm: RSA4096

OpenBenchmarking.orgsign/s, More Is BetterOpenSSL 3.0Algorithm: RSA4096armv8.4-a+svearmv8.4-a11002200330044005500SE +/- 0.53, N = 3SE +/- 0.78, N = 35088.15090.5-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

OpenSSL

Algorithm: RSA4096

OpenBenchmarking.orgverify/s, More Is BetterOpenSSL 3.0Algorithm: RSA4096armv8.4-a+svearmv8.4-a80K160K240K320K400KSE +/- 10.52, N = 3SE +/- 8.85, N = 3356407.8356359.6-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -pthread -O3 -lssl -lcrypto -ldl

Liquid-DSP

Threads: 8 - Buffer Length: 256 - Filter Length: 57

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 2021.01.31Threads: 8 - Buffer Length: 256 - Filter Length: 57armv8.4-a+svearmv8.4-a40M80M120M160M200MSE +/- 26666.67, N = 3SE +/- 12018.50, N = 3167733333176363333-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

Liquid-DSP

Threads: 16 - Buffer Length: 256 - Filter Length: 57

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 2021.01.31Threads: 16 - Buffer Length: 256 - Filter Length: 57armv8.4-a+svearmv8.4-a80M160M240M320M400MSE +/- 20275.88, N = 3SE +/- 30550.50, N = 3335423333352700000-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

Liquid-DSP

Threads: 32 - Buffer Length: 256 - Filter Length: 57

OpenBenchmarking.orgsamples/s, More Is BetterLiquid-DSP 2021.01.31Threads: 32 - Buffer Length: 256 - Filter Length: 57armv8.4-a+svearmv8.4-a150M300M450M600M750MSE +/- 1978807.16, N = 3SE +/- 125476.87, N = 3668636667705233333-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -pthread -lm -lc -lliquid

GROMACS

Implementation: MPI CPU - Input: water_GMX50_bare

OpenBenchmarking.orgNs Per Day, More Is BetterGROMACS 2022.1Implementation: MPI CPU - Input: water_GMX50_barearmv8.4-a+svearmv8.4-a0.51231.02461.53692.04922.5615SE +/- 0.002, N = 3SE +/- 0.001, N = 32.2752.277-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3

ASTC Encoder

Preset: Medium

OpenBenchmarking.orgSeconds, Fewer Is BetterASTC Encoder 3.2Preset: Mediumarmv8.4-a+svearmv8.4-a1.09872.19743.29614.39485.4935SE +/- 0.0080, N = 3SE +/- 0.0125, N = 34.80924.8833-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -flto -pthread

ASTC Encoder

Preset: Thorough

OpenBenchmarking.orgSeconds, Fewer Is BetterASTC Encoder 3.2Preset: Thorougharmv8.4-a+svearmv8.4-a3691215SE +/- 0.0027, N = 3SE +/- 0.0046, N = 39.01319.1435-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -flto -pthread

ASTC Encoder

Preset: Exhaustive

OpenBenchmarking.orgSeconds, Fewer Is BetterASTC Encoder 3.2Preset: Exhaustivearmv8.4-a+svearmv8.4-a816243240SE +/- 0.01, N = 3SE +/- 0.02, N = 335.1935.36-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -flto -pthread

Basis Universal

Settings: ETC1S

OpenBenchmarking.orgSeconds, Fewer Is BetterBasis Universal 1.13Settings: ETC1Sarmv8.4-a+svearmv8.4-a612182430SE +/- 0.04, N = 3SE +/- 0.03, N = 324.8124.821. (CXX) g++ options: -std=c++11 -fvisibility=hidden -fPIC -fno-strict-aliasing -O3 -rdynamic -lm -lpthread

Basis Universal

Settings: UASTC Level 0

OpenBenchmarking.orgSeconds, Fewer Is BetterBasis Universal 1.13Settings: UASTC Level 0armv8.4-a+svearmv8.4-a246810SE +/- 0.002, N = 3SE +/- 0.002, N = 36.8956.9101. (CXX) g++ options: -std=c++11 -fvisibility=hidden -fPIC -fno-strict-aliasing -O3 -rdynamic -lm -lpthread

Basis Universal

Settings: UASTC Level 2

OpenBenchmarking.orgSeconds, Fewer Is BetterBasis Universal 1.13Settings: UASTC Level 2armv8.4-a+svearmv8.4-a48121620SE +/- 0.00, N = 3SE +/- 0.01, N = 313.9813.991. (CXX) g++ options: -std=c++11 -fvisibility=hidden -fPIC -fno-strict-aliasing -O3 -rdynamic -lm -lpthread

Basis Universal

Settings: UASTC Level 3

OpenBenchmarking.orgSeconds, Fewer Is BetterBasis Universal 1.13Settings: UASTC Level 3armv8.4-a+svearmv8.4-a510152025SE +/- 0.00, N = 3SE +/- 0.00, N = 322.6022.631. (CXX) g++ options: -std=c++11 -fvisibility=hidden -fPIC -fno-strict-aliasing -O3 -rdynamic -lm -lpthread

SQLite Speedtest

Timed Time - Size 1,000

OpenBenchmarking.orgSeconds, Fewer Is BetterSQLite Speedtest 3.30Timed Time - Size 1,000armv8.4-a+svearmv8.4-a20406080100SE +/- 0.08, N = 3SE +/- 0.31, N = 380.3680.49-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3 -lz

Google Draco

Model: Lion

OpenBenchmarking.orgms, Fewer Is BetterGoogle Draco 1.5.0Model: Lionarmv8.4-a+svearmv8.4-a11002200330044005500SE +/- 2.40, N = 3SE +/- 2.65, N = 353095354-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3

Google Draco

Model: Church Facade

OpenBenchmarking.orgms, Fewer Is BetterGoogle Draco 1.5.0Model: Church Facadearmv8.4-a+svearmv8.4-a2K4K6K8K10KSE +/- 7.00, N = 3SE +/- 6.64, N = 378437935-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3

Redis

Test: GET

OpenBenchmarking.orgRequests Per Second, More Is BetterRedis 6.0.9Test: GETarmv8.4-a+svearmv8.4-a500K1000K1500K2000K2500KSE +/- 9056.40, N = 3SE +/- 1605.36, N = 32513289.202523377.92-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -MM -MT -g3 -fvisibility=hidden -O3

Redis

Test: SET

OpenBenchmarking.orgRequests Per Second, More Is BetterRedis 6.0.9Test: SETarmv8.4-a+svearmv8.4-a400K800K1200K1600K2000KSE +/- 7427.93, N = 3SE +/- 794.78, N = 31861924.131865840.13-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -MM -MT -g3 -fvisibility=hidden -O3

Caffe

Model: AlexNet - Acceleration: CPU - Iterations: 200

OpenBenchmarking.orgMilli-Seconds, Fewer Is BetterCaffe 2020-02-13Model: AlexNet - Acceleration: CPU - Iterations: 200armv8.4-a+svearmv8.4-a9K18K27K36K45KSE +/- 12.55, N = 3SE +/- 31.22, N = 34393143634-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -fPIC -rdynamic -lglog -lgflags -lprotobuf -lcrypto -lcurl -lpthread -lsz -lz -ldl -lm -llmdb -lopenblas

Caffe

Model: GoogleNet - Acceleration: CPU - Iterations: 200

OpenBenchmarking.orgMilli-Seconds, Fewer Is BetterCaffe 2020-02-13Model: GoogleNet - Acceleration: CPU - Iterations: 200armv8.4-a+svearmv8.4-a30K60K90K120K150KSE +/- 105.70, N = 3SE +/- 49.72, N = 3125125123807-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -fPIC -rdynamic -lglog -lgflags -lprotobuf -lcrypto -lcurl -lpthread -lsz -lz -ldl -lm -llmdb -lopenblas

TNN

Target: CPU - Model: DenseNet

OpenBenchmarking.orgms, Fewer Is BetterTNN 0.3Target: CPU - Model: DenseNetarmv8.4-a+svearmv8.4-a6001200180024003000SE +/- 22.16, N = 3SE +/- 12.63, N = 32346.322730.40-march=armv8.4-a+sve - MIN: 2268.4 / MAX: 2446.6-march=armv8.4-a - MIN: 2665.42 / MAX: 2834.731. (CXX) g++ options: -O3 -fopenmp -pthread -fvisibility=hidden -fvisibility=default -rdynamic -ldl

TNN

Target: CPU - Model: MobileNet v2

OpenBenchmarking.orgms, Fewer Is BetterTNN 0.3Target: CPU - Model: MobileNet v2armv8.4-a+svearmv8.4-a60120180240300SE +/- 0.69, N = 3SE +/- 0.10, N = 3280.24260.78-march=armv8.4-a+sve - MIN: 277.9 / MAX: 282.31-march=armv8.4-a - MIN: 259.13 / MAX: 262.381. (CXX) g++ options: -O3 -fopenmp -pthread -fvisibility=hidden -fvisibility=default -rdynamic -ldl

TNN

Target: CPU - Model: SqueezeNet v2

OpenBenchmarking.orgms, Fewer Is BetterTNN 0.3Target: CPU - Model: SqueezeNet v2armv8.4-a+svearmv8.4-a20406080100SE +/- 0.09, N = 3SE +/- 0.07, N = 376.3071.13-march=armv8.4-a+sve - MIN: 76.07 / MAX: 76.53-march=armv8.4-a - MIN: 70.76 / MAX: 71.581. (CXX) g++ options: -O3 -fopenmp -pthread -fvisibility=hidden -fvisibility=default -rdynamic -ldl

TNN

Target: CPU - Model: SqueezeNet v1.1

OpenBenchmarking.orgms, Fewer Is BetterTNN 0.3Target: CPU - Model: SqueezeNet v1.1armv8.4-a+svearmv8.4-a60120180240300SE +/- 0.14, N = 3SE +/- 0.08, N = 3205.80257.70-march=armv8.4-a+sve - MIN: 205.46 / MAX: 206.28-march=armv8.4-a - MIN: 256.95 / MAX: 258.391. (CXX) g++ options: -O3 -fopenmp -pthread -fvisibility=hidden -fvisibility=default -rdynamic -ldl

Sysbench

Test: CPU

OpenBenchmarking.orgEvents Per Second, More Is BetterSysbench 1.0.20Test: CPUarmv8.4-a+svearmv8.4-a20K40K60K80K100KSE +/- 2.72, N = 3SE +/- 8.50, N = 396666.7696726.40-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O2 -funroll-loops -O3 -rdynamic -ldl -laio -lm

ONNX Runtime

Model: GPT-2 - Device: CPU - Executor: Standard

OpenBenchmarking.orgInferences Per Minute, More Is BetterONNX Runtime 1.11Model: GPT-2 - Device: CPU - Executor: Standardarmv8.4-a+svearmv8.4-a3K6K9K12K15KSE +/- 12.91, N = 3SE +/- 63.90, N = 31231712364-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -ffunction-sections -fdata-sections -march=native -mtune=native -flto -fno-fat-lto-objects -ldl -lrt

ONNX Runtime

Model: bertsquad-12 - Device: CPU - Executor: Standard

OpenBenchmarking.orgInferences Per Minute, More Is BetterONNX Runtime 1.11Model: bertsquad-12 - Device: CPU - Executor: Standardarmv8.4-a+svearmv8.4-a170340510680850SE +/- 0.50, N = 3SE +/- 0.44, N = 3772773-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -ffunction-sections -fdata-sections -march=native -mtune=native -flto -fno-fat-lto-objects -ldl -lrt

ONNX Runtime

Model: fcn-resnet101-11 - Device: CPU - Executor: Standard

OpenBenchmarking.orgInferences Per Minute, More Is BetterONNX Runtime 1.11Model: fcn-resnet101-11 - Device: CPU - Executor: Standardarmv8.4-a+svearmv8.4-a1632486480SE +/- 0.00, N = 3SE +/- 0.00, N = 37373-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -ffunction-sections -fdata-sections -march=native -mtune=native -flto -fno-fat-lto-objects -ldl -lrt

ONNX Runtime

Model: ArcFace ResNet-100 - Device: CPU - Executor: Standard

OpenBenchmarking.orgInferences Per Minute, More Is BetterONNX Runtime 1.11Model: ArcFace ResNet-100 - Device: CPU - Executor: Standardarmv8.4-a+svearmv8.4-a2004006008001000SE +/- 0.17, N = 3SE +/- 0.88, N = 3935938-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -ffunction-sections -fdata-sections -march=native -mtune=native -flto -fno-fat-lto-objects -ldl -lrt

ONNX Runtime

Model: super-resolution-10 - Device: CPU - Executor: Standard

OpenBenchmarking.orgInferences Per Minute, More Is BetterONNX Runtime 1.11Model: super-resolution-10 - Device: CPU - Executor: Standardarmv8.4-a+svearmv8.4-a12002400360048006000SE +/- 2.17, N = 3SE +/- 0.93, N = 354115413-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -ffunction-sections -fdata-sections -march=native -mtune=native -flto -fno-fat-lto-objects -ldl -lrt

WavPack Audio Encoding

WAV To WavPack

OpenBenchmarking.orgSeconds, Fewer Is BetterWavPack Audio Encoding 5.3WAV To WavPackarmv8.4-a+svearmv8.4-a510152025SE +/- 0.03, N = 5SE +/- 0.00, N = 520.5220.49-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -rdynamic

GnuPG

2.7GB Sample File Encryption

OpenBenchmarking.orgSeconds, Fewer Is BetterGnuPG 2.2.272.7GB Sample File Encryptionarmv8.4-a+svearmv8.4-a1326395265SE +/- 0.54, N = 8SE +/- 0.02, N = 358.1457.20-march=armv8.4-a+sve-march=armv8.4-a1. (CC) gcc options: -O3

Kripke

OpenBenchmarking.orgThroughput FoM, More Is BetterKripke 1.2.4armv8.4-a+svearmv8.4-a40M80M120M160M200MSE +/- 298633.53, N = 3SE +/- 226703.11, N = 3192709233204143167-march=armv8.4-a+sve-march=armv8.4-a1. (CXX) g++ options: -O3 -fopenmp


Phoronix Test Suite v10.8.4