nvidia-gpu-compute-oring-agx-32g

ARMv8 Cortex-A78E testing with a NVIDIA Jetson AGX Orin Developer Kit (36.3.0-gcid-36191598 BIOS) and Orin on Ubuntu 22.04 via the Phoronix Test Suite.

Compare your own system(s) to this result file with the Phoronix Test Suite by running the command: phoronix-test-suite benchmark 2408272-NE-NVIDIAGPU37

Jump To Table - Results

baseline

Processor: ARMv8 Cortex-A78E @ 2.20GHz (12 Cores), Motherboard: NVIDIA Jetson AGX Orin Developer Kit (36.3.0-gcid-36191598 BIOS), Memory: 30GB, Disk: 1000GB Samsung SSD 960 EVO 1TB + 64GB G1M15M, Graphics: Orin, Network: Realtek RTL8822CE 802.11ac PCIe

OS: Ubuntu 22.04, Kernel: 5.15.136-tegra (aarch64), Desktop: GNOME Shell 42.9, Display Server: X Server 1.21.1.4, Display Driver: NVIDIA, Vulkan: 1.3.251, Compiler: GCC 11.4.0 + CUDA 12.2, File-System: ext4, Screen Resolution: 6582x1234

Kernel Notes: Transparent Huge Pages: always
Compiler Notes: --build=aarch64-linux-gnu --disable-libquadmath --disable-libquadmath-support --disable-werror --enable-bootstrap --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-fix-cortex-a53-843419 --enable-gnu-unique-object --enable-languages=c,ada,c++,go,d,fortran,objc,obj-c++,m2 --enable-libphobos-checking=release --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-link-serialization=2 --enable-multiarch --enable-nls --enable-objc-gc=auto --enable-plugin --enable-shared --enable-threads=posix --host=aarch64-linux-gnu --program-prefix=aarch64-linux-gnu- --target=aarch64-linux-gnu --with-build-config=bootstrap-lto-lean --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-target-system-zlib=auto -v
Processor Notes: Scaling Governor: tegra194 performance
Graphics Notes: BAR1 / Visible vRAM Size: N/A
Python Notes: Python 3.10.12
Security Notes: gather_data_sampling: Not affected + itlb_multihit: Not affected + l1tf: Not affected + mds: Not affected + meltdown: Not affected + mmio_stale_data: Not affected + retbleed: Not affected + spec_rstack_overflow: Not affected + spec_store_bypass: Mitigation of SSB disabled via prctl + spectre_v1: Mitigation of __user pointer sanitization + spectre_v2: Mitigation of CSV2 but not BHB + srbds: Not affected + tsx_async_abort: Not affected

VkFFT

VkFFT is a Fast Fourier Transform (FFT) Library that is GPU accelerated by means of the Vulkan API. The VkFFT benchmark runs FFT performance differences of many different sizes before returning an overall benchmark score. Learn more via the OpenBenchmarking.org test page.

Libplacebo

Libplacebo is a multimedia rendering library based on the core rendering code of the MPV player. The libplacebo benchmark relies on the Vulkan API and tests various primitives. Learn more via the OpenBenchmarking.org test page.

baseline: The test quit with a non-zero exit status. E: ./libplacebo: 3: ./src/bench: not found

cl-mem

A basic OpenCL memory benchmark. Learn more via the OpenBenchmarking.org test page.

Benchmark: Copy

baseline: The test quit with a non-zero exit status.

Benchmark: Read

baseline: The test quit with a non-zero exit status.

Benchmark: Write

baseline: The test quit with a non-zero exit status.

Betsy GPU Compressor

Betsy is an open-source GPU compressor of various GPU compression techniques. Betsy is written in GLSL for Vulkan/OpenGL (compute shader) support for GPU-based texture compression. Learn more via the OpenBenchmarking.org test page.

Codec: ETC1 - Quality: Highest

baseline: The test quit with a non-zero exit status. E: ./betsy: 3: ./betsy: not found

Codec: ETC2 RGB - Quality: Highest

baseline: The test quit with a non-zero exit status. E: ./betsy: 3: ./betsy: not found

VkResample

VkResample is a Vulkan-based image upscaling library based on VkFFT. The sample input file is upscaling a 4K image to 8K using Vulkan-based GPU acceleration. Learn more via the OpenBenchmarking.org test page.

FAHBench

FAHBench is a Folding@Home benchmark on the GPU. Learn more via the OpenBenchmarking.org test page.

baseline: The test quit with a non-zero exit status.

ViennaCL

ViennaCL is an open-source linear algebra library written in C++ and with support for OpenCL and OpenMP. This test profile makes use of ViennaCL's built-in benchmarks. Learn more via the OpenBenchmarking.org test page.

NCNN

NCNN is a high performance neural network inference framework optimized for mobile and other platforms developed by Tencent. Learn more via the OpenBenchmarking.org test page.

40 Results Shown

VkFFT:
FFT + iFFT R2C / C2R
FFT + iFFT C2C 1D batched in half precision
FFT + iFFT C2C Bluestein in single precision
FFT + iFFT C2C 1D batched in double precision
FFT + iFFT C2C 1D batched in single precision
FFT + iFFT C2C multidimensional in single precision
FFT + iFFT C2C Bluestein benchmark in double precision
FFT + iFFT C2C 1D batched in single precision, no reshuffling
VkResample:
2x - Double
2x - Single
ViennaCL:
CPU BLAS - sCOPY
CPU BLAS - sAXPY
CPU BLAS - sDOT
CPU BLAS - dCOPY
CPU BLAS - dAXPY
CPU BLAS - dDOT
CPU BLAS - dGEMV-N
CPU BLAS - dGEMV-T
CPU BLAS - dGEMM-NN
CPU BLAS - dGEMM-NT
CPU BLAS - dGEMM-TN
CPU BLAS - dGEMM-TT
NCNN:
Vulkan GPU - mobilenet
Vulkan GPU-v2-v2 - mobilenet-v2
Vulkan GPU-v3-v3 - mobilenet-v3
Vulkan GPU - shufflenet-v2
Vulkan GPU - mnasnet
Vulkan GPU - efficientnet-b0
Vulkan GPU - blazeface
Vulkan GPU - googlenet
Vulkan GPU - vgg16
Vulkan GPU - resnet18
Vulkan GPU - alexnet
Vulkan GPU - resnet50
Vulkan GPUv2-yolov3v2-yolov3 - mobilenetv2-yolov3
Vulkan GPU - yolov4-tiny
Vulkan GPU - squeezenet_ssd
Vulkan GPU - regnety_400m
Vulkan GPU - vision_transformer
Vulkan GPU - FastestDet

baseline

Testing initiated at 27 August 2024 10:46 by user mclark.

nvidia-gpu-compute-oring-agx-32g

Statistics

Graph Settings

Multi-Way Comparison

Table

Run Management

baseline

VkFFT

Libplacebo

cl-mem

Betsy GPU Compressor

VkResample

FAHBench

ViennaCL

NCNN

40 Results Shown

baseline