nvidia_opencl_linux

AMD wx4150 on Ubuntu 20.04.2 with ROCM fan speed max

Compare your own system(s) to this result file with the Phoronix Test Suite by running the command: phoronix-test-suite benchmark 2103143-HA-2011304FI79
Jump To Table - Results

View

Do Not Show Noisy Results
Do Not Show Results With Incomplete Data
Do Not Show Results With Little Change/Spread
List Notable Results

Limit displaying results to tests within:

CPU Massive 3 Tests
Creator Workloads 2 Tests
HPC - High Performance Computing 2 Tests
Multi-Core 2 Tests
NVIDIA GPU Compute 5 Tests
OpenCL 7 Tests
Server CPU Tests 2 Tests
Common Workstation Benchmarks 2 Tests

Statistics

Show Overall Harmonic Mean(s)
Show Overall Geometric Mean
Show Geometric Means Per-Suite/Category
Show Wins / Losses Counts (Pie Chart)
Normalize Results
Remove Outliers Before Calculating Averages

Graph Settings

Force Line Graphs Where Applicable
Convert To Scalar Where Applicable
Disable Color Branding
Prefer Vertical Bar Graphs

Multi-Way Comparison

Condense Multi-Option Tests Into Single Result Graphs

Table

Show Detailed System Result Table

Run Management

Highlight
Result
Hide
Result
Result
Identifier
Performance Per
Dollar
Date
Run
  Test
  Duration
nvidia_opencl_linux
December 01 2020
  4 Hours, 2 Minutes
amd_opencl_linux
March 14 2021
  1 Hour, 13 Minutes
amd-opencl-linux
March 14 2021
  1 Hour, 11 Minutes
Invert Hiding All Results Option
  2 Hours, 8 Minutes

Only show results where is faster than
Only show results matching title/arguments (delimit multiple options with a comma):
Do not show results matching title/arguments (delimit multiple options with a comma):


nvidia_opencl_linuxProcessorMotherboardChipsetMemoryDiskGraphicsAudioNetworkMonitorOSKernelDesktopDisplay ServerDisplay DriverOpenCLVulkanCompilerFile-SystemScreen ResolutionOpenGLnvidia_opencl_linuxamd_opencl_linuxamd-opencl-linuxIntel Core i7-4700MQ @ 3.40GHz (4 Cores / 8 Threads)HP 1909 (L70 Ver. 01.42 BIOS)Intel Xeon E3-1200 v3/4th32GB500GB Samsung SSD 860 + 256GB SAMSUNG MZ7PD256 + 500GB Seagate ST500LT012-1DG14 + 256GB SAMSUNG MZMPD256 + 128GB ED2S5NVIDIA Quadro M1000M 2GB (135/405MHz)IDT 92HD91BXXIntel I217-LM + Intel 7260Ubuntu 20.045.4.0-53-generic (x86_64)GNOME Shell 3.36.4X Server 1.20.8NVIDIA 450.80.02OpenCL 1.2 CUDA 11.0.2281.2.131GCC 9.3.0ext41920x1200HP 1909 (L70 Ver. 01.45 BIOS)500GB Samsung SSD 860 + 500GB Seagate ST500LT012-1DG14Intel HD 4600 2GB (1150MHz)HP ZR24w5.6.0-1042-oem (x86_64)X Server 1.20.94.5 Mesa 20.2.6OpenCL 2.0 AMD-APP (3212.0)1.2.1453840x1200OpenBenchmarking.orgCompiler Details- --build=x86_64-linux-gnu --disable-vtable-verify --disable-werror --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-gnu-unique-object --enable-languages=c,ada,c++,go,brig,d,fortran,objc,obj-c++,gm2 --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-multilib --enable-nls --enable-objc-gc=auto --enable-offload-targets=nvptx-none=/build/gcc-9-HskZEa/gcc-9-9.3.0/debian/tmp-nvptx/usr,hsa --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --program-prefix=x86_64-linux-gnu- --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-default-libstdcxx-abi=new --with-gcc-major-version-only --with-multilib-list=m32,m64,mx32 --with-target-system-zlib=auto --with-tune=generic --without-cuda-driver -v Processor Details- Scaling Governor: intel_pstate powersave - CPU Microcode: 0x28 - Thermald 1.9.1OpenCL Details- nvidia_opencl_linux: GPU Compute Cores: 512Python Details- nvidia_opencl_linux: Python 3.8.5Security Details- itlb_multihit: KVM: Mitigation of Split huge pages + l1tf: Mitigation of PTE Inversion; VMX: conditional cache flushes SMT vulnerable + mds: Mitigation of Clear buffers; SMT vulnerable + meltdown: Mitigation of PTI + spec_store_bypass: Mitigation of SSB disabled via prctl and seccomp + spectre_v1: Mitigation of usercopy/swapgs barriers and __user pointer sanitization + spectre_v2: Mitigation of Full generic retpoline IBPB: conditional IBRS_FW STIBP: conditional RSB filling + srbds: Mitigation of Microcode + tsx_async_abort: Not affected Kernel Details- amd_opencl_linux, amd-opencl-linux: Transparent Huge Pages: madviseGraphics Details- amd_opencl_linux, amd-opencl-linux: GLAMOR

nvidia_opencl_linuxamd_opencl_linuxamd-opencl-linuxResult OverviewPhoronix Test Suite100%133%167%200%233%RodiniaSHOC Scalable HeterOgeneous ComputingclpeakDarktablecl-memLuxMarkSmallPT GPU

nvidia_opencl_linuxblender: BMW27 - OpenCLblender: Barbershop - OpenCLcl-mem: Copycl-mem: Readcl-mem: Writeclpeak: Kernel Latencyclpeak: Integer Compute INTclpeak: Single-Precision Floatclpeak: Double-Precision Doubleclpeak: Global Memory Bandwidthclpeak: Transfer Bandwidth enqueueReadBufferclpeak: Transfer Bandwidth enqueueWriteBufferdarktable: Boat - OpenCLdarktable: Masskrug - OpenCLdarktable: Server Rack - OpenCLdarktable: Server Room - OpenCLluxmark: GPU - Hotelluxmark: GPU - Microphoneluxmark: GPU - Luxball HDRrodinia: OpenCL LavaMDrodinia: OpenCL Myocyterodinia: OpenCL Heartwallrodinia: OpenCL Particle Filtershoc: OpenCL - Triadshoc: OpenCL - FFT SPshoc: OpenCL - MD5 Hashshoc: OpenCL - Max SP Flopsshoc: OpenCL - Bus Speed Downloadshoc: OpenCL - Bus Speed Readbackshoc: OpenCL - Texture Read Bandwidthsmallpt-gpu: GPU - 1920 x 1200 - Causticsmallpt-gpu: GPU - 1920 x 1200 - Cornellsmallpt-gpu: GPU - 1920 x 1200 - Caustic3nvidia_opencl_linuxamd_opencl_linuxamd-opencl-linux694.902379.3360.167.463.37.75251.31700.2935.6366.996.6310.9311.03110.8220.3234.594751251237303.95155.3385.97845.30610.5663122.3731.43221130.1212.687812.7640110.99016067567481606756869160675699568.579.674.16.13368.481823.80115.4475.2110.3619.398.0479.7510.2232.90249037575322230.3457.7984.7502217.1392.35029081355.66905.246181.667116157009271615701063161570120368.579.573.96.15367.951824.96115.5475.2610.4119.468.0839.7760.2282.92649137675324229.8597.8104.7505217.3322.34929086055.66585.244781.8607161570693116157070681615707207OpenBenchmarking.org

Blender

Blender is an open-source 3D creation software project. This test is of Blender's Cycles benchmark with various sample files. GPU computing via OpenCL or CUDA is supported. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgSeconds, Fewer Is BetterBlender 2.90Blend File: BMW27 - Compute: OpenCLnvidia_opencl_linux150300450600750SE +/- 10.50, N = 3694.90

OpenBenchmarking.orgSeconds, Fewer Is BetterBlender 2.90Blend File: Barbershop - Compute: OpenCLnvidia_opencl_linux5001000150020002500SE +/- 5.31, N = 32379.33

cl-mem

A basic OpenCL memory benchmark. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgGB/s, More Is Bettercl-mem 2017-01-13Benchmark: Copyamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux1530456075SE +/- 0.06, N = 3SE +/- 0.12, N = 3SE +/- 0.00, N = 368.568.560.11. (CC) gcc options: -O2 -flto -lOpenCL
OpenBenchmarking.orgGB/s, More Is Bettercl-mem 2017-01-13Benchmark: Copyamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux1326395265Min: 68.4 / Avg: 68.5 / Max: 68.6Min: 68.3 / Avg: 68.5 / Max: 68.7Min: 60.1 / Avg: 60.1 / Max: 60.11. (CC) gcc options: -O2 -flto -lOpenCL

OpenBenchmarking.orgGB/s, More Is Bettercl-mem 2017-01-13Benchmark: Readamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux20406080100SE +/- 0.07, N = 3SE +/- 0.03, N = 3SE +/- 0.03, N = 379.579.667.41. (CC) gcc options: -O2 -flto -lOpenCL
OpenBenchmarking.orgGB/s, More Is Bettercl-mem 2017-01-13Benchmark: Readamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux1530456075Min: 79.4 / Avg: 79.47 / Max: 79.6Min: 79.6 / Avg: 79.63 / Max: 79.7Min: 67.3 / Avg: 67.37 / Max: 67.41. (CC) gcc options: -O2 -flto -lOpenCL

OpenBenchmarking.orgGB/s, More Is Bettercl-mem 2017-01-13Benchmark: Writeamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux1632486480SE +/- 0.12, N = 3SE +/- 0.13, N = 3SE +/- 0.03, N = 373.974.163.31. (CC) gcc options: -O2 -flto -lOpenCL
OpenBenchmarking.orgGB/s, More Is Bettercl-mem 2017-01-13Benchmark: Writeamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux1428425670Min: 73.7 / Avg: 73.93 / Max: 74.1Min: 73.8 / Avg: 74.07 / Max: 74.2Min: 63.3 / Avg: 63.33 / Max: 63.41. (CC) gcc options: -O2 -flto -lOpenCL

clpeak

Clpeak is designed to test the peak capabilities of OpenCL devices. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgus, Fewer Is BetterclpeakOpenCL Test: Kernel Latencyamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux246810SE +/- 0.06, N = 5SE +/- 0.06, N = 7SE +/- 0.05, N = 36.156.137.751. (CXX) g++ options: -O3 -rdynamic -lOpenCL
OpenBenchmarking.orgus, Fewer Is BetterclpeakOpenCL Test: Kernel Latencyamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215Min: 5.99 / Avg: 6.15 / Max: 6.37Min: 5.94 / Avg: 6.13 / Max: 6.33Min: 7.68 / Avg: 7.75 / Max: 7.841. (CXX) g++ options: -O3 -rdynamic -lOpenCL

OpenBenchmarking.orgGIOPS, More Is BetterclpeakOpenCL Test: Integer Compute INTamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux80160240320400SE +/- 0.05, N = 3SE +/- 0.11, N = 3SE +/- 1.52, N = 3367.95368.48251.311. (CXX) g++ options: -O3 -rdynamic -lOpenCL
OpenBenchmarking.orgGIOPS, More Is BetterclpeakOpenCL Test: Integer Compute INTamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux70140210280350Min: 367.86 / Avg: 367.95 / Max: 368Min: 368.26 / Avg: 368.48 / Max: 368.59Min: 248.26 / Avg: 251.31 / Max: 252.831. (CXX) g++ options: -O3 -rdynamic -lOpenCL

OpenBenchmarking.orgGFLOPS, More Is BetterclpeakOpenCL Test: Single-Precision Floatamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux400800120016002000SE +/- 0.08, N = 3SE +/- 0.11, N = 3SE +/- 0.31, N = 31824.961823.80700.291. (CXX) g++ options: -O3 -rdynamic -lOpenCL
OpenBenchmarking.orgGFLOPS, More Is BetterclpeakOpenCL Test: Single-Precision Floatamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux30060090012001500Min: 1824.84 / Avg: 1824.96 / Max: 1825.1Min: 1823.66 / Avg: 1823.8 / Max: 1824.02Min: 699.78 / Avg: 700.29 / Max: 700.851. (CXX) g++ options: -O3 -rdynamic -lOpenCL

OpenBenchmarking.orgGFLOPS, More Is BetterclpeakOpenCL Test: Double-Precision Doubleamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux306090120150SE +/- 0.02, N = 3SE +/- 0.11, N = 3SE +/- 0.02, N = 3115.54115.4435.631. (CXX) g++ options: -O3 -rdynamic -lOpenCL
OpenBenchmarking.orgGFLOPS, More Is BetterclpeakOpenCL Test: Double-Precision Doubleamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux20406080100Min: 115.51 / Avg: 115.54 / Max: 115.56Min: 115.29 / Avg: 115.44 / Max: 115.65Min: 35.61 / Avg: 35.63 / Max: 35.661. (CXX) g++ options: -O3 -rdynamic -lOpenCL

OpenBenchmarking.orgGBPS, More Is BetterclpeakOpenCL Test: Global Memory Bandwidthamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux20406080100SE +/- 0.01, N = 3SE +/- 0.06, N = 3SE +/- 0.05, N = 375.2675.2166.991. (CXX) g++ options: -O3 -rdynamic -lOpenCL
OpenBenchmarking.orgGBPS, More Is BetterclpeakOpenCL Test: Global Memory Bandwidthamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux1428425670Min: 75.23 / Avg: 75.26 / Max: 75.27Min: 75.1 / Avg: 75.21 / Max: 75.28Min: 66.89 / Avg: 66.99 / Max: 67.061. (CXX) g++ options: -O3 -rdynamic -lOpenCL

OpenBenchmarking.orgGBPS, More Is BetterclpeakOpenCL Test: Transfer Bandwidth enqueueReadBufferamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215SE +/- 0.04, N = 3SE +/- 0.07, N = 3SE +/- 0.01, N = 310.4110.366.631. (CXX) g++ options: -O3 -rdynamic -lOpenCL
OpenBenchmarking.orgGBPS, More Is BetterclpeakOpenCL Test: Transfer Bandwidth enqueueReadBufferamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215Min: 10.32 / Avg: 10.41 / Max: 10.46Min: 10.29 / Avg: 10.36 / Max: 10.49Min: 6.62 / Avg: 6.63 / Max: 6.641. (CXX) g++ options: -O3 -rdynamic -lOpenCL

OpenBenchmarking.orgGBPS, More Is BetterclpeakOpenCL Test: Transfer Bandwidth enqueueWriteBufferamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux510152025SE +/- 0.01, N = 3SE +/- 0.03, N = 3SE +/- 0.06, N = 319.4619.3910.931. (CXX) g++ options: -O3 -rdynamic -lOpenCL
OpenBenchmarking.orgGBPS, More Is BetterclpeakOpenCL Test: Transfer Bandwidth enqueueWriteBufferamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux510152025Min: 19.44 / Avg: 19.46 / Max: 19.48Min: 19.35 / Avg: 19.39 / Max: 19.45Min: 10.81 / Avg: 10.93 / Max: 111. (CXX) g++ options: -O3 -rdynamic -lOpenCL

Darktable

Darktable is an open-source photography / workflow application this will use any system-installed Darktable program or on Windows will automatically download the pre-built binary from the project. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgSeconds, Fewer Is BetterDarktable 3.0.1Test: Boat - Acceleration: OpenCLamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215SE +/- 0.016, N = 3SE +/- 0.044, N = 3SE +/- 0.011, N = 38.0838.04711.031
OpenBenchmarking.orgSeconds, Fewer Is BetterDarktable 3.0.1Test: Boat - Acceleration: OpenCLamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215Min: 8.05 / Avg: 8.08 / Max: 8.1Min: 7.96 / Avg: 8.05 / Max: 8.1Min: 11.01 / Avg: 11.03 / Max: 11.05

OpenBenchmarking.orgSeconds, Fewer Is BetterDarktable 3.0.1Test: Masskrug - Acceleration: OpenCLamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215SE +/- 0.016, N = 3SE +/- 0.032, N = 3SE +/- 0.011, N = 39.7769.75110.822
OpenBenchmarking.orgSeconds, Fewer Is BetterDarktable 3.0.1Test: Masskrug - Acceleration: OpenCLamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215Min: 9.74 / Avg: 9.78 / Max: 9.8Min: 9.69 / Avg: 9.75 / Max: 9.79Min: 10.8 / Avg: 10.82 / Max: 10.84

OpenBenchmarking.orgSeconds, Fewer Is BetterDarktable 3.0.1Test: Server Rack - Acceleration: OpenCLamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux0.07270.14540.21810.29080.3635SE +/- 0.003, N = 3SE +/- 0.002, N = 15SE +/- 0.001, N = 30.2280.2230.323
OpenBenchmarking.orgSeconds, Fewer Is BetterDarktable 3.0.1Test: Server Rack - Acceleration: OpenCLamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux12345Min: 0.22 / Avg: 0.23 / Max: 0.23Min: 0.2 / Avg: 0.22 / Max: 0.23Min: 0.32 / Avg: 0.32 / Max: 0.32

OpenBenchmarking.orgSeconds, Fewer Is BetterDarktable 3.0.1Test: Server Room - Acceleration: OpenCLamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux1.03372.06743.10114.13485.1685SE +/- 0.030, N = 3SE +/- 0.021, N = 15SE +/- 0.007, N = 32.9262.9024.594
OpenBenchmarking.orgSeconds, Fewer Is BetterDarktable 3.0.1Test: Server Room - Acceleration: OpenCLamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux246810Min: 2.87 / Avg: 2.93 / Max: 2.96Min: 2.66 / Avg: 2.9 / Max: 2.96Min: 4.58 / Avg: 4.59 / Max: 4.61

LuxMark

LuxMark is a multi-platform OpenGL benchmark using LuxRender. LuxMark supports targeting different OpenCL devices and has multiple scenes available for rendering. LuxMark is a fully open-source OpenCL program with real-world rendering examples. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.1OpenCL Device: GPU - Scene: Hotelamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux160320480640800SE +/- 0.67, N = 3SE +/- 3.48, N = 3491490751
OpenBenchmarking.orgScore, More Is BetterLuxMark 3.1OpenCL Device: GPU - Scene: Hotelamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux130260390520650Min: 489 / Avg: 490.33 / Max: 491Min: 745 / Avg: 750.67 / Max: 757

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.1OpenCL Device: GPU - Scene: Microphoneamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux8001600240032004000SE +/- 10.84, N = 3SE +/- 2.67, N = 3376737572512
OpenBenchmarking.orgScore, More Is BetterLuxMark 3.1OpenCL Device: GPU - Scene: Microphoneamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux7001400210028003500Min: 3735 / Avg: 3756.67 / Max: 3768Min: 2509 / Avg: 2511.67 / Max: 2517

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.1OpenCL Device: GPU - Scene: Luxball HDRamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux11002200330044005500SE +/- 3.21, N = 3SE +/- 12.68, N = 3532453223730
OpenBenchmarking.orgScore, More Is BetterLuxMark 3.1OpenCL Device: GPU - Scene: Luxball HDRamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux9001800270036004500Min: 5318 / Avg: 5324 / Max: 5329Min: 3705 / Avg: 3730.33 / Max: 3744

Rodinia

Rodinia is a suite focused upon accelerating compute-intensive applications with accelerators. CUDA, OpenMP, and OpenCL parallel models are supported by the included applications. This profile utilizes select OpenCL, NVIDIA CUDA and OpenMP test binaries at the moment. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenCL LavaMDnvidia_opencl_linux0.8891.7782.6673.5564.445SE +/- 0.052, N = 53.9511. (CXX) g++ options: -O2 -lOpenCL

OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenCL Myocyteamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux50100150200250SE +/- 0.24, N = 3SE +/- 0.41, N = 3SE +/- 0.27, N = 3229.86230.3555.341. (CXX) g++ options: -O2 -lOpenCL
OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenCL Myocyteamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux4080120160200Min: 229.51 / Avg: 229.86 / Max: 230.31Min: 229.61 / Avg: 230.34 / Max: 231.04Min: 55.04 / Avg: 55.34 / Max: 55.881. (CXX) g++ options: -O2 -lOpenCL

OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenCL Heartwallamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux246810SE +/- 0.013, N = 3SE +/- 0.026, N = 3SE +/- 0.057, N = 147.8107.7985.9781. (CXX) g++ options: -O2 -lOpenCL
OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenCL Heartwallamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215Min: 7.78 / Avg: 7.81 / Max: 7.83Min: 7.75 / Avg: 7.8 / Max: 7.83Min: 5.89 / Avg: 5.98 / Max: 6.721. (CXX) g++ options: -O2 -lOpenCL

OpenBenchmarking.orgSeconds, Fewer Is BetterRodinia 3.1Test: OpenCL Particle Filternvidia_opencl_linux1020304050SE +/- 0.03, N = 345.311. (CXX) g++ options: -O2 -lOpenCL

SHOC Scalable HeterOgeneous Computing

The CUDA and OpenCL version of Vetter's Scalable HeterOgeneous Computing benchmark suite. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Triadamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215SE +/- 0.0336, N = 3SE +/- 0.0171, N = 3SE +/- 0.0068, N = 34.75054.750210.56631. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi
OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Triadamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215Min: 4.7 / Avg: 4.75 / Max: 4.81Min: 4.73 / Avg: 4.75 / Max: 4.78Min: 10.55 / Avg: 10.57 / Max: 10.581. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi

OpenBenchmarking.orgGFLOPS, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: FFT SPamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux50100150200250SE +/- 0.19, N = 3SE +/- 0.28, N = 3SE +/- 0.82, N = 3217.33217.14122.371. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi
OpenBenchmarking.orgGFLOPS, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: FFT SPamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux4080120160200Min: 217 / Avg: 217.33 / Max: 217.66Min: 216.84 / Avg: 217.14 / Max: 217.69Min: 120.73 / Avg: 122.37 / Max: 123.211. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi

OpenBenchmarking.orgGHash/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: MD5 Hashamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux0.52881.05761.58642.11522.644SE +/- 0.0007, N = 3SE +/- 0.0003, N = 3SE +/- 0.0007, N = 32.34922.35021.43221. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi
OpenBenchmarking.orgGHash/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: MD5 Hashamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux246810Min: 2.35 / Avg: 2.35 / Max: 2.35Min: 2.35 / Avg: 2.35 / Max: 2.35Min: 1.43 / Avg: 1.43 / Max: 1.431. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi

OpenBenchmarking.orgGFLOPS, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Max SP Flopsamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux200K400K600K800K1000KSE +/- 5904.67, N = 3SE +/- 1706.69, N = 3SE +/- 1.98, N = 3908605.00908135.001130.121. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi
OpenBenchmarking.orgGFLOPS, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Max SP Flopsamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux160K320K480K640K800KMin: 896808 / Avg: 908605 / Max: 914971Min: 904918 / Avg: 908134.67 / Max: 910732Min: 1127.91 / Avg: 1130.12 / Max: 1134.071. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi

OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Bus Speed Downloadamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215SE +/- 0.0024, N = 3SE +/- 0.0012, N = 3SE +/- 0.0011, N = 35.66585.669012.68781. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi
OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Bus Speed Downloadamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux48121620Min: 5.66 / Avg: 5.67 / Max: 5.67Min: 5.67 / Avg: 5.67 / Max: 5.67Min: 12.69 / Avg: 12.69 / Max: 12.691. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi

OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Bus Speed Readbackamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux3691215SE +/- 0.0023, N = 3SE +/- 0.0057, N = 3SE +/- 0.0014, N = 35.24475.246112.76401. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi
OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Bus Speed Readbackamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux48121620Min: 5.24 / Avg: 5.24 / Max: 5.25Min: 5.24 / Avg: 5.25 / Max: 5.26Min: 12.76 / Avg: 12.76 / Max: 12.771. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi

OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Texture Read Bandwidthamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux20406080100SE +/- 0.37, N = 3SE +/- 0.21, N = 3SE +/- 0.57, N = 381.8681.67110.991. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi
OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Texture Read Bandwidthamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux20406080100Min: 81.37 / Avg: 81.86 / Max: 82.58Min: 81.27 / Avg: 81.67 / Max: 82Min: 110.33 / Avg: 110.99 / Max: 112.121. (CXX) g++ options: -O2 -lSHOCCommonMPI -lSHOCCommonOpenCL -lSHOCCommon -lOpenCL -lrt -pthread -lmpi_cxx -lmpi

SmallPT GPU

SmallPT GPU is an OpenCL benchmark that's run with various PTS changes compared to upstream and multiple rendering scenes are available. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgSamples/sec, More Is BetterSmallPT GPU 1.6pts1OpenCL Device: GPU - Resolution: 1920 x 1200 - Scene: Causticamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux300M600M900M1200M1500MSE +/- 25.98, N = 3SE +/- 25.98, N = 3SE +/- 24.25, N = 31615706931161570092716067567481. (CC) gcc options: -O3 -lm -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL
OpenBenchmarking.orgSamples/sec, More Is BetterSmallPT GPU 1.6pts1OpenCL Device: GPU - Resolution: 1920 x 1200 - Scene: Causticamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux300M600M900M1200M1500MMin: 1615706886 / Avg: 1615706931 / Max: 1615706976Min: 1615700882 / Avg: 1615700927 / Max: 1615700972Min: 1606756706 / Avg: 1606756748 / Max: 16067567901. (CC) gcc options: -O3 -lm -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL

OpenBenchmarking.orgSamples/sec, More Is BetterSmallPT GPU 1.6pts1OpenCL Device: GPU - Resolution: 1920 x 1200 - Scene: Cornellamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux300M600M900M1200M1500MSE +/- 24.54, N = 3SE +/- 24.54, N = 3SE +/- 20.78, N = 31615707068161570106316067568691. (CC) gcc options: -O3 -lm -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL
OpenBenchmarking.orgSamples/sec, More Is BetterSmallPT GPU 1.6pts1OpenCL Device: GPU - Resolution: 1920 x 1200 - Scene: Cornellamd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux300M600M900M1200M1500MMin: 1615707025 / Avg: 1615707067.67 / Max: 1615707110Min: 1615701021 / Avg: 1615701063.33 / Max: 1615701106Min: 1606756833 / Avg: 1606756869 / Max: 16067569051. (CC) gcc options: -O3 -lm -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL

OpenBenchmarking.orgSamples/sec, More Is BetterSmallPT GPU 1.6pts1OpenCL Device: GPU - Resolution: 1920 x 1200 - Scene: Caustic3amd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux300M600M900M1200M1500MSE +/- 25.98, N = 3SE +/- 26.27, N = 3SE +/- 23.96, N = 31615707207161570120316067569951. (CC) gcc options: -O3 -lm -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL
OpenBenchmarking.orgSamples/sec, More Is BetterSmallPT GPU 1.6pts1OpenCL Device: GPU - Resolution: 1920 x 1200 - Scene: Caustic3amd-opencl-linuxamd_opencl_linuxnvidia_opencl_linux300M600M900M1200M1500MMin: 1615707162 / Avg: 1615707207 / Max: 1615707252Min: 1615701157 / Avg: 1615701202.67 / Max: 1615701248Min: 1606756954 / Avg: 1606756995.33 / Max: 16067570371. (CC) gcc options: -O3 -lm -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL