OpenCL CUDA NVIDIA GPGPU Linux Tests

All Maxwell and various Kepler graphics cards tested on the NVIDIA Linux driver. Benchmarks by Michael Larabel for a future article on Phoronix.com just delivering various GPGPU benchmarks for reference purposes.

HTML result view exported from: https://openbenchmarking.org/result/1710233-TY-1511113PT43&export=txt&grt&rdt&rro.

OpenCL CUDA NVIDIA GPGPU Linux TestsProcessorMotherboardChipsetMemoryDiskGraphicsAudioNetworkOSKernelDesktopDisplay ServerDisplay DriverOpenGLCompilerFile-SystemScreen ResolutionGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 760eVGA NVIDIA GeForce GTX 1060Intel Core i5-6600K @ 3.50GHz (4 Cores)MSI Z170A GAMING PRO (MS-7984) v1.0Intel Device 191f16384MB256GB TS256GSSD370SeVGA NVIDIA GeForce GTX 950 2048MB (135/405MHz)Intel Device a170Intel Device 15b8Ubuntu 14.043.19.0-33-generic (x86_64)Unity 7.2.5X Server 1.17.1NVIDIA 352.394.3.0GCC 4.8.4 + Clang 3.4-1ubuntu3 + CUDA 7.5ext43840x2160NVIDIA GeForce GTX 980 Ti 6144MB (999/3505MHz)eVGA NVIDIA GeForce GTX 970 4096MB (1163/3505MHz)NVIDIA GeForce GTX 980 4096MB (1126/3505MHz)eVGA NVIDIA GeForce GTX 960 2048MB (1277/3505MHz)NVIDIA GeForce GTX TITAN X 12288MB (1001/3505MHz)NVIDIA GeForce GTX 780 Ti 3072MB (875/3500MHz)NVIDIA GeForce GTX 680 2048MB (1006/3004MHz)eVGA NVIDIA GeForce GTX 750 1024MB (1019/2505MHz)NVIDIA GeForce GTX 760 2048MB (980/3004MHz)2 x Intel Xeon E5-2670 0 @ 3.30GHz (32 Cores)HP 158AIntel Xeon E5/Core500GB Samsung SSD 850 + 4001GB Seagate ST4000DM005-2DP1 + 500GB Seagate ST500DM002-1BD14eVGA NVIDIA GeForce GTX 1060 6GBRealtek ALC262Intel 82579LM Gigabit Connection + Broadcom Limited BCM4360 802.11ac WirelessManjaroLinux 17.0.54.9.53-1-MANJARO (x86_64)KDE Frameworks 5GCC 7.2.0 + Clang 5.0.0 + CUDA 8.0800x600OpenBenchmarking.orgCompiler Details- GeForce GTX 950: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX 980 Ti: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX 970: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX 980: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX 960: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX TITAN X: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX 780 Ti: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX 680: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX 750: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- GeForce GTX 760: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v- eVGA NVIDIA GeForce GTX 1060: --disable-libssp --disable-libstdcxx-pch --disable-libunwind-exceptions --disable-werror --enable-__cxa_atexit --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-default-ssp --enable-gnu-indirect-function --enable-gnu-unique-object --enable-install-libiberty --enable-languages=c,c++,ada,fortran,go,lto,objc,obj-c++ --enable-libmpx --enable-lto --enable-multilib --enable-plugin --enable-shared --enable-threads=posix --mandir=/usr/share/man --with-isl --with-linker-hash-style=gnuProcessor Details- GeForce GTX 950: Scaling Governor: acpi-cpufreq performance- GeForce GTX 980 Ti: Scaling Governor: acpi-cpufreq performance- GeForce GTX 970: Scaling Governor: acpi-cpufreq performance- GeForce GTX 980: Scaling Governor: acpi-cpufreq performance- GeForce GTX 960: Scaling Governor: acpi-cpufreq performance- GeForce GTX TITAN X: Scaling Governor: acpi-cpufreq performance- GeForce GTX 780 Ti: Scaling Governor: acpi-cpufreq performance- GeForce GTX 680: Scaling Governor: acpi-cpufreq performance- GeForce GTX 750: Scaling Governor: acpi-cpufreq performance- GeForce GTX 760: Scaling Governor: acpi-cpufreq performance- eVGA NVIDIA GeForce GTX 1060: Scaling Governor: intel_pstate powersaveOpenCL Details- GeForce GTX 950: GPU Compute Cores: 768- GeForce GTX 980 Ti: GPU Compute Cores: 2816- GeForce GTX 970: GPU Compute Cores: 1664- GeForce GTX 980: GPU Compute Cores: 2048- GeForce GTX 960: GPU Compute Cores: 1024- GeForce GTX TITAN X: GPU Compute Cores: 3072- GeForce GTX 780 Ti: GPU Compute Cores: 2880- GeForce GTX 680: GPU Compute Cores: 1536- GeForce GTX 750: GPU Compute Cores: 512- GeForce GTX 760: GPU Compute Cores: 1152System Details- GeForce GTX 950: GPU Compute Cores: 768.- GeForce GTX 980 Ti: GPU Compute Cores: 2816.- GeForce GTX 970: GPU Compute Cores: 1664.- GeForce GTX 980: GPU Compute Cores: 2048.- GeForce GTX 960: GPU Compute Cores: 1024.- GeForce GTX TITAN X: GPU Compute Cores: 3072.- GeForce GTX 780 Ti: GPU Compute Cores: 2880.- GeForce GTX 680: GPU Compute Cores: 1536.- GeForce GTX 750: GPU Compute Cores: 512.- GeForce GTX 760: GPU Compute Cores: 1152.

OpenCL CUDA NVIDIA GPGPU Linux Testsaskap: Griddingaskap: Degriddingcuda-mini-nbody: Originalcuda-mini-nbody: Cache Blockingcuda-mini-nbody: Loop Unrollingcuda-mini-nbody: SOA Data Layoutcuda-mini-nbody: Flush Denormals To Zerojuliagpu: GPUluxmark: GPU - Hotelluxmark: GPU - Microphoneluxmark: GPU - Luxball HDRmandelbulbgpu: GPUshoc: CUDA - FFT SPshoc: CUDA - MD5 Hashshoc: OpenCL - FFT SPshoc: OpenCL - MD5 Hashshoc: CUDA - Texture Read Bandwidthshoc: OpenCL - Texture Read BandwidthGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 760eVGA NVIDIA GeForce GTX 10603399.145706.07105.3049.8947.54108.50108.4864913682.637692423531337156070.87172.282.3663.222.34326.23239.198320.5017380.6034.5819.7718.4640.9440.85127978049.53185562681380271656708.83311.466.81170.366.79348.92345.555325.129509.1454.3228.5326.4255.8755.80104144917.2313464458973758811317.17263.144.79117.234.77325.16283.366051.271109445.3825.1323.8850.1549.53113830604.27149247761071363616558.77289.635.70140.125.68336.48332.603144.855290.3282.0137.0835.3579.9779.8480042041.738972460547444953399.47212.433.3862.783.36351.31269.988458.7717380.6032.3718.6517.5937.4337.37136037921.43190663601408175614774.13324.097.42173.897.41356.52354.0961.0329.9927.0554.3953.2678839770.139924302963947400001.90126.713.78286.6248074789.035772127455431636512.9774.971.91242.16180.6698.1989.34199.95199.8336136874.00349120060275.53113.641.0854.691.07158.42121.1438310650.504631941425325392138.5078.441.40170.26OpenBenchmarking.org

ASKAP tConvolveCuda

Processing: Gridding

OpenBenchmarking.orgMillion Grid Points Per Second, More Is BetterASKAP tConvolveCuda 2015-11-10Processing: GriddingGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 9502K4K6K8K10KSE +/- 130.14, N = 4SE +/- 12.43, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 14.40, N = 38458.773144.856051.275325.128320.503399.141. (CXX) g++ options: -fPIC -O3 -m64 -lcudadevrt -lcudart_static -lrt -lpthread -ldl

ASKAP tConvolveCuda

Processing: Degridding

OpenBenchmarking.orgMillion Grid Points Per Second, More Is BetterASKAP tConvolveCuda 2015-11-10Processing: DegriddingGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 9504K8K12K16K20KSE +/- 369.80, N = 3SE +/- 34.80, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 369.80, N = 3SE +/- 41.05, N = 317380.605290.3211094.009509.1417380.605706.071. (CXX) g++ options: -fPIC -O3 -m64 -lcudadevrt -lcudart_static -lrt -lpthread -ldl

CUDA Mini-Nbody

Test: Original

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: OriginalGeForce GTX 750GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 9504080120160200SE +/- 0.05, N = 3SE +/- 0.50, N = 3SE +/- 0.35, N = 3SE +/- 0.43, N = 3SE +/- 0.10, N = 3SE +/- 0.13, N = 3SE +/- 0.57, N = 3SE +/- 0.21, N = 3180.6661.0332.3782.0145.3854.3234.58105.30

CUDA Mini-Nbody

Test: Cache Blocking

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: Cache BlockingGeForce GTX 750GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 95020406080100SE +/- 0.00, N = 3SE +/- 0.27, N = 3SE +/- 0.10, N = 3SE +/- 0.01, N = 3SE +/- 0.06, N = 3SE +/- 0.01, N = 3SE +/- 0.21, N = 3SE +/- 0.02, N = 398.1929.9918.6537.0825.1328.5319.7749.89

CUDA Mini-Nbody

Test: Loop Unrolling

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: Loop UnrollingGeForce GTX 750GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 95020406080100SE +/- 0.04, N = 3SE +/- 0.05, N = 3SE +/- 0.25, N = 3SE +/- 0.03, N = 3SE +/- 0.21, N = 3SE +/- 0.02, N = 3SE +/- 0.15, N = 3SE +/- 0.03, N = 389.3427.0517.5935.3523.8826.4218.4647.54

CUDA Mini-Nbody

Test: SOA Data Layout

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: SOA Data LayoutGeForce GTX 750GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 9504080120160200SE +/- 0.04, N = 3SE +/- 0.16, N = 3SE +/- 0.20, N = 3SE +/- 0.08, N = 3SE +/- 0.21, N = 3SE +/- 0.05, N = 3SE +/- 0.11, N = 3SE +/- 0.02, N = 3199.9554.3937.4379.9750.1555.8740.94108.50

CUDA Mini-Nbody

Test: Flush Denormals To Zero

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: Flush Denormals To ZeroGeForce GTX 750GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 9504080120160200SE +/- 0.03, N = 3SE +/- 0.10, N = 3SE +/- 0.08, N = 3SE +/- 0.01, N = 3SE +/- 0.18, N = 3SE +/- 0.07, N = 3SE +/- 0.10, N = 3SE +/- 0.02, N = 3199.8353.2637.3779.8449.5355.8040.85108.48

JuliaGPU

OpenCL Device: GPU

OpenBenchmarking.orgSamples/sec, More Is BetterJuliaGPU 1.2pts1OpenCL Device: GPUGeForce GTX 760GeForce GTX 750GeForce GTX 680GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 95030M60M90M120M150MSE +/- 14125.16, N = 3SE +/- 22546.70, N = 3SE +/- 59682.63, N = 3SE +/- 293396.06, N = 3SE +/- 318277.32, N = 3SE +/- 157475.07, N = 3SE +/- 218639.12, N = 3SE +/- 84325.23, N = 3SE +/- 473156.02, N = 3SE +/- 58084.93, N = 338310650.5036136874.0048074789.0378839770.13136037921.4380042041.73113830604.27104144917.23127978049.5364913682.631. (CC) gcc options: -O3 -march=native -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL -lm

LuxMark

OpenCL Device: GPU - Scene: Hotel

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.0OpenCL Device: GPU - Scene: HotelGeForce GTX 760GeForce GTX 680GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 950400800120016002000SE +/- 0.33, N = 3SE +/- 2.00, N = 3SE +/- 0.00, N = 3SE +/- 0.33, N = 3SE +/- 0.67, N = 3SE +/- 1.20, N = 3SE +/- 0.00, N = 3SE +/- 0.33, N = 3SE +/- 0.00, N = 34635779921906897149213461855769

LuxMark

OpenCL Device: GPU - Scene: Microphone

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.0OpenCL Device: GPU - Scene: MicrophoneGeForce GTX 760GeForce GTX 680GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 95014002800420056007000SE +/- 0.67, N = 3SE +/- 3.06, N = 3SE +/- 12.00, N = 3SE +/- 3.00, N = 3SE +/- 1.15, N = 3SE +/- 0.67, N = 3SE +/- 7.64, N = 3SE +/- 18.50, N = 3SE +/- 4.26, N = 3194121274302636024604776445862682423

LuxMark

OpenCL Device: GPU - Scene: Luxball HDR

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.0OpenCL Device: GPU - Scene: Luxball HDRGeForce GTX 760GeForce GTX 750GeForce GTX 680GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 9503K6K9K12K15KSE +/- 1.45, N = 3SE +/- 11.67, N = 3SE +/- 12.17, N = 3SE +/- 35.97, N = 3SE +/- 4.70, N = 3SE +/- 0.88, N = 3SE +/- 1.20, N = 3SE +/- 24.85, N = 3SE +/- 44.35, N = 3SE +/- 16.67, N = 34253349145549639140815474107139737138025313

MandelbulbGPU

OpenCL Device: GPU

OpenBenchmarking.orgSamples/sec, More Is BetterMandelbulbGPU 1.0pts1OpenCL Device: GPUGeForce GTX 760GeForce GTX 750GeForce GTX 680GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 95016M32M48M64M80MSE +/- 28089.31, N = 3SE +/- 9818.73, N = 3SE +/- 36731.70, N = 3SE +/- 48150.35, N = 3SE +/- 166919.37, N = 3SE +/- 75512.83, N = 3SE +/- 140370.89, N = 3SE +/- 91420.68, N = 3SE +/- 168304.91, N = 3SE +/- 29855.85, N = 325392138.5020060275.5331636512.9747400001.9075614774.1344953399.4763616558.7758811317.1771656708.8337156070.871. (CC) gcc options: -O3 -lm -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL

SHOC Scalable HeterOgeneous Computing

Target: CUDA - Benchmark: FFT SP

OpenBenchmarking.orgGFLOPS, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: CUDA - Benchmark: FFT SPGeForce GTX 750GeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 95070140210280350SE +/- 0.69, N = 3SE +/- 1.19, N = 3SE +/- 1.49, N = 3SE +/- 3.09, N = 3SE +/- 2.44, N = 3SE +/- 0.32, N = 3SE +/- 0.47, N = 3113.64324.09212.43289.63263.14311.46172.281. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

SHOC Scalable HeterOgeneous Computing

Target: CUDA - Benchmark: MD5 Hash

OpenBenchmarking.orgGHash/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: CUDA - Benchmark: MD5 HashGeForce GTX 750GeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 950246810SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 31.087.423.385.704.796.812.361. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

SHOC Scalable HeterOgeneous Computing

Target: OpenCL - Benchmark: FFT SP

OpenBenchmarking.orgGFLOPS, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: FFT SPGeForce GTX 760GeForce GTX 750GeForce GTX 680GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 9504080120160200SE +/- 0.31, N = 3SE +/- 0.08, N = 3SE +/- 0.87, N = 3SE +/- 0.19, N = 3SE +/- 0.19, N = 3SE +/- 1.20, N = 3SE +/- 1.30, N = 3SE +/- 0.52, N = 3SE +/- 0.65, N = 3SE +/- 0.08, N = 378.4454.6974.97126.71173.8962.78140.12117.23170.3663.221. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

SHOC Scalable HeterOgeneous Computing

Target: OpenCL - Benchmark: MD5 Hash

OpenBenchmarking.orgGHash/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: MD5 HashGeForce GTX 760GeForce GTX 750GeForce GTX 680GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 950246810SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 31.401.071.913.787.413.365.684.776.792.341. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

SHOC Scalable HeterOgeneous Computing

Target: CUDA - Benchmark: Texture Read Bandwidth

OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: CUDA - Benchmark: Texture Read BandwidthGeForce GTX 750GeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 95080160240320400SE +/- 0.42, N = 3SE +/- 0.12, N = 3SE +/- 0.14, N = 3SE +/- 1.15, N = 3SE +/- 0.28, N = 3SE +/- 1.22, N = 3SE +/- 0.85, N = 3158.42356.52351.31336.48325.16348.92326.231. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

SHOC Scalable HeterOgeneous Computing

Target: OpenCL - Benchmark: Texture Read Bandwidth

OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Texture Read BandwidthGeForce GTX 760GeForce GTX 750GeForce GTX 680GeForce GTX 780 TiGeForce GTX TITAN XGeForce GTX 960GeForce GTX 980GeForce GTX 970GeForce GTX 980 TiGeForce GTX 95080160240320400SE +/- 0.28, N = 3SE +/- 0.23, N = 3SE +/- 1.02, N = 3SE +/- 0.02, N = 3SE +/- 1.56, N = 3SE +/- 0.56, N = 3SE +/- 0.20, N = 3SE +/- 0.06, N = 3SE +/- 0.21, N = 3SE +/- 0.73, N = 3170.26121.14242.16286.62354.09269.98332.60283.36345.55239.191. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft


Phoronix Test Suite v10.8.5