OpenCL CUDA NVIDIA GPGPU Linux Tests

ttnx

HTML result view exported from: https://openbenchmarking.org/result/1605256-GA-1511113PT74&grs&rdt.

OpenCL CUDA NVIDIA GPGPU Linux TestsProcessorMotherboardChipsetMemoryDiskGraphicsAudioNetworkOSKernelDesktopDisplay ServerDisplay DriverOpenGLCompilerFile-SystemScreen ResolutionGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 760tnxIntel Core i5-6600K @ 3.50GHz (4 Cores)MSI Z170A GAMING PRO (MS-7984) v1.0Intel Device 191f16384MB256GB TS256GSSD370SeVGA NVIDIA GeForce GTX 950 2048MB (135/405MHz)Intel Device a170Intel Device 15b8Ubuntu 14.043.19.0-33-generic (x86_64)Unity 7.2.5X Server 1.17.1NVIDIA 352.394.3.0GCC 4.8.4 + Clang 3.4-1ubuntu3 + CUDA 7.5ext43840x2160NVIDIA GeForce GTX 980 Ti 6144MB (999/3505MHz)eVGA NVIDIA GeForce GTX 970 4096MB (1163/3505MHz)NVIDIA GeForce GTX 980 4096MB (1126/3505MHz)eVGA NVIDIA GeForce GTX 960 2048MB (1277/3505MHz)NVIDIA GeForce GTX TITAN X 12288MB (1001/3505MHz)NVIDIA GeForce GTX 780 Ti 3072MB (875/3500MHz)NVIDIA GeForce GTX 680 2048MB (1006/3004MHz)eVGA NVIDIA GeForce GTX 750 1024MB (1019/2505MHz)NVIDIA GeForce GTX 760 2048MB (980/3004MHz)Intel Xeon E5-1650 v3 @ 3.80GHz (12 Cores)Supermicro X10SRL-FIntel Xeon E7 v3/Xeon8 x 16384 MB DDR4-2133MHz Samsung2 x 512GB Samsung SSD 850eVGA NVIDIA GeForce GTX TITAN XNVIDIA Device 0fb0Intel 82599ES 10-Gigabit SFI/SFP+Ubuntu 16.044.4.0-21-generic (x86_64)GCC 5.3.1 201604131024x768OpenBenchmarking.orgCompiler Details- GeForce GTX 950: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX 980 Ti: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX 970: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX 980: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX 960: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX TITAN X: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX 780 Ti: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX 680: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX 750: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - GeForce GTX 760: --build=x86_64-linux-gnu --disable-browser-plugin --disable-libmudflap --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,c++,java,go,d,fortran,objc,obj-c++ --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-multilib-list=m32,m64,mx32 --with-tune=generic -v - tnx: --build=x86_64-linux-gnu --disable-browser-plugin --disable-vtable-verify --disable-werror --enable-checking=release --enable-clocale=gnu --enable-gnu-unique-object --enable-gtk-cairo --enable-java-awt=gtk --enable-java-home --enable-languages=c,ada,c++,java,go,d,fortran,objc,obj-c++ --enable-libmpx --enable-libstdcxx-debug --enable-libstdcxx-time=yes --enable-multiarch --enable-multilib --enable-nls --enable-objc-gc --enable-plugin --enable-shared --enable-threads=posix --host=x86_64-linux-gnu --target=x86_64-linux-gnu --with-abi=m64 --with-arch-32=i686 --with-arch-directory=amd64 --with-default-libstdcxx-abi=new --with-multilib-list=m32,m64,mx32 --with-tune=generic -v Processor Details- GeForce GTX 950: Scaling Governor: acpi-cpufreq performance- GeForce GTX 980 Ti: Scaling Governor: acpi-cpufreq performance- GeForce GTX 970: Scaling Governor: acpi-cpufreq performance- GeForce GTX 980: Scaling Governor: acpi-cpufreq performance- GeForce GTX 960: Scaling Governor: acpi-cpufreq performance- GeForce GTX TITAN X: Scaling Governor: acpi-cpufreq performance- GeForce GTX 780 Ti: Scaling Governor: acpi-cpufreq performance- GeForce GTX 680: Scaling Governor: acpi-cpufreq performance- GeForce GTX 750: Scaling Governor: acpi-cpufreq performance- GeForce GTX 760: Scaling Governor: acpi-cpufreq performance- tnx: Scaling Governor: intel_pstate powersaveOpenCL Details- GeForce GTX 950: GPU Compute Cores: 768- GeForce GTX 980 Ti: GPU Compute Cores: 2816- GeForce GTX 970: GPU Compute Cores: 1664- GeForce GTX 980: GPU Compute Cores: 2048- GeForce GTX 960: GPU Compute Cores: 1024- GeForce GTX TITAN X: GPU Compute Cores: 3072- GeForce GTX 780 Ti: GPU Compute Cores: 2880- GeForce GTX 680: GPU Compute Cores: 1536- GeForce GTX 750: GPU Compute Cores: 512- GeForce GTX 760: GPU Compute Cores: 1152System Details- GeForce GTX 950: GPU Compute Cores: 768.- GeForce GTX 980 Ti: GPU Compute Cores: 2816.- GeForce GTX 970: GPU Compute Cores: 1664.- GeForce GTX 980: GPU Compute Cores: 2048.- GeForce GTX 960: GPU Compute Cores: 1024.- GeForce GTX TITAN X: GPU Compute Cores: 3072.- GeForce GTX 780 Ti: GPU Compute Cores: 2880.- GeForce GTX 680: GPU Compute Cores: 1536.- GeForce GTX 750: GPU Compute Cores: 512.- GeForce GTX 760: GPU Compute Cores: 1152.

OpenCL CUDA NVIDIA GPGPU Linux Testsshoc: OpenCL - MD5 Hashshoc: CUDA - MD5 Hashcuda-mini-nbody: Originalcuda-mini-nbody: Flush Denormals To Zerocuda-mini-nbody: SOA Data Layoutcuda-mini-nbody: Cache Blockingcuda-mini-nbody: Loop Unrollingluxmark: GPU - Hotelluxmark: GPU - Luxball HDRmandelbulbgpu: GPUjuliagpu: GPUaskap: Degriddingluxmark: GPU - Microphoneshoc: OpenCL - FFT SPshoc: OpenCL - Texture Read Bandwidthshoc: CUDA - FFT SPaskap: Griddingshoc: CUDA - Texture Read BandwidthGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 760tnx2.342.36105.30108.48108.5049.8947.54769531337156070.8764913682.635706.07242363.22239.19172.283399.14326.236.796.8134.5840.8540.9419.7718.4618551380271656708.83127978049.5317380.606268170.36345.55311.468320.50348.924.774.7954.3255.8055.8728.5326.421346973758811317.17104144917.239509.144458117.23283.36263.145325.12325.165.685.7045.3849.5350.1525.1323.8814921071363616558.77113830604.27110944776140.12332.60289.636051.27336.483.363.3882.0179.8479.9737.0835.35897547444953399.4780042041.735290.32246062.78269.98212.433144.85351.317.417.4232.3737.3737.4318.6517.5919061408175614774.13136037921.4317380.606360173.89354.09324.098458.77356.523.7861.0353.2654.3929.9927.05992963947400001.9078839770.134302126.71286.621.91577455431636512.9748074789.03212774.97242.161.071.08180.66199.83199.9598.1989.34349120060275.5336136874.0054.69121.14113.64158.421.40463425325392138.5038310650.50194178.44170.26OpenBenchmarking.org

SHOC Scalable HeterOgeneous Computing

Target: OpenCL - Benchmark: MD5 Hash

OpenBenchmarking.orgGHash/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: MD5 HashGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 760246810SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 32.346.794.775.683.367.413.781.911.071.401. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

SHOC Scalable HeterOgeneous Computing

Target: CUDA - Benchmark: MD5 Hash

OpenBenchmarking.orgGHash/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: CUDA - Benchmark: MD5 HashGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 750246810SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 32.366.814.795.703.387.421.081. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

CUDA Mini-Nbody

Test: Original

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: OriginalGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 7504080120160200SE +/- 0.21, N = 3SE +/- 0.57, N = 3SE +/- 0.13, N = 3SE +/- 0.10, N = 3SE +/- 0.43, N = 3SE +/- 0.35, N = 3SE +/- 0.50, N = 3SE +/- 0.05, N = 3105.3034.5854.3245.3882.0132.3761.03180.66

CUDA Mini-Nbody

Test: Flush Denormals To Zero

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: Flush Denormals To ZeroGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 7504080120160200SE +/- 0.02, N = 3SE +/- 0.10, N = 3SE +/- 0.07, N = 3SE +/- 0.18, N = 3SE +/- 0.01, N = 3SE +/- 0.08, N = 3SE +/- 0.10, N = 3SE +/- 0.03, N = 3108.4840.8555.8049.5379.8437.3753.26199.83

CUDA Mini-Nbody

Test: SOA Data Layout

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: SOA Data LayoutGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 7504080120160200SE +/- 0.02, N = 3SE +/- 0.11, N = 3SE +/- 0.05, N = 3SE +/- 0.21, N = 3SE +/- 0.08, N = 3SE +/- 0.20, N = 3SE +/- 0.16, N = 3SE +/- 0.04, N = 3108.5040.9455.8750.1579.9737.4354.39199.95

CUDA Mini-Nbody

Test: Cache Blocking

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: Cache BlockingGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 75020406080100SE +/- 0.02, N = 3SE +/- 0.21, N = 3SE +/- 0.01, N = 3SE +/- 0.06, N = 3SE +/- 0.01, N = 3SE +/- 0.10, N = 3SE +/- 0.27, N = 3SE +/- 0.00, N = 349.8919.7728.5325.1337.0818.6529.9998.19

CUDA Mini-Nbody

Test: Loop Unrolling

OpenBenchmarking.orgSeconds, Fewer Is BetterCUDA Mini-Nbody 2015-11-10Test: Loop UnrollingGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 75020406080100SE +/- 0.03, N = 3SE +/- 0.15, N = 3SE +/- 0.02, N = 3SE +/- 0.21, N = 3SE +/- 0.03, N = 3SE +/- 0.25, N = 3SE +/- 0.05, N = 3SE +/- 0.04, N = 347.5418.4626.4223.8835.3517.5927.0589.34

LuxMark

OpenCL Device: GPU - Scene: Hotel

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.0OpenCL Device: GPU - Scene: HotelGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 760400800120016002000SE +/- 0.00, N = 3SE +/- 0.33, N = 3SE +/- 0.00, N = 3SE +/- 1.20, N = 3SE +/- 0.67, N = 3SE +/- 0.33, N = 3SE +/- 0.00, N = 3SE +/- 2.00, N = 3SE +/- 0.33, N = 37691855134614928971906992577463

LuxMark

OpenCL Device: GPU - Scene: Luxball HDR

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.0OpenCL Device: GPU - Scene: Luxball HDRGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 7603K6K9K12K15KSE +/- 16.67, N = 3SE +/- 44.35, N = 3SE +/- 24.85, N = 3SE +/- 1.20, N = 3SE +/- 0.88, N = 3SE +/- 4.70, N = 3SE +/- 35.97, N = 3SE +/- 12.17, N = 3SE +/- 11.67, N = 3SE +/- 1.45, N = 35313138029737107135474140819639455434914253

MandelbulbGPU

OpenCL Device: GPU

OpenBenchmarking.orgSamples/sec, More Is BetterMandelbulbGPU 1.0pts1OpenCL Device: GPUGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 76016M32M48M64M80MSE +/- 29855.85, N = 3SE +/- 168304.91, N = 3SE +/- 91420.68, N = 3SE +/- 140370.89, N = 3SE +/- 75512.83, N = 3SE +/- 166919.37, N = 3SE +/- 48150.35, N = 3SE +/- 36731.70, N = 3SE +/- 9818.73, N = 3SE +/- 28089.31, N = 337156070.8771656708.8358811317.1763616558.7744953399.4775614774.1347400001.9031636512.9720060275.5325392138.501. (CC) gcc options: -O3 -lm -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL

JuliaGPU

OpenCL Device: GPU

OpenBenchmarking.orgSamples/sec, More Is BetterJuliaGPU 1.2pts1OpenCL Device: GPUGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 76030M60M90M120M150MSE +/- 58084.93, N = 3SE +/- 473156.02, N = 3SE +/- 84325.23, N = 3SE +/- 218639.12, N = 3SE +/- 157475.07, N = 3SE +/- 318277.32, N = 3SE +/- 293396.06, N = 3SE +/- 59682.63, N = 3SE +/- 22546.70, N = 3SE +/- 14125.16, N = 364913682.63127978049.53104144917.23113830604.2780042041.73136037921.4378839770.1348074789.0336136874.0038310650.501. (CC) gcc options: -O3 -march=native -ftree-vectorize -funroll-loops -lglut -lOpenCL -lGL -lm

ASKAP tConvolveCuda

Processing: Degridding

OpenBenchmarking.orgMillion Grid Points Per Second, More Is BetterASKAP tConvolveCuda 2015-11-10Processing: DegriddingGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN X4K8K12K16K20KSE +/- 41.05, N = 3SE +/- 369.80, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 34.80, N = 3SE +/- 369.80, N = 35706.0717380.609509.1411094.005290.3217380.601. (CXX) g++ options: -fPIC -O3 -m64 -lcudadevrt -lcudart_static -lrt -lpthread -ldl

LuxMark

OpenCL Device: GPU - Scene: Microphone

OpenBenchmarking.orgScore, More Is BetterLuxMark 3.0OpenCL Device: GPU - Scene: MicrophoneGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 76014002800420056007000SE +/- 4.26, N = 3SE +/- 18.50, N = 3SE +/- 7.64, N = 3SE +/- 0.67, N = 3SE +/- 1.15, N = 3SE +/- 3.00, N = 3SE +/- 12.00, N = 3SE +/- 3.06, N = 3SE +/- 0.67, N = 3242362684458477624606360430221271941

SHOC Scalable HeterOgeneous Computing

Target: OpenCL - Benchmark: FFT SP

OpenBenchmarking.orgGFLOPS, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: FFT SPGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 7604080120160200SE +/- 0.08, N = 3SE +/- 0.65, N = 3SE +/- 0.52, N = 3SE +/- 1.30, N = 3SE +/- 1.20, N = 3SE +/- 0.19, N = 3SE +/- 0.19, N = 3SE +/- 0.87, N = 3SE +/- 0.08, N = 3SE +/- 0.31, N = 363.22170.36117.23140.1262.78173.89126.7174.9754.6978.441. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

SHOC Scalable HeterOgeneous Computing

Target: OpenCL - Benchmark: Texture Read Bandwidth

OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: OpenCL - Benchmark: Texture Read BandwidthGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 780 TiGeForce GTX 680GeForce GTX 750GeForce GTX 76080160240320400SE +/- 0.73, N = 3SE +/- 0.21, N = 3SE +/- 0.06, N = 3SE +/- 0.20, N = 3SE +/- 0.56, N = 3SE +/- 1.56, N = 3SE +/- 0.02, N = 3SE +/- 1.02, N = 3SE +/- 0.23, N = 3SE +/- 0.28, N = 3239.19345.55283.36332.60269.98354.09286.62242.16121.14170.261. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

SHOC Scalable HeterOgeneous Computing

Target: CUDA - Benchmark: FFT SP

OpenBenchmarking.orgGFLOPS, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: CUDA - Benchmark: FFT SPGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 75070140210280350SE +/- 0.47, N = 3SE +/- 0.32, N = 3SE +/- 2.44, N = 3SE +/- 3.09, N = 3SE +/- 1.49, N = 3SE +/- 1.19, N = 3SE +/- 0.69, N = 3172.28311.46263.14289.63212.43324.09113.641. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft

ASKAP tConvolveCuda

Processing: Gridding

OpenBenchmarking.orgMillion Grid Points Per Second, More Is BetterASKAP tConvolveCuda 2015-11-10Processing: GriddingGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN X2K4K6K8K10KSE +/- 14.40, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 0.00, N = 3SE +/- 12.43, N = 3SE +/- 130.14, N = 43399.148320.505325.126051.273144.858458.771. (CXX) g++ options: -fPIC -O3 -m64 -lcudadevrt -lcudart_static -lrt -lpthread -ldl

SHOC Scalable HeterOgeneous Computing

Target: CUDA - Benchmark: Texture Read Bandwidth

OpenBenchmarking.orgGB/s, More Is BetterSHOC Scalable HeterOgeneous Computing 2015-11-10Target: CUDA - Benchmark: Texture Read BandwidthGeForce GTX 950GeForce GTX 980 TiGeForce GTX 970GeForce GTX 980GeForce GTX 960GeForce GTX TITAN XGeForce GTX 75080160240320400SE +/- 0.85, N = 3SE +/- 1.22, N = 3SE +/- 0.28, N = 3SE +/- 1.15, N = 3SE +/- 0.14, N = 3SE +/- 0.12, N = 3SE +/- 0.42, N = 3326.23348.92325.16336.48351.31356.52158.421. (CXX) g++ options: -O2 -lSHOCCommon -lcudadevrt -lcudart_static -lrt -lpthread -ldl -lcufft


Phoronix Test Suite v10.8.4