Vulkan Compute

Intel Core i7-2600K testing with a Gigabyte P67A-UD7-B3 (F7 BIOS) and eVGA NVIDIA GeForce RTX 3060 12GB on Gentoo 2.7 via the Phoronix Test Suite.

Compare your own system(s) to this result file with the Phoronix Test Suite by running the command: phoronix-test-suite benchmark 2111059-HA-2111054TJ65
Jump To Table - Results

View

Do Not Show Noisy Results
Do Not Show Results With Incomplete Data
Do Not Show Results With Little Change/Spread
List Notable Results

Limit displaying results to tests within:

NVIDIA GPU Compute 5 Tests
Vulkan Compute 5 Tests

Statistics

Show Overall Harmonic Mean(s)
Show Overall Geometric Mean
Show Geometric Means Per-Suite/Category
Show Wins / Losses Counts (Pie Chart)
Normalize Results
Remove Outliers Before Calculating Averages

Graph Settings

Force Line Graphs Where Applicable
Convert To Scalar Where Applicable
Disable Color Branding
Prefer Vertical Bar Graphs

Additional Graphs

Show Perf Per Core/Thread Calculation Graphs Where Applicable

Multi-Way Comparison

Condense Multi-Option Tests Into Single Result Graphs

Table

Show Detailed System Result Table

Run Management

Highlight
Result
Hide
Result
Result
Identifier
Performance Per
Dollar
Date
Run
  Test
  Duration
RX 480 4GB
October 30 2021
  1 Hour, 27 Minutes
RTX 3060
November 05 2021
  2 Hours, 40 Minutes
RTX 3060 Linux
November 05 2021
  2 Hours, 32 Minutes
Invert Hiding All Results Option
  2 Hours, 13 Minutes

Only show results where is faster than
Only show results matching title/arguments (delimit multiple options with a comma):
Do not show results matching title/arguments (delimit multiple options with a comma):


Vulkan ComputeProcessorMotherboardMemoryDiskGraphicsAudioMonitorChipsetNetworkOSKernelDisplay DriverOpenCLFile-SystemScreen ResolutionDesktopDisplay ServerOpenGLVulkanCompilerRX 480 4GBRTX 3060RTX 3060 LinuxIntel Core i7-2600K (8 Cores)Gigabyte P67A-UD7-B3 (F7 BIOS)0 x 4096 MB224GB INTEL SSDSC2CW240A3 + 1863GB WDC WD20EARX-008FB0AMD Radeon RX 480Realtek HD Audio + AMD HD Audio DeviceS7A950DMicrosoft Windows 7 Professional Build 76016.1 (x86_64)27.20.14501.18003OpenCL 2.1 AMD-APP (3188.4)NTFS1920x1080NVIDIA GeForce RTX 3060 12GBNVIDIA HD Audio + Realtek HD Audio472.12 (30.0.14.7212)OpenCL 3.0 CUDA 11.4.136 + OpenCL 2.1 AMD-APP (3188.4)Intel Core i7-2600K @ 3.80GHz (4 Cores / 8 Threads)Intel 2nd Generation Core DRAM16GB240GB INTEL SSDSC2CW24 + 2000GB Western Digital WD20EARX-008eVGA NVIDIA GeForce RTX 3060 12GB (1882/7500MHz)Realtek ALC889S27A950D2 x Realtek RTL8111/8168/8411Gentoo 2.75.14.15-gentoo (x86_64)Xfce 4.16X Server 1.20.6NVIDIA 470.63.014.6.01.2.175GCC 11.2.0 + Clang 12.0.1 + LLVM 12.0.1ext4OpenBenchmarking.orgEnvironment Details- RX 480 4GB, RTX 3060: windows_tracing_flags=3Security Details- RX 480 4GB: __user pointer sanitization: Disabled- RTX 3060: __user pointer sanitization: Disabled + KPTI Enabled: Yes + PTE Inversion: Yes- RTX 3060 Linux: itlb_multihit: KVM: Mitigation of VMX disabled + l1tf: Mitigation of PTE Inversion; VMX: conditional cache flushes SMT vulnerable + mds: Vulnerable: Clear buffers attempted no microcode; SMT vulnerable + meltdown: Mitigation of PTI + spec_store_bypass: Vulnerable + spectre_v1: Mitigation of usercopy/swapgs barriers and __user pointer sanitization + spectre_v2: Mitigation of Full generic retpoline STIBP: disabled RSB filling + srbds: Not affected + tsx_async_abort: Not affected Compiler Details- RTX 3060 Linux: --bindir=/usr/x86_64-pc-linux-gnu/gcc-bin/11.2.0 --build=x86_64-pc-linux-gnu --datadir=/usr/share/gcc-data/x86_64-pc-linux-gnu/11.2.0 --disable-esp --disable-fixed-point --disable-libada --disable-libssp --disable-libunwind-exceptions --disable-libvtv --disable-systemtap --disable-valgrind-annotations --disable-vtable-verify --disable-werror --enable-__cxa_atexit --enable-checking=release --enable-clocale=gnu --enable-default-pie --enable-default-ssp --enable-languages=c,c++,fortran --enable-libgomp --enable-libstdcxx-time --enable-lto --enable-multilib --enable-nls --enable-obsolete --enable-secureplt --enable-shared --enable-targets=all --enable-threads=posix --host=x86_64-pc-linux-gnu --includedir=/usr/lib/gcc/x86_64-pc-linux-gnu/11.2.0/include --mandir=/usr/share/gcc-data/x86_64-pc-linux-gnu/11.2.0/man --with-multilib-list=m32,m64 --with-python-dir=/share/gcc-data/x86_64-pc-linux-gnu/11.2.0/python --without-isl --without-zstd Processor Details- RTX 3060 Linux: Scaling Governor: intel_cpufreq performance - CPU Microcode: 0x25Graphics Details- RTX 3060 Linux: GLAMOR

RX 480 4GBRTX 3060RTX 3060 LinuxResult OverviewPhoronix Test Suite100%136%171%207%242%RealSR-NCNNWaifu2x-NCNN Vulkanvkpeak

Vulkan Computevkpeak: fp32-scalarvkpeak: fp32-vec4vkpeak: fp64-scalarvkpeak: fp64-vec4vkpeak: int32-scalarvkpeak: int32-vec4realsr-ncnn: 4x - Norealsr-ncnn: 4x - Yeswaifu2x-ncnn: 2x - 3 - Nowaifu2x-ncnn: 2x - 3 - Yesvkfft: vkresample: 2x - Doublevkresample: 2x - Singlevkpeak: fp16-scalarvkpeak: fp16-vec4vkpeak: int16-scalarvkpeak: int16-vec4RX 480 4GBRTX 3060RTX 3060 Linux5621.785522.92372.65372.671192.331192.4825.359185.0792.07911.3526977.639262.87219.08218.166988.836978.9211.74270.3302.2185.6582779247.08127.1867004.1813614.374628.096162.137024.459320.50220.53217.957009.776981.6611.63968.8525.42123701354.86722.9547051.7413697.944636.486165.04OpenBenchmarking.org

vkpeak

Vkpeak is a Vulkan compute benchmark inspired by OpenCL's clpeak. Vkpeak provides Vulkan compute performance measurements for FP16 / FP32 / FP64 / INT16 / INT32 scalar and vec4 performance. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp32-scalarRX 480 4GBRTX 3060RTX 3060 Linux15003000450060007500SE +/- 10.76, N = 3SE +/- 20.38, N = 3SE +/- 0.06, N = 35621.786977.637024.45
OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp32-scalarRX 480 4GBRTX 3060RTX 3060 Linux12002400360048006000Min: 5607.48 / Avg: 5621.78 / Max: 5642.85Min: 6955.24 / Avg: 6977.63 / Max: 7018.31Min: 7024.35 / Avg: 7024.45 / Max: 7024.57

OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp32-vec4RX 480 4GBRTX 3060RTX 3060 Linux2K4K6K8K10KSE +/- 6.80, N = 3SE +/- 27.73, N = 3SE +/- 0.22, N = 35522.929262.879320.50
OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp32-vec4RX 480 4GBRTX 3060RTX 3060 Linux16003200480064008000Min: 5510.26 / Avg: 5522.92 / Max: 5533.57Min: 9234.88 / Avg: 9262.87 / Max: 9318.34Min: 9320.19 / Avg: 9320.5 / Max: 9320.93

OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp64-scalarRX 480 4GBRTX 3060RTX 3060 Linux80160240320400SE +/- 0.01, N = 3SE +/- 0.69, N = 3SE +/- 0.01, N = 3372.65219.08220.53
OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp64-scalarRX 480 4GBRTX 3060RTX 3060 Linux70140210280350Min: 372.64 / Avg: 372.65 / Max: 372.67Min: 218.38 / Avg: 219.08 / Max: 220.46Min: 220.51 / Avg: 220.53 / Max: 220.55

OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp64-vec4RX 480 4GBRTX 3060RTX 3060 Linux80160240320400SE +/- 0.00, N = 2SE +/- 0.18, N = 3SE +/- 0.01, N = 3372.67218.16217.95
OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp64-vec4RX 480 4GBRTX 3060RTX 3060 Linux70140210280350Min: 372.67 / Avg: 372.67 / Max: 372.67Min: 217.8 / Avg: 218.16 / Max: 218.34Min: 217.94 / Avg: 217.95 / Max: 217.96

OpenBenchmarking.orgGIOPS, More Is Bettervkpeak 20210424int32-scalarRX 480 4GBRTX 3060RTX 3060 Linux15003000450060007500SE +/- 0.01, N = 3SE +/- 19.64, N = 3SE +/- 2.65, N = 31192.336988.837009.77
OpenBenchmarking.orgGIOPS, More Is Bettervkpeak 20210424int32-scalarRX 480 4GBRTX 3060RTX 3060 Linux12002400360048006000Min: 1192.3 / Avg: 1192.33 / Max: 1192.35Min: 6949.55 / Avg: 6988.83 / Max: 7008.53Min: 7004.54 / Avg: 7009.77 / Max: 7013.08

OpenBenchmarking.orgGIOPS, More Is Bettervkpeak 20210424int32-vec4RX 480 4GBRTX 3060RTX 3060 Linux15003000450060007500SE +/- 0.02, N = 3SE +/- 0.07, N = 3SE +/- 0.09, N = 31192.486978.926981.66
OpenBenchmarking.orgGIOPS, More Is Bettervkpeak 20210424int32-vec4RX 480 4GBRTX 3060RTX 3060 Linux12002400360048006000Min: 1192.45 / Avg: 1192.48 / Max: 1192.52Min: 6978.83 / Avg: 6978.92 / Max: 6979.05Min: 6981.5 / Avg: 6981.66 / Max: 6981.81

RealSR-NCNN

RealSR-NCNN is an NCNN neural network implementation of the RealSR project and accelerated using the Vulkan API. RealSR is the Real-World Super Resolution via Kernel Estimation and Noise Injection. NCNN is a high performance neural network inference framework optimized for mobile and other platforms developed by Tencent. This test profile times how long it takes to increase the resolution of a sample image by a scale of 4x with Vulkan. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgSeconds, Fewer Is BetterRealSR-NCNN 20200818Scale: 4x - TAA: NoRX 480 4GBRTX 3060RTX 3060 Linux612182430SE +/- 0.23, N = 7SE +/- 0.09, N = 12SE +/- 0.01, N = 325.3611.7411.64
OpenBenchmarking.orgSeconds, Fewer Is BetterRealSR-NCNN 20200818Scale: 4x - TAA: NoRX 480 4GBRTX 3060RTX 3060 Linux612182430Min: 25.07 / Avg: 25.36 / Max: 26.72Min: 11.56 / Avg: 11.74 / Max: 12.68Min: 11.62 / Avg: 11.64 / Max: 11.66

OpenBenchmarking.orgSeconds, Fewer Is BetterRealSR-NCNN 20200818Scale: 4x - TAA: YesRX 480 4GBRTX 3060RTX 3060 Linux4080120160200SE +/- 0.21, N = 3SE +/- 0.06, N = 3SE +/- 0.15, N = 3185.0870.3368.85
OpenBenchmarking.orgSeconds, Fewer Is BetterRealSR-NCNN 20200818Scale: 4x - TAA: YesRX 480 4GBRTX 3060RTX 3060 Linux306090120150Min: 184.83 / Avg: 185.08 / Max: 185.5Min: 70.22 / Avg: 70.33 / Max: 70.43Min: 68.57 / Avg: 68.85 / Max: 69.05

Waifu2x-NCNN Vulkan

Waifu2x-NCNN is an NCNN neural network implementation of the Waifu2x converter project and accelerated using the Vulkan API. NCNN is a high performance neural network inference framework optimized for mobile and other platforms developed by Tencent. This test profile times how long it takes to increase the resolution of a sample image with Vulkan. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgSeconds, Fewer Is BetterWaifu2x-NCNN Vulkan 20200818Scale: 2x - Denoise: 3 - TAA: NoRX 480 4GBRTX 30600.49910.99821.49731.99642.4955SE +/- 0.054, N = 15SE +/- 0.051, N = 152.0792.218
OpenBenchmarking.orgSeconds, Fewer Is BetterWaifu2x-NCNN Vulkan 20200818Scale: 2x - Denoise: 3 - TAA: NoRX 480 4GBRTX 3060246810Min: 2.01 / Avg: 2.08 / Max: 2.84Min: 2.08 / Avg: 2.22 / Max: 2.92

OpenBenchmarking.orgSeconds, Fewer Is BetterWaifu2x-NCNN Vulkan 20200818Scale: 2x - Denoise: 3 - TAA: YesRX 480 4GBRTX 3060RTX 3060 Linux3691215SE +/- 0.051, N = 3SE +/- 0.060, N = 3SE +/- 0.003, N = 311.3525.6585.421
OpenBenchmarking.orgSeconds, Fewer Is BetterWaifu2x-NCNN Vulkan 20200818Scale: 2x - Denoise: 3 - TAA: YesRX 480 4GBRTX 3060RTX 3060 Linux3691215Min: 11.28 / Avg: 11.35 / Max: 11.45Min: 5.54 / Avg: 5.66 / Max: 5.73Min: 5.42 / Avg: 5.42 / Max: 5.43

VkFFT

VkFFT is a Fast Fourier Transform (FFT) Library that is GPU accelerated by means of the Vulkan API. The VkFFT benchmark runs FFT performance differences of many different sizes before returning an overall benchmark score. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgBenchmark Score, More Is BetterVkFFT 1.1.1RTX 3060RTX 3060 Linux6K12K18K24K30KSE +/- 297.81, N = 4SE +/- 187.69, N = 327792237011. (CXX) g++ options: -O3 -pthread
OpenBenchmarking.orgBenchmark Score, More Is BetterVkFFT 1.1.1RTX 3060RTX 3060 Linux5K10K15K20K25KMin: 26907 / Avg: 27792.25 / Max: 28201Min: 23460 / Avg: 23701.33 / Max: 240711. (CXX) g++ options: -O3 -pthread

RX 480 4GB: The test quit with a non-zero exit status.

VkResample

VkResample is a Vulkan-based image upscaling library based on VkFFT. The sample input file is upscaling a 4K image to 8K using Vulkan-based GPU acceleration. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgms, Fewer Is BetterVkResample 1.0Upscale: 2x - Precision: DoubleRTX 3060RTX 3060 Linux80160240320400SE +/- 0.38, N = 15SE +/- 0.13, N = 347.08354.871. (CXX) g++ options: -O3 -pthread
OpenBenchmarking.orgms, Fewer Is BetterVkResample 1.0Upscale: 2x - Precision: DoubleRTX 3060RTX 3060 Linux60120180240300Min: 44.15 / Avg: 47.08 / Max: 49.3Min: 354.61 / Avg: 354.87 / Max: 355.061. (CXX) g++ options: -O3 -pthread

Upscale: 2x - Precision: Double

RX 480 4GB: The test quit with a non-zero exit status.

OpenBenchmarking.orgms, Fewer Is BetterVkResample 1.0Upscale: 2x - Precision: SingleRTX 3060RTX 3060 Linux612182430SE +/- 2.24, N = 15SE +/- 0.00, N = 327.1922.951. (CXX) g++ options: -O3 -pthread
OpenBenchmarking.orgms, Fewer Is BetterVkResample 1.0Upscale: 2x - Precision: SingleRTX 3060RTX 3060 Linux612182430Min: 22.93 / Avg: 27.19 / Max: 44.46Min: 22.95 / Avg: 22.95 / Max: 22.961. (CXX) g++ options: -O3 -pthread

Upscale: 2x - Precision: Single

RX 480 4GB: The test quit with a non-zero exit status.

vkpeak

Vkpeak is a Vulkan compute benchmark inspired by OpenCL's clpeak. Vkpeak provides Vulkan compute performance measurements for FP16 / FP32 / FP64 / INT16 / INT32 scalar and vec4 performance. Learn more via the OpenBenchmarking.org test page.

OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp16-scalarRTX 3060RTX 3060 Linux15003000450060007500SE +/- 21.16, N = 3SE +/- 0.11, N = 37004.187051.74
OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp16-scalarRTX 3060RTX 3060 Linux12002400360048006000Min: 6982.68 / Avg: 7004.18 / Max: 7046.49Min: 7051.62 / Avg: 7051.74 / Max: 7051.97

OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp16-vec4RTX 3060RTX 3060 Linux3K6K9K12K15KSE +/- 39.61, N = 3SE +/- 0.33, N = 313614.3713697.94
OpenBenchmarking.orgGFLOPS, More Is Bettervkpeak 20210424fp16-vec4RTX 3060RTX 3060 Linux2K4K6K8K10KMin: 13573.12 / Avg: 13614.37 / Max: 13693.57Min: 13697.41 / Avg: 13697.94 / Max: 13698.54

OpenBenchmarking.orgGIOPS, More Is Bettervkpeak 20210424int16-scalarRTX 3060RTX 3060 Linux10002000300040005000SE +/- 3.82, N = 3SE +/- 0.06, N = 34628.094636.48
OpenBenchmarking.orgGIOPS, More Is Bettervkpeak 20210424int16-scalarRTX 3060RTX 3060 Linux8001600240032004000Min: 4624.2 / Avg: 4628.09 / Max: 4635.73Min: 4636.38 / Avg: 4636.48 / Max: 4636.6

OpenBenchmarking.orgGIOPS, More Is Bettervkpeak 20210424int16-vec4RTX 3060RTX 3060 Linux13002600390052006500SE +/- 0.07, N = 3SE +/- 0.08, N = 36162.136165.04
OpenBenchmarking.orgGIOPS, More Is Bettervkpeak 20210424int16-vec4RTX 3060RTX 3060 Linux11002200330044005500Min: 6162.06 / Avg: 6162.13 / Max: 6162.26Min: 6164.92 / Avg: 6165.04 / Max: 6165.19

17 Results Shown

vkpeak:
  fp32-scalar
  fp32-vec4
  fp64-scalar
  fp64-vec4
  int32-scalar
  int32-vec4
RealSR-NCNN:
  4x - No
  4x - Yes
Waifu2x-NCNN Vulkan:
  2x - 3 - No
  2x - 3 - Yes
VkFFT
VkResample:
  2x - Double
  2x - Single
vkpeak:
  fp16-scalar
  fp16-vec4
  int16-scalar
  int16-vec4