Caffe

This is a benchmark of the Caffe deep learning framework and currently supports the AlexNet and Googlenet model and execution on both CPUs and NVIDIA GPUs.

To run this test with the Phoronix Test Suite, the basic command is: phoronix-test-suite benchmark caffe.

Test Created

14 November 2015

Last Updated

26 September 2020

Test Maintainer

Michael Larabel 

Test Type

System

Average Install Time

31 Seconds

Average Run Time

2 Minutes, 1 Second

Test Dependencies

C/C++ Compiler Toolchain + CMake + Python + BLAS (Basic Linear Algebra Sub-Routine) + C++ Boost + Linear Algebra Pack + Snappy Compression + GFlags + OpenCV + HDF5

Accolades

100k+ Downloads

Supported Platforms


Public Result UploadsReported Installs*Test Completions*OpenBenchmarking.orgEventsCaffe AlexNet Popularity Statisticspts/caffe2015.112016.012016.032016.072016.092016.112017.012017.032017.052017.072017.092017.112018.012018.032018.052018.072018.092018.112019.012019.032019.052019.092020.012020.072020.102020.122021.022021.042K4K6K8K10K
* Data based on those opting to upload their test results to OpenBenchmarking.org and users enabling the opt-in anonymous statistics reporting while running benchmarks from an Internet-connected platform.
Data current as of Sat, 08 May 2021 20:56:40 GMT.
GoogleNet48.1%AlexNet51.9%Model Option PopularityOpenBenchmarking.org
NVIDIA CUDA 16.5%CPU83.5%Acceleration Option PopularityOpenBenchmarking.org
20046.1%100011.2%10042.6%Iterations Option PopularityOpenBenchmarking.org

Revision History

pts/caffe-1.5.0   [View Source]   Sat, 26 Sep 2020 21:35:45 GMT
Overhaul Caffe test profile with latest Git snapshot, switch to CMake build system, clean up test options, etc.

pts/caffe-1.4.0   [View Source]   Sat, 29 Dec 2018 11:15:41 GMT
Update Caffe to latest Git snapshot to hopefully workaround build problems on newer distros.

pts/caffe-1.3.3   [View Source]   Sun, 01 Apr 2018 18:50:19 GMT
Basic fix for OpenCV 3.4.

pts/caffe-1.3.2   [View Source]   Wed, 04 Jan 2017 11:07:36 GMT
Fix for OpenCV 3.2.

pts/caffe-1.3.1   [View Source]   Wed, 28 Dec 2016 20:36:42 GMT
Don't show title string of "Caffe AlexNet" but "Caffe" with recent test profile versions supporting more than just AlexNet.

pts/caffe-1.3.0   [View Source]   Wed, 28 Dec 2016 20:34:27 GMT
Update to latest Git snapshot to fix OpenCV compatibility.

pts/caffe-1.2.0   [View Source]   Mon, 15 Aug 2016 16:11:16 GMT
Add Googlenet support, decrease CPU only iteration count.

pts/caffe-1.1.1   [View Source]   Sun, 12 Jun 2016 18:32:44 GMT
Add OpenCV and OpenBLAS support.

pts/caffe-1.1.0   [View Source]   Sat, 11 Jun 2016 19:32:55 GMT
Update

pts/caffe-1.0.0   [View Source]   Sat, 14 Nov 2015 15:29:45 GMT
Initial commit of Caffe deep learning framework and with this benchmark using the AlexNet model for benchmarking.

Suites Using This Test

Machine Learning

HPC - High Performance Computing

NVIDIA GPU Compute


Performance Metrics

Analyze Test Configuration:

Caffe 2020-02-13

Model: GoogleNet - Acceleration: CPU - Iterations: 1000

OpenBenchmarking.org metrics for this test profile configuration based on 71 public results since 26 September 2020 with the latest data as of 19 April 2021.

Below is an overview of the generalized performance for components where there is sufficient statistically significant data based upon user-uploaded results. It is important to keep in mind particularly in the Linux/open-source space there can be vastly different OS configurations, with this overview intended to offer just general guidance as to the performance expectations.

Component
Percentile Rank
# Matching Public Results
Milli-Seconds (Average)
99th
5
917555 +/- 627
88th
5
934042 +/- 19623
84th
8
993424 +/- 10271
Mid-Tier
75th
> 1001758
71st
6
1068995 +/- 1687
61st
3
1134771 +/- 2664
54th
3
1220154 +/- 168
Median
50th
1280710
50th
3
1280884 +/- 4454
47th
3
1285610 +/- 1088
30th
3
1718599 +/- 2880
Low-Tier
25th
> 1777493
19th
3
1901288 +/- 1101
OpenBenchmarking.orgDistribution Of Public Results - Model: GoogleNet - Acceleration: CPU - Iterations: 100071 Results Range From 916456 To 3406981 Milli-Seconds91645696626710160781065889111570011655111215322126513313149441364755141456614643771514188156399916138101663621171343217632431813054186286519126761962487201229820621092111920216173122115422261353231116423609752410786246059725104082560219261003026598412709652275946328092742859085290889629587073008518305832931081403157951320776232575733307384335719534070063691215

Based on OpenBenchmarking.org data, the selected test / test configuration (Caffe 2020-02-13 - Model: GoogleNet - Acceleration: CPU - Iterations: 1000) has an average run-time of 1 hour, 15 minutes. By default this test profile is set to run at least 3 times but may increase if the standard deviation exceeds pre-defined defaults or other calculations deem additional runs necessary for greater statistical accuracy of the result.

OpenBenchmarking.orgMinutesTime Required To Complete BenchmarkModel: GoogleNet - Acceleration: CPU - Iterations: 1000Run-Time4080120160200Min: 46 / Avg: 74.87 / Max: 234

Based on public OpenBenchmarking.org results, the selected test / test configuration has an average standard deviation of 0.1%.

OpenBenchmarking.orgPercent, Fewer Is BetterAverage Deviation Between RunsModel: GoogleNet - Acceleration: CPU - Iterations: 1000Deviation246810Min: 0 / Avg: 0.07 / Max: 2

Notable Instruction Set Usage

Notable instruction set extensions supported by this test, based on an automatic analysis by the Phoronix Test Suite / OpenBenchmarking.org analytics engine.

Instruction Set
Support
Instructions Detected
SSE2 (SSE2)
Used by default on supported hardware.
 
PUNPCKLQDQ MOVDQA MOVDQU CVTSS2SD MOVD ADDSD DIVSD CVTTSD2SI MOVUPD CVTPS2PD CVTPD2PS CVTSD2SS PSHUFD XORPD SHUFPD SUBSD MULSD CVTSI2SD MOVAPD UCOMISD UNPCKLPD CVTDQ2PS COMISD CVTDQ2PD SQRTSD ANDPD ANDNPD CMPNLESD ORPD DIVPD MULPD MINSD MINPD MAXPD MAXSD CMPLTPD ADDPD CMPLTSD MOVHPD SUBPD MOVLPD UNPCKHPD PMULUDQ PSRLDQ
Requires passing a supported compiler/build flag (verified with targets: sandybridge, skylake, tigerlake, cascadelake, sapphirerapids, alderlake, znver2, znver3).
Found on Intel processors since Sandy Bridge (2011).
Found on AMD processors since Bulldozer (2011).

 
VZEROUPPER VINSERTF128 VEXTRACTF128 VPERM2F128 VPERMILPS VPERMILPD VBROADCASTSS VBROADCASTSD VMASKMOVPS
Requires passing a supported compiler/build flag (verified with targets: skylake, tigerlake, cascadelake, sapphirerapids, alderlake, znver2, znver3).
Found on Intel processors since Haswell (2013).
Found on AMD processors since Excavator (2016).

 
VPERM2I128 VPERMD VPERMPD VPBROADCASTQ VPBROADCASTD VPERMQ VGATHERQPS VEXTRACTI128 VPMASKMOVD VINSERTI128 VPGATHERDD VPBROADCASTW
FMA (FMA)
Requires passing a supported compiler/build flag (verified with targets: skylake, tigerlake, cascadelake, sapphirerapids, alderlake, znver2, znver3).
Found on Intel processors since Haswell (2013).
Found on AMD processors since Bulldozer (2011).

 
VFMADD132SS VFMADD132SD VFMSUB213PS VFMSUB132SS VFMSUB213PD VFMSUB132SD VFNMADD213SD VFNMADD213SS VFMADD231SS VFNMADD231SS VFMADD213SS VFNMADD132SS VFMADD231SD VFNMADD132SD VFMADD213SD VFMADD132PS VFMADD132PD VFNMADD132PD VFNMADD213PD VFNMADD132PS VFNMADD213PS VFMSUB231SD VFNMADD231SD VFMADD231PD
The test / benchmark does honor compiler flag changes.
Last automated analysis: 30 January 2021

This test profile binary relies on the shared libraries libcaffe.so.1.0.0, libglog.so.0, libgflags.so.2.2, libprotobuf.so.23, libpthread.so.0, libsz.so.2, libz.so.1, libdl.so.2, libm.so.6, liblmdb.so.0, libopenblas.so.0, libc.so.6, libunwind.so.8, libaec.so.0, libgfortran.so.5, liblzma.so.5, libquadmath.so.0.

Recent Test Results

OpenBenchmarking.org Results Compare

1 System - 6 Benchmark Results

Intel Xeon D-1559 - Kontron COMe-bBD7 E2 v1.0.2 - Intel Xeon E7 v4

Ubuntu 20.04 - 5.4.0-72-generic - X Server 1.20.9

1 System - 60 Benchmark Results

AMD Ryzen 7 4800HS - ASUS GA401IU v1.0 - AMD Renoir Root Complex

Fedora 33 - 5.11.8-200.fc33.x86_64 - KDE Plasma 5.21.3

1 System - 6 Benchmark Results

Intel Core i3-8100T - LENOVO 312D - Intel Cannon Lake PCH

Ubuntu 18.04 - 4.15.0-137-generic - GNOME Shell 3.28.4

1 System - 6 Benchmark Results

Intel Core i9-10900K - ASUS ROG STRIX Z490-F GAMING - Intel Comet Lake PCH

Ubuntu 20.04 - 5.8.0-45-generic - GNOME Shell 3.36.4

4 Systems - 210 Benchmark Results

POWER9 - PowerNV T2P9D01 REV 1.01 - 64GB

Ubuntu 20.10 - 5.9.10-050910-generic - X Server

2 Systems - 152 Benchmark Results

AMD EPYC 7601 32-Core - TYAN B8026T70AE24HR - AMD 17h

Ubuntu 20.04 - 5.4.0-47-generic - GNOME Shell 3.36.4

1 System - 6 Benchmark Results

2 x Intel Xeon E5-2680 v2 - Supermicro X9DRW v0123456789 - Intel Xeon E7 v2

Peppermint 10 - 5.0.0-37-generic - LXDE

7 Systems - 349 Benchmark Results

AMD Ryzen 7 5800X 8-Core - ASRock X570 Pro4 - AMD Starship

Ubuntu 20.10 - 5.8.0-26-generic - GNOME Shell 3.38.1

6 Systems - 349 Benchmark Results

AMD Ryzen 9 5900X 12-Core - ASRock X570 Pro4 - AMD Starship

Ubuntu 20.10 - 5.8.0-26-generic - GNOME Shell 3.38.1

1 System - 282 Benchmark Results

AMD Ryzen 9 5900X 12-Core - ASRock X570 Pro4 - AMD Starship

Ubuntu 20.10 - 5.8.0-26-generic - GNOME Shell 3.38.1

5 Systems - 44 Benchmark Results

2 x Intel Xeon Gold 6248R - GIGABYTE MD61-SC2-00 v01000100 - Intel Sky Lake-E DMI3 Registers

Ubuntu 18.04 - 5.3.0-40-generic - GNOME Shell 3.28.4

Most Popular Test Results

OpenBenchmarking.org Results Compare

3 Systems - 46 Benchmark Results

AMD Ryzen Threadripper 3960X 24-Core - MSI Creator TRX40 - AMD Starship

Ubuntu 20.04 - 5.9.0-rc5-14sep-patch - GNOME Shell 3.36.4

3 Systems - 111 Benchmark Results

2 x AMD EPYC 7742 64-Core - AMD DAYTONA_X - AMD Starship

Ubuntu 20.10 - 5.4.0-42-generic - GNOME Shell 3.36.4

4 Systems - 210 Benchmark Results

POWER9 - PowerNV T2P9D01 REV 1.01 - 64GB

Ubuntu 20.10 - 5.9.10-050910-generic - X Server

5 Systems - 44 Benchmark Results

Intel Xeon Gold 6238R - GIGABYTE MD61-SC2-00 v01000100 - Intel Sky Lake-E DMI3 Registers

Ubuntu 18.04 - 5.3.0-40-generic - GNOME Shell 3.28.4

6 Systems - 349 Benchmark Results

AMD Ryzen 9 5900X 12-Core - ASRock X570 Pro4 - AMD Starship

Ubuntu 20.10 - 5.8.0-26-generic - GNOME Shell 3.38.1

3 Systems - 19 Benchmark Results

2 x Intel Xeon Platinum 8280 - GIGABYTE MD61-SC2-00 v01000100 - Intel Sky Lake-E DMI3 Registers

Ubuntu 20.04 - 5.9.0-050900rc4daily20200912-generic - GNOME Shell 3.36.1

2 Systems - 152 Benchmark Results

AMD EPYC 7601 32-Core - TYAN B8026T70AE24HR - AMD 17h

Ubuntu 20.04 - 5.4.0-47-generic - GNOME Shell 3.36.3

4 Systems - 129 Benchmark Results

2 x AMD EPYC 7742 64-Core - AMD DAYTONA_X - AMD Starship

Ubuntu 20.04 - 5.4.0-48-generic - GNOME Shell 3.36.4

3 Systems - 28 Benchmark Results

AMD Ryzen 9 3950X 16-Core - ASUS ROG CROSSHAIR VIII HERO - AMD Starship

Ubuntu 20.04 - 5.4.0-48-generic - GNOME Shell 3.36.4

7 Systems - 349 Benchmark Results

AMD Ryzen 5 5600X 6-Core - ASRock X570 Pro4 - AMD Starship

Ubuntu 20.10 - 5.8.0-26-generic - GNOME Shell 3.38.1

4 Systems - 35 Benchmark Results

AMD EPYC 7F72 24-Core - ASRockRack EPYCD8 - AMD Starship

Ubuntu 20.04 - 5.9.0-050900rc6daily20200921-generic - GNOME Shell 3.36.4

Find More Test Results