jcjohnson / jcjohnson/cnn-benchmarks

Pascal architecture is much slower than Maxwell...

Open
#6 3 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
2.5k
Forks
402
PR merge metrics
No merged PRs in 30d

Description

Hi Justin,

In fact Pascal architecture is much slower than Maxwell. Could you please look at my benchmarks down below? The point is to measure old architecture GPUs with optimized DNN libraries for that specific architecture - CUDA 7.5 cuDNN v4.0.

Measured in CNTK, Caffe under Windows 7 and Ubuntu OS.

GTX 980 Ti - CUDA 7.5 with cuDNN 4.0
Caffe Performance: 5242 imgs/s

GTX 980 Ti - CUDA 8.0 RC with cuDNN 5.1
Caffe Performance: 4183 imgs/s

GTX 1080 - CUDA 7.5 with cuDNN 4.0
Caffe Performance: Not Applicable

GTX 1080 - CUDA 8.0 RC with cuDNN 5.1
Caffe Performance: 4628 imgs/s

Best Regards,
Ondrej

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reviewing the reported CNTK and Caffe benchmark results for GTX 980 Ti and GTX 1080 across CUDA 7.5/8.0 and cuDNN 4.0/5.1 on Windows 7 and Ubuntu. Done means reproducing or explaining the Pascal-versus-Maxwell performance discrepancy with the relevant benchmark configuration documented.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning, performance
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.