arXiv (Cornell University)· 2014· Preprint
cuDNN: Efficient Primitives for Deep Learning
- 1,028citations
- 2014year
Short summary
A new library, cuDNN, provides optimized routines for deep learning computational kernels, analogous to BLAS for HPC, to address the challenge of reoptimizing code for evolving parallel architectures.
AI-generated from the title and abstract; the full text is not read.
TakeawaysIn the app
Key pointsIn the app
Ask the paperIn the app
The rest is in the Pofolia app
Takeaways, key points and questions to the paper; new summaries every day for your field. Free.
Sign in on the web to openField: Hardware and Architecture
Hardware and ArchitectureComputer Science