10% off all books and free delivery over £50
Buy from our bookstore and 25% of the cover price will be given to a school of your choice to buy more books. *15% of eBooks.

Accelerating Deep Neural Networks

View All Editions (1)

The selected edition of this book is not available to buy right now.
Add To Wishlist
Write A Review

About

Accelerating Deep Neural Networks Synopsis

Deep learning models are powerful, but are often large, slow, and expensive to run. This book is a practical guide to accelerating and compressing neural networks using proven techniques such as quantization, pruning, distillation, and fast architectures. It explains how and why these methods work, fostering a comprehensive understanding.

Written for engineers, researchers, and advanced students, the book combines clear theoretical insights with hands-on PyTorch implementations and numerical results. Readers will learn how to reduce inference time and memory usage, lower deployment costs, and select the right acceleration strategy for their task. Whether you're working with large language models, vision systems, or edge devices, this book gives you the tools and intuition needed to build faster, leaner AI systems, without sacrificing performance.

It is perfect for anyone who wants to go beyond intuition and take a principled approach to optimizing AI systems

About This Edition

ISBN: 9781009687089
Publication date:
Author: Ryoma Sato
Publisher: Cambridge University Press
Format: Hardback
Pagination: 311 pages
Genres: Pattern recognition
Information theory
Data science and analysis: general

Frequently asked questions