Accelerating Deep Neural Networks [Kõva köide]

Ryoma Sato (National Institute of Informatics, Chiyoda, Japan)

Formaat: Hardback, 311 pages, Worked examples or Exercises
Ilmumisaeg: 31-May-2026
Kirjastus: Cambridge University Press
ISBN-10: 1009687085
ISBN-13: 9781009687089

Teised raamatud teemal:

Pattern recognition
Information theory
Data analysis: general - (Hetkel poes: 3 nimetust)

Kõva köide
Hind: 55,65 €
See raamat ei ole veel ilmunud. Raamatu kohalejõudmiseks kulub orienteeruvalt 3-4 nädalat peale raamatu väljaandmist.
Kogus:
- - 1
  - 2
  - 3
  - 4
  - 5
  - 6
  - 7
  - 8
  - 9
  - 10
Lisa ostukorvi
Tasuta tarne
Tellimisaeg 2-4 nädalat
Lisa soovinimekirja

Formaat: Hardback, 311 pages, Worked examples or Exercises
Ilmumisaeg: 31-May-2026
Kirjastus: Cambridge University Press
ISBN-10: 1009687085
ISBN-13: 9781009687089

Teised raamatud teemal:

Pattern recognition
Information theory
Data analysis: general - (Hetkel poes: 3 nimetust)

Püsilink: https://www.kriso.ee/db/9781009687089.html

Märksõnad:

Deep learning models are powerful, but are often large, slow, and expensive to run. This book is a practical guide to accelerating and compressing neural networks using proven techniques such as quantization, pruning, distillation, and fast architectures. It explains how and why these methods work, fostering a comprehensive understanding. Written for engineers, researchers, and advanced students, the book combines clear theoretical insights with hands-on PyTorch implementations and numerical results. Readers will learn how to reduce inference time and memory usage, lower deployment costs, and select the right acceleration strategy for their task. Whether you're working with large language models, vision systems, or edge devices, this book gives you the tools and intuition needed to build faster, leaner AI systems, without sacrificing performance. It is perfect for anyone who wants to go beyond intuition and take a principled approach to optimizing AI systems

Arvustused

'This book is a practical guide to DNN and LLM acceleration, bridging the gap between theory and practice. Moving beyond 'black-box' tricks, it pairs the latest techniques-like FlashAttention-with runnable code and empirical data. Readers will gain both the technical tools and the fundamental understanding to optimize models effectively.' Masashi Sugiyama, RIKEN and University of Tokyo 'This book effectively bridges theory and practice in accelerating deep learning. It offers clear insights into modern architectures such as Mamba, while also elucidating fundamental concepts and practical techniques for efficient deep learning. It will be a valuable resource for researchers and graduate students seeking a deep understanding of modern deep learning.' Makoto Yamada, Okinawa Institute of Science and Technology

Muu info

Accelerate AI with theory-backed techniques like quantization, pruning and efficient architectures with code and numerical results.

1. Introduction;
2. Overview of acceleration methods;
3. Quantization
and low precision;
4. Pruning;
5. Distillation;
6. Low-rank approximation;
7.
Fast architectures;
8. Tools for tuning;
9. Efficient training; Conclusion;
References; Index.

Ryoma Sato is Assistant Professor at the National Institute of Informatics, Japan, specializing in graph neural networks, optimal transport, and efficient deep learning. He is the author of 'Theory and Algorithms of Optimal Transport' (2023) and 'Graph Neural Networks' (2024). He is a former IOI Japan representative and ACM-ICPC World Finalist, as well as lead developer of Readable, an AI-powered PDF translation service.

Accelerating Deep Neural Networks [Kõva köide]

Arvustused

Muu info

Konto & seaded

Otsing

Otsingu andmebaas

Filtreeri tulemusi

Teemad Ingliskeelsed raamatud

Vali ostukorv