LIBRISTO
LIBROAMANTO
povinné
Staňte se součástí komunity milovníků knih z celého světa a získejte hromadu výhod. Založit účet zdarma
0
Doprava zdarma se Zásilkovnou nad 1 499 Kč
Kurýr DPD 69 Balíkovna 69 PPL kurýr 74 PPL box 39 Zásilkovna 39 Výdejní místo DPD 49 PPL shop 49 Balíkovna 49

Doprava zdarma při nákupu nad 1 499 Kč přes Zásilkovnu nebo PPL Box.

Silicon, Power, and Intelligence (Volume-II)

Model Compression and Efficient Inference

Jazyk AngličtinaAngličtina
Kniha Brožovaná
Kniha Silicon, Power, and Intelligence (Volume-II) Sanzaya Patel
Libristo kód: 52749714
Nakladatelství Independently published, květen 2026
Modern AI models are powerful. Running them efficiently is the real challenge.As large language mode... Celý popis
? points 86 b Nové Nové
857
Skladem u dodavatele Odesíláme za 14-21 dnů

Až 30 dní na vrácení zboží

Modern AI models are powerful. Running them efficiently is the real challenge.

As large language models grow to billions and even trillions of parameters, the future of artificial intelligence is no longer defined solely by model capability-it is defined by efficiency. Memory bandwidth, latency, power consumption, context length, and deployment costs have become the new battlegrounds of AI engineering.

In Volume II: Model Compression and Efficient Inference, engineer and researcher Sanzaya Patel explores the technologies that are transforming massive neural networks into practical, deployable systems. From quantization and pruning to knowledge distillation, KV-cache optimization, PagedAttention, FlashAttention, and Mixture-of-Experts architectures, this volume provides a comprehensive engineering roadmap for reducing computational cost while preserving intelligence.

Moving beyond theory, the book reveals how modern AI systems overcome memory bottlenecks, optimize data movement, compress model representations, and maximize performance across edge devices, workstations, and large-scale inference infrastructure.

Inside, you'll discover:

The mathematics and engineering of model quantization

How NF4 and low-bit representations revolutionized LLM deployment

Structural and unstructured pruning techniques

Knowledge distillation and edge fine-tuning strategies

The hidden memory crisis caused by KV caches

How PagedAttention transformed LLM memory management

Why FlashAttention became one of the most important breakthroughs in modern AI systems

The architecture and economics of Mixture-of-Experts models

Practical strategies for building faster, smaller, and more efficient AI systems

Designed for engineers, researchers, architects, students, and AI practitioners, this volume bridges machine learning theory, systems engineering, memory architecture, and deployment optimization into a unified framework for modern inference.

The future of AI belongs not to the largest models, but to the most efficient ones.

Learn how modern intelligence is compressed, accelerated, and deployed at scale.

Herečka & Polyglotka
EWA KASP pro
Přehrát video
Ewa Kasp
Libristo má největší výběr cizojazyčné literatury. Proto své knihy kupuji tady.

Informace o knize

Plný název Silicon, Power, and Intelligence (Volume-II)
Jazyk Angličtina
Vazba Kniha - Brožovaná
Datum vydání 2026
Počet stran 372
EAN 9798199263566
Libristo kód 52749714
Nakladatelství Independently published
Váha 862
Rozměry 216 x 280 x 20
Darujte tuto knihu ještě dnes
Je to snadné
1 Přidejte knihu do košíku a zvolte doručit jako dárek 2 Obratem vám zašleme poukaz 3 Kniha dorazí na adresu obdarovaného

Přihlášení

Přihlaste se ke svému účtu. Ještě nemáte Libristo účet? Vytvořte si ho nyní!

 
povinné
povinné

Nemáte účet? Získejte výhody Libristo účtu!

Díky Libristo účtu budete mít vše pod kontrolou.

Vytvořit Libristo účet
Knižní rádce Libroamiko
Ahoj, jsem Libroamiko, můžu pomoct?