LIBRISTO
LIBROAMANTO
obligatorisch
Werden Sie Teil einer Gemeinschaft von Buchliebhabern aus der ganzen Welt und erhalten Sie eine Reihe von Vorteilen. Konto kostenlos anlegen
0
Österreichische Post 5.49 GLS-Kurier 4.99 DPD-Kurier 4.49 DPD-Stelle 3.49

Silicon, Power, and Intelligence (Volume-II)

Model Compression and Efficient Inference

Sprache EnglischEnglisch
Buch Broschur
Buch Silicon, Power, and Intelligence (Volume-II) Sanzaya Patel
Libristo-Code: 52749714
Verlag Independently published, Mai 2026
Modern AI models are powerful. Running them efficiently is the real challenge.As large language mode... Vollständige Beschreibung
? points 77 b Neu Neu
31.29 inkl. MwSt.
Externes Lager Wir versenden in 14-21 Tagen

Bis zu 30 Tage Rückgaberecht

Modern AI models are powerful. Running them efficiently is the real challenge.

As large language models grow to billions and even trillions of parameters, the future of artificial intelligence is no longer defined solely by model capability-it is defined by efficiency. Memory bandwidth, latency, power consumption, context length, and deployment costs have become the new battlegrounds of AI engineering.

In Volume II: Model Compression and Efficient Inference, engineer and researcher Sanzaya Patel explores the technologies that are transforming massive neural networks into practical, deployable systems. From quantization and pruning to knowledge distillation, KV-cache optimization, PagedAttention, FlashAttention, and Mixture-of-Experts architectures, this volume provides a comprehensive engineering roadmap for reducing computational cost while preserving intelligence.

Moving beyond theory, the book reveals how modern AI systems overcome memory bottlenecks, optimize data movement, compress model representations, and maximize performance across edge devices, workstations, and large-scale inference infrastructure.

Inside, you'll discover:

The mathematics and engineering of model quantization

How NF4 and low-bit representations revolutionized LLM deployment

Structural and unstructured pruning techniques

Knowledge distillation and edge fine-tuning strategies

The hidden memory crisis caused by KV caches

How PagedAttention transformed LLM memory management

Why FlashAttention became one of the most important breakthroughs in modern AI systems

The architecture and economics of Mixture-of-Experts models

Practical strategies for building faster, smaller, and more efficient AI systems

Designed for engineers, researchers, architects, students, and AI practitioners, this volume bridges machine learning theory, systems engineering, memory architecture, and deployment optimization into a unified framework for modern inference.

The future of AI belongs not to the largest models, but to the most efficient ones.

Learn how modern intelligence is compressed, accelerated, and deployed at scale.

Schauspielerin & Polyglotte
EWA KASP für
Video abspielen
Ewa Kasp
Libristo bietet die größte Auswahl an fremdsprachiger Literatur an. Deshalb kaufe ich meine Bücher hier ein.

Informationen zum Buch

Vollständiger Name Silicon, Power, and Intelligence (Volume-II)
Sprache Englisch
Einband Buch - Broschur
Datum der Veröffentlichung 2026
Anzahl der Seiten 372
EAN 9798199263566
Libristo-Code 52749714
Gewicht 862
Abmessungen 216 x 280 x 20
Verschenken Sie dieses Buch noch heute
Es ist ganz einfach
1 Legen Sie das Buch in Ihren Warenkorb und wählen Sie den Versand als Geschenk 2 Wir schicken Ihnen umgehend einen Gutschein 3 Das Buch wird an die Adresse des beschenkten Empfängers geliefert

Anmeldung

Melden Sie sich bei Ihrem Konto an. Sie haben noch kein Libristo-Konto? Erstellen Sie es jetzt!

 
obligatorisch
obligatorisch

Sie haben kein Konto? Nutzen Sie die Vorteile eines Libristo-Kontos!

Mit einem Libristo-Konto haben Sie alles unter Kontrolle.

Erstellen Sie ein Libristo-Konto
Buchberater Libroamiko
Hallo, ich bin Libroamiko, kann ich helfen?