• Open Daily: 10am - 10pm
    Alley-side Pickup: 10am - 7pm

    3038 Hennepin Ave Minneapolis, MN
    612-822-4611

Open Daily: 10am - 10pm | Alley-side Pickup: 10am - 7pm
3038 Hennepin Ave Minneapolis, MN
612-822-4611
Silicon, Power, and Intelligence (Volume-II): Model Compression and Efficient Inference

Silicon, Power, and Intelligence (Volume-II): Model Compression and Efficient Inference

Paperback

Series: Silicon, Power, and Intelligence - A Hardware-Aware AI Engineering, Book 2

Technology & EngineeringProgramming

ISBN13: 9798199263566
Publisher: Independently Published
Published: May 30 2026
Pages: 372
Weight: 1.90
Height: 0.77 Width: 8.50 Depth: 11.00
Language: English

Modern AI models are powerful. Running them efficiently is the real challenge.

As large language models grow to billions and even trillions of parameters, the future of artificial intelligence is no longer defined solely by model capability-it is defined by efficiency. Memory bandwidth, latency, power consumption, context length, and deployment costs have become the new battlegrounds of AI engineering.

Also from

Patel, Sanzaya

Also in

Programming