Quantized Model Deployment : INT8 and FP16 Compression for Mobile Acceleration

Language: English

Published by Amazon Digital Services LLC - Kdp Mai 2026, 2026

9798196245466

  • Softcover
  • New
See all details

Seller: AHA-BUCH GmbH, Einbeck, GermanyAHA-BUCH GmbH

5-star seller

AbeBooks seller since August 14, 2006

Softcover

Condition: New

£ 24.81

£ 29.75 shipping 
Ships from Germany to U.S.A.

Quantity: 2 available

Add to basket
Free 30-day returns

Item description from seller

Neuware - What if the only thing standing between your neural network and real-time mobile performance is the precision you refuse to give up Your model ran flawlessly in PyTorch-400MB of FP32 weights, a 350-watt GPU, and all the thermal headroom in the world. Then you deployed it to a phone. It stuttered. It heated up. The OS killed it before it produced a single inference. The market no longer asks whether AI can run on mobile. It asks why your AI is slower and less accurate than the cloud version. The answer is not your architecture. It is your precision.This book is the field manual for engineers who refuse to accept the old compromise of smaller models and weaker accuracy. Inside, you will learn: - Why INT8 and FP16 are not arbitrary format choices, but hardware-mandated keys to dedicated acceleration paths on Snapdragon, Apple Neural Engine, and MediaTek APU - How naïve post-training quantization can crater accuracy by double-digit percentages-and the calibration, range estimation, and outlier handling techniques that prevent it - The exact deployment architecture for TensorFlow Lite, Core ML, ONNX Runtime Mobile, and NNAPI, including operator fusion and numerical equivalence testing - Why quantization is the only optimization that simultaneously improves latency, accuracy, and power consumption-and how to combine it with pruning and knowledge distillation for wearables and IoTStop accepting the compromise between speed and accuracy. Build models that run cooler, faster, and sharper on the devices already in your users' pockets. The precision you can no longer afford is the precision you can finally reclaim.…

Seller Inventory # 9798196245466

Title
Quantized Model Deployment : INT8 and FP16 Compression for Mobile Acceleration
Author
Clara Whiskers
Publisher
Amazon Digital Services LLC - Kdp Mai 2026
Publication year
2026
Condition
Neu
Binding
Taschenbuch
Language
English
ISBN 13
9798196245466
Item weight
381 grams
Dimensions
244x170x12 mm

AHA-BUCH GmbH

Einbeck, Germany

5-star seller

AbeBooks seller since August 14, 2006

Shipping rates from Germany to U.S.A.

Item7 to 10 business days5 to 7 business days
First item£ 29.75£ 38.24
Delivery times are set by sellers and vary by carrier and location. Orders passing through Customs may face delays and buyers are responsible for any associated duties or fees. Sellers may contact you regarding additional charges to cover any increased costs to ship your items.

Payment methods

  • Visa
  • Mastercard
  • American Express
  • Apple Pay
  • Google Pay
  • Bank Wire Transfer
  • Check
  • Paypal

Store description

Das Unternehmen AHA-BUCH GmbH: Seit der Gründung von AHA-BUCH im Juli 2005 ist unser Hauptziel, zufriedenen Kunden so schnell und so preisgünstig wie möglich ihren Bücherwunsch zu erfüllen. Unsere Firma beschäftigt 16 Mitarbeiter, die nur ein Ziel kennen: den Kunden und seine Wünsche! Auf über 3700 m2 Fläche haben wir über 100.000 Bücher, Modernes Antiquariat und Spiele auf Lager.

Specialty

Kinderbücher & Kinderhör Casetten, German Books, Software, Natur & Tiere, Ratgeber, Sachbücher, Englische Bücher, Medizin & Gesundheit, Universität & Studium

Seller's business information

AHA-BUCH GmbH

Garlebsen 48
Einbeck, Germany 37574