Hands on Llm Serving Optimization by Wang Peiheng (9 results)

Author
Title
Refine with Advanced Search

Refine your search

  • Books (9)

to

Custom price range (£)

to

  • Language: English

    Published by O'Reilly Media, 2026

    9798341621497

    • Softcover

    Seller: GreatBookPrices, Columbia, MD, U.S.A.GreatBookPrices

    5-star seller
    Contact seller

    Condition: New

    £ 44.25

    £ 1.97 shipping 
    Ships within U.S.A.

    Quantity: Over 20 available

    Condition: New.

  • Language: English

    Published by O'Reilly Media, 2026

    9798341621497

    • Softcover

    Seller: GreatBookPrices, Columbia, MD, U.S.A.GreatBookPrices

    5-star seller
    Contact seller

    Condition: Used - As new

    £ 48.15

    £ 1.97 shipping 
    Ships within U.S.A.

    Quantity: Over 20 available

    Condition: As New. Unread book in perfect condition.

  • Language: English

    Published by O'Reilly Media, 2026

    9798341621497

    • Softcover

    Seller: Rarewaves USA, HEBRON, KY, U.S.A.Rarewaves USA

    5-star seller
    Contact seller

    Condition: New

    £ 52.32

     Free Shipping 
    Ships within U.S.A.

    Quantity: Over 20 available

    Paperback. Condition: New.

  • Language: English

    Published by O'Reilly Media, 2026

    9798341621497

    • Softcover

    Seller: California Books, Miami, FL, U.S.A.California Books

    5-star seller
    Contact seller

    Condition: New

    £ 53.88

     Free Shipping 
    Ships within U.S.A.

    Quantity: Over 20 available

    Condition: New.

  • Language: English

    Published by O'Reilly Media, US, 2026

    9798341621497

    • Softcover

    Seller: Rarewaves.com USA, London, LONDO, United KingdomRarewaves.com USA

    5-star seller
    Contact seller

    Condition: New

    £ 59.08

     Free Shipping 
    Ships from United Kingdom to U.S.A.

    Quantity: 1 available

    Paperback. Condition: New. Large language models (LLMs) are rapidly becoming the backbone of AI-driven applications. Without proper optimization, however, LLMs can be expensive to run, slow to serve, and prone to performance bottlenecks. As the demand for real-time AI applications grows, along comes Hands-On Serving and Optimizing LLM Models, a comprehensive guide to the complexities of deploying and optimizing LLMs at scale.In this hands-on book, authors Chi Wang and Peiheng Hu take a real-world approach backed by practical examples and code, and assemble essential strategies for designing robust infrastructures that are equal to the demands of modern AI applications. Whether you're building high-performance AI systems or looking to enhance your knowledge of LLM optimization, this indispensable book will serve as a pillar of your success.Learn the key principles for designing a model-serving system tailored to popular business scenariosUnderstand the common challenges of hosting LLMs at scale while minimizing costsPick up practical techniques for optimizing LLM serving performanceBuild a model-serving system that meets specific business requirementsImprove LLM serving throughput and reduce latencyHost LLMs in a cost-effective manner, balancing performance and resource efficiency.

  • Language: English

    Published by O'Reilly Media, 2026

    9798341621497

    • Softcover

    Seller: GreatBookPricesUK, Woodford Green, United KingdomGreatBookPricesUK

    5-star seller
    Contact seller

    Condition: New

    £ 42.92

    £ 15.00 shipping 
    Ships from United Kingdom to U.S.A.

    Quantity: Over 20 available

    Condition: New.

  • Language: English

    Published by O'Reilly Media, 2026

    9798341621497

    • Softcover

    Seller: GreatBookPricesUK, Woodford Green, United KingdomGreatBookPricesUK

    5-star seller
    Contact seller

    Condition: Used - As new

    £ 50.65

    £ 15.00 shipping 
    Ships from United Kingdom to U.S.A.

    Quantity: Over 20 available

    Condition: As New. Unread book in perfect condition.

  • Language: English

    Published by O'Reilly Media, US, 2026

    9798341621497

    • Softcover

    Seller: Rarewaves USA United, HEBRON, KY, U.S.A.Rarewaves USA United

    5-star seller
    Contact seller

    Condition: New

    £ 51.88

    £ 37.36 shipping 
    Ships within U.S.A.

    Quantity: Over 20 available

    Paperback. Condition: New. Large language models (LLMs) are rapidly becoming the backbone of AI-driven applications. Without proper optimization, however, LLMs can be expensive to run, slow to serve, and prone to performance bottlenecks. As the demand for real-time AI applications grows, along comes Hands-On Serving and Optimizing LLM Models, a comprehensive guide to the complexities of deploying and optimizing LLMs at scale.In this hands-on book, authors Chi Wang and Peiheng Hu take a real-world approach backed by practical examples and code, and assemble essential strategies for designing robust infrastructures that are equal to the demands of modern AI applications. Whether you're building high-performance AI systems or looking to enhance your knowledge of LLM optimization, this indispensable book will serve as a pillar of your success.Learn the key principles for designing a model-serving system tailored to popular business scenariosUnderstand the common challenges of hosting LLMs at scale while minimizing costsPick up practical techniques for optimizing LLM serving performanceBuild a model-serving system that meets specific business requirementsImprove LLM serving throughput and reduce latencyHost LLMs in a cost-effective manner, balancing performance and resource efficiency.

  • Language: English

    Published by O'Reilly Media, US, 2026

    9798341621497

    • Softcover

    Seller: Rarewaves.com UK, London, United KingdomRarewaves.com UK

    5-star seller
    Contact seller

    Condition: New

    £ 54.70

    £ 65.00 shipping 
    Ships from United Kingdom to U.S.A.

    Quantity: 1 available

    Paperback. Condition: New. Large language models (LLMs) are rapidly becoming the backbone of AI-driven applications. Without proper optimization, however, LLMs can be expensive to run, slow to serve, and prone to performance bottlenecks. As the demand for real-time AI applications grows, along comes Hands-On Serving and Optimizing LLM Models, a comprehensive guide to the complexities of deploying and optimizing LLMs at scale.In this hands-on book, authors Chi Wang and Peiheng Hu take a real-world approach backed by practical examples and code, and assemble essential strategies for designing robust infrastructures that are equal to the demands of modern AI applications. Whether you're building high-performance AI systems or looking to enhance your knowledge of LLM optimization, this indispensable book will serve as a pillar of your success.Learn the key principles for designing a model-serving system tailored to popular business scenariosUnderstand the common challenges of hosting LLMs at scale while minimizing costsPick up practical techniques for optimizing LLM serving performanceBuild a model-serving system that meets specific business requirementsImprove LLM serving throughput and reduce latencyHost LLMs in a cost-effective manner, balancing performance and resource efficiency.