{"product_id":"9798341621497-hands-on-llm-serving-and-optimization","title":"Hands-On LLM Serving and Optimization","description":"\u003cmeta content=\"text\/html; charset=utf-8\" http-equiv=\"Content-Type\"\u003e\u003cp\u003e\u003cspan\u003eHosting LLMs at Scale\u003cbr\u003eAs the demand for real-time AI applications grows, along comes this comprehensive guide to the complexities of deploying and optimizing LLMs at scale. The authors take a real-world approach backed by practical examples and code, and assemble essential strategies for designing infrastructures that are equal to the demands of modern AI applications.\u003cbr\u003e\u003cp\u003eLarge language models (LLMs) are rapidly becoming the backbone of AI-driven applications. Without proper optimization, however, LLMs can be expensive to run, slow to serve, and prone to performance bottlenecks. As the demand for real-time AI applications grows, along comes Hands-On Serving and Optimizing LLM Models, a comprehensive guide to the complexities of deploying and optimizing LLMs at scale.\u003c\/p\u003e\n\u003cp\u003eIn this hands-on book, authors Chi Wang and Peiheng Hu take a real-world approach backed by practical examples and code, and assemble essential strategies for designing robust infrastructures that are equal to the demands of modern AI applications. Whether you're building high-performance AI systems or looking to enhance your knowledge of LLM optimization, this indispensable book will serve as a pillar of your success.\u003c\/p\u003e\n\u003cul\u003e\n\u003cli\u003eLearn the key principles for designing a model-serving system tailored to popular business scenarios\u003c\/li\u003e\n\u003cli\u003eUnderstand the common challenges of hosting LLMs at scale while minimizing costs\u003c\/li\u003e\n\u003cli\u003ePick up practical techniques for optimizing LLM serving performance\u003c\/li\u003e\n\u003cli\u003eBuild a model-serving system that meets specific business requirements\u003c\/li\u003e\n\u003cli\u003eImprove LLM serving throughput and reduce latency\u003c\/li\u003e\n\u003cli\u003eHost LLMs in a cost-effective manner, balancing performance and resource efficiency\u003c\/li\u003e\n\u003c\/ul\u003e\n\u003cbr\u003e\u003cbr\u003e\u003c\/span\u003e\u003c\/p\u003e","brand":"Rarewaves","offers":[{"title":"Default Title","offer_id":58304904266102,"sku":"9798341621497","price":55.02,"currency_code":"GBP","in_stock":true}],"thumbnail_url":"\/\/cdn.shopify.com\/s\/files\/1\/0092\/7504\/8033\/files\/stand_41436232.jpg?v=1784519790","url":"https:\/\/www.rarewaves.com\/products\/9798341621497-hands-on-llm-serving-and-optimization","provider":"Rarewaves.com","version":"1.0","type":"link"}