Skip to product information
1 of 1

Generative AI on Kubernetes

Generative AI on Kubernetes

Operationalizing Large Language Models

Paperback

Regular price £43.02
Regular price Sale price £43.02

Join our rewards scheme and earn 129 reward points on this purchase!

Earn 129 points on this!

Sign in or Sign up!
View full details
  • Release Date: 13/03/2026
  • Barcode: 9781098171926
  • Genre: Computing & The Internet
  • Sub-Genre: Applications & Software
  • Imprint: O'Reilly Media
  • Publisher: O'Reilly Media
Generative AI on Kubernetes

Generative AI on Kubernetes

Collapsible content

DESCRIPTION

Operationalizing Large Language Models
This book serves as a practical, hands-on guide for MLOps engineers, software developers, Kubernetes administrators, and AI professionals ready to unlock AI innovation with the power of cloud native infrastructure.
Generative AI is revolutionizing industries, and Kubernetes has fast become the backbone for deploying and managing these resource-intensive workloads. This book serves as a practical, hands-on guide for MLOps engineers, software developers, Kubernetes administrators, and AI professionals ready to unlock AI innovation with the power of cloud native infrastructure. Authors Roland Huss and Daniele Zonca provide a clear road map for training, fine-tuning, deploying, and scaling GenAI models on Kubernetes, addressing challenges like resource optimization, automation, and security along the way.

With actionable insights with real-world examples, readers will learn to tackle the opportunities and complexities of managing GenAI applications in production environments. Whether you're experimenting with large-scale language models or facing the nuances of AI deployment at scale, you'll uncover expertise you need to operationalize this exciting technology effectively.

  • Learn to run GenAI models on Kubernetes for efficient scalability
  • Get techniques to train and fine-tune LLMs within Kubernetes environments
  • See how to deploy production-ready AI systems with automation and resource optimization
  • Discover how to monitor and scale GenAI applications to handle real-world demand
  • Uncover the best tools to operationalize your GenAI workloads
  • Learn how to run agent-based and AI-driven applications


DELIVERY & RETURNS

UK Delivery:

  • Free delivery on all orders of £10 or more.
  • £1.49 delivery fee on orders below £10.
  • UK orders are shipped via Royal Mail 2nd Class.

International Delivery:

  • Flat rate delivery charges vary by country.

Dispatch and Delivery Times:

  • All orders are shipped from our warehouse in Northampton, UK within 48 hours of receipt during working hours.
  • UK mainland orders typically arrive within 3-5 working days via Royal Mail 2nd Class.
  • International estimated delivery times:
  • Europe & Channel Islands: 7 to 10 working days
  • USA: 7 to 15 working days
  • Rest of the World: 9 to 21 working days

View our full delivery infomation here.

  • OVER

    2 MILLION PRODUCTS

  • 60 MILLION CUSTOMERS

    ACROSS 190 COUNTRIES