reklama - zainteresowany?

Kubernetes for Generative AI Solutions. A complete guide to designing, optimizing, and deploying Generative AI workloads on Kubernetes - Helion

Kubernetes for Generative AI Solutions. A complete guide to designing, optimizing, and deploying Generative AI workloads on Kubernetes
ebook
Autor: Ashok Srirama, Sukirti Gupta
Tytuł oryginału: Kubernetes for Generative AI Solutions. A complete guide to designing, optimizing, and deploying Generative AI workloads on Kubernetes
ISBN: 9781836209928
Format: ebook
Księgarnia: Helion

Cena książki: 139,00 zł

Książka będzie dostępna od czerwca 2025

Generative AI (GenAI) is revolutionizing industries, from chatbots to recommendation engines to content creation, but deploying these systems at scale poses significant challenges in infrastructure, scalability, security, and cost management.
This book is your practical guide to designing, optimizing, and deploying GenAI workloads with Kubernetes (K8s) the leading container orchestration platform trusted by AI pioneers. Whether you're working with large language models, transformer systems, or other GenAI applications, this book helps you confidently take projects from concept to production. You’ll get to grips with foundational concepts in machine learning and GenAI, understanding how to align projects with business goals and KPIs. From there, you'll set up Kubernetes clusters in the cloud, deploy your first workload, and build a solid infrastructure. But your learning doesn't stop at deployment. The chapters highlight essential strategies for scaling GenAI workloads in production, covering model optimization, workflow automation, scaling, GPU efficiency, observability, security, and resilience.
By the end of this book, you’ll be fully equipped to confidently design and deploy scalable, secure, resilient, and cost-effective GenAI solutions on Kubernetes.

 

Zobacz także:

  • Jak zhakowa
  • Windows Media Center. Domowe centrum rozrywki
  • Ruby on Rails. Ćwiczenia
  • Efekt piaskownicy. Jak szefować żeby roboty nie zabrały ci roboty
  • Przywództwo w świecie VUCA. Jak być skutecznym liderem w niepewnym środowisku

Spis treści

Kubernetes for Generative AI Solutions. A complete guide to designing, optimizing, and deploying Generative AI workloads on Kubernetes eBook -- spis treści

  • 1. GenAI—Intro, Evolution, and Project Lifecycle
  • 2. K8s—Introduction and Integration with GenAI
  • 3. Getting Started with K8s in the Cloud
  • 4. GenAI Model Optimization for Domain-Specific Use Cases (RAG, Fine Tuning, etc.)
  • 5. Getting Started with GenAI on K8s—Chatbot Example
  • 6. Deploying GenAI on K8s—Scaling Best Practices
  • 7. Deploying GenAI on K8s—Cost Optimization Best Practices
  • 8. Deploying GenAI on K8s—Networking Best Practices
  • 9. Deploying GenAI on K8s—Security Best Practices
  • 10. Optimizing GPU Resources in K8s for GenAI Applications
  • 11. GenAIOps: Creating GenAI Automation Pipeline
  • 12. Getting Visibility into GenAI Workloads Resource Utilization
  • 13. High Availability and Disaster Recovery Implementation
  • 14. Wrap Up and Further Readings

Code, Publish & WebDesing by CATALIST.com.pl



(c) 2005-2025 CATALIST agencja interaktywna, znaki firmowe należą do wydawnictwa Helion S.A.