Könyv Kubernetes for Generative AI Solutions Sukirti Gupta

Kubernetes for Generative AI Solutions

Szerző: Sukirti Gupta
Nyelv: Angol
Kötés: Puha kötésű
Elérhetőség: Beszállítói készleten
Küldés 9-15 napon belül
17 930 Ft
Master the complete Generative AI project lifecycle on Kubernetes (K8s) from design and optimization...

Információk a könyvről

Szerző
Nyelv
Angol
Kötés
Könyv - Puha kötésű
Kiadva
2025
oldal
334
EAN
9781836209935
ISBN
1836209932
Enbook ID
48972580
Súly
625
Méretek
191 x 235 x 18

Teljes leírás

Master the complete Generative AI project lifecycle on Kubernetes (K8s) from design and optimization to deployment using best practices, cost-effective strategies, and real-world examples.

Key Features:

- Build and deploy your first Generative AI workload on Kubernetes with confidence

- Learn to optimize costly resources such as GPUs using fractional allocation, Spot Instances, and automation

- Gain hands-on insights into observability, infrastructure automation, and scaling Generative AI workloads

- Purchase of the print or Kindle book includes a free PDF eBook

Book Description:

Generative AI (GenAI) is revolutionizing industries, from chatbots to recommendation engines to content creation, but deploying these systems at scale poses significant challenges in infrastructure, scalability, security, and cost management.

This book is your practical guide to designing, optimizing, and deploying GenAI workloads with Kubernetes (K8s) the leading container orchestration platform trusted by AI pioneers. Whether you're working with large language models, transformer systems, or other GenAI applications, this book helps you confidently take projects from concept to production. You'll get to grips with foundational concepts in machine learning and GenAI, understanding how to align projects with business goals and KPIs. From there, you'll set up Kubernetes clusters in the cloud, deploy your first workload, and build a solid infrastructure. But your learning doesn't stop at deployment. The chapters highlight essential strategies for scaling GenAI workloads in production, covering model optimization, workflow automation, scaling, GPU efficiency, observability, security, and resilience.

By the end of this book, you'll be fully equipped to confidently design and deploy scalable, secure, resilient, and cost-effective GenAI solutions on Kubernetes.

What You Will Learn:

- Explore GenAI deployment stack, agents, RAG, and model fine-tuning

- Implement HPA, VPA, and Karpenter for efficient autoscaling

- Optimize GPU usage with fractional allocation, MIG, and MPS setups

- Reduce cloud costs and monitor spending with Kubecost tools

- Secure GenAI workloads with RBAC, encryption, and service meshes

- Monitor system health and performance using Prometheus and Grafana

- Ensure high availability and disaster recovery for GenAI systems

- Automate GenAI pipelines for continuous integration and delivery

Who this book is for:

This book is for solutions architects, product managers, engineering leads, DevOps teams, GenAI developers, and AI engineers. It's also suitable for students and academics learning about GenAI, Kubernetes, and cloud-native technologies. A basic understanding of cloud computing and AI concepts is needed, but no prior knowledge of Kubernetes is required.

Table of Contents

- GenAI-Intro, Evolution, and Project Lifecycle

- K8s-Introduction and Integration with GenAI

- Getting Started with K8s in the Cloud

- GenAI Model Optimization for Domain-Specific Use Cases (RAG, Fine Tuning, etc.)

- Getting Started with GenAI on K8s-Chatbot Example

- Deploying GenAI on K8s-Scaling Best Practices

- Deploying GenAI on K8s-Cost Optimization Best Practices

- Deploying GenAI on K8s-Networking Best Practices

- Deploying GenAI on K8s-Security Best Practices

- Optimizing GPU Resources in K8s for GenAI Applications

- GenAIOps: Creating GenAI Automation Pipeline

- Getting Visibility into GenAI Workloads Resource Utilization

- High Availability and Disaster Recovery Implementation

- Wrap Up and Further Readings

Érdekelheti

8 032 Ft

Designing in Ethics

Jeroen Van Den Hoven
48 938 Ft

The World of Ice

Robert Michael Ballantyne
3 811 Ft
22 521 Ft

Liar Liar

James Patterson
5 873 Ft

States of Emergency

Stephen Morton
53 613 Ft

Terrible Two

Mac Barnett
5 138 Ft
13 384 Ft
12 961 Ft

Lotus Shoes

Jane Yang
5 209 Ft
11 452 Ft
70 101 Ft
19 484 Ft

Attunement

Alberto Perez-Gomez
15 303 Ft

Conspiracy Nation

Peter Knight
44 588 Ft

Syntax of Dutch

Hans Broekhuis
58 257 Ft

Mansfield Park

Jane Austen
11 309 Ft
13 794 Ft

Azok a vásárlók, akik ezt a könyvet megvásárolták, a következőket is megvásárolták

3 544 Ft

Un anarquista

Diego Ameixeiras
2 627 Ft
3 317 Ft
7 160 Ft

Pád

Zbyšek Lorenz
2 783 Ft
3 602 Ft

Jalta Global

MR Daniel Raza
3 758 Ft