Sign In Start Free Trial
Account

Add to playlist

Create a Playlist

Modal Close icon
You need to login to use this feature.
  • Book Overview & Buying Kubernetes for Generative AI Solutions
  • Table Of Contents Toc
Kubernetes for Generative AI Solutions

Kubernetes for Generative AI Solutions

By : Ashok Srirama, Sukirti Gupta
close
close
Kubernetes for Generative AI Solutions

Kubernetes for Generative AI Solutions

By: Ashok Srirama, Sukirti Gupta

Overview of this book

Generative AI (GenAI) is revolutionizing industries, from chatbots to recommendation engines to content creation, but deploying these systems at scale poses significant challenges in infrastructure, scalability, security, and cost management. This book is your practical guide to designing, optimizing, and deploying GenAI workloads with Kubernetes (K8s) the leading container orchestration platform trusted by AI pioneers. Whether you're working with large language models, transformer systems, or other GenAI applications, this book helps you confidently take projects from concept to production. You’ll get to grips with foundational concepts in machine learning and GenAI, understanding how to align projects with business goals and KPIs. From there, you'll set up Kubernetes clusters in the cloud, deploy your first workload, and build a solid infrastructure. But your learning doesn't stop at deployment. The chapters highlight essential strategies for scaling GenAI workloads in production, covering model optimization, workflow automation, scaling, GPU efficiency, observability, security, and resilience. By the end of this book, you’ll be fully equipped to confidently design and deploy scalable, secure, resilient, and cost-effective GenAI solutions on Kubernetes.
Table of Contents (21 chapters)
close
close
Lock Free Chapter
1
Part 1:GenAI and Kubernetes Foundation
5
Part 2: Productionalizing GenAI Workloads Using K8s
13
Part 3: Operating GenAI Workloads on K8s
16
Chapter 13: High Availability and Disaster Recovery for GenAI Applications

Index

As this ebook edition doesn't have fixed pagination, the page numbers below are hyperlinked for reference only, based on the printed edition of this book.

A

A/B testing

reference link 159

adapter-based tuning 77

add-on software components

CNI plugin 40

CoreDNS 40

CSI plugin 40

device plugins 40

monitoring plugins 40

Advanced Encryption Standard (AES) encryption algorithm 189

agent 5

Alpine Linux

reference link 176

Amazon Bedrock 69

URL 111

Amazon CloudWatch 240

Amazon CloudWatch Logs 240

Amazon EBS CSI driver add-on 89

Amazon EC2

URL 49

Amazon EC2 Capacity Blocks for ML

reference link 218

Amazon EC2 placement groups

cluster 169

partition 169

reference link 169

spread 169

Amazon EC2 Trn1 instances

reference link 199

Amazon EC2 Trn2 instances 199

Amazon EC2 UltraClusters

reference link 218

Amazon ECR

URL 177

Amazon EKS 46

Amazon EKS Blueprints...

CONTINUE READING
83
Tech Concepts
36
Programming languages
73
Tech Tools
Icon Unlimited access to the largest independent learning library in tech of over 8,000 expert-authored tech books and videos.
Icon Innovative learning tools, including AI book assistants, code context explainers, and text-to-speech.
Icon 50+ new titles added per month and exclusive early access to books as they are being written.
Kubernetes for Generative AI Solutions
notes
bookmark Notes and Bookmarks search Search in title playlist Add to playlist download Download options font-size Font size

Change the font size

margin-width Margin width

Change margin width

day-mode Day/Sepia/Night Modes

Change background colour

Close icon Search
Country selected

Close icon Your notes and bookmarks

Confirmation

Modal Close icon
claim successful

Buy this book with your credits?

Modal Close icon
Are you sure you want to buy this book with one of your credits?
Close
YES, BUY

Submit Your Feedback

Modal Close icon
Modal Close icon
Modal Close icon