EasterBlack-owned or founded brands at TargetGroceryClothing, Shoes & AccessoriesBabyHomeFurnitureKitchen & DiningOutdoor Living & GardenToysElectronicsVideo GamesMovies, Music & BooksSports & OutdoorsBeautyPersonal CareHealthPetsHousehold EssentialsArts, Crafts & SewingSchool & Office SuppliesParty SuppliesLuggageGift IdeasGift CardsClearanceTarget New ArrivalsTarget Finds#TargetStyleTop DealsTarget Circle DealsWeekly AdShop Order PickupShop Same Day DeliveryRegistryRedCardTarget CircleFind Stores

Kubernetes for Generative AI Solutions - by Ashok Srirama & Sukirti Gupta (Paperback)

Kubernetes for Generative AI Solutions - by  Ashok Srirama & Sukirti Gupta (Paperback) - 1 of 1
$49.99 when purchased online
Target Online store #3991

About this item

Highlights

  • Master the complete Generative AI project lifecycle on Kubernetes (K8s) from design and optimization to deployment using best practices, cost-effective strategies, and real-world examples.Key Features: - Build and deploy your first Generative AI workload on Kubernetes with confidence- Learn to optimize costly resources such as GPUs using fractional allocation, Spot Instances, and automation- Gain hands-on insights into observability, infrastructure automation, and scaling Generative AI workloads- Purchase of the print or Kindle book includes a free PDF eBookBook Description: Generative AI (GenAI) is revolutionizing industries, from chatbots to recommendation engines to content creation, but deploying these systems at scale poses significant challenges in infrastructure, scalability, security, and cost management.This book is your practical guide to designing, optimizing, and deploying GenAI workloads with Kubernetes (K8s) the leading container orchestration platform trusted by AI pioneers.
  • Author(s): Ashok Srirama & Sukirti Gupta
  • 334 Pages
  • Computers + Internet, Systems Architecture

Description



Book Synopsis



Master the complete Generative AI project lifecycle on Kubernetes (K8s) from design and optimization to deployment using best practices, cost-effective strategies, and real-world examples.

Key Features:

- Build and deploy your first Generative AI workload on Kubernetes with confidence

- Learn to optimize costly resources such as GPUs using fractional allocation, Spot Instances, and automation

- Gain hands-on insights into observability, infrastructure automation, and scaling Generative AI workloads

- Purchase of the print or Kindle book includes a free PDF eBook

Book Description:

Generative AI (GenAI) is revolutionizing industries, from chatbots to recommendation engines to content creation, but deploying these systems at scale poses significant challenges in infrastructure, scalability, security, and cost management.

This book is your practical guide to designing, optimizing, and deploying GenAI workloads with Kubernetes (K8s) the leading container orchestration platform trusted by AI pioneers. Whether you're working with large language models, transformer systems, or other GenAI applications, this book helps you confidently take projects from concept to production. You'll get to grips with foundational concepts in machine learning and GenAI, understanding how to align projects with business goals and KPIs. From there, you'll set up Kubernetes clusters in the cloud, deploy your first workload, and build a solid infrastructure. But your learning doesn't stop at deployment. The chapters highlight essential strategies for scaling GenAI workloads in production, covering model optimization, workflow automation, scaling, GPU efficiency, observability, security, and resilience.

By the end of this book, you'll be fully equipped to confidently design and deploy scalable, secure, resilient, and cost-effective GenAI solutions on Kubernetes.

What You Will Learn:

- Explore GenAI deployment stack, agents, RAG, and model fine-tuning

- Implement HPA, VPA, and Karpenter for efficient autoscaling

- Optimize GPU usage with fractional allocation, MIG, and MPS setups

- Reduce cloud costs and monitor spending with Kubecost tools

- Secure GenAI workloads with RBAC, encryption, and service meshes

- Monitor system health and performance using Prometheus and Grafana

- Ensure high availability and disaster recovery for GenAI systems

- Automate GenAI pipelines for continuous integration and delivery

Who this book is for:

This book is for solutions architects, product managers, engineering leads, DevOps teams, GenAI developers, and AI engineers. It's also suitable for students and academics learning about GenAI, Kubernetes, and cloud-native technologies. A basic understanding of cloud computing and AI concepts is needed, but no prior knowledge of Kubernetes is required.

Table of Contents

- Gen AI - Intro, Evolution & Project Lifecycle

- K8s- Introduction & Integration with Gen AI

- Getting Started with K8s in Cloud

- Gen AI Model optimization for domain specific use cases (RAG, Fine tuning etc.)

- Getting started with Gen AI on K8s: Chatbot example

- Deploying GenAI on K8s - Scaling Best practices

- Deploying GenAI on K8s - Cost Optimization Best practices

- Deploying GenAI on K8s - Networking Best practices

- Deploying GenAI on K8s - Security Best practices

- Optimizing GPU Resources in K8s for GenAI Applications

- GenAIOps: Creating GenAI Automation Pipeline

- Getting visibility into GenAI Workloads Resource Utilization

- High Availability and Disaster Recovery Implementation

- Wrap up and further readings

Dimensions (Overall): 9.25 Inches (H) x 7.5 Inches (W) x .7 Inches (D)
Weight: 1.27 Pounds
Suggested Age: 22 Years and Up
Number of Pages: 334
Genre: Computers + Internet
Sub-Genre: Systems Architecture
Publisher: Packt Publishing
Theme: Distributed Systems & Computing
Format: Paperback
Author: Ashok Srirama & Sukirti Gupta
Language: English
Street Date: June 6, 2025
TCIN: 1004856135
UPC: 9781836209935
Item Number (DPCI): 247-05-8798
Origin: Made in the USA or Imported
If the item details above aren’t accurate or complete, we want to know about it.

Shipping details

Estimated ship dimensions: 0.7 inches length x 7.5 inches width x 9.25 inches height
Estimated ship weight: 1.27 pounds
We regret that this item cannot be shipped to PO Boxes.
This item cannot be shipped to the following locations: American Samoa (see also separate entry under AS), Guam (see also separate entry under GU), Northern Mariana Islands, Puerto Rico (see also separate entry under PR), United States Minor Outlying Islands, Virgin Islands, U.S., APO/FPO

Return details

This item can be returned to any Target store or Target.com.
This item must be returned within 90 days of the date it was purchased in store, shipped, delivered by a Shipt shopper, or made ready for pickup.
See the return policy for complete information.

Related Categories

Get top deals, latest trends, and more.

Privacy policy

Footer

About Us

About TargetCareersNews & BlogTarget BrandsBullseye ShopSustainability & GovernancePress CenterAdvertise with UsInvestorsAffiliates & PartnersSuppliersTargetPlus

Help

Target HelpReturnsTrack OrdersRecallsContact UsFeedbackAccessibilitySecurity & FraudTeam Member Services

Stores

Find a StoreClinicPharmacyTarget OpticalMore In-Store Services

Services

Target Circle™Target Circle™ CardTarget Circle 360™Target AppRegistrySame Day DeliveryOrder PickupDrive UpFree 2-Day ShippingShipping & DeliveryMore Services
PinterestFacebookInstagramXYoutubeTiktokTermsCA Supply ChainPrivacyCA Privacy RightsYour Privacy ChoicesInterest Based AdsHealth Privacy Policy