ODBIERZ TWÓJ BONUS :: »

Kubernetes for Generative AI Solutions. A complete guide to designing, optimizing, and deploying Generative AI workloads on Kubernetes Ashok Srirama, Sukirti Gupta

Język publikacji: angielski
Kubernetes for Generative AI Solutions. A complete guide to designing, optimizing, and deploying Generative AI workloads on Kubernetes Ashok Srirama, Sukirti Gupta - okladka książki

Kubernetes for Generative AI Solutions. A complete guide to designing, optimizing, and deploying Generative AI workloads on Kubernetes Ashok Srirama, Sukirti Gupta - okladka książki

Autorzy:
Ashok Srirama, Sukirti Gupta
Serie wydawnicze:
Hands-on
Ocena:
Generative AI (GenAI) is revolutionizing industries, from chatbots to recommendation engines to content creation, but deploying these systems at scale poses significant challenges in infrastructure, scalability, security, and cost management.
This book is your practical guide to designing, optimizing, and deploying GenAI workloads with Kubernetes (K8s) the leading container orchestration platform trusted by AI pioneers. Whether you're working with large language models, transformer systems, or other GenAI applications, this book helps you confidently take projects from concept to production. You’ll get to grips with foundational concepts in machine learning and GenAI, understanding how to align projects with business goals and KPIs. From there, you'll set up Kubernetes clusters in the cloud, deploy your first workload, and build a solid infrastructure. But your learning doesn't stop at deployment. The chapters highlight essential strategies for scaling GenAI workloads in production, covering model optimization, workflow automation, scaling, GPU efficiency, observability, security, and resilience.
By the end of this book, you’ll be fully equipped to confidently design and deploy scalable, secure, resilient, and cost-effective GenAI solutions on Kubernetes.

Wybrane bestsellery

O autorach książki

Ashok Srirama is a Principal Specialist Solutions Architect at Amazon Web Services (AWS) with over 19 years of IT experience, specializing in cloud architecture, distributed systems, Kubernetes, and Generative AI. Recognized with the prestigious AWS Gold Jacket and Kubestronaut accreditation, he has authored numerous technical publications, presented at 25+ tech summits, and created AWS solutions for enterprise container deployments.
Sukirti Gupta has over 15 years of experience spanning Cloud Computing, Kubernetes, Generative AI, and Data Center Architecture. Sukirti currently leads go to market strategy for AWS (Amazon Web Services), supporting customers with their GenAI journey and has played pivotal roles at AWS, AMD, and Intel Corporation.

Zobacz pozostałe książki z serii Hands-on

Packt Publishing - inne książki

Zamknij

Przenieś na półkę
Dodano produkt na półkę
Usunięto produkt z półki
Przeniesiono produkt do archiwum
Przeniesiono produkt do biblioteki

Zamknij

Wybierz metodę płatności

Sposób płatności