Generative AI on Kubernetes: Operationalizing Large Language Models - Marjane Mall - Image 1
248
00DH
496.00 DH
-50%

Generative AI on Kubernetes: Operationalizing Large Language Models

Livraison

Détails
Frais de livraison à partir de :
Livraison entre le Mercredi 9 septembre 2026 et le Jeudi 10 septembre 2026

À propos de cet article :

Marque : GENERIC
Vendu par HEAVENBOOKS.MA

Generative AI is revolutionizing industries, and Kubernetes has fast become the backbone for deploying and managing these resource-intensive workloads. This book serves as a practical, hands-on guide for MLOps engineers, software developers, Kubernetes administrators, and AI professionals ready to combine AI innovation with the power of cloud nativ...

1

Mode de paiement

Paiement par carte bancaire
Carte marocaines
Paiement à la livraison
Paiement en espèce à la livraison
Politique de retours
Note de politique de retour

Description produit

Marque
GENERIC
Titre
Generative AI on Kubernetes: Operationalizing Large Language Models
Éditeur
O'Reilly Media
Type de produit
paperback
Présentation du livre
paperback
Date de sortie
4/7/2026 12:00:00 AM
Langue d'origine
English
ISBN
1098171926
Nombre de pages
404 pages
Langue
English
Résumé
Generative AI is revolutionizing industries, and Kubernetes has fast become the backbone for deploying and managing these resource-intensive workloads. This book serves as a practical, hands-on guide for MLOps engineers, software developers, Kubernetes administrators, and AI professionals ready to combine AI innovation with the power of cloud native infrastructure. Authors Roland Huß and Daniele Zonca provide a clear road map for training, fine-tuning, deploying, and scaling GenAI models on Kubernetes, addressing challenges like resource optimization, automation, and security along the way.With actionable insights with real-world examples, readers will learn to tackle the opportunities and complexities of managing GenAI applications in production environments. Whether you're experimenting with large-scale language models or facing the nuances of AI deployment at scale, you'll uncover expertise you need to operationalize this exciting technology effectively.Learn how to deploy LLMs more efficiently with optimized inference runtimesGet hands-on with GPU scheduling, including hardware detection and multinode scalingMonitor and understand LLM-specific metrics like Time to First Token and token throughputKnow when to fine-tune a model or when retrieval augmentation is the better choiceDiscover how to evaluate models with standardized benchmarks before committing GPU resourcesLearn to run agentic applications with secure tool integration, identity management, and persistent state Read more
Auteur
Roland Huß, Daniele Zonca
Date de parution
4/7/2026 12:00:00 AM