Inicio > > Redes y comunicaciones informáticas > Kubernetes for Generative AI Solutions
Kubernetes for Generative AI Solutions

Kubernetes for Generative AI Solutions

Ashok Srirama / Sukirti Gupta

69,13 €
IVA incluido
Disponible
Editorial:
Packt Publishing
Año de edición:
2025
Materia
Redes y comunicaciones informáticas
ISBN:
9781836209935
69,13 €
IVA incluido
Disponible

Selecciona una librería:

  • Librería Samer Atenea
  • Librería Aciertas (Toledo)
  • Kálamo Books
  • Librería Perelló (Valencia)
  • Librería Elías (Asturias)
  • Donde los libros
  • Librería Kolima (Madrid)
  • Librería Proteo (Málaga)

Master the complete Generative AI project lifecycle on Kubernetes (K8s) from design and optimization to deployment using best practices, cost-effective strategies, and real-world examples.Key Features:- Build and deploy your first Generative AI workload on Kubernetes with confidence- Learn to optimize costly resources such as GPUs using fractional allocation, Spot Instances, and automation- Gain hands-on insights into observability, infrastructure automation, and scaling Generative AI workloads- Purchase of the print or Kindle book includes a free PDF eBookBook Description:Generative AI (GenAI) is revolutionizing industries, from chatbots to recommendation engines to content creation, but deploying these systems at scale poses significant challenges in infrastructure, scalability, security, and cost management.This book is your practical guide to designing, optimizing, and deploying GenAI workloads with Kubernetes (K8s) the leading container orchestration platform trusted by AI pioneers. Whether you’re working with large language models, transformer systems, or other GenAI applications, this book helps you confidently take projects from concept to production. You’ll get to grips with foundational concepts in machine learning and GenAI, understanding how to align projects with business goals and KPIs. From there, you’ll set up Kubernetes clusters in the cloud, deploy your first workload, and build a solid infrastructure. But your learning doesn’t stop at deployment. The chapters highlight essential strategies for scaling GenAI workloads in production, covering model optimization, workflow automation, scaling, GPU efficiency, observability, security, and resilience.By the end of this book, you’ll be fully equipped to confidently design and deploy scalable, secure, resilient, and cost-effective GenAI solutions on Kubernetes.What You Will Learn:- Explore GenAI deployment stack, agents, RAG, and model fine-tuning- Implement HPA, VPA, and Karpenter for efficient autoscaling- Optimize GPU usage with fractional allocation, MIG, and MPS setups- Reduce cloud costs and monitor spending with Kubecost tools- Secure GenAI workloads with RBAC, encryption, and service meshes- Monitor system health and performance using Prometheus and Grafana- Ensure high availability and disaster recovery for GenAI systems- Automate GenAI pipelines for continuous integration and deliveryWho this book is for:This book is for solutions architects, product managers, engineering leads, DevOps teams, GenAI developers, and AI engineers. It’s also suitable for students and academics learning about GenAI, Kubernetes, and cloud-native technologies. A basic understanding of cloud computing and AI concepts is needed, but no prior knowledge of Kubernetes is required.Table of Contents- GenAI-Intro, Evolution, and Project Lifecycle- K8s-Introduction and Integration with GenAI- Getting Started with K8s in the Cloud- GenAI Model Optimization for Domain-Specific Use Cases (RAG, Fine Tuning, etc.)- Getting Started with GenAI on K8s-Chatbot Example- Deploying GenAI on K8s-Scaling Best Practices- Deploying GenAI on K8s-Cost Optimization Best Practices- Deploying GenAI on K8s-Networking Best Practices- Deploying GenAI on K8s-Security Best Practices- Optimizing GPU Resources in K8s for GenAI Applications- GenAIOps: Creating GenAI Automation Pipeline- Getting Visibility into GenAI Workloads Resource Utilization- High Availability and Disaster Recovery Implementation- Wrap Up and Further Readings

Artículos relacionados

  • Next Generation Search Engines
    Recent technological progress in computer science, Web technologies, and the constantly evolving information available on the Internet has drastically changed the landscape of search and access to information. Current search engines employ advanced techniques involving machine learning, social networks, and semantic analysis. Next Generation Search Engines: Advanced Models for ...
    Disponible

    256,63 €

  • Collaboration and the Semantic Web
    Collaborative working has been increasingly viewed as a good practice for organizations to achieve efficiency. Organizations that work well in collaboration may have access to new sources of funding, deliver new, improved, and more integrated services, make savings on shared costs, and exchange knowledge, information and expertise. Collaboration and the Semantic Web: Social Net...
    Disponible

    229,92 €

  • Resource Allocation in Next-Generation Broadband Wireless Access Networks
    With the growing popularity of wireless networks in recent years, the need to increase network capacity and efficiency has become more prominent in society. This has led to the development and implementation of heterogeneous networks. Resource Allocation in Next-Generation Broadband Wireless Access Networks is a comprehensive reference source for the latest scholarly research o...
    Disponible

    249,42 €

  • Advanced Topics in Information Technology Standards and Standardization Research, Volume 1
    Kai Jakobs
    ...
    Disponible

    118,72 €

  • Data Warehouses and OLAP
    ...
    Disponible

    118,72 €

  • Selected Readings on Database Technologies and Applications
    Terry Halpin
    Education and research in the field of database technology can prove problematic without the proper resources and tools on the most relevant issues, trends, and advancements. Selected Readings on Database Technologies and Applications supplements course instruction and student research with quality chapters focused on key issues concerning the development, design, and analysis ...
    Disponible

    256,64 €