1 sessions
- WorkshopAI/MLKubernetesGenerative AI400 – ExpertData EngineerData ScientistDeveloper / EngineerWynnAmazon Elastic Compute Cloud (Amazon EC2)Amazon Elastic InferenceAmazon Elastic Kubernetes Service (Amazon EKS)1:00 p.m. Wednesday, Dec 04Wednesday, Dec 04Some of the most innovative AWS customers are using Amazon EKS to train, fine-tune, and serve cutting-edge generative AI models. In this workshop, learn how to serve fine-tuned models with vLLM and Ray Serve for a high-performance solution and be cost effective with Karpenter for right-sized accelerated compute. Through real-world, hands-on examples, learn some key patterns that can help you implement LLMOps for scaling inference workloads, integrating RAG frameworks, and optimizing performance within the Kubernetes ecosystem. You must bring your laptop to participate.
, Sr. WW Spec. SA, Containers, Amazon Web Services, In.c
, Sr. Specialist SA, Containers, Amazon Web Services
- Wednesday, Dec 41:00 PM - 3:00 PM PSTWynn | Convention Promenade | Margaux 2