2 sessions
- Chalk talkAI/MLComputeCross-Industry SolutionsGenerative AICost OptimizationInnovation & Transformation300 – AdvancedAcademic / ResearcherData ScientistDeveloper / EngineerMGM GrandAmazon BedrockAmazon Elastic Compute Cloud (Amazon EC2)AWS Inferentia3:00 p.m. Monday, Dec 02Monday, Dec 02Learn how to accelerate the development and deployment of large language models (LLMs) with Ray, AWS Trainium, and AWS Inferentia. This chalk talk delves into how Ray’s unified compute framework seamlessly integrates with powerful AWS AI chips to optimize performance and cost efficiency. Learn practical techniques for scaling LLMs using tensor parallelism and harnessing the full potential of Trainium for training and Inferentia for inference.
, Principal SA - Accelerated Computing, AWS
, Principal Specialist, Accelerated Computing, AWS
- Monday, Dec 23:00 PM - 4:00 PM PSTMGM Grand | Level 3 | 302
- Chalk talkAI/MLComputeCross-Industry SolutionsGenerative AICost OptimizationInnovation & Transformation300 – AdvancedAcademic / ResearcherData ScientistDeveloper / EngineerCaesars ForumAmazon BedrockAmazon Elastic Compute Cloud (Amazon EC2)AWS Inferentia10:00 a.m. Wednesday, Dec 04Wednesday, Dec 04Learn how to accelerate the development and deployment of large language models (LLMs) with Ray, AWS Trainium, and AWS Inferentia. This chalk talk delves into how Ray’s unified compute framework seamlessly integrates with powerful AWS AI chips to optimize performance and cost efficiency. Learn practical techniques for scaling LLMs using tensor parallelism and harnessing the full potential of Trainium for training and Inferentia for inference.
, Principal SA - Accelerated Computing, AWS
, Principal Specialist, Accelerated Computing, AWS
- Wednesday, Dec 410:30 AM - 11:30 AM PSTCaesars Forum | Level 1 | Alliance 305