Class Central is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

YouTube

Uber's GenAI Leap: Batch Predictions Using Ray and vLLM - Ray Summit 2024

Anyscale via YouTube

Overview

Explore Uber's innovative approach to large-scale Generative AI batch prediction in this Ray Summit 2024 presentation. Learn how Uber integrates Ray and vLLM within their Michelangelo machine learning platform to enhance GenAI application development. Discover how this new method addresses limitations in traditional Spark-based approaches, particularly for GPU-intensive tasks. Gain insights into the architecture of Uber's new system, its integration with Kubernetes and Michelangelo's LLM evaluation workflow, and its application to various Uber services. Understand the benchmarking results and lessons learned from developing and implementing this solution. Acquire valuable knowledge for scaling Generative AI capabilities, leveraging Ray and vLLM to improve prediction tasks, reduce latency, and enhance overall GenAI performance.

Syllabus

Uber's GenAI Leap: Batch Predictions Using Ray and vLLM | Ray Summit 2024

Taught by

Anyscale

Reviews

Start your review of Uber's GenAI Leap: Batch Predictions Using Ray and vLLM - Ray Summit 2024

Never Stop Learning.

Get personalized course recommendations, track subjects and courses with reminders, and more.

Someone learning on their laptop while sitting on the floor.