Online Model Serving with Ray Serve (Example) thumbnail

Online Model Serving with Ray Serve (Example)

Workload Example

An example of online model serving implemented with Ray Serve.

About this course

5 lessons
1 module
Tags
Workload ExamplesBeginnerModel ServingWorkloadsLegacy Import
Level: Beginner
Online Inference Services
FastAPI Integration

Course Modules

0

Workload

In this module, you’ll deploy a Hugging Face sentiment model as a scalable online inference service using Ray Serve with a FastAPI HTTP endpoint. You’ll learn how to run and scale the deployment with replicas, send test client requests, and properly shut down the Serve app and Ray cluster.

  • Overview of Ray Serve
  • Example Overview
  • Ray Serve Deployment
+2 more lessons