Policy Learning Workload (Ray Train) thumbnail

Policy Learning Workload (Ray Train)

Workload Example

A policy learning workload example utilizing Ray Train.

About this course

6 lessons
1 module
Tags
Workload ExamplesAdvancedModel TrainingWorkloadsLegacy Import
Level: Advanced
Reinforcement Learning Workloads
Policy Gradient Methods

Course Modules

0

Workload

Learn how to build and run an end-to-end diffusion-policy training workload for the `Pendulum-v1` control task using a real offline dataset, from data generation/preprocessing with Ray Data to distributed training on an Anyscale cluster with Ray Train V2. You’ll accomplish migrating a local PyTorch + Gymnasium workflow into a scalable, fault-tolerant Ray pipeline with minimal code changes.

  • Introduction
  • Imports and the Dataset
  • Diffusion Policy
+3 more lessons