Tabular Data Workload Example (Ray Train) thumbnail

Tabular Data Workload Example (Ray Train)

Workload Example

A tabular data workload example built with Ray Train.

About this course

6 lessons
1 module
Tags
Workload ExamplesIntermediateModel TrainingData ProcessingWorkloadsLegacy Import
Level: Intermediate
Tabular Data Workloads
XGBoost and Tree-Based Models

Course Modules

0

Workload

In this Workload module, you’ll learn how to scale a tabular XGBoost forest-cover classification pipeline from local training to a distributed Ray Train V2 job on an Anyscale cluster. You’ll ingest the UCI Cover Type dataset, persist train/validation Parquet splits to shared storage, and train/evaluate the model using Ray Datasets and distributed execution.

  • Introduction
  • Imports and the Dataset
  • Distributed Training
+3 more lessons