MODAL
Powered by

Scaling Reinforcement Learning on Modal

Modal

34:31

Watch

Join us for a webinar with Joy Liu from the Modal Training Team. We’ll explore how engineers on Modal run inference, training workloads, and sandboxed environments across the entire RL pipeline. We'll dive into:

  • Taking advantage of multi-node training and Flash proxy for high-performance inference on Modal
  • A case study of RL fine-tuning of Qwen3-4B using LLM-as-a-judge reward signal
  • Scaffolding SLIME and other RL frameworks effectively on Modal

Scaling Reinforcement Learning on Modal

34:31

Watch