Open sourceActive

Unsloth Studio

Unsloth Studio is an open-source web UI for training and running models with no-code and observability.

Open source page

Product features

Product features

Basics/Tips

🏃 From RLHF, PPO to GRPO and RLVR

GRPO notebooks:

How GRPO Trains a Model

🤞 Luck (well Patience) Is All You Need

💡 Reinforcement Learning (RL) Guide

📋 Reward Functions / Verifiers

RL on unsupported models:

💻 Training with GRPO

❓ What is Reinforcement Learning (RL)?

🦥 What Unsloth offers for RL

🦥 What you will learn