rl-handbook. Code companion for the RL Post-Training Handbook - training reasoning models on a single GPU

github.com/Infatoshi/rl-handbook

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.