trlx. A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)

github.com/mymusise/trlx

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.