This is your work, valued

OpenRLHF

Expert
@OpenRLHF

Open-sourced Reinforcment Learning from Human Feedback