PaLM-rlhf-pytorch. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

github.com/zhiyuanhubj/PaLM-rlhf-pytorch

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.