T5-RLHF-TF. Implementation of Reinforcement Learning from Human Feedback for Summarization Task in TensorFlow

github.com/ayulockin/T5-RLHF-TF

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.