DR-IRL. Inverse Reinforcement Learning with Dynamic Reward Scaling for LLM Alignment

github.com/Rosy0912/DR-IRL

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.