Rare find

r1_reward. ✨✨ [ICLR 2026] R1-Reward: Training Multimodal Reward Model Through Stable Reinforcement Learning

github.com/yfzhang114/r1_reward

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.