This is your work, valued
SophiaVL-R1. SophiaVL-R1: Reinforcing MLLMs Reasoning with Thinking Reward
Reagent. Agent-RRM: Exploring Reasoning Reward Model for Agents
R1-Collection. A collection of R1-based repos.