Rare find

R1-ShareVL. [NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward

github.com/HJYao00/R1-ShareVL

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.