Rare find

DFT. [ICLR 2026] On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification.

github.com/yongliang-wu/DFT

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.