Rare find

Building-Math-Agents-with-Multi-Turn-Iterative-Preference-Learning. This is an official implementation of the paper ``Building Math Agents with Multi-Turn Iterative Preference Learning'' with multi-turn DPO and KTO.

github.com/WeiXiongUST/Building-Math-Agents-with-Multi-Turn-Iterative-Preference-Learning

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.