KRPO_LLMs_RL. The code repository for paper "Kalman Filter Enhanced Group Relative Policy Optimization for Language Model Reasoning"

github.com/billhhh/KRPO_LLMs_RL

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.