RL-SaLLM-F. [AAMAS'25] Code for "Online Preference-based Reinforcement Learning with Self-augmented Feedback from Large Language Model"

github.com/TU2021/RL-SaLLM-F

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.