Rare find

A2PR. [ICML 2024] The offical implementation of A2PR, a simple way to achieve SOTA in offline reinforcement learning with an adaptive advantage-guided policy regularization method, in Pytorch

github.com/ltlhuuu/A2PR

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.