This is your work, valued
Reinforcement Learning
pomdp-baselines. Simple (but often Strong) Baselines for POMDPs in PyTorch, ICML 2022
347OrganSegRSTN_PyTorch. PyTorch implementation of OrganSegRSTN - CVPR 2018
107Memory-RL. When Do Transformers Shine in RL? Decoupling Memory from Credit Assignment, NeurIPS 2023 (oral)
74Elastic-Boundary-Projection. Elastic Boundary Projection for 3D Medical Image Segmentation - CVPR 2019
65f-IRL. Inverse Reinforcement Learning via State Marginal Matching, CoRL 2020
45Meta-SAC. Auto-tune the Entropy Temperature of Soft Actor-Critic via Metagradient - 7th ICML AutoML workshop 2020
33self-predictive-rl. Bridging State and History Representations: Understanding Self-Predictive RL, ICLR 2024
27llm-reasoning-uft. Code for Offline Learning and Forgetting for Reasoning with Large Language Models, TMLR 2025
14neubay. JAX code for "Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism", ICML 2026
8Anonymous-Forum. PKU Network Practice
1