Robotics Graduate Student at CMU.
Awesome-Flow-RL-Papers. A collection of paper/projects that trains flow matching model/policies via RL.
pi0-online-rl. Tonghe's personal implementation of fine-tuning LeRobot pi0 policy in ManiSkill3 with online RL.
vllm_serving. Serving VLM for robot agent.