PhD student working on safe RL in REALM lab @ MIT, advised by Prof. Chuchu Fan.
PPO-for-Beginners. A simple and well styled PPO implementation. Based on my Medium series: https://medium.com/@eyyu/coding-ppo-from-scratch-with-pytorch-part-1-4-613dfc1b14c8.
1.3kpocar. Code implementation for the NeurIPS 2022 paper "Policy Optimization with Advantage Regularization for Long-Term Fairness in Decision Systems".
9DeMAC. The Decentralized Multi-Agent Coordination (DeMAC) Framework. A lightweight tool designed to easily coordinate multiple agents with decentralized policies in a shared multi-agent environment.
5BitFit. BitFit was created as a project for the UCSD course CSE 110 "Software Engineering" taught by Professor Gary Gillespie during Spring Quarter 2020.
1Ramdroids_2019_2020. Java
1