This is your work, valued
PhD. Student, Tsinghua-Berkeley Shenzhen Institute, Tsinghua University
MEEE. Code to reproduce the experiments in Sample Efficient Reinforcement Learning via Model-Ensemble Exploration and Exploitation (MEEE).
433Cross-Labeling-Supervision-for-Semi-Supervised-Learning. A PyTorch implementation of CLS
3i-Razor. Python
2tf2-mbpo. A tensorflow 2 implementation of the paper "When to Trust Your Model: Model-Based Policy Optimization"
1