UT Austin ML/RL PhD student
deep_control. Deep Reinforcement Learning for Continuous Control in PyTorch
106super_sac. A general model-free off-policy actor-critic implementation. Continuous and Discrete Soft Actor-Critic with multimodal observations, data augmentation, offline learning and behavioral cloning.
42AdvancedPacmanDQNs. Extensions of Deep Q Learning
27dmc_remastered. A version of the DeepMind Control Suite with randomly generated graphics, for measuring visual generalization in continuous control.
9keras-rl. Deep Reinforcement Learning for Keras.
9cc-afbc. Advantage-Filtered Behavioral Cloning for Offline Continuous Control
4lunarDQfD. example of keras-rl dqfd implementation.
4articanon. Neural Text Generation
4supersonic. Multiworker PPO with random network distillation in eager execution Tensorflow
1