Lead ML Research Scientist at RevEng.ai. Formerly at Rowden Technologies and before that a postdoc in the WMG data science group, focusing on deep RL research.
settlers_of_catan_RL. Learning to play Settlers of Catan with Deep RL - custom training environment and implementation of PPO
102big2_PPOalgorithm. Application of proximal policy optimization algorithm to the card game Big 2 using Tensorflow
84dexterous-gym. Challenging dexterous manipulation environments for RL that extend the hand manipulation environments introduced in OpenAI's Gym
55multi_action_head_PPO. PPO with multi-head/autoregressive action outputs
47SCDM. Solving Complex Dexterous Manipulation Tasks with Trajectory Optimisation and Reinforcement Learning
23PlanGAN. Python
18intrinsically_motivated_collective_motion. Main code for the collective motion model I've been working on during my PhD. Includes some apps made using D3.js to help describe the model.
3big2_server. Web-app to play against Big 2 agent trained with deep RL.
3MonteCarloTreeSearch. Somewhat General MCTS class. c++ standard template library practice.
1