This is your work, valued
Investigating Intelligence. Primary focus on Deep Reinforcement Learning. PhD Student at University of Oxford.
transformers-metarl. Transformers are Meta-Reinforcement Learners - International Conference on Machine Learning (ICML) 2022
69BAL-PM. Deep Bayesian Active Learning for Preference Modeling in Large Language Models (NeurIPS 2024)
7humanoid-run-ppo. Code for the paper "Learning Humanoid Robot Running Skills through Proximal Policy Optimization"
5deep-rl-humanoid-motions-masters. Repository for code from the Master's Thesis "Imitation Learning and Meta-Reinforcement Learning for Optimizing Humanoid Robot Motions".
5ml-room. Implementations of machine learning algorithms from scratch.
3neural-networks-generate-lyrics. LSTM-based model for generate music lyrics
3stable-pg-llm. Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
3TD-VCL. Temporal-Difference Variational Continual Learning (NeurIPS 2025)
3deep-rl-undergrad-thesis. A Deep Reinforcement Learning Method for Humanoid Kick Motion - Bachelor's Thesis
2ab-test-RL. Using reinforcement learning for AB test
2kaggle-eef-house-prediction. Code from some models used in First Place's solution of Kaggle Data Science Challenge for EEF for ITA/Unifesp SJC students.
1