This is your work, valued
I am a PhD student at Berkeley AI Research. My research interests focus on deep reinforcement learning for robotics.
dsrl_pi0. Official implementation for pi0 steering via DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)
★ 283Cal-QL. official implementation for our paper Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning (NeurIPS 2023)
★ 124V-GPS. official implementation for our paper Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance (CoRL 2024)
★ 55ecg-pytorch-sample. A PyTorch implementation for training deep learning models for 12-lead ECGs (2D-CNN, 1D-CNN, Transformer)
★ 13openpi. fork from https://github.com/Physical-Intelligence/openpi
★ 3dopamine. forked from https://github.com/google/dopamine
★ 3minibullet. Jupyter Notebook
★ 1pybullet_meshes. HTML
★ 1RL4VLA. Python
★ 279RoboBrain2.5. RoboBrain 2.5: Advanced version of RoboBrain. Depth in Sight, Time in Mind. 🎉🎉🎉
★ 1.1kLW-BenchHub. LW-BenchHub is a unified benchmark hub built on Isaac Lab–Arena for embodied AI, providing consistent interfaces, realistic environments, multi-robot support, and large-scale evaluation. It includes Lightwheel-Libero-Tasks and Lightwheel-RoboCasa-Tasks with 268 tasks, reproducible RL configs, and a full pipeline for benchmarking robot policies.
★ 181RLDX-1. Python
★ 316Isaac-GR00T. NVIDIA Isaac GR00T N1.7 - A Foundation Model for Generalist Robots.
★ 7.7krobocasa-gr1-tabletop-tasks. Simulation benchmarks of GR1 Tabletop Tasks for GR00T N1
★ 146Entropy-Mechanism-of-RL. The Entropy Mechanism of Reinforcement Learning for Large Language Model Reasoning.
★ 446rtk. CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
★ 74kVAGEN. World model reasoning RL for multi-turn VLM agents
★ 489EmbodiedBench. [ICML 2025 Oral] Official repo of EmbodiedBench, a comprehensive benchmark designed to evaluate MLLMs as embodied agents.
★ 322dsrl_pi0. Official implementation for pi0 steering via DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)
★ 283dsrl. Official implementation for DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)
★ 208aloha_sim. A collection of tabletop tasks in Mujoco
★ 316horizon-reduction. The official implementation of "Horizon Reduction Makes RL Scalable"
★ 200AgiBot-World. [IROS 2025 Best Paper Award Finalist & IEEE TRO 2026] The Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems
★ 3.1kgym-aloha. A gym environment for ALOHA
★ 216manipulator_gym. Gym-like environment to interact with manipulator robots
★ 43mujoco_playground. An open-source library for GPU-accelerated robot learning and sim-to-real transfer.
★ 2.1kV-GPS. official implementation for our paper Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance (CoRL 2024)
★ 56openpi. Python
★ 13kbigvision-palivla. Python
★ 15tpu_pod_commander. TPU pod commander is a package for managing and launching jobs on Google Cloud TPU pods.
★ 20hato. 🕊️ HATO: Learning Visuotactile Skills with Two Multifingered Hands [ICRA 2025]
★ 171jax-smi. JAX Synergistic Memory Inspector
★ 186scalax. A simple library for scaling up JAX programs
★ 148AI-Scientist. The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery 🧑🔬
★ 14kAI-Researcher. Python
★ 397mujoco_menagerie. A collection of high-quality models for the MuJoCo physics engine, curated by Google DeepMind.
★ 3.8kact. Python
★ 2.1kLatte. [TMLR 2025] Latte: Latent Diffusion Transformer for Video Generation.
★ 1.9kIDQL. Repo for Implicit Diffusion Q-Learning
★ 126SEINE. [ICLR 2024] SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction
★ 967LaVie. [IJCV 2024] LaVie: High-Quality Video Generation with Cascaded Latent Diffusion Models
★ 952generative-models. Generative Models by Stability AI
★ 27kML-from-scratch-seminar. This repository is part of a "Machine Learning from Scratch" seminar at Harvard Medical School.
★ 311gpt_paper_assistant. GPT4 based personalized ArXiv paper assistant bot
★ 547Real-Time-Latent-Consistency-Model. App showcasing multiple real-time diffusion models pipelines with Diffusers
★ 916susie. Code for subgoal synthesis via image editing
★ 160mind-reader. [NeurIPS2022] Mind Reader: Reconstructing complex images from brain activities
★ 63dynalang. Code for "Learning to Model the World with Language." ICML 2024 Oral.
★ 421VideoCrafter. VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
★ 5.1kAnimateDiff. Official implementation of AnimateDiff.
★ 12kDREAMPlace. Deep learning toolkit-enabled VLSI placement
★ 1kcircuit_training. Python
★ 1.7kcva6. The CORE-V CVA6 is a highly configurable, 6-stage RISC-V core for both application and embedded applications. Application class configurations are capable of booting Linux.
★ 3kmaskplace. [NeurIPS 2022 Spotlight] MaskPlace: Fast Chip Placement via Reinforced Visual Representation Learning
★ 71x-magical. [CoRL 2021] A robotics benchmark for cross-embodiment imitation.
★ 60VideoLDM. Unofficial PyTorch implementation of the VideoLDM.
★ 165awesome_cs-ja_phd_life. collection of articles about PhD life written in 🇯🇵
★ 342lang-segment-anything. SAM with text prompt
★ 2.6ksegment-anything. The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 55kGrounded-Segment-Anything. Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
★ 18kAwesome-Multimodal-Large-Language-Models. :sparkles::sparkles:Latest Advances on Multimodal Large Language Models
★ 18kbridge_data_v2. Python
★ 287Awesome-Out-Of-Distribution-Detection. Out-of-distribution detection, robustness, and generalization resources. The repository contains a curated list of papers, tutorials, books, videos, articles and open-source libraries etc
★ 1kCoDeF. [CVPR'24 Highlight] Official PyTorch implementation of CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
★ 4.8krecurrent-interface-network-pytorch. Implementation of Recurrent Interface Network (RIN), for highly efficient generation of images and video without cascading networks, in Pytorch
★ 210instruction-tuned-sd. Code for instruction-tuning Stable Diffusion.
★ 250instruct-pix2pix. Python
★ 6.9kLLM-groundedDiffusion. LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models (LLM-grounded Diffusion: LMD, TMLR 2024)
★ 483make-a-stable-diffusion-video. 🤗 Diffusers: State-of-the-art diffusion models for image and audio generation in PyTorch - fork with video pseudo3d
★ 99Text2Video-Zero. [ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators
★ 4.2kframe-interpolation. FILM: Frame Interpolation for Large Motion, In ECCV 2022.
★ 3.1kHIQL. HIQL: Offline Goal-Conditioned RL with Latent States as Actions (NeurIPS 2023)
★ 98genaug. main augmentation script for real world robot dataset.
★ 40Cal-QL. official implementation for our paper Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning (NeurIPS 2023)
★ 124furniture-bench. FurnitureBench: Real-World Furniture Assembly Benchmark (RSS 2023)
★ 236StructDiffusion. StructDiffusion: Language-Guided Creation of Physically-Valid Structures using Unseen Objects
★ 59Text-To-Video-Finetuning. Finetune ModelScope's Text To Video model using Diffusers 🧨
★ 700modelscope. ModelScope: bring the notion of Model-as-a-Service to life.
★ 9.1kcog-text2video. Python
★ 62roboverse. A set of environments utilizing pybullet for simulation of robotic manipulation tasks.
★ 3cliport. CLIPort: What and Where Pathways for Robotic Manipulation
★ 547PVDM. [CVPR'23] Video Probabilistic Diffusion Models in Projected Latent Space
★ 322Tune-A-Video. [ICCV 2023] Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation
★ 4.4kmake-a-video-pytorch. Implementation of Make-A-Video, new SOTA text to video generator from Meta AI, in Pytorch
★ 2knlp-phd-global-equality. A repo for open resources & information for people to succeed in PhD in CS & career in AI / NLP
★ 1.1kSynthER. Synthetic Experience Replay
★ 114Everything-LLMs-And-Robotics. The world's largest GitHub Repository for LLMs + Robotics
★ 849awesome-offline-rl. An index of algorithms for offline reinforcement learning (offline-rl)
★ 1.1kopen_llama. OpenLLaMA, a permissively licensed open source reproduction of Meta AI’s LLaMA 7B trained on the RedPajama dataset
★ 7.5kcommunity-events. Place where folks can contribute to 🤗 community events
★ 427Personalize-SAM. Personalize Segment Anything Model (SAM) with 1 shot in 10 seconds
★ 1.7kIF. Python
★ 7.8kmakeavid-sd-tpu. Make-A-Video Latent Diffusion Model
★ 19Awesome-Video-Diffusion. A curated list of recent diffusion models for video generation, editing, and various other applications.
★ 5.7kEasyLM. Large language models (LLMs) made easy, EasyLM is a one stop solution for pre-training, finetuning, evaluating and serving LLMs in JAX/Flax.
★ 2.5kmcvd-pytorch. Official implementation of MCVD: Masked Conditional Video Diffusion for Prediction, Generation, and Interpolation (https://arxiv.org/abs/2205.09853)
★ 370Segment-Everything-Everywhere-All-At-Once. [NeurIPS 2023] Official implementation of the paper "Segment Everything Everywhere All at Once"
★ 4.8km3ae_public. Multimodal Masked Autoencoders (M3AE): A JAX/Flax Implementation
★ 110consistency_models. Official repo for consistency models.
★ 6.5keinops. Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)
★ 9.6kfew-shot-diffusion-models. Few-Shot Diffusion Models
★ 114FiLM-pytorch. PyTorch implementation of FiLM: Visual Reasoning with a General Conditioning Layer
★ 70film. FiLM: Visual Reasoning with a General Conditioning Layer
★ 461diffusion-models-class. Materials for the Hugging Face Diffusion Models Course
★ 4.3kvdm. Jupyter Notebook
★ 331imagen-pytorch. Implementation of Imagen, Google's Text-to-Image Neural Network, in Pytorch
★ 8.4kdiffusers. 🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34kv-diffusion-jax. v objective diffusion inference code for JAX.
★ 218ControlNet. Let us control diffusion models!
★ 34kDiT. Official PyTorch Implementation of "Scalable Diffusion Models with Transformers"
★ 8.7kdenoising-diffusion-flax. Implementing the Denoising Diffusion Probabilistic Model in Flax
★ 161CQL. Conservative Q Learning on top of SAC
★ 142online-dt. Online Decision Transformer
★ 276Off2OnRL. Python
★ 61Transformers-Tutorials. This repository contains demos I made with the Transformers library by HuggingFace.
★ 12kdiffusion_policy. [RSS 2023] Diffusion Policy Visuomotor Policy Learning via Action Diffusion
★ 4.4kwalk-these-ways. Sim-to-real RL training and deployment tools for the Unitree Go1 robot.
★ 1.4kD4RL. A collection of reference environments for offline reinforcement learning
★ 1walk_in_the_park. Python
★ 284Dropout-Q-Functions-for-Doubly-Efficient-Reinforcement-Learning. Source files to replicate experiments in my ICLR 2022 paper.
★ 75rlpd. Python
★ 412brc_cql_example. Python
★ 3tuning_playbook. A playbook for systematically maximizing the performance of deep learning models.
★ 30kstatlearn-s23. Course materials for Advanced Topics in Statistical Learning, Spring 2023
★ 51implicit_q_learning. Python
★ 332batch_rl. Offline Reinforcement Learning (aka Batch Reinforcement Learning) on Atari 2600 games
★ 1reincarnating_rl. [NeurIPS 2022] Open source code for reusing prior computational work in RL.
★ 100dopamine. Dopamine is a research framework for fast prototyping of reinforcement learning algorithms.
★ 11kmj_envs. A collection of MuJoCo based environments.
★ 19hand_dapg. Repository to accompany RSS 2018 paper on dexterous hand manipulation
★ 323interbotix_ros_toolboxes. Support-level ROS Packages for Interbotix Robots
★ 45JaxCQL. Conservative Q learning in Jax
★ 58pytorch_sac. PyTorch implementation of Soft Actor-Critic (SAC)
★ 600roboverse. A set of environments utilizing pybullet for simulation of robotic manipulation tasks.
★ 29minimal-isaac-gym. A Minimal Example of Isaac Gym with DQN and PPO.
★ 112trax. Trax — Deep Learning with Clear Code and Speed
★ 8.3kcleanrl. High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)
★ 10kjaxrl. JAX (Flax) implementation of algorithms for Deep Reinforcement Learning with continuous action spaces.
★ 757D4RL-Evaluations. Python
★ 203D4RL. A collection of reference environments for offline reinforcement learning
★ 1.7kBPref. Official codebase for "B-Pref: Benchmarking Preference-BasedReinforcement Learning" contains scripts to reproduce experiments.
★ 136min-decision-transformer. Minimal implementation of Decision Transformer: Reinforcement Learning via Sequence Modeling in PyTorch for mujoco control tasks in OpenAI gym
★ 294transformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163kdiffuser. Code for the paper "Planning with Diffusion for Flexible Behavior Synthesis"
★ 1.3kdiayn-sac. DIYAN implementation based on SAC (PyTorch Ver.)
★ 6sac. Soft Actor-Critic
★ 1.3kurl_benchmark. Python
★ 368amorpheus. My Body Is A Cage
★ 41abci_tutorial. Python
★ 10metamorph. Code for "MetaMorph: Learning Universal Controllers with Transformers", Gupta et al, ICLR 2022
★ 131Gymnasium-Robotics. A collection of robotics simulation environments for reinforcement learning
★ 950brax. Massively parallel rigidbody physics simulation on accelerator hardware.
★ 3.2kminimalRL. Implementations of basic RL algorithms with minimal lines of codes! (pytorch based)
★ 3.2kpytorch-rl. This repository contains model-free deep reinforcement learning algorithms implemented in Pytorch
★ 453alf. Agent Learning Framework https://alf.readthedocs.io
★ 365phd-master-application-docs. A collection of the application documents I used to apply to universities in the US.
★ 489jonbarron.github.io. HTML
★ 3.6kSSL_ECG. Python
★ 5moco. PyTorch implementation of MoCo: https://arxiv.org/abs/1911.05722
★ 5.1kecg-selfsupervised. Self-supervised representation learning from 12-lead ECG data
★ 97Minigrid. Simple and easily configurable grid world environments for reinforcement learning
★ 2.5kswarm_learning. Scripts for figures and calculations of the manuscript by Warnat-Herresthal el al. 2020
★ 197rlkit. Collection of reinforcement learning algorithms
★ 2.9klifelong_rl. Pytorch implementations of RL algorithms, focusing on model-based, lifelong, reset-free, and offline algorithms. Official codebase for Reset-Free Lifelong Learning with Skill-Space Planning.
★ 110Deep-Reinforcement-Learning-Algorithms-with-PyTorch. PyTorch implementations of deep reinforcement learning algorithms and environments
★ 5.9kreward_machines. Python
★ 77Cartoon-StyleGAN. Fine-tuning StyleGAN2 for Cartoon Face Generation
★ 656dads. Code for 'Dynamics-Aware Unsupervised Discovery of Skills' (DADS). Enables skill discovery without supervision, which can be combined with model-based control.
★ 201SSL-FEW-SHOT. SSL-FEW-SHOT
★ 171music. Mutual Information State Intrinsic Control (ICLR 2021 Spotlight)
★ 39learn2learn. A PyTorch Library for Meta-learning Research
★ 2.9kcactus-protonets. Code for Unsupervised Learning via Meta-Learning.
★ 67UMTRA-Release. Unsupervised Meta Learning for Image Classification (UMTRA) algorithm
★ 21lwm. Latent World Models For Intrinsically Motivated Exploration | Official repository
★ 23pfrl. PFRL: a PyTorch-based deep reinforcement learning library
★ 1.3krlrd. PyTorch implementation of our paper Reinforcement Learning with Random Delays (ICLR 2020)
★ 44ViT-pytorch. Pytorch reimplementation of the Vision Transformer (An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale)
★ 2.2kprml. TeX
★ 378Transformer-MM-Explainability. [ICCV 2021- Oral] Official PyTorch implementation for Generic Attention-model Explainability for Interpreting Bi-Modal and Encoder-Decoder Transformers, a novel method to visualize any Transformer-based network. Including examples for DETR, VQA.
★ 911pytorchvideo. A deep learning library for video understanding research.
★ 3.6kattention-transfer. Improving Convolutional Networks via Attention Transfer (ICLR 2017)
★ 1.5kWhite-box-Cartoonization. Official tensorflow implementation for CVPR2020 paper “Learning to Cartoonize Using White-box Cartoon Representations”
★ 4ktoonify. Jupyter Notebook
★ 617AIMedical2021-2nd. Python
★ 11pytorch-wavenet. An implementation of WaveNet with fast generation
★ 1kresnet1d. PyTorch implementations of several SOTA backbone deep neural networks (such as ResNet, ResNeXt, RegNet) on one-dimensional (1D) signal/time-series data.
★ 546react-media-recorder. react-media-recorder is a react component with render prop that can be used to record audio/video streams using MediaRecorder API.
★ 588Recorderjs. A plugin for recording/exporting the output of Web Audio API nodes
★ 4.2kpretrained-models.pytorch. Pretrained ConvNets for pytorch: NASNet, ResNeXt, ResNet, InceptionV4, InceptionResnetV2, Xception, DPN, etc.
★ 9.1kTimeSformer-pytorch. Implementation of TimeSformer from Facebook AI, a pure attention-based solution for video classification
★ 729FavoritePapers.
★ 460eeg-gcnn. Resources for the paper titled "EEG-GCNN: Augmenting Electroencephalogram-based Neurological Disease Diagnosis using a Domain-guided Graph Convolutional Neural Network". Accepted for publication (with an oral spotlight!) at ML4H Workshop, NeurIPS 2020.
★ 184Belief-Propagation. Overview and implementation of Belief Propagation and Loopy Belief Propagation algorithms: sum-product, max-product, max-sum
★ 182albumentations. Fast and flexible image augmentation library. Paper about the library: https://www.mdpi.com/2078-2489/11/2/125
★ 15kanime-face-detector. A Faster-RCNN based anime face detector implementation using tensorflow.
★ 252lbpcascade_animeface. A Face detector for anime/manga using OpenCV
★ 2ksyncnet_trainer. Disentangled Speech Embeddings using Cross-Modal Self-Supervision
★ 167CartoonGAN-Test-Pytorch-Torch. Pytorch and Torch testing code of CartoonGAN [Chen et al., CVPR18]
★ 690CartoonGan-tensorflow. Generate your own cartoon-style images with CartoonGAN (CVPR 2018), powered by TensorFlow 2.0 Alpha.
★ 936face-alignment. :fire: 2D and 3D Face alignment library build using pytorch
★ 7.5kWav2Lip. This repository contains the codes of "A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild", published at ACM Multimedia 2020. For HD commercial model, please try out Sync Labs
★ 13k3D-Grad-CAM. This repo contains Grad-CAM for 3D volumes.
★ 84tmc_wrs_docker. Shell
★ 65Deep-Reinforcement-Learning-Book. 書籍「つくりながら学ぶ!深層強化学習」のサポートリポジトリです
★ 3573D-ResNets-PyTorch. 3D ResNets for Action Recognition (CVPR 2018)
★ 4kml-engineer-roadmap. WIP: Roadmap to becoming a machine learning engineer in 2020
★ 2.2ksystem-design-primer. Learn how to design large-scale systems. Prep for the system design interview. Includes Anki flashcards.
★ 360kcoding-interview-university. A complete computer science study plan to become a software engineer.
★ 357k