This is your work, valued
Awesome-RL-VLA. A Survey on Reinforcement Learning of Vision-Language-Action Models for Robotic Manipulation
★ 810SafeBimanual. [CoRL2025]: SafeBimanual:Diffusion-based Trajectory Optimization for Safe Bimanual Manipulation
★ 7Bimanual_ur5e_action_control_for_IL. Python
★ 3Nora_lerobot. Python
★ 1EE6222-Machine-Vision-Assignment2. NTU-EEE-Msc CCA EE6222 Assignment2: Human-Action-Recognition-In-The-Dark-master
★ 1SoftVTBench. visuotactile benchmark for the deformable object manipulation
★ 48Dream-Tac. Official Code for Dream-Tac
★ 34UniHand. [TPAMI 2026] Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views (with Visual Imitation Learning for Robots)
★ 47LIBERO-plus. Official repository of LIBERO-plus, a generalized benchmark for in-depth robustness analysis of vision-language-action models.
★ 400N0-Foundation. 𝒩0-Foundation: Towards the Age of Tactile Intelligence
★ 44N0-TWAM. N₀-TWAM: A Tactile-Native World Action Model for Contact-Rich Manipulation
★ 45N0-VTLA. N0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens
★ 27E2HiL-project-a1x. Python
★ 4Hy-Embodied-0.5-VLA. From Vision-Language-Action Models to a Real-World Robot Learning Stack
★ 239Xiaomi-Robotics-1. Code for Xiaomi-Robotics-1
★ 278rlt-openpi. Open-source implementation of "RL Token: Bootstrapping Online RL with Vision-Language-Action Models"
★ 77StarWAM. A Generalizable Codebase for World-Action Models
★ 197HY-World-2.0. HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds
★ 2.5kDiffusionNFT. [ICLR 2026 Oral] DiffusionNFT: Online Diffusion Reinforcement with Forward Process
★ 999articraft. Superseded by https://github.com/articraftresearch/Articraft. This repository is kept for reference.
★ 1.4knewton. An open-source, GPU-accelerated physics simulation engine built upon NVIDIA Warp, specifically targeting roboticists and simulation researchers.
★ 5.3kDiT4DiT. This is the official code repo for DiT4DiT, a Vision-Action-Model (VAM) framework that combines video generation model with flow-matching-based action prediction for generalizable robotic manipulation.
★ 418iMaC. iMac: Translating Actions into Motion and Contact Images for Embodied World Models
★ 33Awesome-Progress-Reward-Model.
★ 27Awesome-Video-Action-Model.
★ 74RL-100. Implementation of RL-100, Performant Robotic Manipulation with Real-World Reinforcement Learning
★ 66Awesome-WAM. A curated, continuously updated reading list, paper blogs, and resources for World Action Models (WAMs) in embodied AI.
★ 1.2kVLA-JEPA. [ECCV 2026] VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
★ 514RISE. [RSS 2026] Code for RISE: Self-Improving Robot Policy with Compositional World Model
★ 330Awesome-World-Models. A comprehensive list of papers for the definition of World Models and using World Models for General Video Generation, Embodied AI, and Autonomous Driving, including papers, codes, and related websites.
★ 1.9kfaster. Python
★ 30vla_foundry. Python
★ 420awesome-embodied-vla-va-vln. A curated list of state-of-the-art research in embodied AI, focusing on vision-language-action (VLA) models, vision-language navigation (VLN), and related multimodal learning approaches.
★ 3.4kBeing-H. Being-H is BeingBeyond's family of human-centric embodied foundation models.
★ 1.1kcap-x. A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation
★ 683FastWAM. Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?
★ 1.2kABot-PhysWorld. Python
★ 370Awesome-RL-for-LRMs. A Survey of Reinforcement Learning for Large Reasoning Models
★ 2.5kEmbodied-AI-Paper-TopConf. [Actively Maintained🔥] A list of Embodied AI papers accepted by top conferences (ICLR, NeurIPS, ICML, RSS, CoRL, ICRA, IROS, CVPR, ICCV, ECCV).
★ 729Awesome-RL-for-Video-Generation. A curated list of papers on reinforcement learning for video generation
★ 575dsrl_pi0. Official implementation for pi0 steering via DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)
★ 283ULIP. Python
★ 611robometer. Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
★ 236Genie-Envisioner-V1. Python
★ 565robocasa. RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots
★ 1.6kEvo-RL. We release Evo-RL, the opensource real-world offline RL on So-101 and AgileX PiPER for easier reproduction.
★ 734VLAExplain. VLA model interpretability tools
★ 177pi-StepNFT. Python
★ 58dev-browser. A Claude Skill to give your agent the ability to use a web browser
★ 6.5kcover-vla. This is the official codebase for paper: Scaling Verification Can Be More Effective than Scaling Policy Learning for Vision-Language-Action Alignment
★ 60agentskills. Specification and documentation for Agent Skills
★ 24kmiles. Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.
★ 1.8kmicrogpt-rust. An atomic way to train and inference a GPT in Rust and Go. Inspired by @karpathy.
★ 5AnyTouch2. [ICLR 2026] AnyTouch 2: General Optical Tactile Representation Learning For Dynamic Tactile Perception
★ 54code. VLS: Steering Pretrained Robot Policies via Vision–Language Models
★ 66lingbot-va. [RSS 2026] Causal video-action world model for generalist robot control
★ 1.7kworld-gymnast. World-Gymnast: Training Robots with Reinforcement Learning in a World Model
★ 47GreenVLA. Green-VLA: Staged Vision-Language-Action Model for Generalist Robots
★ 137dmpo-release. DMPO: Diffusion Model Policy Optimization
★ 64vla-scratch. Python
★ 348starVLA. StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing
★ 3.4kAwesome-Touch. Tactile Sensing • Data Collection • IL/RL/VLA/WM • Manipulation • Simulation • Open Source
★ 751DiffTactile. [ICLR 2024] DiffTactile: A Physics-based Differentiable Tactile Simulator for Contact-rich Robotic Manipulation
★ 315agent-lightning. The absolute trainer to light up AI agents.
★ 17kvlagents. VLA policy server for simulation evaluation and real world rollout/deployment
★ 21every-embodied. 仅需Python基础,从0构建自己的具身智能机器人;从0逐步构建VLA/OpenVLA/SmolVLA/Pi0, 深入理解具身智能
★ 2.9kbridger. Python
★ 20octopi. Python
★ 77VLA-Touch. Implementation of RA-L (2026) paper: VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback
★ 91DeepTutor. DeepTutor: Lifelong Personalized Tutoring. https://deeptutor.info/.
★ 32kAwesome-Force-Tactile-VLA. A paper list of multimodal VLAs
★ 119Dream-VLX. Dream-VL and Dream-VLA, a diffusion VLM and a diffusion VLA.
★ 114RoboCOIN. RoboCoin + LeRobot integration
★ 207VisualGRPO. E-GRPO: High Entropy Steps Drive Effective Reinforcement Learning for Flow Models
★ 45learning_thousand_tasks. This repository contains the implementation of all methods evaluated in the paper "Learning a Thousand Tasks in a Day". We provide model architectures, training scripts, and deployment examples.
★ 147World-In-Your-Hands. Python
★ 126TwinAligner. [arxiv 2025] TwinAligner: Visual-Dynamic Alignment Empowers Physics-aware Real2Sim2Real for Robotic Manipulation
★ 74Real2Edit2Real. [CVPR 2026] Real2Edit2Real: Generating Robotic Demonstrations via a 3D Control Interface
★ 83REALM. [IEEE RA-L 2026] REALM: A Real-to-Sim Validated Benchmark for Generalization in Robotic Manipulation
★ 64Xbotics-Embodied-Guide. Xbotics 社区具身智能学习指南:我们把“具身综述→学习路线→仿真学习→开源实物→人物访谈→公司图谱”串起来,帮助新手和实战者快速定位路径、落地项目与参与开源。
★ 1.2kMIND-V. Python
★ 37agentlace. Connect agent policies for distributed ML applications
★ 86MM-ACT. [CVPR'2026] "MM-ACT: Learn from Multimodal Parallel Generation to Act"
★ 117Awesome-World-Models. A Curated List of Awesome Works in World Modeling, Aiming to Serve as a One-stop Resource for Researchers, Practitioners, and Enthusiasts Interested in World Modeling.
★ 3.3ktrl. Train transformer language models with reinforcement learning.
★ 19kMiMo-Embodied. MiMo-Embodied
★ 399Depth-Anything-3. Depth Anything 3
★ 6ksiiRL. siiRL: Shanghai Innovation Institute RL Framework for Advanced LLMs and Multi-Agent Systems
★ 368sam3. The repository provides code for running inference and finetuning with the Meta Segment Anything Model 3 (SAM 3), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 11ksam-3d-objects. SAM 3D Objects
★ 7.2kinverse-mixed-strategy. (ICRA 2025) Inverse Mixed Strategy Games with Generative Trajectory Models
★ 18DeepThinkVLA. DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
★ 527nora-1.5. NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards
★ 110G0. Galaxea's VLA
★ 3hiors. Human-in-the-loop Online Rejection Sampling for Robotic Manipulation
★ 27MoMaGen. Python
★ 67X-VLA. [ICLR 2026] The offical Implementation of "Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model"
★ 699Cal-QL. official implementation for our paper Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning (NeurIPS 2023)
★ 124Awesome-RL-VLA. A Survey on Reinforcement Learning of Vision-Language-Action Models for Robotic Manipulation
★ 812Ctrl-World. ICLR 2026 Paper: Ctrl-World
★ 538SimPublisher. A cross-environment tool to publish objects from simulation for Augmented Reality and Human Robot Interaction.
★ 43Taccel. Taccel: Scaling-up Vision-based Tactile Robotics with High-performance GPU Simulation
★ 197RDT2. Official code of RDT 2
★ 798SoM. [arXiv 2023] Set-of-Mark Prompting for GPT-4V and LMMs
★ 1.6kdppo. Official implementation of Diffusion Policy Policy Optimization, arxiv 2024
★ 842VLA-Adapter. VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model
★ 2.3kVLAC. VLAC: A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
★ 322DriveMoE. [CVPR 2026] Drive-π0 and DriveMoE on End-to-end Autonomous Driving
★ 230vjepa2. PyTorch code and models for VJEPA2 self-supervised learning from video.
★ 4.4kLiveCodeBench. Official repository for the paper "LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code"
★ 919lerobot-sim2real. LeRobot sim2real code. Train in fast simulation and deploy visual policies zero shot to the real world
★ 389BEHAVIOR-1K. BEHAVIOR-1K: a platform for accelerating Embodied AI research. Join our Discord for support: https://discord.gg/bccR5vGFEx
★ 1.6kdetectron2. Detectron2 is a platform for object detection, segmentation and other visual recognition tasks.
★ 35kGalaxeaDP. Galaxea's first diffusion policy release
★ 41RLinf. RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI
★ 4.4kInternManip. An All-in-one robot manipulation learning suite for policy models training and evaluation on various datasets and benchmarks.
★ 175Dex1B. Implementation of Dex1B: Learning with 1B Demonstrations for Dexterous Manipulation, from Ye et al of UCSD
★ 26H_RDT. H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation
★ 156flowmatching_policy_rl. Python
★ 23qc. Python
★ 397EmbodieDreamer.
★ 33Tool_as_Interface. Python
★ 55mast3r. Grounding Image Matching in 3D with MASt3R
★ 3.1karp. Autoregressive Policy for Robot Learning (RA-L 2025)
★ 154E2FGVI. Official code for "Towards An End-to-End Framework for Flow-Guided Video Inpainting" (CVPR2022)
★ 1.2krovi-aug. Augment robotics demonstration datasets with different robots and viewpoints
★ 42Entropy-Mechanism-of-RL. The Entropy Mechanism of Reinforcement Learning for Large Language Model Reasoning.
★ 446roboengine. Official Reporsitory of "RoboEngine: Plug-and-Play Robot Data Augmentation with Semantic Robot Segmentation and Background Generation"
★ 165dex-retargeting. Various retargeting optimizers to translate human hand motion to robot hand motion.
★ 1.2kViTaL. Accompanying codebase for paper"Touch begins where vision ends: Generalizable policies for contact-rich manipulation"
★ 110Awesome-Image-Harmonization. A curated list of papers, code and resources pertaining to image harmonization.
★ 536COCOCO. Video-Inpaint-Anything: This is the inference code for our paper CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility.
★ 325robot-colosseum. A Benchmark for Evaluating Generalization for Robotic Manipulation
★ 151YOTO. [TPAMI2026 / RSS2025] Code for my paper "You Only Teach Once: Learn One-Shot Bimanual Robotic Manipulation from Video Demonstrations"
★ 147Inpaint-Anything. Inpaint anything using Segment Anything and inpainting models.
★ 7.7kVideo-Depth-Anything. [CVPR 2025 Highlight] Video Depth Anything: Consistent Depth Estimation for Super-Long Videos
★ 2knora. NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks
★ 222MinerU. Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
★ 76kManiGaussian_Bimanual. [IROS 2025] ManiGaussian++: General Robotic Bimanual Manipulation with Hierarchical Gaussian World Model
★ 46Isaac-GR00T. NVIDIA Isaac GR00T N1.7 - A Foundation Model for Generalist Robots.
★ 7.7kEmbodied-AI-Guide. [Lumina具身智能社区] 具身智能技术指南 Embodied-AI-Guide
★ 15kspoc-robot-training. SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World
★ 156vggt. [CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
★ 14kPredDRL. Python
★ 3gym-hil. Human in the loop Reinforcement Learning suite
★ 230VLA-Diffusion-Policy-Robotics. Awesome collection of resources and papers on Diffusion Models for Robotic Manipulation.
★ 817Awesome-VLA-RL. This repository summarizes recent advances in the VLA + RL paradigm and provides a taxonomic classification of relevant works.
★ 427PRIME. Scalable RL solution for advanced reasoning of language models
★ 1.9kflow_grpo. [NeurIPS 2025] An official implementation of Flow-GRPO: Training Flow Matching Models via Online RL
★ 2.4kReinFlow. [NeurIPS 2025] Flow x RL. "ReinFlow: Fine-tuning Flow Policy with Online Reinforcement Learning". Support VLAs e.g., Pi0, Pi0.5, GR00TN1.5. Fully open-sourced.
★ 351hume. 🦾 A Dual-System VLA with System2 Thinking
★ 148meshcat. Remotely-controllable 3D viewer, built on top of three.js
★ 336RL4VLA. Python
★ 279boring.notch. TheBoringNotch: Not so boring notch That Rocks 🎸🎶
★ 10khil-serl-sim. This warehouse has been improved based on hil-serl, adding simulation training functionality
★ 84verl. verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
★ 23kSimpleVLA-RL. [ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
★ 1.8kAutoBio. Preliminary version of AutoBio (https://arxiv.org/abs/2505.14030)
★ 87RoboMani-PaperList. A Paper List for Robotic Manipulation.
★ 31conrft. This is the official implementation of the paper "ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy".
★ 363PARL. A high-performance distributed training framework for Reinforcement Learning
★ 3.5krobotwin. Python
★ 5Chemistry3D. Python
★ 74umi-on-legs. UMI on Legs: Making Manipulation Policies Mobile with Manipulation-Centric Whole-body Controllers
★ 530RAGEN. RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.
★ 2.8kRoboVerse. RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning
★ 1.8kcleanrl. High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)
★ 10khumanoid-bench. Python
★ 779CLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34kGR-1. Code for "Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation"
★ 308auto_eval. AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World | CoRL 2025
★ 99TesserAct. ICCV 2025 | TesserAct: Learning 4D Embodied World Models
★ 404collector. [ICRA 2024] AirExo: Low-Cost Exoskeletons for Learning Whole-Arm Manipulation in the Wild
★ 50LAPA. [ICLR 2025] LAPA: Latent Action Pretraining from Videos
★ 562Metaworld. Collections of robotics environments geared towards benchmarking multi-task and meta reinforcement learning
★ 1.9kpyroki. A Modular Toolkit for Robot Kinematic Optimization
★ 1.7kXLeRobot. XLeRobot: Practical Dual-Arm Mobile Home Robot for $660
★ 5.4kvlarl. Single-file implementation to advance vision-language-action (VLA) models with reinforcement learning.
★ 447Improved-3D-Diffusion-Policy. [IROS 2025] Generalizable Humanoid Manipulation with 3D Diffusion Policies. Part 1: Train & Deploy of iDP3
★ 552DemoGen. [RSS25] Official implementation of DemoGen: Synthetic Demonstration Generation for Data-Efficient Visuomotor Policy Learning
★ 253discover-hidden-visual-concepts. Official implementation of `Discovering Hidden Visual Concepts Beyond Linguistic Input in Infant Learning`, CVPR 2025
★ 12openvla-oft. Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
★ 1.3kmshab. A Benchmark for Low-Level Manipulation in Home Rearrangement Tasks
★ 196vres. Vacation Research Experience Topic: Developing Bimanual Teleoperation Device for Dual Robotic Manipulator Arms and Robot Learning
★ 2triad_openvr. This is an enhanced wrapper for the already excellent pyopenvr library by cmbruns. The goal of this library is to create easy to use python functions for any SteamVR tracked system.
★ 180r3m. Pre-training Reusable Representations for Robotic Manipulation Using Diverse Human Video Data
★ 377vr-teleoperation. Code for bimanual teleoperation with Meta Quest3
★ 223d-foundation-policy. Jupyter Notebook
★ 114PointFlowMatch. Jupyter Notebook
★ 85EAGLE. Official Implementation of EAGLE-1 (ICML'24), EAGLE-2 (EMNLP'24), and EAGLE-3 (NeurIPS'25).
★ 2.5kTSP3D. [CVPR 2025, All Strong Accept] TSP3D: Text-guided Sparse Voxel Pruning for Efficient 3D Visual Grounding
★ 253calvin. CALVIN - A benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks
★ 964brs-ctrl. Official Hardware Codebase for the Paper "BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities"
★ 144brs-algo. Official Algorithm Codebase for the Paper "BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities"
★ 171furniture-bench. FurnitureBench: Real-World Furniture Assembly Benchmark (RSS 2023)
★ 236Slow_Thinking_with_LLMs. A series of technical report on Slow Thinking with LLM
★ 767STP. stp files of Tabletop Teleoperation structures
★ 3Product_User_Guide. Config files for my GitHub profile.
★ 38anyplace. Official implementation of "AnyPlace: Learning Generalized Object Placement for Robot Manipulation"
★ 97RLBench. A large-scale benchmark and learning environment.
★ 1.8kLIBERO. Benchmarking Knowledge Transfer in Lifelong Robot Learning
★ 2.1kdift. [NeurIPS'23] Emergent Correspondence from Image Diffusion
★ 773V-GPS. official implementation for our paper Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance (CoRL 2024)
★ 56nerfies.github.io. JavaScript
★ 4.3kAwesome-Robotics-Diffusion. A curated list of recent robot learning papers incorporating diffusion models for robotics tasks.
★ 351