This is your work, valued
LIBERO-plus. Official repository of LIBERO-plus, a generalized benchmark for in-depth robustness analysis of vision-language-action models.
★ 399Penguin-VL. Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders [Technical Report]
★ 206MergeVLA. [CVPR 2026] MergeVLA: Cross-Skill Model Merging Toward a Generalist Vision-Language-Action Agent
★ 36LangForce. [ICML 2026] This repo is the official implementation of "LangForce : Bayesian Decomposition of Vision Language Action Models via Latent Action Queries"
★ 73PromptTA. Source-free Domain Generalization
★ 16VLA-ADP. Action aware Dynamic Pruning for Efficient Vision Language Action Manipulation
★ 25le-wm. Official code base for LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
★ 4.2kFluxVLA. An all-in-one VLA engineering platform for embodied AI — from data to real-robot deployment.
★ 578VLD. Python
★ 11SVIP. Learning Source-free Domain Adaptation for Visible-Infrared Person Re-Identification (NeruIPS, 2025, Code)-SVIP
★ 3TTAC2. [TPAMI 2024] The official implementation of "Revisiting Realistic Test-Time Training: Sequential Inference and Adaptation by Anchored Clustering Regularized Self-Training"
★ 13NeurIPS25-CogVLA. [NeurIPS 2025] CogVLA: Cognition-Aligned Vision-Language-Action Models via Instruction-Driven Routing & Sparsification
★ 185starVLA. StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing
★ 3.3kX-VLA. [ICLR 2026] The offical Implementation of "Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model"
★ 697TGA-ZSR. (TPAMI 2026) Complementary Text-Guided Attention for Zero-Shot Adversarial Robustness & & (NeurIPS 2024) Text-Guided Attention is All You Need for Zero-Shot Robustness in Vision-Language Models
★ 24VLA-Adapter-Real. The real-world system implementation of VLA-Adapter.
★ 11AnyTouch2. [ICLR 2026] AnyTouch 2: General Optical Tactile Representation Learning For Dynamic Tactile Perception
★ 53RLinf. RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI
★ 4.3kCorrWiseLosses. Code for CVPR 2022 paper "Comparing Correspondences: Video Prediction with Correspondence-wise Losses"
★ 25F2BA. A tale of works on the complexity of first-order bilevel optimization.
★ 25SimpleVLA-RL. [ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
★ 1.8kLanguageBind. 【ICLR 2024🔥】 Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment
★ 883LLaMA2-Accessory. An Open-source Toolkit for LLM Development
★ 2.8kLLaMA-Adapter. [ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parameters
★ 5.9kImageBind-LoRA. Fine-tuning "ImageBind One Embedding Space to Bind Them All" with LoRA
★ 195VLA-Adapter. VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model
★ 2.3khand-on-rl. Python
★ 113AI-Notes. Bilibili东川路第一可爱猫猫虫的AI笔记
★ 277SpeechCLIP. SpeechCLIP: Integrating Speech with Pre-Trained Vision and Language Model, Accepted to IEEE SLT 2022
★ 120Generalization-Causality. 关于domain generalization,domain adaptation,causality,robutness,prompt,optimization,generative model各式各样研究的阅读笔记
★ 1.2kimagecorruptions. Python package to corrupt arbitrary images.
★ 477CBSA. [NeurIPS25 Spotlight] Official Implementation for CBSA (Contract-and-Broadcast Self-Attention)
★ 36GAMP.
★ 1HOS-Net. [AAAI-2024] High-Order Structure Based Middle-Feature Learning for Visible-Infrared Person Re-Identification
★ 252024-ICML-TAC. Code for the paper "Image Clustering with External Guidance" (ICML 2024)
★ 71Awesome-Touch. Tactile Sensing • Data Collection • IL/RL/VLA/WM • Manipulation • Simulation • Open Source
★ 750Awesome-Robotics-Foundation-Models.
★ 1.4ktvl. [ICML 2024] A Touch, Vision, and Language Dataset for Multimodal Alignment
★ 101VHKOR. Visual-Haptic-Kinesthetic Object Recognition with Multimodal Transformer
★ 3mvitac. Self-Supervised Visual-Tactile Representation Learning via Multimodal Contrastive Training
★ 27DGTR. Official Code for Dexterous Grasp Transformer (CVPR 2024)
★ 67TTST. [IEEE TIP 2024] TTST: A Top-k Token Selective Transformer for Remote Sensing Image Super-Resolution
★ 141pytorch-CycleGAN-and-pix2pix. Image-to-Image Translation in PyTorch
★ 25kSGPA. Example code of Sparse Gaussian Process Attention (ICLR 2023)
★ 26Non-Local-Sparse-Attention. PyTorch code for our paper "Image Super-Resolution with Non-Local Sparse Attention" (CVPR2021).
★ 183LVM. Python
★ 1.8krebiber. A simple tool to update bib entries with their official information (e.g., DBLP or the ACL anthology).
★ 3kLReID-KRKC. The official implementation of AAAI2023 Oral paper "Lifelong Person Re-Identification via Knowledge Refreshing and Knowledge Consolidation"
★ 26LifelongReID. offical implement of our Lifelong Person Re-Identification via Adaptive Knowledge Accumulation in CVPR2021
★ 90PythonGuide. 「Python学习+面试指南」一份涵盖大部分Python相关行业的程序员所需要掌握的核心知识。准备Python面试,来看PythonGuide!
★ 108awesome-machine-learning-1. Learning Resources And Links Of Machine Learning(updating)
★ 817research-method. 论文写作与资料分享
★ 3.3kMFH. [ICLR 23 oral] The Modality Focusing Hypothesis: Towards Understanding Crossmodal Knowledge Distillation
★ 44DDPM. PyTorch DDPM implementation
★ 868SegDiff. Python
★ 170MedSegDiff. Using Diffusion Models to Segment/Reconstruct Organs from Medical Images [AAAI Most influential Paper]
★ 1.4kUSL-VI-ReID. The implementation of cvpr 2023 paper "Unsupervised Visible-Infrared Person Re-Identification via Progressive Graph Matching and Alternate Learning"
★ 58UNIReID. Towards Modality-Agnostic Person Re-identification with Descriptive Query CVPR2023
★ 31visual_prompting. Exploring Visual Prompts for Adapting Large-Scale Models
★ 292Visualizer. assistant tools for attention visualization in deep learning
★ 1.3kPythonExercise. Python 编程练习题 100 例(源码),实例在 Python 3.6 环境下测试通过。
★ 623AISystem. AISystem 主要是指AI系统,包括AI芯片、AI编译器、AI推理和训练框架等AI全栈底层技术
★ 17kVehicleX. VehicleX: Simulating Content Consistent Vehicle Datasets with Attribute Descent (ECCV 2020, TPAMI 2023)
★ 168VVeRi-901-Trial. Python
★ 3PhD-Learning. Jianan Zhao, Fengliang Qi, GuangYu Ren, Lin Xu*. PhD Learning: Learning with Pompeiu-hausdorff Distances for Video-based Person Re-Identification. IEEE Conference on Computer Vision and Pattern Recognition (CVPR). 2021.
★ 17CF-AAN. Video-based Person Re-identification without Bells and Whistles (CVPRw 2021)
★ 27CAViT. Python
★ 16awesome-reid-dataset. Collection of public available person re-identification datasets
★ 1.1kmulti-modal-vehicle-Re-ID. The multi-modal datasets and the codes of HAMNet.
★ 20Diffusion-based-Segmentation. This is the official Pytorch implementation of the paper "Diffusion Models for Implicit Image Segmentation Ensembles".
★ 318LLCM. [CVPR 2023] Diverse Embedding Expansion Network and Low-Light Cross-Modality Benchmark for Visible-Infrared Person Re-identification
★ 146FloW-Dataset. A dataset for floating waste detection in inland waters.
★ 55BTRDA. Python
★ 5Hypergraphs-Image-Inpainting. (WACV 2021) Hyperrealistic Image Inpainting with Hypergraphs
★ 93PyGCL. PyGCL: A PyTorch Library for Graph Contrastive Learning
★ 962LCCRF. Pytorch code for "Leaning Compact and Representative Features for Cross-Modality Person Re-Identification"(World Wide Web,CCF-B).
★ 9DCRN. [AAAI 2022] An official source code for paper Deep Graph Clustering via Dual Correlation Reduction.
★ 199Hetero-center-triplet-loss-for-VT-Re-ID. Python
★ 45SPConv.pytorch. [ IJCAI-20 ] Split to Be Slim: An Overlooked Redundancy in Vanilla Convolution
★ 127MITML. Python
★ 39GroupWhitening. This project is the PyTorch implementation of our accepted CVPR 2021 paper, Group Whitening: Balancing Learning Efficiency and Representational Capacity
★ 9dwt-domain-adaptation. Code for paper "Unsupervised Domain Adaptation using Feature-Whitening and Consensus Loss" (CVPR 2019)
★ 65PytorchWCT. This is the Pytorch implementation of Universal Style Transfer via Feature Transforms.
★ 369MSG-Transformer. MSG-Transformer: Exchanging Local Spatial Information by Manipulating Messenger Tokens (CVPR 2022)
★ 80MixMatch-pytorch. Code for "MixMatch - A Holistic Approach to Semi-Supervised Learning"
★ 652TBH. Auto-Encoding Twin-Bottleneck Hashing for CVPR2020
★ 90deep-cross-modal-hashing. Deep learning cross modal hashing in PyTorch
★ 109MyReidProject. Python
★ 30SAT. Official Pytorch code for Structure-Aware Transformer.
★ 266PMT. Code for paper (Lu, H., Zou, X., & Zhang, P. Learning progressive modality-shared transformers for effective visible-infrared person re-identification. In Proceedings of the AAAI conference on artificial intelligence (Vol. 37, No. 2, pp. 1835-1843).)
★ 54DCR-ReID_TCSVT2023. Python
★ 25pytorch_ConvUnitOptimization. This is the official repository of "Improving generalization of Batch Whitening by Convolutional Unit Optimization", ICCV 2021.
★ 8IterNorm-pytorch. This is the pytorch re-implementation of the IterNorm
★ 41Switchable-Whitening. Code for Switchable Whitening (ICCV2019)
★ 137Deep-Feature-Consistent-VAE. Pytorch Implementation of Hou, Shen, Sun, Qiu, "Deep Feature Consistent Variational Autoencoder", 2016
★ 511785-Flow-VAE-dientanglement. Large-scale image generation on ImageNet
★ 1VQ-VAE. This is my implementation of Vector Quantized VAE trained on ImageNet data
★ 5ResNetVAE. Variational AutoEncoder + ResNet Transfer Learning
★ 233rq-vae-transformer. The official implementation of Autoregressive Image Generation using Residual Quantization (CVPR '22)
★ 1kenhancing-transformers. An unofficial implementation of both ViT-VQGAN and RQ-VAE in Pytorch
★ 324transformer-vae. A library for making Transformer Variational Autoencoders. (Extends the Huggingface/transformers library.)
★ 144Machine-Learning-with-Graphs.
★ 101ADCA.
★ 59GCMT. Graph Consistency based Mean-Teaching for Unsupervised Domain Adaptive Person Re-Identification
★ 6SYSU-MM01. Introduction and evaluation code for RGB-IR re-id dataset SYSU-MM01.
★ 134Cross-Modal-Re-ID-baseline. Pytorch Code for Cross-Modality Person Re-Identification (Visible Thermal/Infrared Re-ID)
★ 401HiCMD. [CVPR2020] Hi-CMD: Hierarchical Cross-Modality Disentanglement for Visible-Infrared Person Re-Identification
★ 82CutMix-PyTorch. Official Pytorch implementation of CutMix regularizer
★ 1.3ksimpleHRNET. Pose Detection (Simple HRNET)
★ 1hrnet. Python
★ 1HRNet. Implementation of HRNet
★ 1hrnet. Python
★ 1HRNet. A compact HRNet implementation in PyTorch
★ 11