This is your work, valued
TPN. Code for ICLR19 paper: Learning to Propagate Labels: Transductive Propagation Network for Few-shot Learning.
★ 242TPN-pytorch. Pytorch Code for ICLR19 paper: Learning to Propagate Labels: Transductive Propagation Network for Few-shot Learning.
★ 176SCOT. CVPR 2020, Semantic Correspondence as an Optimal Transport Problem, Pytorch Implementation.
★ 104Few-shot-Meta-learning-papers. Recent few-shot meta-learning papers
★ 853D-Medical-Generative-Survey.
★ 27LSMI-Sinkhorn. Code for ECML/PKDD paper: "LSMI-Sinkhorn: Semi-supervised Mutual Information Estimation with Optimal Transport"
★ 17tri-M-ICCV. A multi-mode modulator for multi-domain few-shot classification (ICCV)
★ 10DiffNCuts. This is the implementation of our ECCV 2024 paper "Unsupervised Dense Prediction using Differentiable Normalized Cuts" by Yanbin Liu and Stephen Gould.
★ 9Semi-supervised_Neural_Network. forked from https://github.com/jibancanyang/Semi-supervised_Neural_Network
★ 4NewMachineConfig. Usually used cmds for new machine configurations
★ 3MAML-Pytorch-Multi-GPUs. Python
★ 2prototypical-network-pytorch-1. A re-implementation of "Prototypical Networks for Few-shot Learning"
★ 2trl. Train transformer language models with reinforcement learning.
★ 19kskills. Give your agents the power of the Hugging Face ecosystem
★ 11kawesome-courses. I will collect a list of (open) courses first for AI, CS and Math. High-quality knowledge matters! Try to update everyday!
★ 9SmartBench. SmartBench
★ 3Deep-Learning-Based-Anomaly-Detection.
★ 360SSG. [CVPR 2026 Oral] Guiding a Diffusion Model by Swapping Its Tokens
★ 43camouflaged-vlm. Open-Vocabulary Camouflaged Object Segmentation with Cascaded Vision Language Models
★ 16AI-Research-SKILLs. Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepower. Maintained by Orchestra Research.
★ 11kOTT-Vid. Official code for "OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models"
★ 9Awesome-Attention-Sink. 🚀 First survey on Attention Sink in Transformers — 200+ papers on utilization, interpretation, and mitigation.
★ 137FasterVLM. Official code for paper: [CLS] Attention is All You Need for Training-Free Visual Token Pruning: Make VLM Inference Faster.
★ 114LLaVA-PruMerge. LLaVA-PruMerge: Adaptive Token Reduction for Efficient Large Multimodal Models
★ 173VLMEvalKit. Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
★ 4.3kDVS. ICML 2026: Information-geometric adaptive sampling for graph diffusion
★ 18drifting. Python
★ 485IMBA-Loss. [ICCV 2025] Official Implementation of the Paper "Imbalance in Balance: Online Concept Balancing in Generation Models".
★ 10mllms_know. [ICLR'25] Official code for the paper 'MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs'
★ 382LLMxFM_ICRL. Code for the "Can LLMs Reason Over Non-Text Modalities in a Training-Free Manner? A Case Study with In-Context Representation Learning" paper
★ 10evo-memory. Code to train and evaluate Neural Attention Memory Models to obtain universally-applicable memory systems for transformers.
★ 360MetaSPO. Python
★ 83Binoculars. [ICML 2024] Binoculars: Zero-Shot Detection of LLM-Generated Text
★ 400fast-detect-gpt. Code base for ICLR 2024 "Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature".
★ 419tuning-contribution. Official codebase for TuCo: Measuring the Contribution of Fine-Tuning to Individual Responses of LLMs (ICML 2025).
★ 5prompts.chat. f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
★ 167kchatgpt-prompts-for-academic-writing. This list of writing prompts covers a range of topics and tasks, including brainstorming research ideas, improving language and style, conducting literature reviews, and developing research plans.
★ 4.9kClavaDDPM. [NeurIPS 2024] Official implementation of "ClavaDDPM:Multi-relational Data Synthesis with Cluster-guided Diffusion Models"
★ 21GruM. Official Code Repository for the paper "Graph Generation with Diffusion Mixture" (ICML 2024).
★ 35denoising-diffusion-pytorch. Implementation of Denoising Diffusion Probabilistic Model in Pytorch
★ 11kICML2025-TCPA. Python
★ 23Academic-LaTeX-Writing-Submission-Checklist-. This checklist is designed to help you systematically prepare and polish academic papers for top conferences and journals (e.g., ICML, NeurIPS, CVPR). It incorporates widely recommended best practices, formatting standards, and common reviewer expectations.
★ 226segment-anything-annotator. We developed a python UI based on labelme and segment-anything for pixel-level annotation. It support multiple masks generation by SAM(box/point prompt), efficient polygon modification and category record. We will add more features (such as incorporating CLIP-based methods for category proposal and VOS methods for video datasets
★ 385Segment-Anything-CLIP. Connecting segment-anything's output masks with the CLIP model; Awesome-Segment-Anything-Works
★ 209SAMRS. The official repo for [NeurIPS'23] "SAMRS: Scaling-up Remote Sensing Segmentation Dataset with Segment Anything Model"
★ 385RSRefSeg. This is the pytorch implement of the paper "RSRefSeg: Referring Remote Sensing Image Segmentation with Foundation Models"
★ 78RobustSAM. RobustSAM: Segment Anything Robustly on Degraded Images (CVPR 2024 Highlight)
★ 368SkySense-O. [CVPR 2025] This is a model aggregated with CLIP and SAM version of SkySense for remote sensing interpretation described in SkySense-O: Towards Open-World Remote Sensing Interpretation with Vision-Centric Visual-Language Modeling.
★ 271GroundingDINO. [ECCV 2024] Official implementation of the paper "Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection"
★ 10kGeoPix. [GRSM] Project Page for "GeoPix: Multi-Modal Large Language Model for Pixel-level Image Understanding in Remote Sensing"
★ 73GeoPixel. GeoPixel: A Pixel Grounding Large Multimodal Model for Remote Sensing is specifically developed for high-resolution remote sensing image analysis, offering advanced multi-target pixel grounding capabilities.
★ 152SegEarth-OV. [CVPR 2025 Oral] SegEarth-OV: Towards Training-Free Open-Vocabulary Segmentation for Remote Sensing Images
★ 273open-sentinel-map. Dataset download instructions.
★ 40BayesianLM. [NeurIPS 2024] "Bayesian-Guided Label Mapping for Visual Reprogramming". Official Website: https://github.com/tmlr-group/BayesianLM
★ 9BayesianLM. [NeurIPS 2024 Oral] "Bayesian-Guided Label Mapping for Visual Reprogramming"
★ 12SAVs. Official Codebase for "Generative Multimodal Model Features Are Discriminative Vision-Language Classifiers"
★ 26OpenCity. "OpenCity: Open Spatio-Temporal Foundation Models for Traffic Prediction"
★ 176synthehicle. [WACVW 2023] A massive synthetic dataset for 3D multi-target multi-camera tracking and segmentation.
★ 54torch-LLM4SGG. Official PyTorch implementation Source code for LLM4SGG: Large Language Models for Weakly Supervised Scene Graph Generation, accepted at CVPR 2024
★ 116multimodal-icl. Code repo for paper Can Multimodal Large Language Models Truly Perform Multimodal In-Context Learning? WACV 2025
★ 8DiSA. Official Implementation of Diffusion Step Annealing (DiSA) in Autoregressive Image Generation
★ 143fewshotgraph. Awesome Few-Shot Learning on Graphs
★ 25imu2clip. Code repository for IMU2CLIP(https//arxiv.org/pdf/2210.14395.pdf)
★ 105COMODO. [UbiComp/IMWUT '26] Official Repo for COMODO: Cross-Modal Video-to-IMU Distillation for Efficient Egocentric Human Activity Recognition
★ 24SelfPAB. Large-Scale Pre-Training for Dual-Accelerometer Human Activity Recognition
★ 18Self-Supervised-Learning-HAR. Python
★ 11causal. notebooks on methods for causal inference
★ 39EuroSAT. EuroSAT: Land Use and Land Cover Classification with Sentinel-2
★ 567gallop. Adaptation of vision-language models (CLIP) to downstream tasks using local and global prompts.
★ 52SymDPO. We have developed Symbol Demonstration Direct Preference Optimization (SymDPO) and validating its effectiveness across multiple benchmarks.
★ 23CASS. [CVPR 2025] Official Pytorch Code for Distilling Spectral Graph for Object-Context Aware Open-Vocabulary Semantic Segmentation
★ 50reevo. [NeurIPS 2024] ReEvo: Large Language Models as Hyper-Heuristics with Reflective Evolution
★ 290EoH. Evolution of Heuristics
★ 352CMuST. [NeurIPS 2024 Oral] Repository of the CMuST paper: "Get Rid of Isolation: A Continuous Multi-task Spatio-Temporal Learning Framework"
★ 15AustraliaFires. Exploratory analysis of wild-fires in Australia & a machine learning approach for wildfire modeling in Google Earth Engine
★ 19WeCLIP. CVPR2024
★ 111Time-Series-Library. A Library for Advanced Deep Time Series Models for General Time Series Analysis.
★ 13kvmf-lib. Python
★ 41SeafloorAI. Python
★ 14FastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40kllama-models. Utilities intended for use with Llama models.
★ 7.7kCVinW_Readings. A collection of papers on the topic of ``Computer Vision in the Wild (CVinW)''
★ 1.4kGeoChat. [CVPR 2024 🔥] GeoChat, the first grounded Large Vision Language Model for Remote Sensing
★ 734DeepOSM. Train a deep learning net with OpenStreetMap features and satellite imagery.
★ 1.3kSSP. Our Implement for SSP
★ 8DINOv. [CVPR 2024] Official implementation of the paper "Visual In-context Learning"
★ 542DiffNCuts. This is the implementation of our ECCV 2024 paper "Unsupervised Dense Prediction using Differentiable Normalized Cuts" by Yanbin Liu and Stephen Gould.
★ 9GraphAdapter. The efficient tuning method for VLMs
★ 83SUMformer. Rethinking Urban Mobility Prediction: A Super-Multivariate Time Series Forecasting Approach (TITS)
★ 20UniST. Official implementation for "UniST: A Prompt-Empowered Universal Model for Urban Spatio-Temporal Prediction" (KDD 2024)
★ 228lmms-eval. One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
★ 4.3kssl_eval_protocols. Code for the paper "A Closer Look at Benchmarking Self-Supervised Pre-training with Image Classification"
★ 8ai-for-grant-writing. A curated list of resources for using LLMs to develop more competitive grant applications.
★ 4.2kSlowFast. PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.
★ 7.4kAI_for_Science_paper_collection. List the AI for Science papers accepted by top conferences
★ 175Multi-Level-Logit-Distillation. Code for 'Multi-level Logit Distillation' (CVPR2023)
★ 70atlas. [NeurIPS'24] Official PyTorch implementation for paper "Knowledge Composition using Task Vectors with Learned Anisotropic Scaling"
★ 28RepDistiller. [ICLR 2020] Contrastive Representation Distillation (CRD), and benchmark of recent knowledge distillation methods
★ 2.4kNOAH. [TPAMI] Searching prompt modules for parameter-efficient transfer learning.
★ 241leopart. Python
★ 100low_cost_robot. Python
★ 3.4kaction_seg_ot. [CVPR 2024] - Official code for the paper "Temporally Consistent Unbalanced Optimal Transport for Unsupervised Action Segmentation"
★ 54mamba. Mamba SSM architecture
★ 19kchronos-forecasting. Chronos: Pretrained Models for Time Series Forecasting
★ 5.7kdemistfy_correspondence. Code for the ECCV22 paper Demystifying Unsupervised Semantic Correspondence Estimation
★ 14Candidate-Reranking-CIR. The official implementation for Candidate Set Re-ranking for Composed Image Retrieval (TMLR) 01/2024
★ 20Neural-Network-Diffusion. We introduce a novel approach for parameter generation, named neural network parameter diffusion (p-diff), which employs a standard latent diffusion model to synthesize a new set of parameters
★ 886POT. POT : Python Optimal Transport
★ 2.8kVLGuard. [ICML 2024] Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language Models.
★ 90Depth-Anything. [CVPR 2024] Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data. Foundation Model for Monocular Depth Estimation
★ 8.2ksnntorch. Deep and online learning with spiking neural networks in Python
★ 2kMonocular-Depth-Estimation-Toolbox. Monocular Depth Estimation Toolbox based on MMSegmentation.
★ 970LLaVA. [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
★ 25kQ-Instruct. ②[CVPR 2024] Low-level visual instruction tuning, with a 200K dataset and a model zoo for fine-tuned checkpoints.
★ 238AesBench. An expert benchmark aiming to comprehensively evaluate the aesthetic perception capacities of MLLMs.
★ 261Amodal-Completion-in-the-Wild. Python
★ 87FusedGW-Entity-Alignment. "A Fused Gromov-Wasserstein Framework for Unsupervised Knowledge Graph Entity Alignment" in ACL 2023
★ 13LLM-Shearing. [ICLR 2024] Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning
★ 643DejaVu. Python
★ 359weight-selection. Python
★ 193Awesome-Pruning. A curated list of neural network pruning resources.
★ 2.5kmeta-omnium. Implementation of "Meta Omnium: A Benchmark for General-Purpose Learning-to-Learn"
★ 25OneLLM. [CVPR 2024] OneLLM: One Framework to Align All Modalities with Language
★ 666yet-another-gpt-tutorial-v2. Jupyter Notebook
★ 14objaverse-xl. 🪐 Objaverse-XL is a Universe of 10M+ 3D Objects. Contains API Scripts for Downloading and Processing!
★ 1.3k3dv_tutorial. An Invitation to 3D Vision: A Tutorial for Everyone
★ 1.8kopen_clip. An open source implementation of CLIP.
★ 14kImageBind. ImageBind One Embedding Space to Bind Them All
★ 9.1kLLMsPracticalGuide. A curated list of practical guide resources of LLMs (LLMs Tree, Examples, Papers)
★ 10kFoolyourVLLMs. [ICML 2024] Fool Your (Vision and) Language Model With Embarrassingly Simple Permutations
★ 15guijiejie.github.io. HTML
★ 12dl_tutorials_10weeks. Deep Learning Tutorials for 10 Weeks
★ 420intro-to-linear-algebra. Jupyter Notebook
★ 15rl_tutorial. Yet Another Reinforcement Learning Tutorial
★ 72yet-another-gpt-tutorial. Jupyter Notebook
★ 42yet-another-rl-tutorial. Jupyter Notebook
★ 58cs231n.github.io. Public facing notes page
★ 11kSyntheticTumors. [CVPR 2023] Label-Free Liver Tumor Segmentation
★ 380Contrastive-Correspondence. Codes for "Learning Contrastive Representation for Semantic Correspondence"
★ 10SimpleCRF. matlab and python wrap of crf and dense crf, both 2d and 3d are supported
★ 171PyTorch-Gumbel-Sigmoid. Implementation of the Gumbel-Sigmoid distribution in PyTorch.
★ 20pytorch-gradual-warmup-lr. Gradually-Warmup Learning Rate Scheduler for PyTorch
★ 987TokenCut. (CVPR 2022) Pytorch implementation of "Self-supervised transformers for unsupervised object discovery using normalized cut"
★ 339fiftyone. Refine high-quality datasets and visual AI models
★ 11kawesome-tips.
★ 4.7kCutLER. Code release for "Cut and Learn for Unsupervised Object Detection and Instance Segmentation" and "VideoCutLER: Surprisingly Simple Unsupervised Video Instance Segmentation"
★ 1.1kSegment-and-Track-Anything. An open-source project dedicated to tracking and segmenting any objects in videos, either automatically or interactively. The primary algorithms utilized include the Segment Anything Model (SAM) for key-frame segmentation and Associating Objects with Transformers (AOT) for efficient tracking and propagation purposes.
★ 3.1kTimeCycle. Learning Correspondence from the Cycle-consistency of Time (CVPR 2019)
★ 723OpenSeeD. [ICCV 2023] Official implementation of the paper "A Simple Framework for Open-Vocabulary Segmentation and Detection"
★ 762codon. A high-performance, zero-overhead, extensible Python compiler with built-in NumPy support
★ 17kdeep-spectral-segmentation. [CVPR 2022] Deep Spectral Methods: A Surprisingly Strong Baseline for Unsupervised Semantic Segmentation and Localization
★ 239transformer. Transformer: PyTorch Implementation of "Attention Is All You Need"
★ 4.6kTRASHED2022-how-to-record-lectures. Python
★ 1learned_optimization. Python
★ 812einops. Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)
★ 9.6kstable-diffusion. A latent text-to-image diffusion model
★ 73knerf-pytorch. A PyTorch re-implementation of Neural Radiance Fields
★ 911giraffe. This repository contains the code for the CVPR 2021 paper "GIRAFFE: Representing Scenes as Compositional Generative Neural Feature Fields"
★ 1.2keditnerf. Editing a Conditional Radiance Field
★ 258loss-landscape. Code for visualizing the loss landscape of neural nets
★ 3.2kBag_of_Tricks_for_Image_Classification_with_Convolutional_Neural_Networks. experiments on Paper <Bag of Tricks for Image Classification with Convolutional Neural Networks> and other useful tricks to improve CNN acc
★ 741nerf-pytorch. A PyTorch implementation of NeRF (Neural Radiance Fields) that reproduces the results.
★ 6.1kvit-pytorch. Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
★ 25kToMe. A method to increase the speed and lower the memory footprint of existing vision transformers.
★ 1.2kclrs. Jupyter Notebook
★ 539Awesome-Multi-label-Image-Recognition. Awesome Multi-label Image Recognition Paper List
★ 165cvpr20_LESA. Shared Attention for Multi-label Zero-shot Learning accepted @ CVPR20
★ 32LaSO. LaSO: Label-Set Operations networks for multi-label few-shot learning - official implementation
★ 87FewShotMultiLabel. Code for AAAI2021 paper: Few-Shot Learning for Multi-label Intent Detection.
★ 111optnet. OptNet: Differentiable Optimization as a Layer in Neural Networks
★ 587jaxopt. Hardware accelerated, batchable and differentiable optimizers in JAX.
★ 1.1k3D-Medical-Generative-Survey.
★ 27CLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34ktreeOT. Python
★ 8learning_minimal. Learning to Solve Hard Minimal Problems
★ 145VS-Survey.
★ 205jaxgptoolbox. Geometry processing utilities compatible with jax for autodifferentiation.
★ 90FractalDB-Pretrained-ResNet-PyTorch. Pre-training without Natural Images (ACCV 2020 Best Paper Honorable Mention Award)
★ 221FROT. Feature Robust Optimal Transport
★ 8neurite. Neural networks toolbox focused on medical image analysis
★ 376Sharp-MAML. Sharp-MAML: Sharpness-Aware Model-Agnostic Meta Learning
★ 34AdaContrast. Official repository of CVPR '22 paper "Contrastive Test-Time Adaptation"
★ 106head-pulse-track. Detecting Pulse from Head Motions in Video
★ 21Latex-Template-Rice-USTC. A Latex template including book, thesis, and assignment for Rice and USTC university.
★ 4ddn. Deep Declarative Networks
★ 252Efficient-3DCNNs. PyTorch Implementation of "Resource Efficient 3D Convolutional Neural Networks", codes and pretrained models.
★ 821HowToCook. Programmer's guide about how to cook at home.
★ 101kLASEM. Code for the ICML 2021 paper "Sharing Less is More: Lifelong Learning in Deep Networks with Selective Layer Transfer"
★ 12ott. Optimal transport tools implemented with the JAX framework, to solve large scale matching problems of any flavor.
★ 752pytorch-meta-dataset. A non-official 100% PyTorch implementation of META-DATASET benchmark for few-shot classification
★ 59generalized-category-discovery. Code for our CVPR 2022 paper 'Generalized Category Discovery'. Project page: https://www.robots.ox.ac.uk/~vgg/research/gcd/
★ 241Few_Shot_Distribution_Calibration. [ICLR2021 Oral] Free Lunch for Few-Shot Learning: Distribution Calibration
★ 475MixCo-Mixup-Contrast. (NeurIPS 2020 Workshop on SSL) Official Implementation of "MixCo: Mix-up Contrastive Learning for Visual Representation"
★ 58tri-M-ICCV. A multi-mode modulator for multi-domain few-shot classification (ICCV)
★ 10eg3d. Python
★ 3.3kstylegan3. Official PyTorch implementation of StyleGAN3
★ 6.9kMedicalNet. Many studies have shown that the performance on deep learning is significantly affected by volume of training data. The MedicalNet project provides a series of 3D-ResNet pre-trained models and relative code.
★ 2.2kMedicalZooPytorch. A pytorch-based deep learning framework for multi-modal 2D/3D medical image segmentation
★ 1.9kACSConv. [IEEE JBHI'21] Reinventing 2D Convolutions for 3D Images - 1 line of code to convert pretrained 2D models to 3D!
★ 173pytorch-fid. Compute FID scores with PyTorch.
★ 3.9ksimple-pytorch-3dgan. A simple and unofficial 3D-GAN implementation using PyTorch [NeurIPS 2016]
★ 923D-GAN-pytorch. A pytorch implementation of 3D-GAN
★ 393dbraingen. Official Pytorch Implementation of "Generation of 3D Brain MRI Using Auto-Encoding Generative Adversarial Network" (accepted by MICCAI 2019)
★ 135medical-datasets. tracking medical datasets, with a focus on medical imaging
★ 923MedMNIST. [pip install medmnist] 18x Standardized Datasets for 2D and 3D Biomedical Image Classification
★ 1.4kLiftedGAN. (CVPR 2021) Lifting 2D StyleGAN for 3D-Aware Face Generation
★ 793DStyleGAN. 3D StyleGAN2 for Medical Images
★ 73cST-ML. The codes and data of paper "cST-ML: Continuous Spatial-Temporal Meta-Learning for Traffic Dynamics Prediction"
★ 10pytorch-CycleGAN-and-pix2pix. Image-to-Image Translation in PyTorch
★ 25kReg-GAN. Python
★ 190Domain-Consensus-Clustering. [CVPR2021] Domain Consensus Clustering for Universal Domain Adaptation
★ 112Solacex.
★ 1