This is your work, valued
MTMT. Code for the CVPR 2020 paper "A Multi-task Mean Teacher for Semi-supervised Shadow Detection"
★ 106ViSha. Code for the CVPR 2021 paper "Triple-cooperative Video Shadow Detection"
★ 28MedRPG. Python
★ 22liver_segmentation. Unet,Unet++,FPN,DAF implementations for liver segmentation
★ 15MICCAI2019. Python
★ 6ICBNet. Python
★ 2Video_segmentation_pytorch. Better for single label video segmentation task (e.g. Video saliency detection, Video shadow detection)
★ 1Peak-End-Net. [ACM MM 2026] Peak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic Assessment
★ 26wolfcha. AI-powered Werewolf (Mafia) social deduction game where every player is controlled by top LLMs like DeepSeek, Qwen, Gemini, and more
★ 664ADE-CoT. [CVPR 26] From Scale to Speed: Adaptive Test-Time Scaling for Image Editing
★ 46BlockPilot. Python
★ 61DreamX-World. DreamX-World: A General-Purpose Interactive World Model
★ 740APPO. Python
★ 72roleagent. Python
★ 79UniMRG. [ICML 2026] The official implementation of paper "Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation"
★ 90TransitLM. TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation
★ 125MACE-Dance. [SIGGRAPH 2026] MACE-Dance: Motion-Appearance Cascaded Experts for Music-Driven Dance Video Generation
★ 107LLaTiSA. This is the official repository of "LLaTiSA: Towards Difficulty-Stratified Time Series Reasoning from Visual Perception to Semantics".
★ 78EMF. [2026 CVPR]Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation
★ 109DCW. [CVPR 2026] Elucidating the SNR-t Bias of Diffusion Probabilistic Models
★ 120FASA-ICLR2026. [ICLR 2026] FASA: FREQUENCY-AWARE SPARSE ATTENTION
★ 20AR-MAP-Main. Python
★ 24Pos2Distill-. [EMNLP25] Official code for "POSITION BIAS MITIGATES POSITION BIAS: Mitigate Position Bias Through Inter-Position Knowledge Distillation"
★ 38HS-STaR. [EMNLP’ 25] Official code for "HS-STaR: Hierarchical Sampling for Self-Taught Reasoners via Difficulty Estimation and Budget Reallocation"
★ 37LD-RPS. [ICCV25] LD-RPS
★ 48UPRE-ICCV2025. [ICCV2025] UPRE: Zero-Shot Domain Adaptation for Object Detection via Unified Prompt and Representation Enhancement
★ 29SkillClaw. Let Skills Evolve Collectively with Agentic Evolver
★ 2.3kOmni-WorldBench. A comprehensive benchmark specifically designed to evaluate the interactive response capabilities of world models in 4D settings.
★ 106MobilityBench. [KDD 2026 Oral] MobilityBench: A Scalable Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios
★ 157Code2World. Code2World: A GUI World Model via Renderable Code Generation
★ 323Q-Hawkeye. Python
★ 61SpatialGenEval. [ICLR2026] Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models
★ 132MathForge. [ICLR 2026] Harder Is Better: Boosting Mathematical Reasoning via Difficulty-Aware GRPO and Multi-Aspect Question Reformulation
★ 128SocioReasoner. [ICLR26] Official implementation of the paper "Urban Socio-Semantic Segmentation with Vision-Language Reasoning"
★ 175Thinking-with-Map. [ACL 2026 Findings] Thinking with Map: Reinforced Parallel Map-Augmented Agent for Geolocalization
★ 176Taming-Hallucinations. JavaScript
★ 55BenchX. BenchX: A Unified Benchmark Framework for Medical Vision-Language Pretraining on Chest X-Rays
★ 49IC-Light. More relighting!
★ 8.5kx-flux. Python
★ 2.2kguided-diffusion. Python
★ 7.4kdenoising-diffusion-pytorch. Implementation of Denoising Diffusion Probabilistic Model in Pytorch
★ 11kprompt-to-prompt. Jupyter Notebook
★ 3.5kMMN-VSOD. Python
★ 15lang-segment-anything. SAM with text prompt
★ 2.6kmamba. Mamba SSM architecture
★ 19kMedKLIP. The official code for MedKLIP: Medical Knowledge Enhanced Language-Image Pre-Training in Radiology. We propose to leverage medical specific knowledge enhancing language-image pre-training method, significantly advancing the ability of pre-trained models to handle unseen diseases on zero-shot classification and grounding tasks.
★ 180chexpert-labeler. CheXpert NLP tool to extract observations from radiology reports.
★ 409meru. Code for the paper "Hyperbolic Image-Text Representations", Desai et al, ICML 2023
★ 204MedRPG. Python
★ 22Grounded-Segment-Anything. Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
★ 18kvgtr. [ICME'22] Visual Grounding with Transformers
★ 28R2Gen. Python
★ 207TBraTS. 【MICCAI 2022】"TBraTS: Trusted Brain Tumor Segmentation"
★ 57FDRNet. Code for our ICCV 2021 paper "Mitigating Intensity Bias in Shadow Detection via Feature Decomposition and Reweighting"
★ 19mmsegmentation. OpenMMLab Semantic Segmentation Toolbox and Benchmark.
★ 9.9kBasicSR. Open Source Image and Video Restoration Toolbox for Super-resolution, Denoise, Deblurring, etc. Currently, it includes EDSR, RCAN, SRResNet, SRGAN, ESRGAN, EDVR, BasicVSR, SwinIR, ECBSR, etc. Also support StyleGAN2, DFDNet.
★ 8.4kSID. Official implementation for ICCV19 "Shadow Removal via Shadow Image Decomposition"
★ 111exposure-fusion-shadow-removal. We propose a new method for effective shadow removal by regarding it as an exposure fusion problem.
★ 197Implicit-Internal-Video-Inpainting. [ICCV 2021]: IIVI: Internal Video Inpainting by Implicit Long-range Propagation
★ 259pytorch-toolbelt. PyTorch extensions for fast R&D prototyping and Kaggle farming
★ 1.6krvos. RVOS: End-to-End Recurrent Network for Video Object Segmentation (CVPR 2019)
★ 277detectron2. Detectron2 is a platform for object detection, segmentation and other visual recognition tasks.
★ 35kViSha. Code for the CVPR 2021 paper "Triple-cooperative Video Shadow Detection"
★ 28STM-Training. training script for space time memory network
★ 119FEELVOS. FEELVOS implementation in PyTorch; FEELVOS: Fast End-to-End Embedding Learning for Video Object Segmentation
★ 69MedicalNet. Many studies have shown that the performance on deep learning is significantly affected by volume of training data. The MedicalNet project provides a series of 3D-ResNet pre-trained models and relative code.
★ 2.2kDEXTR-PyTorch. Deep Extreme Cut http://www.vision.ee.ethz.ch/~cvlsegmentation/dextr
★ 846Awesome-Crowd-Counting. Awesome Crowd Counting
★ 2.6kMask-ShadowGAN. Mask-ShadowGAN: Learning to Remove Shadows from Unpaired Data | ICCV 2019
★ 150tutorials. MONAI Tutorials
★ 2.5kGridDehazeNet. This repo contains the official training and testing codes for our paper: GridDehazeNet: Attention-Based Multi-Scale Network for Image Dehazing.
★ 1R3Net. Code for the IJCAI 2018 paper "R^3Net: Recurrent Residual Refinement Network for Saliency Detection"
★ 121BDRAR. Code for the ECCV 2018 paper "Bidirectional Feature Pyramid Network with Recurrent Attention Residual Modules for Shadow Detection"
★ 129MTMT. Code for the CVPR 2020 paper "A Multi-task Mean Teacher for Semi-supervised Shadow Detection"
★ 106pytorch-3dunet. 3D U-Net model for volumetric semantic segmentation written in pytorch
★ 2.4kMICCAI2019. Python
★ 6liver_segmentation. Unet,Unet++,FPN,DAF implementations for liver segmentation
★ 15