This is your work, valued
SPFSplat. [ICCV 2025 Highlight] No Pose at All: Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views
★ 153NAS3R. [CVPR 2026] From None to All: Self-Supervised 3D Reconstruction via Novel View Synthesis
★ 115SPFSplatV2. Official implementation of SPFSplatV2: Efficient Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views
★ 53DiffSynth-Studio. Enjoy the magic of Diffusion models!
★ 13kLatentForcing. Python
★ 144Gen3R. [CVPR 2026] Gen3R: 3D Scene Generation Meets Feed-Forward Reconstruction
★ 366VIST3A. [ICLR 2026 oral] Official code for VIST3A: Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator
★ 140Sana. SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
★ 8.6kvggt-omega. [CVPR 2026 Oral] VGGT Omega
★ 3.8kijepa. Official codebase for I-JEPA, the Image-based Joint-Embedding Predictive Architecture. First outlined in the CVPR paper, "Self-supervised learning from images with a joint-embedding predictive architecture."
★ 3.5kml-lgtm. Less Gaussians, Texture More: 4K Feed-Forward Textured Splatting
★ 157stable-virtual-camera. Stable Virtual Camera: Generative View Synthesis with Diffusion Models
★ 1.6kPixelDiT. [CVPR 2026 Best Paper Finalist] Pixel Diffusion Transformers for Image Generation
★ 906lagernvs. Official code for "LagerNVS Latent Geometry for Fully Neural Real-time Novel View Synthesis" (CVPR 2026)
★ 402diffusion-forcing-transformer. [ICML 2025] Official PyTorch Implementation of "History-Guided Video Diffusion"
★ 705AnySplat. [SIGGRAPH Asia 2025 (ACM TOG)] AnySplat: Feed-forward 3D Gaussian Splatting from Unconstrained Views
★ 901UniScene3D. [ECCV 2026] Offical Repository of RGB-Pointmap Pretraining for Unified 3D Scene Understanding
★ 127scenetok. [CVPR '26] SceneTok: A Compressed, Diffusable Token Space for 3D Scenes
★ 205dreamzero. Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals
★ 2.5klingbot-depth. Masked Depth Modeling for Spatial Perception
★ 1.5klingbot-world. Advancing Open-source World Models
★ 4.3kNAS3R. [CVPR 2026] From None to All: Self-Supervised 3D Reconstruction via Novel View Synthesis
★ 115VicaSplat. "VicaSplat: A Single Run is All You Need for 3D Gaussian Splatting and Camera Estimation from Unposed Video Frames"
★ 93rerun. Visualize, query, and stream to train on multimodal robotics data.
★ 11kviser. Web-based 3D visualization in Python
★ 2.7kECC. The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
★ 236kHunyuanWorld-Mirror. [ICML 2026] WorldMirror: Fast and Universal 3D reconstruction model for versatile tasks
★ 1.2ktraj-dist-rs. trajectory similarity high performance solution in rust with python binding
★ 10MVSNet. MVSNet (ECCV2018) & R-MVSNet (CVPR2019)
★ 1.5krobustmvd. Repository for the Robust Multi-View Depth Benchmark
★ 113map-anything. MapAnything: Universal Feed-Forward Metric 3D Reconstruction
★ 3.6kARKitScenes. This repo accompanies the research paper, ARKitScenes - A Diverse Real-World Dataset for 3D Indoor Scene Understanding Using Mobile RGB-D Data and contains the data, scripts to visualize and process assets, and training code described in our paper.
★ 929co3d. Tooling for the Common Objects In 3D dataset.
★ 1.2kBlendedMVS. BlendedMVS: A Large-scale Dataset for Generalized Multi-view Stereo Networks
★ 680wildrgbd. Python
★ 101improved-diffusion. Release for Improved Denoising Diffusion Probabilistic Models
★ 3.8kguided-diffusion. Python
★ 7.4kglide-text2im. GLIDE: a diffusion-based text-conditional image synthesis model
★ 3.7kJiT. PyTorch implementation of JiT https://arxiv.org/abs/2511.13720
★ 2.5kWan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kREPA. [ICLR'25 Oral] Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think
★ 1.7kOpen-DiffusionGS. Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction (ICCV 2025)
★ 858VACE. [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editing
★ 3.9kLGM. [ECCV 2024 Oral] LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation.
★ 2.1kTripoSR. TripoSR: Fast 3D Object Reconstruction from a Single Image
★ 6.8k3DGPT. Python
★ 853infinigen. Infinite Photorealistic Worlds using Procedural Generation
★ 7.2kWorldGen. 🌍 WorldGen - Generate Any 3D Scene in Seconds
★ 2kSPFSplatV2. Official implementation of SPFSplatV2: Efficient Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views
★ 53lightly-studio. Curate, Annotate, and Manage Your Data in LightlyStudio.
★ 871LotteryCodec. Here is the resources and code for the LotteryCodec.
★ 28SPFSplat. Python
★ 2TrajDL. A python toolkit for deep learning in trajectory data mining
★ 13SPFSplat. [ICCV 2025 Highlight] No Pose at All: Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views
★ 153us-visa-bot. US Visa Bot
★ 269RCDS-parallel-programming-in-python. Jupyter Notebook
★ 16Hypo3D. ICML 2025 Hypo3D: Exploring Hypothetical Reasoning in 3D
★ 46GeometryCrafter. [ICCV 2025] GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors
★ 453Geo4D. [ICCV 2025 Highlight] Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction
★ 437DiffGS. [NeurIPS'2024]: DiffGS: Functional Gaussian Splatting Diffusion
★ 289PyTorch-VAE. A Collection of Variational Autoencoders (VAE) in PyTorch.
★ 7.7kDiffSplat. [ICLR 2025] Official implementation of "DiffSplat: Repurposing Image Diffusion Models for Scalable 3D Gaussian Splat Generation".
★ 534latentsplat. [ECCV 2024] Implementation of latentSplat: Autoencoding Variational Gaussians for Fast Generalizable 3D Reconstruction
★ 239vggt. [CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
★ 14kdiff-gaussian-rasterization. Differentiable gaussian rasterization with depth, alpha, normal map and extra per-Gaussian attributes, also support camera pose gradient
★ 317generative-models. Generative Models by Stability AI
★ 27kmast3r. Grounding Image Matching in 3D with MASt3R
★ 3.1kdust3r. DUSt3R: Geometric 3D Vision Made Easy
★ 7.3kHiNet. Official PyTorch implementation of "HiNet: Deep Image Hiding by Invertible Network" (ICCV 2021)
★ 205DeepMIH. Official repository of "DeepMIH: Deep Invertible Network for Multiple Image Hiding", TPAMI 2022.
★ 132bidastereo. [ECCV 2024] Match-Stereo-Videos: Bidirectional Alignment for Consistent Dynamic Stereo Matching.
★ 128InstantSplat. InstantSplat: Sparse-view SfM-free Gaussian Splatting in Seconds
★ 1.7kCF-3DGS. Python
★ 785real-state-10k. The real state 10k dataset from https://google.github.io/realestate10k
★ 613dgs_render_python. Python
★ 625diff-gaussian-rasterization. Cuda
★ 1.5kgaussian-splatting. Original reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"
★ 23kmvsplat. 🌊 [ECCV'24 Oral] MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images
★ 1.3kvggsfm. VGGSfM: Visual Geometry Grounded Deep Structure From Motion
★ 1.4kAwesome-LLM-3D. Awesome-LLM-3D: a curated list of Multi-modal Large Language Model in 3D world Resources
★ 2.2kpoint-e. Point cloud diffusion for 3D model synthesis
★ 6.9kstable-diffusion. A latent text-to-image diffusion model
★ 73kVAR. [NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". An *ultra-simple, user-friendly yet state-of-the-art* codebase for autoregressive image generation!
★ 8.7kCVinW_Readings. A collection of papers on the topic of ``Computer Vision in the Wild (CVinW)''
★ 1.4kLAVIS. LAVIS - A One-stop Library for Language-Vision Intelligence
★ 11kAwesome-Video-Diffusion. A curated list of recent diffusion models for video generation, editing, and various other applications.
★ 5.7kOpen-Sora-Plan. This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
★ 12kCtrl-Adapter. Official implementation of Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model (ICLR 2025 Oral)
★ 470stable-diffusion. Jupyter Notebook
★ 1.5kControlNet. Let us control diffusion models!
★ 34kdiffusers. 🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34kComfyUI. The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
★ 123kstable-diffusion-webui. Stable Diffusion web UI
★ 164kcnnimageretrieval-pytorch. CNN Image Retrieval in PyTorch: Training and evaluating CNNs for Image Retrieval in PyTorch
★ 1.5kdeep-image-retrieval. End-to-end learning of deep visual representations for image retrieval
★ 680awesome-domain-adaptation. A collection of AWESOME things about domain adaptation
★ 5.5kHierarchical-Localization. Visual localization made easy with hloc
★ 4.2kvisuallocalizationbenchmark. Python
★ 356pixloc. Back to the Feature: Learning Robust Camera Localization from Pixels to Pose (CVPR 2021)
★ 793MVSNet_pytorch. PyTorch Implementation of MVSNet
★ 682SegFormer. Official PyTorch implementation of SegFormer
★ 3.6kCIRKD. Official implementations of CIRKD: Cross-Image Relational Knowledge Distillation for Semantic Segmentation and implementations on Cityscapes, ADE20K, COCO-Stuff., Pascal VOC and CamVid.
★ 215NIID-Net. Code for "NIID-Net: Adapting Surface Normal Knowledge for Intrinsic Image Decomposition in Indoor Scenes" TVCG
★ 49deep-iterative-surface-normal-estimation. Code repository for our paper: Deep Iterative Surface Normal Estimation
★ 35IRS. IRS: A Large Synthetic Indoor Robotics Stereo Dataset for Disparity and Surface Normal Estimation
★ 56RGBD2Normal. Code for Deep Surface Normal Estimation with Hierarchical RGB-D Fusion (CVPR2019)
★ 76DPT. Dense Prediction Transformers
★ 2.3kRGBD_Semantic_Segmentation_PyTorch. [ECCV 2020] PyTorch Implementation of some RGBD Semantic Segmentation models.
★ 327Depth2HHA-python. Use python3 to convert depth image into hha image
★ 196moco. PyTorch implementation of MoCo: https://arxiv.org/abs/1911.05722
★ 5.1kpytorch-OpCounter. Count the MACs / FLOPs of your PyTorch model.
★ 5.1kyolact. A simple, fully convolutional model for real-time instance segmentation.
★ 5.2kSwin-Transformer-Semantic-Segmentation. This is an official implementation for "Swin Transformer: Hierarchical Vision Transformer using Shifted Windows" on Semantic Segmentation.
★ 1.3k21LPCVC-UAV_VIdeo_Track-Sample-Solution. Python
★ 25pointnet.pytorch. pytorch implementation for "PointNet: Deep Learning on Point Sets for 3D Classification and Segmentation" https://arxiv.org/abs/1612.00593
★ 2.4kstructure_knowledge_distillation. The official code for the paper 'Structured Knowledge Distillation for Semantic Segmentation'. (CVPR 2019 ORAL) and extension to other tasks.
★ 741DSRL. Implementation of CVPR 2020 Dual Super-Resolution Learning for Semantic Segmentation
★ 106opencv. Open Source Computer Vision Library
★ 90knetwork-slimming. Network Slimming (Pytorch) (ICCV 2017)
★ 919Deformable-ConvNets. Deformable Convolutional Networks
★ 4.1kUPSNet. UPSNet: A Unified Panoptic Segmentation Network
★ 645Deformable-Convolution-V2-PyTorch. Deformable ConvNets V2 (DCNv2) in PyTorch
★ 1.5kpytorch-superpoint. Superpoint Implemented in PyTorch: https://arxiv.org/abs/1712.07629
★ 933Class-balanced-loss-pytorch. Pytorch implementation of the paper "Class-Balanced Loss Based on Effective Number of Samples"
★ 803DALI. A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.
★ 5.7kalbumentations. Fast and flexible image augmentation library. Paper about the library: https://www.mdpi.com/2078-2489/11/2/125
★ 15kopenseg.pytorch. The official Pytorch implementation of OCNet, OCRNet, and SegFix.
★ 1.2kGCNv2_SLAM. Real-time SLAM system with deep features
★ 871semantic-segmentation. Nvidia Semantic Segmentation monorepo
★ 1.8kawesome-semantic-segmentation-pytorch. Semantic Segmentation on PyTorch (include FCN, PSPNet, Deeplabv3, Deeplabv3+, DANet, DenseASPP, BiSeNet, EncNet, DUNet, ICNet, ENet, OCNet, CCNet, PSANet, CGNet, ESPNet, LEDNet, DFANet)
★ 3.1kpytorch-distributed. A quickstart and benchmark for pytorch distributed training.
★ 1.7kapex. A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch
★ 9kDANet. Dual Attention Network for Scene Segmentation (CVPR2019)
★ 2.5knvidia-docker. Build and run Docker containers leveraging NVIDIA GPUs
★ 18kpytorch-sync-batchnorm-example. How to use Cross Replica / Synchronized Batchnorm in Pytorch
★ 248HRNet-Semantic-Segmentation. The OCR approach is rephrased as Segmentation Transformer: https://arxiv.org/abs/1909.11065. This is an official implementation of semantic segmentation for HRNet. https://arxiv.org/abs/1908.07919
★ 3.3ksemantic-segmentation-pytorch. Pytorch implementation for Semantic Segmentation/Scene Parsing on MIT ADE20K dataset
★ 5.1kflops-counter.pytorch. Flops counter for neural networks in pytorch framework
★ 3kdeform_conv_pytorch. PyTorch Implementation of Deformable Convolution
★ 289Attention-Gated-Networks. Use of Attention Gates in a Convolutional Neural Network / Medical Image Classification and Segmentation
★ 2.1kRACNN-pytorch. This is a third party implementation of RA-CNN in pytorch.
★ 205fast-reid. SOTA Re-identification Methods and Toolbox
★ 4kNon-local_pytorch. Implementation of Non-local Block.
★ 1.6kWS-DAN.PyTorch. A PyTorch implementation of WS-DAN (Weakly Supervised Data Augmentation Network) for FGVC (Fine-Grained Visual Classification)
★ 410LearnToPayAttention. PyTorch implementation of ICLR 2018 paper Learn To Pay Attention (and some modification)
★ 169carla. Open-source simulator for autonomous driving research.
★ 14kawesome-autonomous-vehicles. Curated List of Self-Driving Cars and Autonomous Vehicles Resources
★ 2.4kAutonomousVehiclePaper. 无人驾驶相关论文速递
★ 496CAL. [CoRL'18] Conditional Affordance Learning
★ 69research. dataset and code for 2016 paper "Learning a Driving Simulator"
★ 4.1kmmdetection. OpenMMLab Detection Toolbox and Benchmark
★ 33kclass-balanced-loss. Class-Balanced Loss Based on Effective Number of Samples. CVPR 2019
★ 616DCL. Destruction and Construction Learning for Fine-grained Image Recognition
★ 592nitrain. Train AI models efficiently on medical images using any framework
★ 1.9kself-ensemble-visual-domain-adapt. Code repository for the small image experiments our paper 'Self-ensembling for Domain Adaptation'
★ 193senet.pytorch. PyTorch implementation of SENet
★ 2.3kreid-strong-baseline. Bag of Tricks and A Strong Baseline for Deep Person Re-identification
★ 2.4kfast-MPN-COV. @CVPR2018: Efficient unrolling iterative matrix square-root normalized ConvNets, implemented by PyTorch (and code of B-CNN,Compact bilinear pooling etc.) for training from scratch & finetuning.
★ 278siamese-triplet. Siamese and triplet networks with online pair/triplet mining in PyTorch
★ 3.2kpytorch-cifar. 95.47% on CIFAR10 with PyTorch
★ 6.4kdeep-residual-networks. Deep Residual Learning for Image Recognition
★ 6.7kcvpr18-inaturalist-transfer. Large Scale Fine-Grained Categorization and Domain-Specific Transfer Learning. CVPR 2018
★ 196inat_comp. iNaturalist competition details
★ 813pytorch-vgg-cifar10. This is the PyTorch implementation of VGG network trained on CIFAR10 dataset
★ 360pytorch-cifar100. Practice on cifar100(ResNet, DenseNet, VGG, GoogleNet, InceptionV3, InceptionV4, Inception-ResNetv2, Xception, Resnet In Resnet, ResNext,ShuffleNet, ShuffleNetv2, MobileNet, MobileNetv2, SqueezeNet, NasNet, Residual Attention Network, SENet, WideResNet)
★ 4.8kMobileNet-Caffe. Caffe Implementation of Google's MobileNets (v1 and v2)
★ 1.3kcaffe. Caffe: a fast open framework for deep learning.
★ 35kcaffe-mobilenet. A caffe implementation of mobilenet's depthwise convolution layer.
★ 145models. Models and examples built with TensorFlow
★ 78kexamples. A set of examples around pytorch in Vision, Text, Reinforcement Learning, etc.
★ 24kpytorch-cnn-visualizations. Pytorch implementation of convolutional neural network visualization techniques
★ 8.2ktensorflow. An Open Source Machine Learning Framework for Everyone
★ 197kJDDA-Master. Joint Domain Alignment and Discriminative Feature Learning for Unsupervised Deep Domain Adaptation(AAAI-2019)
★ 83Deep-Mutual-Learning. TensorFlow Implementation of Deep Mutual Learning
★ 325DANN. pytorch implementation of Domain-Adversarial Training of Neural Networks
★ 951adversarial-feature-augmentation. Code for the paper "Adversarial Feature Augmentation for Unsupervised Domain Adaptation", CVPR 2018
★ 131domain_adversarial_neural_network. Domain Adaptation Representation Learning Algorithm (as published in JMLR 2016)
★ 142fastswa-semi-sup. Improving Consistency-Based Semi-Supervised Learning with Weight Averaging
★ 190STL10. Python utilities for reading the STL-10 dataset: http://cs.stanford.edu/~acoates/stl10/
★ 125realistic-ssl-evaluation. Open source release of the evaluation benchmark suite described in "Realistic Evaluation of Deep Semi-Supervised Learning Algorithms"
★ 460transferlearning. Transfer learning / domain adaptation / domain generalization / multi-task learning etc. Papers, codes, datasets, applications, tutorials.-迁移学习
★ 14karcface-pytorch. Python
★ 1.9kVAT-pytorch. Virtual Adversarial Training (VAT) implementation for PyTorch
★ 296ladder. Ladder network is a deep learning algorithm that combines supervised and unsupervised learning
★ 518mean-teacher. A state-of-the-art semi-supervised method for image recognition
★ 1.7kCutout. 2.56%, 15.20%, 1.30% on CIFAR10, CIFAR100, and SVHN https://arxiv.org/abs/1708.04552
★ 558mixup-cifar10. mixup: Beyond Empirical Risk Minimization
★ 1.2kESPNet. ESPNet: Efficient Spatial Pyramid of Dilated Convolutions for Semantic Segmentation
★ 542espnet. End-to-End Speech Processing Toolkit
★ 9.9kpytorch-generative-model-collections. Collection of generative models in Pytorch version.
★ 2.6kpretrained-models.pytorch. Pretrained ConvNets for pytorch: NASNet, ResNeXt, ResNet, InceptionV4, InceptionResnetV2, Xception, DPN, etc.
★ 9.1kpytorch-playground. Base pretrained models and datasets in pytorch (MNIST, SVHN, CIFAR10, CIFAR100, STL10, AlexNet, VGG16, VGG19, ResNet, Inception, SqueezeNet)
★ 2.7kresnet-protofiles. Caffe Protofiles for MSRA ResNet: train prototxt
★ 214ResNet-on-Cifar10. Reimplementation ResNet on cifar10 with caffe
★ 129sphereface. Implementation for <SphereFace: Deep Hypersphere Embedding for Face Recognition> in CVPR'17.
★ 1.6kcaffe-windows. Configure Caffe in one hour for Windows users.
★ 1.3karcface-caffe. insightface-caffe
★ 279GoogleNet-BN. GoogleNet-BN,namely InceptionNetV2 based on pytorch.
★ 14hourglass-facekeypoints-detection. face keypoints deteciton based on stackedhourglass
★ 254caffe_to_torch_to_pytorch. Python
★ 150cbp. Multimodal Compact Bilinear Pooling for Torch7
★ 70SENet. Squeeze-and-Excitation Networks
★ 3.6kCliqueNet. Convolutional Neural Networks with Alternately Updated Clique (to appear in CVPR 2018)
★ 327