This is your work, valued
MVSNet_pytorch. PyTorch Implementation of MVSNet
★ 682GwcNet. Group-wise Correlation Stereo Network, CVPR 2019
★ 338Learning-Monocular-Depth-by-Stereo. Learning Monocular Depth by Distilling Cross-domain Stereo Networks, ECCV18
★ 98LIGA-Stereo. Code for LIGA-Stereo Detector, ICCV'21
★ 96github_bot_3d_papers. Call Arxiv API and automatically update paper list
★ 18mmdetection_kitti. 2D detection on KITTI dataset. see configs/kitti
★ 15food_management_api. a simple food management system for own usage. 家庭食物管理系统,避免食物浪费
★ 1awesome-world-action-models.
★ 305Awesome-Video-Action-Model.
★ 74wall-x. Building General-Purpose Robots Based on Embodied Foundation Model
★ 1.2kpeft. 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
★ 21kPi3. [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning
★ 2.1kEpona. Official Code for Epona: Autoregressive Diffusion World Model for Autonomous Driving (ICCV 2025)
★ 376taehv. Tiny AutoEncoder for Hunyuan Video (and other video models)
★ 4513dgrut. Ray tracing and hybrid rasterization of Gaussian particles
★ 2.3kTrajectoryCrafter. [ICCV 2025, Oral] TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
★ 861Wan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kGoalFlow. Repo of "GoalFlow: Goal-Driven Flow Matching for Multimodal Trajectories Generation in End-to-End Autonomous Driving"
★ 410lobehub. 🤯 LobeHub is your Chief Agent Operator, organizing your agents into 7×24 operations by hiring, scheduling, and reporting on your entire AI team.
★ 81kEasy3D. A lightweight, easy-to-use, and efficient library for processing and rendering 3D data (C++ & Python)
★ 1.6kskorch. A scikit-learn compatible neural network library that wraps PyTorch
★ 6.2kgpt-researcher. An autonomous agent that conducts deep research on any data using any LLM providers
★ 29kJanus. Janus-Series: Unified Multimodal Understanding and Generation Models
★ 18kawesome-deepseek-integration. Integrate the DeepSeek API into popular software
★ 38kHunyuan3D-2. High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models.
★ 14kIGCS-GITC. Camera tools based on Frans Bouma's IGCS framework.
★ 82MagicDrive-V2. [ICCV 2025] Official implementation of the paper “MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control”
★ 719DrivingWorld. Code for "DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT"
★ 245SS3DM-Exporter. [NeurIPS 2024] Data exporter for SS3DM: Benchmarking Street-View Surface Reconstruction with a Synthetic 3D Mesh Dataset
★ 16gs2mesh. [ECCV 2024] Official implementation of the paper "GS2Mesh: Surface Reconstruction from Gaussian Splatting via Novel Stereo Views"
★ 353mvsplat. 🌊 [ECCV'24 Oral] MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images
★ 1.3kfilebrowser. File Browser provides a file managing interface within a specified directory and it can be used to upload, delete, preview and edit your files.
★ 36klibcom. Image composition toolbox: everything you want to know about image composition/compositing or object/subject insertion/addition/compositing.
★ 734drivestudio. A 3DGS framework for omni urban scene reconstruction and simulation.
★ 1.2kGeoMaster. GeoMaster: Advanced Geometry Enhancement for High-Resolution 3D Modeling
★ 88omg-tools. Optimal Motion Generation-tools: motion planning made easy
★ 595ReconX. [TIP 2026] ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model
★ 709webdataset. A high-performance Python-based I/O system for large (and small) deep learning problems, with strong support for PyTorch.
★ 3.2kviser. Web-based 3D visualization in Python
★ 2.7kRealityCapture-to-Postshot. Process for processing Reality Capture output into Colmap format to allow importing into Postshot.
★ 36GaussianSplats3D. Three.js-based implementation of 3D Gaussian splatting
★ 2.8kMetalSplatter. Render Gaussian Splats using Metal on Apple platforms (iOS/iPhone/iPad, macOS, and visionOS)
★ 708VEnhancer. Official codes of VEnhancer: Generative Space-Time Enhancement for Video Generation
★ 577msplat. A modular differential gaussian rasterization library.
★ 214hierarchical-3d-gaussians. Official implementation of the SIGGRAPH 2024 paper "A Hierarchical 3D Gaussian Representation for Real-Time Rendering of Very Large Datasets"
★ 1.5kmar. PyTorch implementation of MAR+DiffLoss https://arxiv.org/abs/2406.11838
★ 1.9kNDC. PyTorch implementation of Neural Dual Contouring.
★ 280wayve_scenes. Codebase for the WayveScenes101 Dataset
★ 194street-gaussians-ns. Unofficial implementation of "Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting", ECCV2024.
★ 389DriveDreamer2. [AAAI 2025] DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation
★ 267CAD. CAD: Photorealistic 3D Generation via Adversarial Distillation (CVPR 2024)
★ 129MoneyPrinterTurbo. 利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.
★ 101kLightDiff. About [CVPR 2024] The official implementation of paper " Light the Night: A Multi-Condition Diffusion Framework for Unpaired Low-Light Enhancement in Autonomous Driving"
★ 85Open-Sora-Plan. This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
★ 12ksd-forge-layerdiffuse. [WIP] Layer Diffusion for WebUI (via Forge)
★ 4.1kgenerative-ai-for-beginners. 21 Lessons, Get Started Building with Generative AI
★ 114kEfficientSAM. EfficientSAM: Leveraged Masked Image Pretraining for Efficient Segment Anything
★ 2.5kDepth-Anything. [CVPR 2024] Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data. Foundation Model for Monocular Depth Estimation
★ 8.2kGaussianAvatar. [CVPR 2024] The official repo for "GaussianAvatar: Towards Realistic Human Avatar Modeling from a Single Video via Animatable 3D Gaussians"
★ 600supersplat. 3D Gaussian Splat Editor
★ 9.8kstreet_gaussians. [ECCV 2024] Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting
★ 1.4kdreamgaussian4d. [arXiv 2023] DreamGaussian4D: Generative 4D Gaussian Splatting
★ 619ra-neus. RaNeuS
★ 77SyncMVD. Official PyTorch & Diffusers implementation of "Text-Guided Texturing by Synchronized Multi-View Diffusion"
★ 181MERFStudio. A Unified Framework for Real-Time Rendering on the Web
★ 88StreamDiffusion. StreamDiffusion: A Pipeline-Level Solution for Real-Time Interactive Generation
★ 11kground_normal_filter. Official implementation of: Towards Accurate Ground Plane Normal Estimation from Ego-Motion
★ 67awesome-3D-gaussian-splatting. Curated list of papers and resources focused on 3D Gaussian Splatting, intended to keep pace with the anticipated surge of research in the coming months.
★ 8.8kAdaBins. Official implementation of Adabins: Depth Estimation using adaptive bins
★ 784awesome-knowledge-driven-AD. A curated list of awesome knowledge-driven autonomous driving (continually updated)
★ 500mvs-texturing. Algorithm to texture 3D reconstructions from multi-view stereo images
★ 1.1kChatSim. [CVPR2024 Highlight] Editable Scene Simulation for Autonomous Driving via LLM-Agent Collaboration
★ 430ganerf. GANeRF: Leveraging Discriminators to Optimize Neural Radiance Fields
★ 84neurad-studio. [CVPR2024] NeuRAD: Neural Rendering for Autonomous Driving
★ 492Awesome-Image-Harmonization. A curated list of papers, code and resources pertaining to image harmonization.
★ 536RoMe. Python
★ 279Awesome-LLM-Robotics. A comprehensive list of papers using large language/multi-modal models for Robotics/RL, including papers, codes, and related websites
★ 4.4kDriveLM. [ECCV 2024 Oral] DriveLM: Driving with Graph Visual Question Answering
★ 1.3kLAVIS. LAVIS - A One-stop Library for Language-Vision Intelligence
★ 11kICCV-2023-25-Papers. ICCV 2023-2025 Papers: Discover cutting-edge research from ICCV 2023-25, the leading computer vision conference. Stay updated on the latest in computer vision and deep learning, with code included. ⭐ support visual intelligence development!
★ 968LISA. Project Page for "LISA: Reasoning Segmentation via Large Language Model"
★ 2.7kArcNerf. Nerf and extensions in all
★ 107transferlearning. Transfer learning / domain adaptation / domain generalization / multi-task learning etc. Papers, codes, datasets, applications, tutorials.-迁移学习
★ 14kmaplab. A Modular and Multi-Modal Mapping Framework
★ 2.9kLIO-SAM. LIO-SAM: Tightly-coupled Lidar Inertial Odometry via Smoothing and Mapping
★ 1LandMark. Python
★ 484Awesome-Occupancy-Prediction-Autonomous-Driving. Awesome papers about Multi-Camera Semantic Occupancy Prediction, such as TPVFormer, OccFormer, Occ3D, OpenOccupancy
★ 268omnimotion. Python
★ 2.3kcolmap--Important-code-is-to-parse-line-by-line. 教你一点点掌握视觉三维重建-colmap 重要代码逐行解析(本人利用下班和周末时间update,so 速度会慢)
★ 211VMA. A general map auto annotation framework based on MapTR, with high flexibility in terms of spatial scale and element type
★ 325Strivec. Official code for the paper: Strivec (ICCV2023)
★ 146SAM3D. [SCIS] SAM3D: Zero-Shot 3D Object Detection via Segment Anything Model
★ 228raytracing. A CUDA Mesh RayTracer with BVH acceleration, with python bindings and a GUI.
★ 129NeRF-Texture. [SIGGRAPH 2023, TPAMI 2024] Code for NeRF-Texture: Texture Synthesis with Neural Radiance Fields
★ 221openMVS. open Multi-View Stereo reconstruction library
★ 4.1kExplicit-Visual-Prompt. [CVPR 2023 & TPAMI 2025] Explicit Visual Prompting for Low-Level Structure Segmentations
★ 231neuralsim. neuralsim: 3D surface reconstruction and simulation based on 3D neural rendering.
★ 671DragGAN. Official Code for DragGAN (SIGGRAPH 2023)
★ 36kConsistentNeRF. ConsistentNeRF Enhances Neural Radiance Fields with 3D Consistency for Sparse View Synthesis
★ 75threestudio. A unified framework for 3D content generation.
★ 7kpixel-perfect-sfm. Pixel-Perfect Structure-from-Motion with Featuremetric Refinement (ICCV 2021, Best Student Paper Award)
★ 1.5kHierarchical-Localization. Visual localization made easy with hloc
★ 4.2kagi2nerf. Simple tool for converting Agisoft XML files to NERF JSON files for https://github.com/NVlabs/instant-ngp
★ 118NeRO. [SIGGRAPH2023] NeRO: Neural Geometry and BRDF Reconstruction of Reflective Objects from Multiview Images
★ 599SegmentAnything3D. [ICCV'23 Workshop] SAM3D: Segment Anything in 3D Scenes
★ 1.4kACMH. A simple yet effective PatchMatch MVS method, which is the base model of ACMM and ACMP.
★ 99bitsandbytes. Accessible large language models via k-bit quantization for PyTorch.
★ 8.4k