This is your work, valued
PatchmatchNet. Official code of PatchmatchNet (CVPR 2021 Oral)
★ 554IterMVS. Official code of IterMVS (CVPR 2022)
★ 169img2threejs. Rebuild the object in a reference image as a code-only, procedural, quality-gated, animation-ready Three.js model. Token-efficient image-to-3D.
★ 8.8kAwesome-Interactive-World-Model. A Comprehensive Survey of Interactive Video World Models
★ 215STEM2M. Morphology prediction of small nanoparticles from single HAADF-STEM images
★ 2OSI-Bench. Official repo of From Indoor to Open World: Revealing the Spatial Reasoning Gap in MLLMs
★ 24DepthLM_Official. [ICLR 2026 Oral (top 1.2%)] Official implementation of DepthLM
★ 365diffmvs. [T-PAMI 2025] DiffMVS & CasDiffMVS
★ 174nwm. Official code for the CVPR 2025 paper "Navigation World Models".
★ 662ViewCrafter. [TPAMI 2025] ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis
★ 1.6kTRELLIS. Official repo for paper "Structured 3D Latents for Scalable and Versatile 3D Generation" (CVPR'25 Spotlight).
★ 13kPyramid-Flow. [ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
★ 3.2kscrstudio. [CVPR 2025] A unified framework for Scene Coordinate Regression-based visual localization
★ 142HSR. HSR: Holistic 3D Human-Scene Reconstruction from Monocular Videos (ECCV 2024)
★ 34unilm. Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
★ 22kdepthsplat. [CVPR'25] DepthSplat: Connecting Gaussian Splatting and Depth
★ 1.2kopenvla. OpenVLA: An open-source vision-language-action model for robotic manipulation.
★ 6.7kdust3r. DUSt3R: Geometric 3D Vision Made Easy
★ 7.3kmast3r. Grounding Image Matching in 3D with MASt3R
★ 3.1kPandora. Pandora: Towards General World Model with Natural Language Actions and Video States
★ 538glace. [CVPR 2024] GLACE: Global Local Accelerated Coordinate Encoding
★ 106home-robot. Mobile manipulation research tools for roboticists
★ 1.2kopen_clip. An open source implementation of CLIP.
★ 14kAwesome-LLM-Robotics. A comprehensive list of papers using large language/multi-modal models for Robotics/RL, including papers, codes, and related websites
★ 4.4kdeep-video-mvs. Code for "DeepVideoMVS: Multi-View Stereo on Video with Recurrent Spatio-Temporal Fusion" (CVPR 2021)
★ 233AdaBins. Official implementation of Adabins: Depth Estimation using adaptive bins
★ 784VolRecon. Official code of VolRecon (CVPR 2023)
★ 155IterMVS. Official code of IterMVS (CVPR 2022)
★ 169StableViewSynthesis. Python
★ 214MiDaS. Code for robust monocular depth estimation described in "Ranftl et. al., Towards Robust Monocular Depth Estimation: Mixing Datasets for Zero-shot Cross-dataset Transfer, TPAMI 2022"
★ 5.4kPatchmatchNet. Official code of PatchmatchNet (CVPR 2021 Oral)
★ 554