This is your work, valued
stereo_toolbox. 🔧 A comprehensive stereo matching toolbox for efficient development and research.
★ 155ADL. [CVPR 2024] Adaptive Multi-Modal Cross-Entropy Loss for Stereo Matching
★ 63MIDAS. [ICCV 2025] MIDAS: Modeling Ground-Truth Distributions with Dark Knowledge for Domain Generalized Stereo Matching
★ 6BaCon-Stereo.
★ 1xxxupeng.github.io. CSS
★ 1S3-Stereo.
★ 1notes.
★ 1Three.js-Object-Sculptor-Codex-Plugin. Codex plugin that turns attached object images into code-only, animation-ready procedural Three.js models.
★ 1.5kCLIProxyAPIPlus. CCS-maintained fork of CLIProxyAPIPlus (MIT snapshot, Apr 2026) with daily auto-sync from router-for-me/CLIProxyAPI. See plans for context.
★ 211Research-Paper-Writing-Skills. Skill package for ML/CV/NLP paper writing, curated and adapted from Prof. Peng Sida's open notes for Codex, Claude Code, and Gemini.
★ 5.7kpaper-framework-figure-studio-pro. A multi-round co-design skill for publication-ready paper framework diagrams and method overview figures.
★ 1.8kAuto-claude-code-research-in-sleep. ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works with Claude Code, Codex, OpenClaw, or any LLM agent.
★ 14klingbot-map. A feed-forward 3D foundation model for reconstructing scenes from streaming data
★ 16klingbot-world-v2. Infinite Worlds with Versatile Interactions
★ 1.4ksjh-skills. Shell
★ 36lingbot-vision. Self-supervised learning for spatial perception
★ 884FlashVDM. Unleashing Vecset Diffusion Model for Fast Shape Generation / within 1 Second (ICCV'25 Highlight)
★ 333laion-3d. Collect large 3d dataset and build models
★ 298DreamX-World. DreamX-World: A General-Purpose Interactive World Model
★ 738SOMA-X. SOMA: Unifying Parametric Human Body Models
★ 714github-readme-stats. :zap: Dynamically generated stats for your github readmes
★ 80kSana. SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
★ 8.6kABot-Earth-0.5. Generative 3D Earth Model by AMap-cvlab
★ 190VisionReward. [AAAI 2026] VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation
★ 422ImageReward. [NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences for Text-to-image Generation
★ 1.7kDeepPaperNote. DeepPaperNote is an agent skill for deep-reading a single paper and generating high-quality Obsidian-style research notes. Works with Claude Code, Codex, Cursor, Copilot, Gemini CLI, and more.
★ 566vipe. ViPE: Video Pose Engine for Geometric 3D Perception
★ 2.1kMatrix-Game. Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
★ 2.3kzjuthesis. Zhejiang University Graduation Thesis LaTeX Template
★ 1superpowers. An agentic skills framework & software development methodology that works.
★ 264kzju-connect. ZJU RVPN 客户端的 Go 语言实现
★ 709how-to-use-claude-the-smart-way.
★ 11IndustryBench. A multi-lingual benchmark for evaluating industrial domain knowledge of LLMs.
★ 155talkie. talkie is a vintage language model from 1930
★ 959CarlaAir. CarlaAir: Fly Drones Inside a CARLA World!! A Unified Infrastructure for Air-Ground Embodied Intelligence
★ 1.1kVeOmni. VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
★ 2.1kMLFocalLengths. Estimating the Focal Length of a Monocular Image
★ 39GeoCalib. GeoCalib: Learning Single-image Calibration with Geometric Optimization (ECCV 2024)
★ 879TeleBoost. A Systematic Alignment Framework for High-Fidelity, Controllable, and Robust Video Generation.
★ 10StereoWorld. [CVPR 2026] Stereo World Model
★ 83claw-code. An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
★ 195klearn-coding-agent. Research on Coding Agents
★ 12kPhyGenesis. JavaScript
★ 33cosmos-predict2.5. Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the future state of the world in the form of video.
★ 1.3kfantasy-world. [ICLR 2026] FantasyWorld: Geometry-Consistent World Modeling via Unified Video and 3D Prediction
★ 443cosmos-predict1. Cosmos-Predict1 is a collection of general-purpose world foundation models for Physical AI that can be fine-tuned into customized world models for downstream applications.
★ 465seoul-world-model. Seoul World Model: Grounding World Simulation Models in a Real-World Metropolis
★ 622OpenEMMA. OpenEMMA, a permissively licensed open source "reproduction" of Waymo’s EMMA model.
★ 946font.
★ 312RAD. [NeurIPS 2025] RAD: Training an End-to-End Driving Policy via Large-Scale 3DGS-based Reinforcement Learning
★ 265ReconDreamer-RL. Python
★ 98openclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 385kawesome-openclaw-skills. The awesome collection of OpenClaw skills. 5,400+ skills filtered and categorized from the official OpenClaw Skills Registry.🦞
★ 52kgpt_academic. 为GPT/GLM等LLM大语言模型提供实用化交互接口,特别优化论文阅读/润色/写作体验,模块化设计,支持自定义快捷按钮&函数插件,支持Python和C++等项目剖析&自译解功能,PDF/LaTex论文翻译&总结功能,支持并行问询多种LLM模型,支持chatglm3等本地模型。接入通义千问, deepseekcoder, 讯飞星火, 文心一言, llama2, rwkv, claude2, moss等。
★ 71klingbot-va. [RSS 2026] Causal video-action world model for generalist robot control
★ 1.7klingbot-depth. Masked Depth Modeling for Spatial Perception
★ 1.5klingbot-vla. A Pragmatic VLA Foundation Model
★ 1.7klearning_research. 本人的科研经验
★ 14kAwesome-Foundation-Model-Based-Depth. A curated list of papers & resources on depth estimation & depth completion with foundation models.
★ 10lingbot-world. Advancing Open-source World Models
★ 4.3kAwesome-Foundation-Models. A curated list of foundation models for vision and language tasks
★ 1.2kAwesome-CV-Foundational-Models.
★ 549S3-Stereo.
★ 1Awesome-World-Models. A comprehensive list of papers for the definition of World Models and using World Models for General Video Generation, Embodied AI, and Autonomous Driving, including papers, codes, and related websites.
★ 1.9kAwesome-World-Model. Collect some World Models for Autonomous Driving (and Robotic, etc.) papers.
★ 2.2kGANet. GA-Net: Guided Aggregation Net for End-to-end Stereo Matching
★ 561iREPA. [ICLR 2026] Official implementation for What matters for Representation Alignment: Global Information or Spatial Structure?
★ 258scared_toolkit. Unofficial SCARED Dataset Toolkit
★ 57sam-3d-body. The repository provides code for running inference with the SAM 3D Body Model (3DB), links for downloading the trained model checkpoints and datasets, and example notebooks that show how to use the model.
★ 3.4ksam-3d-objects. SAM 3D Objects
★ 7.2kZeroStereo. [ICCV 2025] ZeroStereo: Zero-Shot Stereo Matching from Single Images
★ 61SVP. Python
★ 2gemini-cli. An open-source AI agent that brings the power of Gemini directly into your terminal.
★ 106kawesome-dust3r. 🌟A curated list of DUSt3R-related papers and resources, tracking recent advancements using this geometric foundation model.
★ 802xxxupeng.github.io. CSS
★ 1flowseek. Source code for ICCV 2025 paper "FlowSeek: Optical Flow Made Easier with Depth Foundation Models and Motion Bases"
★ 154llm_interview_note. 主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题
★ 15ktiny-llm-zh. 从零实现一个小参数量中文大语言模型。
★ 1.1kFastVGGT. [ICLR 2026] FastVGGT: Fast Visual Geometry Transformer
★ 806DMS. Using Stable Diffusion Model for generating multi-baseline images for autonomous driving scenes like KITTI.
★ 9stereo_toolbox. 🔧 A comprehensive stereo matching toolbox for efficient development and research.
★ 155BridgeDepth. [ICCV 2025 Highlight] BridgeDepth: Bridging Monocular and Stereo Reasoning with Latent Alignment
★ 158Amodal-Depth-Anything. [ICCV 2025] Amodal Depth Anything: Amodal Depth Estimation in the Wild
★ 43Geo4D. [ICCV 2025 Highlight] Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction
★ 437unimatch. [TPAMI'23] Unifying Flow, Stereo and Depth Estimation
★ 1.4kSMoE-Stereo. [ICCV 2025 Highlight] 🌟🌟🌟 Learning Robust Stereo Matching in the Wild with Selective Mixture-of-Experts
★ 187HunyuanWorld-1.0. Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels with Hunyuan3D World Model
★ 2.9kdynamic-3dgs-papers. Dynamic Scene Representation Gaussian Splatting
★ 40Resume-Matcher. The #1 AI Harness for Building Resumes, PDFs, Cover Letters & more, locally with 100+ LLMs support.
★ 28kCLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34kGwcNet. Group-wise Correlation Stereo Network, CVPR 2019
★ 338SeeingThroughFog. Python
★ 378read-frog. 🐸 Read Frog - Language Learning & Translate | 🐸 陪读蛙 - 语言学习与翻译
★ 8.8kaanet. [CVPR'20] AANet: Adaptive Aggregation Network for Efficient Stereo Matching
★ 559vggt. [CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
★ 14kfluentui-system-icons. Fluent System Icons are a collection of familiar, friendly and modern icons from Microsoft.
★ 11kHunyuan3D-2.1. From Images to High-Fidelity 3D Assets with Production-Ready PBR Material
★ 3.8kFoundationStereo. [CVPR 2025 Best Paper Nomination] FoundationStereo: Zero-Shot Stereo Matching
★ 2.8kLLaNA. Official code repository of LLaNA: Large Language and NeRF Assistant
★ 19MonSter. 【CVPR 2025 Highlight】MonSter: Marry Monodepth to Stereo Unleashes Power
★ 699DEFOM-Stereo. [CVPR 2025] DEFOM-Stereo: Depth foundation model based stereo matching
★ 286MaterialMVP. MaterialMVP: Illumination-Invariant Material Generation via Multi-view PBR Diffusion
★ 206MIDAS. [ICCV 2025] MIDAS: Modeling Ground-Truth Distributions with Dark Knowledge for Domain Generalized Stereo Matching
★ 6REFRAME. This repository is the official implementation for the paper “REFRAME: Reflective Surface Real-Time Rendering for Mobile Devices”.
★ 22Task-Conditional-Adapter. [MM 2024] Task-Conditional Adapter for Multi-Task Dense Prediction
★ 4stereoanywhere. [CVPR 2025] Stereo Anywhere: Robust Zero-Shot Deep Stereo Matching Even Where Either Stereo or Mono Fail
★ 278Awesome-Multimodal-Large-Language-Models. :sparkles::sparkles:Latest Advances on Multimodal Large Language Models
★ 18kholopix50k. Holopix50k: A Large-Scale In-the-wild Stereo Image Dataset
★ 161Janus. Janus-Series: Unified Multimodal Understanding and Generation Models
★ 18kADL. [CVPR 2024] Adaptive Multi-Modal Cross-Entropy Loss for Stereo Matching
★ 63AlchemyCat. Alchemy Cat —— 🔥Config System for SOTA
★ 111Awesome-Deep-Stereo-Matching. A curated list of awesome Deep Stereo Matching resources
★ 597awesome-4d-generation. List of papers on 4D Generation.
★ 327awesome-3d-diffusion. A collection of papers on diffusion models for 3D generation.
★ 1.3kGeoWizard. [ECCV'24] GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Image
★ 938cs-self-learning. 计算机自学指南
★ 75kBrachioGraph. BrachioGraph is an ultra-cheap (total cost of materials: €14) plotter that can be built with minimal skills.
★ 748