This is your work, valued
Be the change you want to see.
Awesome-World-Model. Collect some World Models for Autonomous Driving (and Robotic, etc.) papers.
โ 2.2kopenclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. ๐ฆ
โ 385kflow_matching. A PyTorch library for implementing flow matching algorithms, featuring continuous and discrete flow matching implementations. It includes practical examples for both text and image modalities.
โ 4.7kQwen3-VL. Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
โ 20kEnd-to-end-Autonomous-Driving. [IEEE T-PAMI 2024] All you need for End-to-end Autonomous Driving
โ 3.7kHands-on-RL. https://hrl.boyuai.com/
โ 4.9kLLM-Trading-Lab. This repo powers my experiment where ChatGPT manages a real-money micro-cap stock portfolio.
โ 7.5kMoving-Least-Squares. Numpy & PyTorch implementation of three algorithms of image deformation using moving least squares. http://dl.acm.org/citation.cfm?doid=1179352.1141920
โ 364LoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
โ 14kLLaVA. [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
โ 25kGrounded-Segment-Anything. Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
โ 18kMDP-Diffusion. Implementation of MDP: A Generalized Framework for Text-Guided Image Editing by Manipulating the Diffusion Path
โ 67LaneSegNet. [ICLR 2024] Map Learning with Lane Segment for Autonomous Driving
โ 379faiss. A library for efficient similarity search and clustering of dense vectors.
โ 41kpollinations. Your Friendly Open-Source Gen-AI Platform
โ 4.9kBaichuan-13B. A 13B large language model developed by Baichuan Intelligent Technology
โ 2.9kNeuS. Code release for NeuS
โ 1.8kTransDepth. Code for Transformers Solve Limited Receptive Field for Monocular Depth Prediction
โ 174mmsegmentation. OpenMMLab Semantic Segmentation Toolbox and Benchmark.
โ 9.9kDenseMatchingBenchmark. Dense Matching Benchmark
โ 174lanedet. An open source lane detection toolbox based on PyTorch, including SCNN, RESA, UFLD, LaneATT, CondLane, etc.
โ 625fiery. PyTorch code for the paper "FIERY: Future Instance Segmentation in Bird's-Eye view from Surround Monocular Cameras"
โ 609Point2Skeleton. Point2Skeleton: Learning Skeletal Representations from Point Clouds (CVPR2021)
โ 241medial-axis-transform. Medial axis transform using Delaunay triangulations
โ 12manhattan_sdf. Code for "Neural 3D Scene Reconstruction with the Manhattan-world Assumption" CVPR 2022 Oral
โ 532runx. Deep Learning Experiment Management
โ 641awesome-mvs. A curated list of tutorials, papers, software related to multi-view stereo.
โ 279MonoScene. [CVPR 2022] "MonoScene: Monocular 3D Semantic Scene Completion": 3D Semantic Occupancy Prediction from a single image
โ 818SurroundOcc. [ICCV 2023] SurroundOcc: Multi-camera 3D Occupancy Prediction for Autonomous Driving
โ 1.1ksimplerecon. [ECCV 2022] SimpleRecon: 3D Reconstruction Without 3D Convolutions
โ 1.4kgmflow. [CVPR'22 Oral] GMFlow: Learning Optical Flow via Global Matching
โ 796openpose. OpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation
โ 34kunimatch. [TPAMI'23] Unifying Flow, Stereo and Depth Estimation
โ 1.4kACMMP. Multi-Scale Geometric Consistency Guided and Planar Prior Assisted Multi-View Stereo (TPAMI 2022)
โ 208TransMVSNet. (CVPR 2022) TransMVSNet: Global Context-aware Multi-view Stereo Network with Transformers.
โ 284ControlNet. Let us control diffusion models!
โ 34ksd-webui-controlnet. WebUI extension for ControlNet
โ 18kMVSFormer. Codes of MVSFormer: Multi-View Stereo by Learning Robust Image Features and Temperature-based Depth (TMLR2023)
โ 211OpenChatKit. Python
โ 9ktransformers. ๐ค Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
โ 163ktabby. Self-hosted AI coding assistant
โ 34kpycolmap. (Deprecated) Python bindings for COLMAP
โ 998loguru. Python logging made (stupidly) simple
โ 24khttp-server. A simple, zero-configuration, command-line http server
โ 14kSphereFormer. The official implementation for "Spherical Transformer for LiDAR-based 3D Recognition" (CVPR 2023).
โ 368MapTR. [ICLR'23 Spotlight & ECCV'24 & IJCV'24] MapTR: Structured Modeling and Learning for Online Vectorized HD Map Construction
โ 1.5ksegment-anything. The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
โ 55kdeepv2d_pytorch. Python
โ 6DeepSpeed. DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
โ 43kroomGPT. Upload a photo of your room to generate your dream room with AI.
โ 11kGKT. Efficient and Robust 2D-to-BEV Representation Learning via Geometry-guided Kernel Transformer
โ 247translating-images-into-maps. Official PyTorch code for 'Translating Images Into Maps' ICRA 2022 (Outstanding Paper Award)
โ 433HDMapNet. Python
โ 764resa. Implementation of our paper 'RESA: Recurrent Feature-Shift Aggregator for Lane Detection' in AAAI2021.
โ 188Transformer-in-Computer-Vision. A paper list of some recent Transformer-based CV works.
โ 1.5kFeatDepth. This is the offical codes for the methods described in the "Feature-metric Loss for Self-supervised Learning of Depth and Egomotion".
โ 249lietorch. Cuda
โ 851DPT. Dense Prediction Transformers
โ 2.3kDROID-SLAM. Python
โ 2.6kPENet_ICRA2021. ICRA 2021 "Towards Precise and Efficient Image Guided Depth Completion"
โ 361LapDepth-release. Monocular Depth Estimation Using Laplacian Pyramid-Based Depth Residuals
โ 334fast-depth. ICRA 2019 "FastDepth: Fast Monocular Depth Estimation on Embedded Systems"
โ 1ktsdf-fusion. Fuse multiple depth frames into a TSDF voxel volume.
โ 818cascade-stereo. cascade-stereo
โ 486nvidia-docker. Build and run Docker containers leveraging NVIDIA GPUs
โ 18ktensorflow. An Open Source Machine Learning Framework for Everyone
โ 1.2kMVSNet. MVSNet (ECCV2018) & R-MVSNet (CVPR2019)
โ 1.5kPatchmatchNet. Official code of PatchmatchNet (CVPR 2021 Oral)
โ 554openseg.pytorch. The official Pytorch implementation of OCNet, OCRNet, and SegFix.
โ 1.2ktf-hrnet. tensorflow implementation for "High-Resolution Representations for Labeling Pixels and Regions"
โ 88RDN-TensorFlow. A TensorFlow implementation of CVPR 2018 paper "Residual Dense Network for Image Super-Resolution".
โ 142aicoco. โ็ฑๅฏๅฏ-็ฑ็ๆดปโๅพฎๅๅ ๅฎน็ฒพ้
โ 575HR-Depth. [AAAI 2021] HR-Depth : High Resolution Self-Supervised Depth Estimation
โ 253Insta-DM. Learning Monocular Depth in Dynamic Scenes via Instance-Aware Projection Consistency (AAAI 2021)
โ 228DeepSFM. This is a PyTorch implementation of the ECCV2020 paper "DeepSFM: Structure From Motion Via Deep Bundle Adjustment".
โ 343GSLAM. A General Simultaneous Localization and Mapping Framework which supports feature based or direct method and different sensors including monocular camera, RGB-D sensors or any other input types can be handled.
โ 1.1kopenvslam. OpenVSLAM: A Versatile Visual SLAM Framework
โ 3kDeepV2D. Python
โ 673VNL_Monocular_Depth_Prediction. Monocular Depth Prediction
โ 470bts. From Big to Small: Multi-Scale Local Planar Guidance for Monocular Depth Estimation
โ 675opencv_extra. OpenCV extra data
โ 972tensorpack. A Neural Net Training Interface on TensorFlow, with focus on speed + flexibility
โ 6.3kStructuredForests. A Python Implementation for Piotr's ICCV Paper "Structured Forests for Fast Edge Detection".
โ 127edges. Structured Edge Detection Toolbox
โ 838LSD. a Line Segment Detector
โ 144SideWindowFilter. Side window is better than Full window
โ 287PlanarReconstruction. [CVPR'19] Single-Image Piece-wise Planar 3D Reconstruction via Associative Embedding
โ 374Indoor-SfMLearner. [ECCV'20] Patch-match and Plane-regularization for Unsupervised Indoor Depth Estimation
โ 151ArtLine. A Deep Learning based project for creating line art portraits.
โ 3.6kViP-DeepLab. Python
โ 227ViT-pytorch. Pytorch reimplementation of the Vision Transformer (An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale)
โ 2.2kSOLO. SOLO and SOLOv2 for instance segmentation, ECCV 2020 & NeurIPS 2020.
โ 1.8kdetr. End-to-End Object Detection with Transformers
โ 15kSFSegNets. [ECCV-2020-oral]-Semantic Flow for Fast and Accurate Scene Parsing
โ 387self-supervised-depth-completion. ICRA 2019 "Self-supervised Sparse-to-Dense: Self-supervised Depth Completion from LiDAR and Monocular Camera"
โ 654utm. Bidirectional UTM-WGS84 converter for python
โ 530pykitti. Python tools for working with KITTI data.
โ 1.2kpacknet-sfm. TRI-ML Monocular Depth Estimation Repository
โ 1.3kShameCom. ๆถ้ๆ กๆๆฑก็นๅ ฌๅธๆ็ป็ป๏ผๅธฎๅฉๅญฆๅผๅญฆๅฆน้ฟ้ทใไบ่็ฝไธๆพ้ๅฟ๏ผ
โ 5.6kpangolin_tutorial. Pangolinไธญๆๆ ทไพ
โ 130pangolin. Python binding of 3D visualization library Pangolin
โ 305PANet. PANet for Instance Segmentation and Object Detection
โ 1.3kmoco. PyTorch implementation of MoCo: https://arxiv.org/abs/1911.05722
โ 5.1kSC-SfMLearner-Release. Unsupervised Scale-consistent Depth Learning from Video (IJCV2021 & NeurIPS 2019)
โ 745DDAD. Dense Depth for Autonomous Driving (DDAD) dataset.
โ 558stereo-from-mono. [ECCV 2020] Learning stereo from single images using monocular depth estimation networks
โ 408awesome-point-cloud-analysis. A list of papers and datasets about point cloud analysis (processing)
โ 4.2kawesome-image-registration. image registration related books, papers, videos, and toolboxes
โ 1.5kamdim-public. Public repo for Augmented Multiscale Deep InfoMax representation learning
โ 402awesome-self-supervised-learning. A curated list of awesome self-supervised methods
โ 6.4kfair_self_supervision_benchmark. Scaling and Benchmarking Self-Supervised Visual Representation Learning
โ 588facebook360_dep. Facebook360 Depth Estimation Pipeline - https://facebook.github.io/facebook360_dep
โ 257monodepth2. [ICCV 2019] Monocular depth estimation from a single image
โ 4.5kLKVOLearner. Learning Depth from Monocular Videos using Direct Methods, CVPR 2018
โ 234models. Models and examples built with TensorFlow
โ 78kSfmLearner-Pytorch. Pytorch version of SfmLearner from Tinghui Zhou et al.
โ 1kdeepcluster. Deep Clustering for Unsupervised Learning of Visual Features
โ 1.7kpytorch_Realtime_Multi-Person_Pose_Estimation. Python
โ 1.4kProposalFlow. Code Release for "Proposal Flow" CVPR 2016.
โ 34inverse-compositional-STN. Inverse Compositional Spatial Transformer Networks :performing_arts: (CVPR 2017 oral)
โ 319weakalign. End-to-end weakly-supervised semantic alignment
โ 209shadowsocks-windows. A C# port of shadowsocks
โ 60k