This is your work, valued
AuraFusion360_official. [CVPR2025] Official Implementation of AuraFusion360
★ 83NYCU_DLP2023. Jupyter Notebook
★ 11ComputerAnimationHW2. C++
★ 2pdm-f23. NYCU Perception and Decision Making 2023 Fall
★ 2diptych-prompt-unofficail. Python
★ 2semantic-segmentation-pytorch. Pytorch implementation for Semantic Segmentation/Scene Parsing on MIT ADE20K dataset
★ 1google-research. Google Research
★ 1DCGAN_Style_Transfer. AI capstone final project
★ 1INP20200FALL_GAME1A2B_PART1. C++
★ 1PopMarkWebGL. HTML
★ 1CV-for-UAV-Autopilot. Python
★ 1TensoRF. [ECCV 2022] Tensorial Radiance Fields, a novel approach to model and reconstruct radiance fields
★ 1github-workshop-2. HTML
★ 1UnderTheLake. VR horro game using HTC Vive Pro
★ 1CS_Union_loyalty_card. Python
★ 1Forest_Fire_Detection. Jupyter Notebook
★ 1kkennethwu.github.io. HTML
★ 1AI_Capstone_Project1. Jupyter Notebook
★ 1onnxruntime-genai. Generative AI extensions for onnxruntime
★ 1.1kApollo-11. Original Apollo 11 Guidance Computer (AGC) source code for the command and lunar modules.
★ 72konnxruntime-qnn. onnxruntime-qnn is the Qualcomm AI Runtime (QAIRT) execution provider for onnxruntime. It provides onnxruntime hardware acceleration and advanced functionalities on Qualcomm devices.
★ 41onnxruntime. ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator
★ 21kSkyfall-GS. [ECCV 2026] Skyfall-GS: Synthesizing Immersive 3D Urban Scenes from Satellite Imagery
★ 938gemini-cli. An open-source AI agent that brings the power of Gemini directly into your terminal.
★ 106kn8n. Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.
★ 199kGaussianLSS. Official PyTorch implementation of "GaussianLSS - Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting" (CVPR 2025).
★ 176LongSplat. [ICCV 2025] LongSplat: Robust Unposed 3D Gaussian Splatting for Casual Long Videos
★ 797depthsplat. [CVPR'25] DepthSplat: Connecting Gaussian Splatting and Depth
★ 1.2kgpt-oss. gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
★ 20kOpenSplat. Production-grade 3D gaussian splatting with CPU/GPU support for Windows, Mac and Linux 🚀
★ 2.1kGrendel-GS. [ICLR 2025 Oral] On Scaling Up 3D Gaussian Splatting Training
★ 680speedy-splat. Python
★ 347EDGS. [CVPR 2026] A PyTorch implementation of the paper "EDGS: Eliminating Densification for Efficient Convergence of 3DGS"
★ 727ever_training. Original reference implementation of "EVER: Exact Volumetric Ellipsoid Rendering for Real-time View Synthesis"
★ 310AnySplat. [SIGGRAPH Asia 2025 (ACM TOG)] AnySplat: Feed-forward 3D Gaussian Splatting from Unconstrained Views
★ 901on-the-fly-nvs. Official implementation of On-the-fly Reconstruction for Large-Scale Novel View Synthesis from Unposed Images. A. Meuleman, I. Shah, A. Lanvin, B. Kerbl, G. Drettakis, ACM TOG (proc. SIGGRAPH) 2025
★ 572tiny-llm. A course of learning LLM inference serving on Apple Silicon for systems engineers: build a tiny vLLM + Qwen.
★ 4.4kFluxText. Implementation of "FLUX-Text: A Simple and Advanced Diffusion Transformer Baseline for Scene Text Editing"
★ 854poster-design. 迅排设计 - 美观且功能强大的图片编辑器、在线海报设计,仿稿定设计,适用于多种场景:海报生成、电商产品图、文章长图、视频/公众号封面等。A beautiful online image designer, suitable for various scenarios like generate posters, making design easier!
★ 4.8kgemini-code. Gemini 2.5 Pro code assistant
★ 548PosterCraft. [ICLR'26] Rethinking High-Quality Aesthetic Poster Generation in a Unified Framework
★ 541lang-segment-anything. SAM with text prompt
★ 2.6kvcard. TypeScript
★ 315GEN3C. [CVPR 2025 Highlight] GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control
★ 1.4kYoLLaVA. 🌋👵🏻 Yo'LLaVA: Your Personalized Language and Vision Assistant (NeurIPS 2024)
★ 123FrugalNeRF. [CVPR 2025] FrugalNeRF: Fast Convergence for Extreme Few-shot Novel View Synthesis without Learned Priors
★ 15asuka-misato. [CVPR 2025 Highlight] Towards Enhanced Image Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency
★ 79CUT3R. Official implementation of Continuous 3D Perception Model with Persistent State
★ 1.5kdcseg. DCSEG: Decoupled 3D Open-Set Segmentation using Gaussian Splatting
★ 13LAPIS.
★ 20ai-toolkit. The ultimate training toolkit for finetuning diffusion models
★ 11kKD_llama_1B-Knowledge-Distillation-for-LLaMA-3.2-1B. This project demonstrates knowledge distillation from LLaMA-3.2-3B-Instruct (teacher) to LLaMA-3.2-1B-Instruct (student). The goal is to transfer knowledge from a larger model into a smaller one to improve efficiency while maintaining performance.
★ 3AEGIS. [ACM SIGCHI 2025] The official repo for “AEGIS: Human Attention-based Explainable Guidance for Intelligent Vehicle Systems”
★ 19GlobustVP. [CVPR 2025 Best Paper Award Candidate & Oral] Convex Relaxation for Robust Vanishing Point Estimation in Manhattan World
★ 148openai-python. The official Python library for the OpenAI API
★ 31kGeoWizard. [ECCV'24] GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Image
★ 938Step1X-Edit. A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gemini 2 Flash.
★ 2.2kawesome-gpt4o-images. Awesome curated collection of images and prompts generated by GPT-4o and gpt-image-1. Explore AI generated visuals created with ChatGPT and Sora, showcasing OpenAI’s advanced image generation capabilities.
★ 8.1kAREdit. Training-Free Text-Guided Image Editing Using Visual Autoregressive Model
★ 76perception_models. State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!
★ 2.3khqq. Official implementation of Half-Quadratic Quantization (HQQ)
★ 948ml-depth-pro. Depth Pro: Sharp Monocular Metric Depth in Less Than a Second.
★ 5.6kAnyText2. Official implementation code of the paper <AnyText2: Visual Text Generation and Editing With Customizable Attributes>
★ 212AnyText. Official implementation code of the paper <AnyText: Multilingual Visual Text Generation And Editing>
★ 4.9kDiffSynth-Studio. Enjoy the magic of Diffusion models!
★ 13kTransPixeler. CVPR2025
★ 926awesome-ai-system-prompts. 🧠 Curated collection of system prompts for top AI tools. Perfect for AI agent builders and prompt engineers. Incuding: ChatGPT, Claude, Perplexity, Manus, Claude-Code, Loveable, v0, Grok, same new, windsurf, notion, and MetaAI.
★ 6.1kMinorityPrompt. Official PyTorch Implementation of "Minority-Focused Text-to-Image Generation via Prompt Optimization" (CVPR 2025 Oral)
★ 28unidisc. UniDisc: A discrete diffusion model for joint multimodal generation, enabling controllable and efficient text-image synthesis, editing, and inpainting.
★ 142arxiv-mcp-server. A Model Context Protocol server for searching and analyzing arXiv papers
★ 3kSee3D. [CVPR'25 Highlight] You See it, You Got it: Learning 3D Creation on Pose-Free Videos at Scale
★ 723Murre. Code for "Multi-view Reconstruction via SfM-guided Monocular Depth Estimation". CVPR 2025 (Oral Presentation)
★ 379AutoPrompt. A framework for prompt tuning using Intent-based Prompt Calibration
★ 3kcse234-w25-PA. Python
★ 52Torch-Pruning. [CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.
★ 3.3knerficg. The ICG Neural Radiance Fields and Novel View Synthesis Framework.
★ 38star-vector. StarVector is a foundation model for SVG generation that transforms vectorization into a code generation task. Using a vision-language modeling architecture, StarVector processes both visual and textual inputs to produce high-quality SVG code with remarkable precision.
★ 4.5kMV-Adapter. [ICCV 2025] Official impl. of "MV-Adapter: Multi-view Consistent Image Generation Made Easy"
★ 1.3kDiptychPrompting. Python
★ 63stable-virtual-camera. Stable Virtual Camera: Generative View Synthesis with Diffusion Models
★ 1.6khackmd-mcp. A Model Context Protocol server for integrating HackMD's note-taking platform with AI assistants.
★ 65vggt. [CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
★ 14kRF-Solver-Edit. [🚀ICML 2025] "Taming Rectified Flow for Inversion and Editing" Using FLUX and HunyuanVideo for image and video editing!
★ 638MIDI-3D. [CVPR 2025] MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
★ 937personalize-anything. [AAAI 2026] Personalize Anything for Free with Diffusion Transformer
★ 362python-genai. Google Gen AI Python SDK provides an interface for developers to integrate Google's generative models into their Python applications.
★ 3.9kdeprecated-generative-ai-python. This SDK is now deprecated, use the new unified Google GenAI SDK.
★ 2.3kAuraFusion360_official. [CVPR2025] Official Implementation of AuraFusion360
★ 83octotools. OctoTools: An agentic framework with extensible tools for complex reasoning
★ 1.5kHunyuanVideo-I2V. HunyuanVideo-I2V: A Customizable Image-to-Video Model based on HunyuanVideo
★ 1.8kHunyuanVideo. HunyuanVideo: A Systematic Framework For Large Video Generation Model
★ 12kFLARE. Python
★ 721fast3r. [CVPR 2025] Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass
★ 1.6kAnyV2V. Code and data for "AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks" [TMLR 2024]
★ 655InstantSplat. InstantSplat: Sparse-view SfM-free Gaussian Splatting in Seconds
★ 1.7kFlowEdit. Official implementation of the paper: "FlowEdit: Inversion-Free Text-Based Editing Using Pre-Trained Flow Models"
★ 1ksyncd. SynCD: Generating Multi-Image Synthetic Data for Text-to-Image Customization (ICCV 2025)
★ 155FlashMLA. FlashMLA: Efficient Multi-head Latent Attention Kernels
★ 13kwonderland. Python
★ 166FLUX-Controlnet-Inpainting. Python
★ 794goku. [CVPR2025 Highlight] Video Generation Foundation Models: https://saiyan-world.github.io/goku/
★ 2.9kEZIGen. [BMVC 2025] Official implementation for paper EZIGen: Enhancing zero-shot personalized image generation with precise subject encoding and decoupled guidance
★ 107prompt-to-prompt. Jupyter Notebook
★ 3.5kAwesome-Controllable-Diffusion. Papers and resources on Controllable Generation using Diffusion Models, including ControlNet, DreamBooth, IP-Adapter.
★ 505accelerate. 🚀 A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (including fp8), and easy-to-configure FSDP and DeepSpeed support
★ 9.8kstable-point-aware-3d. SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
★ 1.1kdiptych-prompt-unofficail. Python
★ 2DepthLab. Official implementation of "DepthLab: From Partial to Complete"
★ 551GPA. Online photography assistance, tailored for food photograph.
★ 22flux. Official inference repo for FLUX.1 models
★ 26klambda-eclipse-inference. [TMLR] Official PyTorch implementation of "λ-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent Space"
★ 53StableCascade. Official Code for Stable Cascade
★ 6.5kInstantStyle. InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation 🔥
★ 2kIP-Adapter. The image prompt adapter is designed to enable a pretrained text-to-image diffusion model to generate images with image prompt.
★ 6.6kTRELLIS. Official repo for paper "Structured 3D Latents for Scalable and Versatile 3D Generation" (CVPR'25 Spotlight).
★ 13kIC-Light. More relighting!
★ 8.5kgenesis-world. Simulation platform for general-purpose robotics & embodied AI learning.
★ 30kleetcode-master. 《代码随想录》LeetCode 刷题攻略:200道经典题目刷题顺序,共60w字的详细图解,视频难点剖析,50余张思维导图,支持C++,Java,Python,Go,JavaScript等多语言版本,从此算法学习不再迷茫!🔥🔥 来看看,你会发现相见恨晚!🚀
★ 62krcg. PyTorch implementation of RCG https://arxiv.org/abs/2312.03701
★ 941SimpleTuner. A general fine-tuning kit geared toward image/video/audio diffusion models.
★ 2.9kSyncNoise. SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing
★ 19LeetCode. This repository contains the solutions and explanations to the algorithm problems on LeetCode. Only medium or above are included. All are written in C++/Python and implemented by myself. The problems attempted multiple times are labelled with hyperlinks.
★ 6.2kAnything-3D. Segment-Anything + 3D. Let's lift anything to 3D.
★ 1.6kVideoGS. [SIGGRAPH Asia 2024] V^3: Viewing Volumetric Videos on Mobiles via Streamable 2D Dynamic Gaussians
★ 107Bend. A massively parallel, high-level programming language
★ 20kgraphrag. A modular graph-based Retrieval-Augmented Generation (RAG) system
★ 35kDepth-Anywhere. Python
★ 104Stable-Diffusion-Inpaint. Stable diffusion for inpainting
★ 229GSD. Geometry-Aware Score Distillation via 3D Consistent Noising and Gradient Consistency Modeling
★ 29bitsandbytes. Accessible large language models via k-bit quantization for PyTorch.
★ 8.4kGScream. Official code for ECCV2024 paper: GScream: Learning 3D Geometry and Feature Consistent Gaussian Splatting for Object Removal
★ 104DNGaussian. [CVPR'24] DNGaussian: Optimizing Sparse-View 3D Gaussian Radiance Fields with Global-Local Depth Normalization
★ 365DiffIR2VR-Zero. Python
★ 178MVIP-NeRF. official repo for MVIP-NeRF (CVPR 2024)
★ 232d-gaussian-splatting. [SIGGRAPH'24] 2D Gaussian Splatting for Geometrically Accurate Radiance Fields
★ 3.3kdiff-surfel-rasterization. A differentiable rasterizer used in the project "2D Gaussian Splatting"
★ 186efficient-kan. An efficient pure-PyTorch implementation of Kolmogorov-Arnold Network (KAN).
★ 4.7kcoremltools. Core ML tools contain supporting tools for Core ML model conversion, editing, and validation.
★ 5.4klitert-torch. Support PyTorch model conversion with LiteRT.
★ 1.1kcanvas-confetti. 🎉 performant confetti animation in the browser
★ 13kaimet. AIMET is a library that provides advanced quantization and compression techniques for trained neural network models.
★ 2.7kpykan. Kolmogorov Arnold Networks
★ 16kInfusion. Official implementation for paper: InFusion: Inpainting 3D Gaussians via Learning Depth Completion from Diffusion Prior
★ 560CustomNeRF. [CVPR 2024] Customize your NeRF: Adaptive Source Driven 3D Scene Editing via Local-Global Iterative Training
★ 44IOPaint. Image inpainting tool powered by SOTA AI Model. Remove any unwanted object, defect, people from your pictures or erase and replace(powered by stable diffusion) any thing on your pictures.
★ 23kUniDepth. Universal Monocular Metric Depth Estimation
★ 1.2kInpaint-Anything. Inpaint anything using Segment Anything and inpainting models.
★ 7.7kGaussianPro. [ICML2024] Official code for GaussianPro: 3D Gaussian Splatting with Progressive Propagation
★ 842DIP_Final. Water Segmentation
★ 2CompletionFormer. [CVPR2023] CompletionFormer: Depth Completion with Convolutions and Vision Transformers
★ 282servstat. Server resource and GPU process monitor.
★ 32garfield. [CVPR'24] Group Anything with Radiance Fields
★ 466SimpleNeRF. Official code release accompanying the paper "SimpleNeRF: Regularizing Sparse Input Neural Radiance Fields with Simpler Solutions"
★ 12awesome-nerf-editing. 🧙🏻♂️A list of papers curated for you to dive into the Awesome Radiance Field-based 3D Editing.
★ 499AnyDoor. Official implementations for paper: Anydoor: zero-shot object-level image customization
★ 4.2kSIGNeRF. SIGNeRF: Scene Integrated Generation for Neural Radiance Fields
★ 128SegAnyGAussians. The official implementation of Segment Any 3D GAussians (AAAI-25)
★ 989clipseg. This repository contains the code of the CVPR 2022 paper "Image Segmentation Using Text and Image Prompts".
★ 1.3kDINO. [ICLR 2023] Official implementation of the paper "DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection"
★ 2.8kByteTrack. [ECCV 2022] ByteTrack: Multi-Object Tracking by Associating Every Detection Box
★ 6.6kmip-splatting. [CVPR'24 Best Student Paper] Mip-Splatting: Alias-free 3D Gaussian Splatting
★ 1.5krepaint123. Official implementation of Repaint123: Fast and High-quality One Image to 3D Generation with Progressive Controllable 2D Repainting (ECCV 2024)
★ 276LoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
★ 14kdeep_sort. Simple Online Realtime Tracking with a Deep Association Metric
★ 6.2ktalktopapers. Jupyter Notebook
★ 210paper-reading. 深度学习经典、新论文逐段精读
★ 34kGaussianSplats3D. Three.js-based implementation of 3D Gaussian splatting
★ 2.8kgaussian-grouping. [ECCV'2024] Gaussian Grouping for open-world Anything reconstruction, segmentation and editing.
★ 1kcalibur. Conversion between different conventions of camera matrices and transform matrices.
★ 101GaussianEditor. [CVPR 2024] GaussianEditor: Swift and Controllable 3D Editing with Gaussian Splatting
★ 1.4kAwesome-CV. :page_facing_up: Awesome CV is LaTeX template for your outstanding job application
★ 28kinterview. 📚 C/C++ 技术面试基础知识总结,包括语言、程序库、数据结构、算法、系统、网络、链接装载库等知识及面试经验、招聘、内推等信息。This repository is a summary of the basic knowledge of recruiting job seekers and beginners in the direction of C/C++ technology, including language, program library, data structure, algorithm, system, network, link loading library, interview experience, recruitment, recommendation, etc.
★ 38kacademicpages.github.io. Github Pages template based upon HTML and Markdown for personal, portfolio-based websites.
★ 17kscene-representation-networks. Official Pytorch implementation of Scene Representation Networks: Continuous 3D-Structure-Aware Neural Scene Representations
★ 440f3rm. F3RM: Feature Fields for Robotic Manipulation. Official repo for the paper "Distilled Feature Fields Enable Few-Shot Language-Guided Manipulation" (CoRL 2023).
★ 220awesome-3D-gaussian-splatting. Curated list of papers and resources focused on 3D Gaussian Splatting, intended to keep pace with the anticipated surge of research in the coming months.
★ 8.8kscalingup. [CoRL 2023] This repository contains data generation and training code for Scaling Up & Distilling Down
★ 414Instant-angelo. Instant-angelo: Build high-fidelity Digital Twin within 20 Minutes!
★ 462Tri-MipRF. [ICCV'23 Oral, Best Paper Finalist]Tri-MipRF: Tri-Mip Representation for Efficient Anti-Aliasing Neural Radiance Fields
★ 457gsplat. CUDA accelerated rasterization of gaussian splatting
★ 5.5kLichtFeld-Studio. Train, inspect, edit, automate, and export 3D Gaussian Splatting scenes from a single native application.
★ 3.5klama. 🦙 LaMa Image Inpainting, Resolution-robust Large Mask Inpainting with Fourier Convolutions, WACV 2022
★ 10krobust_cvd. Robust Consistent Video Depth Estimation
★ 309Applied-Deep-Learning. Applied Deep Learning Course
★ 3.6kZoeDepth. Metric depth estimation from a single image
★ 2.8kstable-diffusion. A latent text-to-image diffusion model
★ 73kdiffusers. 🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34kjonbarron.github.io. HTML
★ 3.6kgdown. Google Drive public file downloader when curl/wget fails.
★ 5.3kinstant-ngp. Instant neural graphics primitives: lightning fast NeRF and more
★ 18kneuralangelo. Official implementation of "Neuralangelo: High-Fidelity Neural Surface Reconstruction" (CVPR 2023)
★ 4.6ktorch_efficient_distloss. Efficient distortion loss with O(n) realization.
★ 132ViP-NeRF. Official code release accompanying the paper - "ViP-NeRF: Visibility Prior for Sparse Input Neural Radiance Fields"
★ 69deformable-sprites. Python
★ 113vision_transformer. Jupyter Notebook
★ 13kgaussian-splatting. Original reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"
★ 23kNCTU_DLP. Jupyter Notebook
★ 22vision-nerf. Official PyTorch Implementation of paper "Vision Transformer for NeRF-Based View Synthesis from a Single Input Image", WACV 2023.
★ 114FreeNeRF. [CVPR23] FreeNeRF: Improving Few-shot Neural Rendering with Free Frequency Regularization
★ 477fourier-feature-networks. Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional Domains
★ 1.4knerfacc. A General NeRF Acceleration Toolbox in PyTorch.
★ 1.5ksdfstudio. A Unified Framework for Surface Reconstruction
★ 2.1khumanrf. Official code for "HumanRF: High-Fidelity Neural Radiance Fields for Humans in Motion"
★ 496EndoNeRF. Neural Rendering for Stereo 3D Reconstruction of Deformable Tissues in Robotic Surgery
★ 260awesome. 😎 Awesome lists about all kinds of interesting topics
★ 491knerf. Code release for NeRF (Neural Radiance Fields)
★ 11kTensoRF. [ECCV 2022] Tensorial Radiance Fields, a novel approach to model and reconstruct radiance fields
★ 1.2knerfstudio. A collaboration friendly studio for NeRFs
★ 12kawesome-computer-vision. A curated list of awesome computer vision resources
★ 23knerf_pl. NeRF (Neural Radiance Fields) and NeRF in the Wild using pytorch-lightning
★ 2.8kawesome-NeRF. A curated list of awesome neural radiance fields papers
★ 6.8knerf-pytorch. A PyTorch implementation of NeRF (Neural Radiance Fields) that reproduces the results.
★ 6.1kgoogle-research. Google Research
★ 38kjax3d. Python
★ 763ExpirationExpert. Swift
★ 1VoxFormer. Official PyTorch implementation of VoxFormer [CVPR 2023 Highlight]
★ 1.2k