This is your work, valued
ganavatar. [3DV'24] GAN-Avatar: Controllable Personalized GAN-based Human Head Avatar
★ 82physhead. [CVPR 2026] PhysHead: Simulation-Ready Gaussian Head Avatars
★ 45DynHair. Head Avatars with Dynamic Explicit Hair [ECCV 2026]
★ 11GNM. An open ecosystem of parametric human models and perception stacks, starting with GNM Head.
★ 1.3kNeuralFur. NeuralFur: Animal Fur Reconstruction From Multi-View Images [3DV 2026]
★ 55anyup. [ICLR '26 Oral] Official repository of the paper "AnyUp: Universal Feature Upsampling".
★ 571AnimateDiff. Official implementation of AnimateDiff.
★ 12kVACE. [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editing
★ 3.9khallo. Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
★ 8.7knglod. Neural Geometric Level of Detail: Real-time Rendering with Implicit 3D Shapes (CVPR 2021 Oral)
★ 921InST. Official implementation of the paper “Inversion-Based Style Transfer with Diffusion Models” (CVPR 2023)
★ 588ComfyUI-TeleStyle. Python
★ 42TeleStyle. state-of-the-art content-preserving style transfer
★ 235Qwen-Image. Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.
★ 8.2kDiffSynth-Studio. Enjoy the magic of Diffusion models!
★ 13kPi3. [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning
★ 2.1kDAMA-code. [PhysHuman @ CVPR2026] DAMA: Disentangled Body-Anchored Gaussians for Controllable Multi-Layered Avatars
★ 12MultiTalk. [NeurIPS 2025] Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
★ 3kQwen3-Omni. Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.
★ 3.9kconfingy. An implicit configuration system for Python
★ 55ml-lito. [ICLR 2026] LiTo: Surface Light Field Tokenization
★ 459dotfiles. Shell
★ 8LVSM. [ICLR 2025 Oral] Official code for "LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias"
★ 550recastnavigation. Industry-standard navigation-mesh toolset for games
★ 7.8kDepth-Anything-3. Depth Anything 3
★ 6kgenwarp. Python
★ 310FramePack. Lets make video diffusion practical!
★ 17kcroco. Python
★ 508LEXIS. [arXiv 2026] LEXIS: LatEnt ProXimal Interaction Signatures for 3D HOI from an Image
★ 14claude-hud. A Claude Code plugin that shows what's happening - context usage, active tools, running agents, and todo progress
★ 27koctformer. OctFormer: Octree-based Transformers for 3D Point Clouds [SIGGRAPH 2023]
★ 360HY-World-2.0. HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds
★ 2.4klyra. Project Lyra: Open Generative 3D World Models
★ 2.2kSpeechDrivenTongueAnimation. ML-driven tongue animation (CVPR'22)
★ 52lagernvs. Official code for "LagerNVS Latent Geometry for Fully Neural Real-time Novel View Synthesis" (CVPR 2026)
★ 402autoresearch. AI agents running research on single-GPU nanochat training automatically
★ 92kpersonaplex. PersonaPlex code.
★ 10knova3r. [ICLR 2026] NOVA3R: Non-pixel-aligned Visual Transformer for Amodal 3D Reconstruction
★ 155tongue. We propose a framework that accurately derives the 3D tongue shape from single images. A high detailed 3D point cloud of the tongue surface and a full head topology along with the tongue expression can be estimated from the image domain.
★ 74physhead. [CVPR 2026] PhysHead: Simulation-Ready Gaussian Head Avatars
★ 47LitePT. [CVPR 2026 Highlight] LitePT: Lighter Yet Stronger Point Transformer
★ 381nerfify-code. [CVPR 2026 Highlight] Offical code for- NERFIFY: A Multi-Agent Framework for Turning NeRF Papers into Code
★ 28DMM. An Implicit Parametric Morphable Dental Model
★ 67pixie. Feed-forward model for predicting 3D physics with 3DGS + NeRF
★ 298meshgraphnets. Rewrite deepmind/meshgraphnets into pytorch
★ 104nerfbaselines. Reproducible evaluation of NeRF and 3DGS methods
★ 3653DGen-Playground. 3D generation made easy!
★ 458large-steps-pytorch. Implementation of "Large Steps in Inverse Rendering of Geometry"
★ 443NeMO. Official Repository for "Finding NeMO: A Geometry-Aware Representation of Template Views for Few-Shot Perception"
★ 36research-skills. Claude Code skills for AI/ML researchers
★ 40tofu. Official code for ICCV 2021 oral paper: Topologically Consistent Multi-View Face Inference Using Volumetric Sampling
★ 163CARI4D. [CVPR 2026] CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction
★ 204LichtFeld-Studio. Train, inspect, edit, automate, and export 3D Gaussian Splatting scenes from a single native application.
★ 3.5kDreamDojo. Official Codebase for "DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos" (ICML 2026)
★ 1kawesome-vlm-architectures. Famous Vision Language Models and Their Architectures
★ 1.3kSenseNova-SI. [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models
★ 293infinigen. Infinite Photorealistic Worlds using Procedural Generation
★ 7.2kMoLingo. [ CVPR 2026 ] MoLingo: Motion-Language Alignment for Text-to-Motion Generation
★ 85stable-worldmodel. A platform for reproducible world model research and evaluation
★ 2.1kLTX-Video. Official repository for LTX-Video
★ 11kFastGen. NVIDIA FastGen: Fast Generation from Diffusion Models
★ 894spear. SPEAR: A Simulator for Photorealistic Embodied AI Research
★ 543cookbook. Examples and guides for using the Gemini API
★ 18kkauldron. Modular, scalable library to train ML models
★ 287SO-ARM100. Standard Open Arm 100
★ 6.9kDiSECt. Differentiable Cutting Simulator
★ 140cosmos-predict2.5. Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the future state of the world in the form of video.
★ 1.3kLongLive. Long Video Gen Infrastructure
★ 2.5kgemma. Gemma open-weight LLM library, from Google DeepMind
★ 5.6kpixel3dmm. [Official Code] Pixel3DMM: Versatile Screen-Space Priors for Single-Image 3D Face Reconstruction
★ 512IsaacLab. Unified framework for robot learning built on NVIDIA Isaac Sim
★ 7.8knexels. The official repository for Nexels: Neurally-Textured Surfels for Real-Time Novel View Synthesis with Sparse Geometries
★ 129UniAnimate-DiT. UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
★ 854mesh-splatting. Python
★ 697cwm. Research code artifacts for Code World Model (CWM) including inference tools, reproducibility, and documentation.
★ 886MPMAvatar. MPMAvatar: Learning 3D Gaussian Avatars with Accurate and Robust Physics-Based Dynamics (NeurIPS 2025)
★ 85TensorRT-LLM. TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
★ 14kAniMer. AniMer: Animal Pose and Shape Estimation Using Family Aware Transformer (CVPR2025)
★ 43DoRA. [SIGGRAPH 2025 (Journal Track)] Facial Appearance Capture at Home with Patch-Level Reflectance Prior.
★ 75densemarks. [ICLR 2026] DenseMarks: dense per-pixel embeddings replacing standard face landmarks
★ 15FaceOLAT. Repository for Dataprocessing of FaceOLAT Dataset
★ 15beta-splatting. [SIGGRAPH'25] Official implementation for the paper "Deformable Beta Splatting"
★ 305tridi. [ICCV'25] Method for generating static human-object interactions
★ 44mujoco_warp. GPU-optimized version of the MuJoCo physics simulator, designed for NVIDIA hardware.
★ 1.4knewton. An open-source, GPU-accelerated physics simulation engine built upon NVIDIA Warp, specifically targeting roboticists and simulation researchers.
★ 5.3kInfiniHuman. [SIGGRAPHASIA2025] InfiniHuman: Infinite 3D Human Creation with Precise Control
★ 85Phy-SIC. [SIGGGRAPHASIA2025] PhySIC: Physically Plausible 3D Human-Scene Interaction and Contact from a Single Image
★ 60BEHAVIOR-1K. BEHAVIOR-1K: a platform for accelerating Embodied AI research. Join our Discord for support: https://discord.gg/bccR5vGFEx
★ 1.6ktext-to-text-transfer-transformer. Code for the paper "Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer"
★ 6.5kviser. Web-based 3D visualization in Python
★ 2.7kIntrinsiX. IntrinsiX: High-Quality PBR Generation using Image Priors
★ 62Wan2.2. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kVILA. VILA is a family of state-of-the-art vision language models (VLMs) for diverse multimodal AI tasks across the edge, data center, and cloud.
★ 3.8kspark. :sparkles: An advanced 3D Gaussian Splatting renderer for THREE.js
★ 3.5krapier. 2D and 3D physics engines focused on performance.
★ 5.6kMoGe. [CVPR'25 Oral] MoGe: Unlocking Accurate Monocular Geometry Estimation for Open-Domain Images with Optimal Training Supervision
★ 2.7kmap-anything. MapAnything: Universal Feed-Forward Metric 3D Reconstruction
★ 3.6kmega-sam. Code for the project "MegaSaM: Accurate, Fast and Robust Structure and Motion from Casual Dynamic Videos"
★ 1.3kIm2Haircut. Im2Haircut: Single-view Strand-based Hair Reconstruction for Human Avatars [ICCV 2025]
★ 58GEN3C. [CVPR 2025 Highlight] GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control
★ 1.4kanimal_papers. Awesome papers for markerless animal motion capture and 3D reconstruction.
★ 339ollama. Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
★ 177kDAViD. Python
★ 394HyperGaussians. HyperGaussians: High-Dimensional Gaussian Splatting for High-Fidelity Animatable Face Avatars
★ 60awesome-neural-physics. TeX
★ 65posescript. Code for the PoseScript (ECCV 22) and PoseFix (ICCV 23) papers.
★ 205OmniAvatar. Python
★ 1.9klaplacian_blend. Tutorial on Laplacian Image Blending
★ 27PhysRig. Physics-based rigging with MPM for realistic character animation. ICCV 2025.
★ 88metaquery. Official Implementation of Paper Transfer between Modalities with MetaQueries
★ 325diffusion-posterior-sampling. Official pytorch repository for "Diffusion Posterior Sampling for General Noisy Inverse Problems"
★ 668OmniGen2. OmniGen2: Exploration to Advanced Multimodal Generation. https://arxiv.org/abs/2506.18871
★ 4.1kvggt-blender. Blender addon for vggt 3D reconstruction
★ 783dgrut. Ray tracing and hybrid rasterization of Gaussian particles
★ 2.3kgen-omnimatte-public. Generative Omnimatte (CVPR 2025)
★ 186Go-with-the-Flow. The official implementation of CVPR'25 Oral paper "Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise"
★ 1.1kHunyuan3D-2.1. From Images to High-Fidelity 3D Assets with Production-Ready PBR Material
★ 3.8kHOIDiNi. Python
★ 78cosmos-transfer1-diffusion-renderer. Cosmos-Transfer1-DiffusionRenderer: High-quality video de-lighting and re-lighting based on Cosmos video diffusion framework
★ 832awesome-llm-apps. 100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.
★ 129kNeuralHaircut. Neural Haircut: Prior-Guided Strand-Based Hair Reconstruction. ICCV 2023
★ 15SMPLSim. Simulating SMPL humanoid, supporting PHC/PHC-MJX/PULSE/SimXR code bases.
★ 348spiritlm. Inference code for the paper "Spirit-LM Interleaved Spoken and Written Language Model".
★ 929WorldGen. 🌍 WorldGen - Generate Any 3D Scene in Seconds
★ 2kInterMimic. [CVPR 2025 Highlight] InterMimic: Towards Universal Whole-Body Control for Physics-Based Human-Object Interactions
★ 520labelme. Image annotation with Python. Supports polygon, rectangle, circle, line, point, and AI-assisted annotation.
★ 16ksmalst. Python
★ 181DecoupledGaussian. Jupyter Notebook
★ 56CogVideo. text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
★ 13kdifflocks. Code for our CVPR'25 paper - "DiffLocks: Generating 3D Hair from a Single Image using Diffusion Models"
★ 180FloED-main. Coherent Video Inpainting Using Optical Flow-Guided Efficient Diffusion
★ 199LoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
★ 14kHunyuan3D-2. High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models.
★ 14kWan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kSMALify. This repository contains an implementation for performing 3D animal (quadruped) reconstruction from a monocular image or video. The system adapts the pose (limb positions) and shape (animal type/height/weight) parameters for the SMAL deformable quadruped model, as well as camera parameters until the projected SMAL model aligns with 2D keypoints and silhouette segmentations extracted from the input frame(s).
★ 150BitNet. Official inference framework for 1-bit LLMs
★ 40kDifix3D. [CVPR 2025 Oral & Best Paper Finalist] Difix3D+: Improving 3D Reconstructions with Single-Step Diffusion Models
★ 1.2kPOCO. [3DV 2024] POCO: 3D Pose and Shape Estimation using Confidence
★ 75InteractVLM. [CVPR 2025] InteractVLM: 3D Interaction Reasoning from 2D Foundational Models
★ 139FRESA. [CVPR 2025 Highlight] FRESA: Feedforward Reconstruction of Personalized Skinned Avatars from Few Images
★ 39cosmos-predict1. Cosmos-Predict1 is a collection of general-purpose world foundation models for Physical AI that can be fine-tuned into customized world models for downstream applications.
★ 465stable-virtual-camera. Stable Virtual Camera: Generative View Synthesis with Diffusion Models
★ 1.6kcube. Roblox Foundation Model for 3D Intelligence
★ 1.2kAwesome-Video-Diffusion. A curated list of recent diffusion models for video generation, editing, and various other applications.
★ 5.7kGEM. GEM - Gaussian Eigen Models for Human Heads [CVPR 2025]
★ 54molecular-plus. particle solver for Blender 4.2+ / 5.1+
★ 392radfoam. Original implementation of "Radiant Foam: Real-Time Differentiable Ray Tracing"
★ 665vggt. [CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
★ 14kcursor.
★ 33kdust3r. DUSt3R: Geometric 3D Vision Made Easy
★ 7.3kGRM. Large Gaussian Reconstruction Model for Efficient 3D Reconstruction and Generation
★ 641TRELLIS. Official repo for paper "Structured 3D Latents for Scalable and Versatile 3D Generation" (CVPR'25 Spotlight).
★ 13k3dgs-mcmc. [NeurIPS 2024 Spotlight] Implementation of the paper "3D Gaussian Splatting as Markov Chain Monte Carlo"
★ 677MASt3R-SLAM. [CVPR 2025] MASt3R-SLAM: Real-Time Dense SLAM with 3D Reconstruction Priors
★ 3.1kHunyuanVideo. HunyuanVideo: A Systematic Framework For Large Video Generation Model
★ 12kTransformerEngine. A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.
★ 3.5kGSOPs. Gaussian Splatting Operators for SideFX Houdini
★ 611nersemble-data. Python
★ 94goku. [CVPR2025 Highlight] Video Generation Foundation Models: https://saiyan-world.github.io/goku/
★ 2.9kRTXMG. NVIDIA RTX Mega Geometry SDK
★ 242oumi. Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!
★ 9.4kTripoSR. TripoSR: Fast 3D Object Reconstruction from a Single Image
★ 6.8kjoker. Official Code Base of the Paper: "Joker: Conditional 3D Head Synthesis with Extreme Facial Expressions"
★ 54headcraft. HeadCraft (3DV 2025)
★ 89Janus. Janus-Series: Unified Multimodal Understanding and Generation Models
★ 18kPhysAvatar. Official repo for PhysAvatar: Learning the Physics of Dressed 3D Avatars from Visual Observations, ECCV 2024
★ 146ml-hugs. Official repository of HUGS: Human Gaussian Splats (CVPR 2024)
★ 398FORCE_dataset. Dataset for FORCE - Physics-aware Human-object Interaction (3DV 2025)
★ 24CameraHMR. Python
★ 253flowface. Repository for the paper "3D Face Tracking from 2D Video through Iterative Dense UV to Image Flow", CVPR 2024
★ 41ava-256. Train universal codec avatars
★ 224PhyRecon. Official implementation of NeurIPS24 paper "PhyRecon: Physically Plausible Neural Scene Reconstruction"
★ 173nvdiffrec. Official code for the CVPR 2022 (oral) paper "Extracting Triangular 3D Models, Materials, and Lighting From Images".
★ 2.3kcosmos. NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.
★ 11kgenesis-world. Simulation platform for general-purpose robotics & embodied AI learning.
★ 30ksmplfitter. The Fast Way From Vertices to Parametric 3D Humans
★ 187tungsten. High performance physically based renderer in C++11
★ 1.8kGen3Diffusion. [T-PAMI2025] Gen-3Diffusion: Realistic Image-to-3D Generation via 2D & 3D Diffusion Synergy
★ 31repulsive-shells. C++
★ 90gghead. [SIGGRAPH ASIA '24] GGHead: Fast and Generalizable 3D Gaussian Heads
★ 186DRTK. Differentiable Rendering Toolkit
★ 158Regional-Prompting-FLUX. Training-free Regional Prompting for Diffusion Transformers 🔥
★ 696VHAP. A complete head tracking pipeline from videos to NeRF/3DGS-ready datasets.
★ 377echoscene. [ECCV 2024] EchoScene: Indoor Scene Generation via Information Echo over Scene Graph Diffusion.
★ 103hitchhiking-rotations. Learning with 3D rotations, a hitchhiker’s guide to SO(3) - ICML 2024
★ 276diffusion-e2e-ft. [WACV'25 Oral] Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think
★ 520SynthMoCap. SynthMoCap Datasets
★ 472nvTorchCam. Python
★ 156voxel-challenge. Python
★ 247taichi-splatting. Python
★ 120RADIO. Official repository for "AM-RADIO: Reduce All Domains Into One"
★ 1.9kperm. [ICLR 2025] Official implementation of "Perm: A Parametric Representation for Multi-Style 3D Hair Modeling"
★ 120motionfix. MotionFix: Text-Driven 3D Human Motion Editing [SIGGRAPH ASIA 2024]
★ 161HAAR. HAAR: Text-Conditioned Generative Model of 3D Strand-based Human Hairstyles (CVPR 2024)
★ 92foundpose. FoundPose: Unseen Object Pose Estimation with Foundation Features, ECCV 2024
★ 1353dgs-render-blender-addon. 3DGS Render by KIRI Engine
★ 1.1kGaussianHaircut. Gaussian Haircut: Human Hair Reconstruction with Strand-Aligned 3D Gaussians
★ 285GPU-Puzzles. Solve puzzles. Learn CUDA.
★ 12kFrosting. [ECCV 2024 - ORAL] Official PyTorch implementation of Gaussian Frosting: Editable Complex Radiance Fields with Real-Time Rendering
★ 308theseus. A library for differentiable nonlinear optimization
★ 2kslang-gaussian-rasterization. Slang
★ 376PuzzleAvatar. [SIGGRAPH Asia 2024] PuzzleAvatar: Assembling 3D Avatars from Personal Albums
★ 320