This is your work, valued
ComfyUI-3D-Pack. An extensive node suite that enables ComfyUI to process 3D inputs (Mesh & UV Texture, etc) using cutting edge algorithms (3DGS, NeRF, etc.)
★ 3.8kComfyUI-AnimateAnyone-Evolved. Improved AnimateAnyone implementation that allows you to use the opse image sequence and reference image to generate stylized video
★ 561UnitySkinnedMeshSlicer_3A. Concise basic functions for slice skinned mesh in Unity3D; Accurately, Alternately, Asynchronously ; )
★ 70Comfy3D_Pre_Builds. All the pre built dependencies for Comfy3D for all supported platforms
★ 62MagnetsGiant_Unity-DOTS-Based-DotsEffect. A giant made of magnets that no one can destory!!!
★ 20FlexiSyncMVD. Python
★ 19Generative_Models_Collection. Generative Adversarial Network related code and info collection
★ 19DC_SSDAE. Deep Compression Single-Step Diffusion Autoencoder
★ 3magic-animate. MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model
★ 2Weighted-UFC-Fight-Predictor. Use enhanced methods of feature engineering and construct deep machine learning model to predict the UFC fight statistics result in detail
★ 2ComfyUI-Manager. JavaScript
★ 2Hunyuan3D-2. High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models.
★ 2ML_PhysicalCharacter2D. Make DeepMimic work in unity for 2D character using 3D physics
★ 2Path-to-AMI. Research Articles that documenting Useful Ideas & Interesting Findings on my Journey to build Agile Machine Intelligence (prefer calling it AMI instead of AGI)
★ 1ComfyUI-KJNodes. Various custom nodes for ComfyUI
★ 1dreamgaussian. Generative Gaussian Splatting for Efficient 3D Content Creation
★ 1uMyo_python_tools. Python stuff for working with uMyo via USB base station
★ 1uMyo. Get this open source EMG sensor!
★ 1DPPO-tf2. Simple DPPO & PPO implement use tensorflow v2
★ 1Blender-For-UnrealEngine-Addons. I have created this addons for export asset from Blender to Unreal Engine
★ 2.6kXNALara-io-Tools. Blender Import/Export add-on for XPS Models and Poses [Blender 4.x port]
★ 87three.js. JavaScript 3D Library.
★ 114kBabylon.js. Babylon.js is a powerful, beautiful, simple, and open game and rendering engine packed into a friendly JavaScript framework.
★ 26kllm-viz. 3D Visualization of an GPT-style LLM
★ 5.5ktransformers.js. State-of-the-art Machine Learning for the web. Run 🤗 Transformers directly in your browser, with no need for a server!
★ 16konnx. Open standard for machine learning interoperability
★ 21konnxruntime. ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator
★ 21kmlflow. The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models and data.
★ 27kcaptum. Model interpretability and understanding for PyTorch
★ 5.7kshap. A game theoretic approach to explain the output of any machine learning model.
★ 26kopik. Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
★ 21ktorchlens. Capture every activation and gradient of any PyTorch model — forward and backward — with automatic graph visualization, rich metadata, and live interventions. Works on any architecture, including dynamic and recurrent ones.
★ 652tensorspace. Neural network 3D visualization framework, build interactive and intuitive model in browsers, support pre-trained deep learning models from TensorFlow, Keras, TensorFlow.js
★ 5.2krust-gpu. 🐉 Making Rust a first-class language and ecosystem for GPU shaders 🚧
★ 3.2knanite-at-home. Nanite made from scratch in Rust as my master thesis
★ 48nanite-webgpu. UE5's Nanite implementation using WebGPU. Includes the meshlet LOD hierarchy, software rasterizer and billboard impostors. Culling on both per-instance and per-meshlet basis.
★ 1.1kOLMo. Modeling, training, eval, and inference code for OLMo
★ 6.6khidden-networks. Python
★ 197netron. Visualizer for neural network, deep learning and machine learning models
★ 33ktorchvista. Interactive Pytorch forward pass visualization in notebooks
★ 759torchview. torchview: visualize pytorch models
★ 1.1kpytorchviz. A small package to create visualizations of PyTorch execution graphs
★ 3.5kMotionStream. MotionStream: Real-Time Video Generation with Interactive Motion Controls
★ 575memobase. User Profile-Based Long-Term Memory for AI Chatbot Applications.
★ 2.8kairi. 💖🧸 Self hosted, you-owned Grok Companion, a container of souls of waifu, cyber livings to bring them into our worlds, wishing to achieve Neuro-sama's altitude. Capable of realtime voice chat, Minecraft, Factorio playing. Web / macOS / Windows supported.
★ 45kkokoro_training. Training code for kokoro tts model
★ 45StyleTTS2. StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
★ 6.3kgodot. Godot Engine – Multi-platform 2D and 3D game engine
★ 115kPrompt-Baking. Jupyter Notebook
★ 5SLM-Bench. Python
★ 3SparCL. SparCL: Sparse Continual Learning on the Edge @ NeurIPS 22
★ 31LoRI. [COLM 2025] LoRI: Reducing Cross-Task Interference in Multi-Task Low-Rank Adaptation
★ 173chatterbox. SoTA open-source TTS
★ 26kdia. A TTS model capable of generating ultra-realistic dialogue in one pass.
★ 19kindex-tts. An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
★ 22kOrpheus-TTS. Towards Human-Sounding Speech
★ 6.3kEmotiVoice. EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine
★ 8.5kStableAvatar. We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-processing, conditioned on a reference image and audio.
★ 1.3kMultiTalk. [NeurIPS 2025] Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
★ 3kInfiniteTalk. Unlimited-length talking video generation that supports image-to-video and video-to-video generation
★ 7.5kton. Main TON monorepo
★ 4.1knobodywho. NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.
★ 1kLakonLab. Official implementation of AsymFlow, pi-Flow, GMFlow
★ 457flash-attention. Fast and memory-efficient exact attention
★ 25kBitNet. Official inference framework for 1-bit LLMs
★ 40kMobileLLM-R1. MobileLLM-R1
★ 86MobileLLM. MobileLLM Optimizing Sub-billion Parameter Language Models for On-Device Use Cases. In ICML 2024.
★ 1.5kai-toolkit. The ultimate training toolkit for finetuning diffusion models
★ 11kLightX2V-Qwen-Image-Lightning. Qwen-Image-Lightning: Speed up Qwen-Image model with distillation
★ 1.3kstreaming-vlm. StreamingVLM: Real-Time Understanding for Infinite Video Streams
★ 1kDC-VideoGen. DC-VideoGen: Efficient Video Generation with Deep Compression Video Autoencoder
★ 192EqM. Python
★ 209LiftFeat. Code for "LiftFeat: 3D Geometry-Aware Local Feature Matching", ICRA2025
★ 251efficientvit. Efficient vision foundation models for high-resolution generation and perception.
★ 3.3kRIPE. [ICCV 2025] Keypoint detection and description learned from image pairs only - no depth, no pose, no artificial augmentation required.
★ 119SSDD. Official implementation for SSDD Single-Step Diffusion Decoder for Efficient Image Tokenization.
★ 65LightX2V. Lightweight Image Video Action Generation Inference Framework
★ 2.5knunchaku. [ICLR2025 Spotlight] SVDQuant: Absorbing Outliers by Low-Rank Components for 4-Bit Diffusion Models
★ 3.9kcwm. Research code artifacts for Code World Model (CWM) including inference tools, reproducibility, and documentation.
★ 886fuwari. ✨A static blog template built with Astro.
★ 4.8kRandom_Pruning. [ICLR 2022] The Unreasonable Effectiveness of Random Pruning: Return of the Most Naive Baseline for Sparse Training by Shiwei Liu, Tianlong Chen, Xiaohan Chen, Li Shen, Decebal Constantin Mocanu, Zhangyang Wang, Mykola Pechenizkiy
★ 79DA-2. Official implementation of "DA²: Depth Anything in Any Direction"
★ 264DC-Gen. DC-Gen: Post-Training Diffusion Acceleration with Deeply Compressed Latent Space
★ 401hard-ash. Python
★ 1ShinkaEvolve. ShinkaEvolve: Towards Open-Ended and Sample-Efficient Program Evolution 🧬
★ 1.3kFlyModel. Unofficial Python implementation of "Algorithmic insights on continual learning from fruit flies"
★ 5SDMContinualLearner. Jupyter Notebook
★ 21Orientation_Matters. Official repository for the paper "Orientation Matters: Making 3D Generative Models Orientation-Aligned" (NeurIPS 2025)
★ 116SceneMover. Project of Siggraph Asia 2020 paper: Scene Mover: Automatic Move Planning for Scene Arrangement by Deep Reinforcement Learning
★ 98satsom. Saturation Self-Organizing Map
★ 6SageAttention. [ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, image, and video models.
★ 3.5kUniRig. [SIGGRAPH 2025] One Model to Rig Them All: Diverse Skeleton Rigging with UniRig
★ 1.7kElevate3D. Python
★ 192HunyuanWorld-1.0. Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels with Hunyuan3D World Model
★ 2.9kPartPacker. Efficient Part-level 3D Object Generation via Dual Volume Packing
★ 824MindJourney. [NeurIPS 2025] Source codes for the paper "MindJourney: Test-Time Scaling with World Models for Spatial Reasoning"
★ 151PartCrafter. [NeurIPS 2025] PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion Transformers
★ 2.5kBella. Bella is best
★ 6.4kHiDream-I1. Python
★ 2.5kcupy. NumPy & SciPy for GPU
★ 12kanima-x. Official Implementation of [AnimaX: Animating the Inanimate in 3D with Joint Video-Pose Diffusion Models]
★ 295brian2. Brian is a free, open source simulator for spiking neural networks.
★ 1.2kdendrify. Introducing dendrites to spiking neural networks. Designed for the Brian 2 simulator.
★ 44transformations. Homogeneous transformation matrices and quaternions.
★ 104Hunyuan3D-2.1. From Images to High-Fidelity 3D Assets with Production-Ready PBR Material
★ 3.8kstochastic_transition. Python
★ 1ComfyUI_InstantID. Python
★ 1.8kalgovivo. An energy-based formulation for soft-bodied virtual creatures
★ 342morphology-adaptive. Morphology-adaptive muscle-driven locomotion via attention mechanisms
★ 154x-flux-comfyui. Python
★ 1.7kComfyUI-IPAdapter-Flux. Python
★ 474Anymate. Python
★ 142worker-comfyui. ComfyUI as a serverless API on Runpod
★ 721docling. Get your documents ready for gen AI
★ 64kcrawl4ai. 🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN
★ 76kfirecrawl. The API to search, scrape, and interact with the web at scale. 🔥
★ 158kreader. Convert any URL to an LLM-friendly input with a simple prefix https://r.jina.ai/
★ 12kdeep-research. Use any LLMs (Large Language Models) for Deep Research. Support SSE API and MCP server.
★ 4.7kMV-Painter. Python
★ 246triton-windows. Fork of the Triton language and compiler for Windows support and easy installation
★ 2kDirect3D-S2. [NeurIPS 2025] Direct3D‑S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention
★ 1.3kComfyUI-MVAdapter. Custom nodes for using MV-Adapter in ComfyUI.
★ 473continuous-thought-machines. Continuous Thought Machines, because thought takes time and reasoning is a process.
★ 2kStep1X-3D. Step1X-3D: Towards High-Fidelity and Controllable Generation of Textured 3D Assets
★ 879liquid-s4. Liquid Structural State-Space Models
★ 397PyNN. A Python package for simulator-independent specification of neuronal network models.
★ 309bindsnet. Simulation of spiking neural networks (SNNs) using PyTorch.
★ 1.7knest-simulator. The NEST simulator
★ 657Absolute-Zero-Reasoner. Official Repository of Absolute Zero Reasoner
★ 1.9klocal-deep-research. ~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.
★ 8.8kpython-markdownify. Convert HTML to Markdown
★ 2.2kDeepGit. Deep research agent to help you find the best GitHub repositories 🕵️!
★ 895deep-searcher. Open Source Deep Research Alternative to Reason and Search on Private Data. Written in Python.
★ 8kdeep-research-web-ui. (Supports DeepSeek R1) An AI-powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models.
★ 2.2knode-DeepResearch. Keep searching, reading webpages, reasoning until it finds the answer (or exceeding the token budget)
★ 5.2kdeep-research. An AI-powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models. The goal of this repo is to provide the simplest implementation of a deep research agent - e.g. an agent that can refine its research direction overtime and deep dive into a topic.
★ 19klocal-deep-researcher. Fully local web research and report writing assistant
★ 9.3kopen_deep_research. Python
★ 12kEdgeRunner. [ICLR 2025] EdgeRunner: Auto-regressive Auto-encoder for Efficient Mesh Generation
★ 311Graphics-LPIPS. Python
★ 25NR-3DQA. Point cloud version for "No-Reference Quality Assessment for 3D Colored Point Cloud and Mesh models".
★ 33TripoSG. TripoSG: High-Fidelity 3D Shape Synthesis using Large-Scale Rectified Flow Models
★ 1.7ksleep-time-compute. accompanying material for sleep-time compute paper
★ 137liquid_time_constant_networks. Code Repository for Liquid Time-Constant Networks (LTCs)
★ 1.9kneo4j. Graphs for Everyone
★ 17kknowledge-graph-of-thoughts. Official Implementation of "Affordable AI Assistants with Knowledge Graph of Thoughts"
★ 232SAMPart3D. SAMPart3D: Segment Any Part in 3D Objects
★ 564HoloPart. HoloPart: Generative 3D Part Amodal Segmentation
★ 663Second-Me. Train your AI self, amplify you, bridge the world
★ 16kAI-Scientist-ICLR2025-Workshop-Experiment. Python
★ 302AI-Scientist-v2. The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
★ 6.9kd3. Bring data to life with SVG, Canvas and HTML. :bar_chart::chart_with_upwards_trend::tada:
★ 113kd3-force. Force-directed graph layout using velocity Verlet integration.
★ 2kTriplaneTurbo. [CVPR2025] Progressive Rendering Distillation: Adapting Stable Diffusion for Instant Text-to-Mesh Generation without 3D Data
★ 69PBR3DGen. [AAAI 2026] PBR3DGen: A VLM-guided mesh generation with High-Quality PBR texture
★ 51zstd. Zstandard - Fast real-time compression algorithm
★ 27kTexGaussian. [CVPR 2025] TexGaussian: Generating High-quality PBR Material via Octree-based 3D Gaussian Splatting
★ 73TripoSF. SparseFlex: High-Resolution and Arbitrary-Topology 3D Shape Modeling
★ 756Orient-Anything. Orient Anything, ICML 2025
★ 392Stable3DGen. A Modular Framework for 3D Generation and Beyond [WIP]
★ 1.3kTreeMeshGPT. [CVPR 2025] TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing
★ 192py7zr. 7zip in python3 with ZStandard, PPMd, LZMA2, LZMA1, Delta, BCJ, BZip2, and Deflate compressions, and AES encryption.
★ 555DeepMesh. [ICCV 2025] Official code of DeepMesh: Auto-Regressive Artist-mesh Creation with Reinforcement Learning
★ 732FlashVDM. Unleashing Vecset Diffusion Model for Fast Shape Generation / within 1 Second (ICCV'25 Highlight)
★ 333panda3d. Powerful, mature open-source cross-platform game engine for Python and C++, developed by Disney and CMU
★ 5.2kdm_nevis. NEVIS'22: Benchmarking the next generation of never-ending learners
★ 101MIDI-3D. [CVPR 2025] MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
★ 937DiscoPOP. Code for Discovering Preference Optimization Algorithms with and for Large Language Models
★ 196GPTSwarm. 🐝 The First Self-Improving agents with RL / Prompting Optimization
★ 1kllm_debate. Code release for "Debating with More Persuasive LLMs Leads to More Truthful Answers"
★ 131ADAS. [ICLR 2025] Automated Design of Agentic Systems
★ 1.6kWan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kAgentLaboratory. Agent Laboratory is an end-to-end autonomous research workflow meant to assist you as the human researcher toward implementing your research ideas
★ 5.8kMeshLib. Mesh processing library
★ 800pytorch_volumetric. Volumetric structures such as voxels and SDFs implemented in pytorch
★ 251recurrent-pretraining. Pretraining and inference code for a large-scale depth-recurrent language model
★ 903fpsample. Python efficient farthest point sampling (FPS) library. Compatible with numpy.
★ 183pyvista. 3D visualization and mesh analysis for science and engineering
★ 3.8kmeshio. :spider_web: input/output for many mesh formats
★ 2.3kPOP. Official implementation of the ICCV 2021 paper: "The Power of Points for Modeling Humans in Clothing".
★ 1933DShape2VecSet. Python
★ 548Dora. [CVPR 2025] Official repository for "Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders"
★ 586posthog. :hedgehog: PostHog is the leading platform for building self-driving products. Our developer tools – AI observability, analytics, session replay, flags, experiments, error tracking, logs, and more – capture all the context agents need to diagnose problems, uncover opportunities, and ship fixes. Steer it all from Slack, web, desktop, or the MCP.
★ 37kMCMat. Code of MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation;
★ 41ITEM3D. Our paper "Directional Texture Editing for 3D Models" has been accepted by Computer Graphics Forum (CGF), 2024
★ 5Janus. Janus-Series: Unified Multimodal Understanding and Generation Models
★ 18kMeshAnythingV2. [ICCV 2025] From anything to mesh like human artists. Official impl. of "MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization"
★ 1kleapfusion-hunyuan-image2video. A novel approach to hunyuan image-to-video sampling
★ 302CaPa. Official Repository of **CaPa**: Carve-n-Paint Synthesis for Efficient 4K Textured Mesh Generation
★ 165pybind11. Seamless operability between C++11 and Python
★ 18kgpytoolbox. A collection of utility functions to prototype geometry processing research in python
★ 277Hunyuan3D-2. High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models.
★ 14klarge_concept_model. Large Concept Models: Language modeling in a sentence representation space
★ 2.4kSlow_Thinking_with_LLMs. A series of technical report on Slow Thinking with LLM
★ 767Qwen3. Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.
★ 27kopenr. OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
★ 1.8kLlamaV-o1. [ACL 2025 🔥] Rethinking Step-by-step Visual Reasoning in LLMs
★ 307MiniMax-01. The official repo of MiniMax-Text-01 and MiniMax-VL-01, large-language-model & vision-language-model based on Linear Attention
★ 3.4kLlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74kSkyThought. Sky-T1: Train your own O1 preview model within $450
★ 3.4kloss-of-plasticity. Demonstrations of Loss of Plasticity and Implementation of Continual Backpropagation
★ 388metamotivo. The first behavioral foundation model to control a virtual physics-based humanoid agent for a wide range of whole-body tasks.
★ 778BrainSimIII. C#
★ 164semanticscholar. Unofficial Python client library for Semantic Scholar APIs.
★ 476stable-point-aware-3d. SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
★ 1.1kstreaming-drl. Deep reinforcement learning without experience replay, target networks, or batch updates.
★ 293scholarly. Retrieve author and publication information from Google Scholar in a friendly, Pythonic way without having to worry about CAPTCHAs!
★ 1.9karxiv.py. Python wrapper for the arXiv API
★ 1.5kVAR. [NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". An *ultra-simple, user-friendly yet state-of-the-art* codebase for autoregressive image generation!
★ 8.7kLatentSync. Taming Stable Diffusion for Lip Sync!
★ 5.9kreact-use-websocket. React Hook for WebSocket communication
★ 1.9kAwesome-Multimodal-Next-Token-Prediction. [Survey] Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey
★ 477sqlitebrowser. Official home of the DB Browser for SQLite (DB4S) project. Previously known as "SQLite Database Browser" and "Database Browser for SQLite". Website at:
★ 24kLayoutVLM. Official code for "LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models" (CVPR 2025)
★ 176awesome-vlm-architectures. Famous Vision Language Models and Their Architectures
★ 1.3kIDArb. Official repo for "IDArb: Intrinsic Decomposition for arbitrary number of input views and illuminations"
★ 102Comfy3D-WinPortable. 🧊ComfyUI-3D-Pack pre-built for Windows. | Comfy3D 整合包
★ 271Neural-LightRig. [CVPR2025] Neural LightRig: Unlocking Accurate Object Normal and Material Estimation with Multi-Light Diffusion
★ 192ComfyUI-IF_Trellis. ComfyUI TRELLIS is a large 3D asset generation in various formats, such as Radiance Fields, 3D Gaussians, and meshes. The cornerstone of TRELLIS is a unified Structured LATent (SLAT) representation that allows decoding to different output formats and Rectified Flow Transformers tailored for SLAT as the powerful backbones.
★ 450