This is your work, valued
text-diffusion. Python
★ 13video-search. Python
★ 3Evol-Instruct-portuguese. Python
★ 2sentiment. simple sentiment analysis with python
★ 1little-book-rl. The Little Book of Reinforcement Learning
★ 1.4kVibeThinker. Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B
★ 1.5kgoat. Generalized Optimal Transport Attention with Trainable Priors
★ 70TradingAgents. TradingAgents: Multi-Agents LLM Financial Trading Framework
★ 95kdatabase-CEPS. Este repositório contém informações detalhadas sobre os CEPs de todas as cidades do Brasil que estão cadastradas nos Correios.
★ 36ANE. Training neural networks on Apple Neural Engine via reverse-engineered private APIs
★ 7.2kTRELLIS.2. Native and Compact Structured Latents for 3D Generation
★ 9.5kresume-template. The best LaTeX resume template.
★ 1.4kAReaL. The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
★ 5.6kOpen-DiffusionGS. Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction (ICCV 2025)
★ 858dllm. dLLM: Simple Diffusion Language Modeling
★ 2.7kalphageometry. Python
★ 4.9kopensloth. Python
★ 243duo. [ICML 2025] The Diffusion Duality
★ 237DreamOn. Diffusion Language Models For Code Infilling Beyond Fixed-size Canvas
★ 118oconfronto. Medieval Browser RPG game made with vanilla PHP
★ 13min-max-gpt. Minimal (400 LOC) implementation Maximum (multi-node, FSDP) GPT training
★ 132mdlm. [NeurIPS 2024] Simple and Effective Masked Diffusion Language Model
★ 707ManimML. ManimML is a project focused on providing animations and visualizations of common machine learning concepts with the Manim Community Library.
★ 3.5kdistilabel. Distilabel is a framework for synthetic data and AI feedback for engineers who need fast, reliable and scalable pipelines based on verified research papers.
★ 3.3klangflow. Langflow is a powerful tool for building and deploying AI-powered agents and workflows.
★ 153kdiffusion-4k. [CVPR 2025] Diffusion-4K: Ultra-High-Resolution Image Synthesis with Latent Diffusion Models
★ 364ReDi. [NeurIPS'25 Spotlight] Boosting Generative Image Modeling via Joint Image-Feature Synthesis
★ 121flowseq. An official pytorch implementation of EACL2024 short paper "Flow Matching for Conditional Text Generation in a Few Sampling Steps"
★ 34SimpleAR. Pytorch implementation for the paper titled "SimpleAR: Pushing the Frontier of Autoregressive Visual Generation"
★ 431Dream. Dream 7B, a large diffusion language model
★ 1.3ksemanticist. (ICCV 2025) "Principal Components" Enable A New Language of Images
★ 86latent-gemma. Jupyter Notebook
★ 27imm. Official implementation of Inductive Moment Matching
★ 585RDLM. Official Code Repository for the paper "Continuous Diffusion Model for Language Modeling" (NeurIPS 2025).
★ 74nano-mdm. Tiny re-implementation of MDM in style of LLaDA and nano-gpt speedrun
★ 57CrossFlow. [CVPR2025] PyTorch-based reimplementation of CrossFlow, as proposed in 'Flowing from Words to Pixels: A Noise-Free Framework for Cross-Modality Evolution'
★ 345native-sparse-attention-pytorch. Implementation of the sparse attention pattern proposed by the Deepseek team in their "Native Sparse Attention" paper
★ 811Liger-Kernel. Efficient Triton Kernels for LLM Training
★ 6.5kmicro_diffusion. Official repository for our work on micro-budget training of large-scale diffusion models.
★ 1.6kllama-box. LM inference server implementation based on *.cpp.
★ 292catvton-flux. Python
★ 614torchtune. PyTorch native post-training library
★ 5.8kSO-ARM100. Standard Open Arm 100
★ 6.9kInstantMesh. InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models
★ 4.5kOpenLRM. An open-source impl. of Large Reconstruction Models
★ 1.2kLP-3DGS. C++
★ 62splatt3r. Official repository for Splatt3R: Zero-shot Gaussian Splatting from Uncalibrated Image Pairs
★ 797ComfyUI-Fluxtapoz. Nodes for image juxtaposition for Flux in ComfyUI
★ 1.4kinstant-ngp. Instant neural graphics primitives: lightning fast NeRF and more
★ 18knGPT-pytorch. Quick implementation of nGPT, learning entirely on the hypersphere, from NvidiaAI
★ 300mini-omni. open-source multimodal large language model that can hear, talk while thinking. Featuring real-time end-to-end speech input and streaming audio output conversational capabilities.
★ 3.6kai-toolkit. The ultimate training toolkit for finetuning diffusion models
★ 11kstable-diffusion.cpp. Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
★ 6.6kfish-speech. SOTA Open Source TTS
★ 32kMVDream. Multi-view Diffusion for 3D Generation
★ 987MVDream-threestudio. 3D generation code for MVDream
★ 549transformers_zamba2. Python
★ 49Zamba2. PyTorch implementation of models from the Zamba2 series.
★ 193DistillKit. An Open Source Toolkit For LLM Distillation
★ 993NeuralFlow. Visualize the intermediate output of Mistral 7B
★ 398framed-multi30k. Jupyter Notebook
★ 2ChunkLlama. [ICML'24] Data and code for our paper "Training-Free Long-Context Scaling of Large Language Models"
★ 451Glyph-ByT5. [ECCV2024] This is an official inference code of the paper "Glyph-ByT5: A Customized Text Encoder for Accurate Visual Text Rendering" and "Glyph-ByT5-v2: A Strong Aesthetic Baseline for Accurate Multilingual Visual Text Rendering""
★ 625mamba2-torch. Python
★ 53matmulfreellm. Implementation for MatMul-free LM.
★ 3.1kRETRO-pytorch. Implementation of RETRO, Deepmind's Retrieval based Attention net, in Pytorch
★ 879neuron_as_deep_net. Code behind the work "Single Cortical Neurons as Deep Artificial Neural Networks", published in Neuron 2021
★ 165minRF. Minimal implementation of scalable rectified flow transformers, based on SD3's approach
★ 641awesome-kan. A comprehensive collection of KAN(Kolmogorov-Arnold Network)-related resources, including libraries, projects, tutorials, papers, and more, for researchers and developers in the Kolmogorov-Arnold Network field.
★ 3.3kfast-kan. FastKAN: Very Fast Implementation of Kolmogorov-Arnold Networks (KAN)
★ 490gym-chess. A simple chess environment for openai/gym
★ 162efficient-kan. An efficient pure-PyTorch implementation of Kolmogorov-Arnold Network (KAN).
★ 4.7kFourierKAN. Python
★ 755pykan. Kolmogorov Arnold Networks
★ 16koutlines. Structured Outputs
★ 15kinfini-transformer. PyTorch implementation of Infini-Transformer from "Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention" (https://arxiv.org/abs/2404.07143)
★ 300spacebyte. A byte-level decoder architecture that matches the performance of tokenized Transformers.
★ 67InfiniTransformer. Unofficial PyTorch/🤗Transformers(Gemma/Llama3) implementation of Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
★ 375AutoCompressors. [EMNLP 2023] Adapting Language Models to Compress Long Contexts
★ 337attorch. A subset of PyTorch's neural network modules, written in Python using OpenAI's Triton.
★ 604llama-cpp-python. Python bindings for llama.cpp
★ 11kLLMLingua. [EMNLP'23, ACL'24] To speed up LLMs' inference and enhance LLM's perceive of key information, compress the prompt and KV-Cache, which achieves up to 20x compression with minimal performance loss.
★ 6.5kconditional-flow-matching. TorchCFM: a Conditional Flow Matching library
★ 2.6kflow-matching. Annotated Flow Matching paper
★ 235BitMat. An efficent implementation of the method proposed in "The Era of 1-bit LLMs"
★ 155smaller-transformers. Load What You Need: Smaller Multilingual Transformers for Pytorch and TensorFlow 2.0.
★ 108unilm. Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
★ 22k1.58BitNet. Experimental BitNet Implementation
★ 73Triton-Puzzles. Puzzles for learning Triton
★ 2.5kquiet-star. Code for Quiet-STaR
★ 739Open-Sora. Open-Sora: Democratizing Efficient Video Production for All
★ 29kBitNet. Implementation of "BitNet: Scaling 1-bit Transformers for Large Language Models" in pytorch
★ 1.9kneedle-in-a-haystack. Doing simple retrieval from LLM models at various context lengths to measure accuracy
★ 2.4kSimpleBitNet. Simple Adaptation of BitNet
★ 31GaLore. GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
★ 1.7krebased. Official implementation of the paper "Linear Transformers with Learnable Kernel Functions are Better In-Context Models"
★ 169based. Code for exploring Based models from "Simple linear attention language models balance the recall-throughput tradeoff"
★ 256axolotl. Go ahead and axolotl questions
★ 12kDRUGS. Stop messing around with finicky sampling parameters and just use DRµGS!
★ 364TOVA. Token Omission Via Attention
★ 131laserRMT. This is our own implementation of 'Layer Selective Rank Reduction'
★ 240llm-autoeval. Automatically evaluate your LLMs in Google Colab
★ 695unsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kbovespaStockRatings. Crawler for Fundamental analysis platform for BOVESPA stocks, generating a score for each share according to the selected criteria on the indicators.
★ 210PowerInfer. High-speed Large Language Model Serving for Local Deployment
★ 9.7kPallaidium. PALLAIDIUM — a generative AI movie studio, seamlessly integrated into the Blender Video Editor (VSE), enabling end-to-end production from script to screen and back.
★ 1.5kpause-transformer. Yet another random morning idea to be quickly tried and architecture shared if it works; to allow the transformer to pause for any amount of time on any token
★ 53minimal-text-diffusion. A minimal implementation of diffusion models for text generation
★ 418openWakeWord. An open-source audio wake word (or phrase) detection framework with a focus on performance and simplicity.
★ 2.6kRealtimeSTT. A robust, efficient, low-latency speech-to-text library with advanced voice activity detection, wake word activation and instant transcription.
★ 10kwanda. A simple and effective LLM pruning approach.
★ 868lm-evaluation-harness. A framework for few-shot evaluation of language models.
★ 13ktest. Measuring Massive Multitask Language Understanding | ICLR 2021
★ 1.6kinstruct-eval. This repository contains code to quantitatively evaluate instruction-tuned models such as Alpaca and Flan-T5 on held-out tasks.
★ 552LLM-Pruner. [NeurIPS 2023] LLM-Pruner: On the Structural Pruning of Large Language Models. Support Llama-3/3.1, Llama-2, LLaMA, BLOOM, Vicuna, Baichuan, TinyLlama, etc.
★ 1.1kDoLa. Official implementation for the paper "DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models"
★ 557dream-textures. Stable Diffusion built-in to Blender
★ 8.2kCoDeF. [CVPR'24 Highlight] Official PyTorch implementation of CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
★ 4.8khugging-chat-api. HuggingChat Python API🤗
★ 939candle. Minimalist ML framework for Rust
★ 21kllama2.rs. A fast llama2 decoder in pure Rust.
★ 1.1kvllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88ksoftmax_variants. PyTorch code for softmax variants: center loss, cosface loss, large-margin gaussian mixture, COCOLoss, ring loss
★ 254MeZO. [NeurIPS 2023] MeZO: Fine-Tuning Language Models with Just Forward Passes. https://arxiv.org/abs/2305.17333
★ 1.2kSophia. The official implementation of “Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training”
★ 1kgptq. Code for the ICLR 2023 paper "GPTQ: Accurate Post-training Quantization of Generative Pretrained Transformers".
★ 2.3kCodeCapybara. Open-source Self-Instruction Tuning Code LLM
★ 172DeepFakeTorch. Python
★ 35StableLM. StableLM: Stability AI Language Models
★ 16kRWKV-LM. RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.
★ 15kRRHF. [NIPS2023] RRHF & Wombat
★ 805EditAnything. Edit anything in images powered by segment-anything, ControlNet, StableDiffusion, etc. (ACM MM)
★ 3.4kGrounded-Segment-Anything. Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
★ 18kSD-CN-Animation. This script allows to automate video stylization task using StableDiffusion and ControlNet.
★ 819tomesd. Speed up Stable Diffusion with this one simple trick!
★ 1.4kText2Video-Zero. [ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators
★ 4.2kalign_sd. Better Aligning Text-to-Image Models with Human Preference. ICCV 2023
★ 292modern-srwm. Official repository for the paper "A Modern Self-Referential Weight Matrix That Learns to Modify Itself" (ICML 2022 & NeurIPS 2021 Deep RL Workshop) and "Accelerating Neural Self-Improvement via Bootstrapping" (ICLR 2023 Workshop)
★ 177modelscope. ModelScope: bring the notion of Model-as-a-Service to life.
★ 9.1kalpaca-lora. Instruct-tune LLaMA on consumer hardware
★ 19kstanford_alpaca. Code and documentation to train Stanford's Alpaca models, and generate the data.
★ 30kggml. Tensor library for machine learning
★ 15kgigagan-pytorch. Implementation of GigaGAN, new SOTA GAN out of Adobe. Culmination of nearly a decade of research into GANs
★ 1.9kqdrant-client. Python client for Qdrant vector search engine
★ 1.3kqdrant. Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/
★ 34kSpikeGPT. Implementation of "SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks"
★ 911torchscale. Foundation Architecture for (M)LLMs
★ 3.1kdenoising-diffusion-gan. Tackling the Generative Learning Trilemma with Denoising Diffusion GANs https://arxiv.org/abs/2112.07804
★ 759automl. Google Brain AutoML
★ 6.5kUniversal-Guided-Diffusion. Jupyter Notebook
★ 511ControlNet. Let us control diffusion models!
★ 34kwhisper.cpp. Port of OpenAI's Whisper model in C/C++
★ 52klanguage-model-agents. Experiments with generating opensource language model assistants
★ 97peft. 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
★ 21kLAVIS. LAVIS - A One-stop Library for Language-Vision Intelligence
★ 11kkernl. Kernl lets you run PyTorch transformer models several times faster on GPU with a single line of code, and is designed to be easily hackable.
★ 1.6kGLM-130B. GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)
★ 7.7kinstruct-pix2pix. Python
★ 6.9ksd-leap-booster. Fast finetuning using a booster model that puts the initial state to a local minimum
★ 113muse-maskgit-pytorch. Implementation of Muse: Text-to-Image Generation via Masked Generative Transformers, in Pytorch
★ 918reward-modeling. Python
★ 98trlx. A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
★ 4.8kStructured-Diffusion-Guidance. Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis
★ 321CfC. Closed-form Continuous-time Neural Networks
★ 1.1kCrossAttentionControl. Unofficial implementation of "Prompt-to-Prompt Image Editing with Cross Attention Control" with Stable Diffusion
★ 1.3kstablediffusion-infinity. Outpainting with Stable Diffusion on an infinite canvas
★ 3.9kstable-diffusion. Jupyter Notebook
★ 1.5kunified-io-inference. Jupyter Notebook
★ 231stable-diffusion-webui. Stable Diffusion web UI
★ 164kwaifu-diffusion. stable diffusion finetuned on weeb stuff
★ 1.9kdpm-solver. Official code for "DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 Steps" (Neurips 2022 Oral)
★ 1.9kstable-diffusion. Optimized Stable Diffusion modified to run on lower GPU VRAM
★ 3.1kdiffusers-interpret. Diffusers-Interpret 🤗🧨🕵️♀️: Model explainability for 🤗 Diffusers. Get explanations for your generated images.
★ 278glid-3-xl-stable. stable diffusion training
★ 295stable-diffusion. A latent text-to-image diffusion model
★ 73kPalette-Image-to-Image-Diffusion-Models. Unofficial implementation of Palette: Image-to-Image Diffusion Models by Pytorch
★ 1.8kCEIA-ProjetoImaginie. Repositório dos códigos desenvolvidos no projeto Algoritmos Inteligentes para captação e retenção de usuários em plataformas educacionais.
★ 1darts. A python library for user-friendly forecasting and anomaly detection on time series.
★ 9.5kconvert-labse-tf-pt. Convert LaBSE model from TF Hub to PyTorch.
★ 15smaller-labse. Applying "Load What You Need: Smaller Versions of Multilingual BERT" to LaBSE
★ 20GFPGAN. GFPGAN aims at developing Practical Algorithms for Real-world Face Restoration.
★ 38kYourTTS. YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone
★ 1.1kblitz-bayesian-deep-learning. A simple and extensible library to create Bayesian Neural Network layers on PyTorch.
★ 983pytorch-spiking. Spiking neuron integration for PyTorch
★ 44notebooks. Notebooks using the Hugging Face libraries 🤗
★ 4.6ktransformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163kserver-express-typescript-template. JavaScript
★ 2react-template. TypeScript
★ 2previsao-do-tempo. Site de previsão do tempo construído para demonstrar conhecimentos em Typescript, React, Redux, Python e Docker
★ 3Cirq. Python framework for creating, editing, and running Noisy Intermediate-Scale Quantum (NISQ) circuits.
★ 5k