This is your work, valued
ftrl_proximal_lr. Multithreaded Asynchronous FTRL Proximal Implementation
★ 126awesome-machine-learning. A curated list of awesome Machine Learning frameworks, libraries and software.
★ 2FlashQLA. high-performance linear attention kernel library built on TileLang
★ 621TileKernels. A kernel library written in tilelang
★ 1.7kml-intern. 🤗 ml-intern: an open-source ML engineer that reads papers, trains models, and ships ML models
★ 11kCubeSandbox. Instant, Concurrent, Secure & Lightweight Sandbox for AI Agents.
★ 11kFlashKDA. FlashKDA: high-performance Kimi Delta Attention kernels
★ 1.1kAwesome-LLM-Scientific-Discovery. [EMNLP2025] From Automation to Autonomy: A Survey on Large Language Models in Scientific Discovery
★ 423pinchtab. High-performance browser automation bridge and multi-instance orchestrator with advanced stealth injection and real-time dashboard.
★ 9.8kSteptronOss. A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modular configuration across SFT, RLVR, and evaluation workflows.
★ 579NextStep-1. [🚀 ICLR 2026 Oral] NextStep-1: SOTA Autogressive Image Generation with Continuous Tokens. A research project developed by the StepFun’s Multimodal Intelligence team.
★ 693slime. slime is an LLM post-training framework for RL Scaling.
★ 7.7kDistRL-open. Python
★ 23vibetensor. Our first fully AI generated deep learning system
★ 634cosmos-reason2. Cosmos-Reason2 models understand the physical common sense and generate appropriate embodied decisions in natural language through long chain-of-thought reasoning processes.
★ 432cosmos-rl. Cosmos-RL is a flexible and scalable Reinforcement Learning framework specialized for Physical AI applications.
★ 468SimpleVLA-RL. [ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
★ 1.8kopenpi. Python
★ 13kpipelining-sft. Simple and efficient DeepSeek V3 SFT using pipeline parallel and expert parallel, with both FP8 and BF16 trainings
★ 118action_piece. Python
★ 66linear-attention-and-beyond-slides. TeX
★ 119gpt-oss. gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
★ 20kalphagenome. This API provides programmatic access to the AlphaGenome model developed by Google DeepMind.
★ 2kLiger-Kernel. Efficient Triton Kernels for LLM Training
★ 6.5kByteCheckpoint. ByteCheckpoint: An Unified Checkpointing Library for LFMs
★ 289Triton-distributed. Distributed Compiler based on Triton for Parallel Systems
★ 1.5kDeepResearch. Tongyi Deep Research, the Leading Open-source Deep Research Agent
★ 20kMegakernels. Kernels, of the mega variety :)
★ 788reverse-engineering-gemma-3n. Reverse Engineering Gemma 3n: Google's New Edge-Optimized Language Model
★ 279Spurious_Rewards. Python
★ 361ChinaTextbook. 所有小初高、大学PDF教材。
★ 76kNeMo-Aligner. Scalable toolkit for efficient model alignment
★ 852transfusion-pytorch. Pytorch implementation of Transfusion, "Predict the Next Token and Diffuse Images with One Multi-Modal Model", from MetaAI
★ 1.4krllm. Democratizing Reinforcement Learning for LLMs
★ 5.7kAReaL. The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
★ 5.6kunderstand-r1-zero. Understanding R1-Zero-Like Training: A Critical Perspective
★ 1.3kDAPO. An Open-source RL System from ByteDance Seed and Tsinghua AIR
★ 1.8ktorchtitan. A PyTorch native platform for training generative AI models
★ 5.6kopen-r1. Fully open reproduction of DeepSeek-R1
★ 26kflux. A fast communication-overlapping library for tensor/expert parallelism on GPUs.
★ 1.4kOLMo-core. PyTorch building blocks for the OLMo ecosystem
★ 1.4ksmollm. Everything about the SmolLM and SmolVLM family of models
★ 3.9k3FS. A high-performance distributed file system designed to address the challenges of AI training and inference workloads.
★ 10kDualPipe. A bidirectional pipeline parallelism algorithm for computation-communication overlap in DeepSeek V3/R1 training.
★ 3kAwesome-System2-Reasoning-LLM. Latest Advances on System-2 Reasoning
★ 1.4kSpargeAttn. [ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
★ 1kDeepEP. DeepEP: an efficient expert-parallel communication library
★ 9.9kvattention. Dynamic Memory Management for Serving LLMs without PagedAttention
★ 506FlashMLA. FlashMLA: Efficient Multi-head Latent Attention Kernels
★ 13kOpen-Reasoner-Zero. Official Repo for Open-Reasoner-Zero
★ 2.1kRAGEN. RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.
★ 2.8kRecFlow. Python
★ 74BladeDISC. BladeDISC is an end-to-end DynamIc Shape Compiler project for machine learning workloads.
★ 931PQCache. [SIGMOD 2025] PQCache: Product Quantization-based KVCache for Long Context LLM Inference
★ 91Awesome-Multimodal-Large-Language-Models. :sparkles::sparkles:Latest Advances on Multimodal Large Language Models
★ 18ktrl. Train transformer language models with reinforcement learning.
★ 19kSlow_Thinking_with_LLMs. A series of technical report on Slow Thinking with LLM
★ 767LRURec. [WSDM 2024] Official PyTorch Implementation of Linear Recurrent Units for Sequential Recommendation (LRURec)
★ 70DreamRec. Python
★ 134moe_attention. Official repository for the paper "SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention"
★ 101MoH. MoH: Multi-Head Attention as Mixture-of-Head Attention
★ 310olmes. Reproducible, flexible LLM evaluations
★ 390open-instruct. AllenAI's post-training codebase
★ 3.8kLLaMA-O1. Large Reasoning Models
★ 803nanoGPT-mup. The simplest, fastest repository for training/finetuning medium-sized GPTs.
★ 200mup. maximal update parametrization (µP)
★ 1.7kDeepSeek-Coder. DeepSeek Coder: Let the Code Write Itself
★ 24kopen_llama. OpenLLaMA, a permissively licensed open source reproduction of Meta AI’s LLaMA 7B trained on the RedPajama dataset
★ 7.5kOLMo. Modeling, training, eval, and inference code for OLMo
★ 6.6kopenr. OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
★ 1.8kO1-Journey. O1 Replication Journey
★ 2kbaselines. OpenAI Baselines: high-quality implementations of reinforcement learning algorithms
★ 17kquiet-star. Code for Quiet-STaR
★ 739triton. Development repository for the Triton language and compiler
★ 20kalphageometry. Python
★ 4.9kxformers. Hackable and optimized Transformers building blocks, supporting a composable construction.
★ 11kgrade-school-math. Python
★ 1.5kDCFormer. Python
★ 226math. The MATH Dataset (NeurIPS 2021)
★ 1.4kgpu.cpp. A lightweight library for portable low-level GPU computation using WebGPU.
★ 4kTransformerEngine. A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.
★ 3.5kdlrover. DLRover: An Automatic Distributed Deep Learning System
★ 1.7kOpenMoE. A family of open-sourced Mixture-of-Experts (MoE) Large Language Models
★ 1.7kgraphrag. A modular graph-based Retrieval-Augmented Generation (RAG) system
★ 35kmegablocks. Python
★ 1.6klingvo. Lingvo
★ 2.9krq-vae-transformer. The official implementation of Autoregressive Image Generation using Residual Quantization (CVPR '22)
★ 1kIndex-1.9B. A lightweight multilingual LLM
★ 1kk2-train. Python
★ 60MOSS-RLHF. Secrets of RLHF in Large Language Models Part I: PPO
★ 1.4knni. An open source AutoML toolkit for automate machine learning lifecycle, including feature engineering, neural architecture search, model compression and hyper-parameter tuning.
★ 14kgenerative-recommenders. Repository hosting code for "Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations" (https://arxiv.org/abs/2402.17152).
★ 2kllm.c. LLM training in simple, raw C/CUDA
★ 31kragas. Supercharge Your LLM Application Evaluations 🚀
★ 15kLLaVA. [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
★ 25ksemantic-kernel. Integrate cutting-edge LLM technology quickly and easily into your apps
★ 28kFlagEmbedding. Retrieval and Retrieval-augmented LLMs
★ 12ktext-embeddings-inference. A blazing fast inference solution for text embeddings models
★ 5kpromptflow. Build high-quality LLM apps - from prototyping, testing to production deployment and monitoring.
★ 11kPokemonRedExperiments. Playing Pokemon Red with Reinforcement Learning
★ 7.9kstreaming-llm. [ICLR 2024] Efficient Streaming Language Models with Attention Sinks
★ 7.3kDocsGPT. Private AI platform for agents, assistants and enterprise search. Built-in Agent Builder, Deep research, Document analysis, Multi-model support, and API connectivity for agents.
★ 18kToRA. ToRA is a series of Tool-integrated Reasoning LLM Agents designed to solve challenging mathematical reasoning problems by interacting with tools [ICLR'24].
★ 1.1kLanguageAgentTreeSearch. [ICML 2024] Official repository for "Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models"
★ 848dspy. DSPy: The framework for programming—not prompting—language models
★ 36kexllama. A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.
★ 2.9kgorilla. Gorilla: Training and Evaluating LLMs for Function Calls (Tool Calls)
★ 13kLightLLM. LightLLM is a Python-based LLM (Large Language Model) inference and serving framework, notable for its lightweight design, easy scalability, and high-speed performance.
★ 4.2kquivr. Opiniated RAG for integrating GenAI in your apps 🧠 Focus on your product rather than the RAG. Easy integration in existing products with customisation! Any LLM: GPT4, Groq, Llama. Any Vectorstore: PGVector, Faiss. Any Files. Anyway you want.
★ 39kPlatypus. Code for fine-tuning Platypus fam LLMs using LoRA
★ 625AgentBench. A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
★ 3.6kllama2.c. Inference Llama 2 in one file of pure C
★ 20kAutoGPT. AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
★ 186kInternLM. Official release of InternLM series (InternLM, InternLM2, InternLM2.5, InternLM3).
★ 7.3killa-builder. Low-code platform allows you to build business apps, enables you to quickly create internal tools such as dashboard, crud app, admin panel, crm, cms, etc. Supports PostgreSQL, MySQL, Supabase, GraphQL, MongoDB, MSSQL, Rest API, Hugging Face, Redis, etc. Automate workflows with schedule or webhook. Open source Retool.
★ 12kGrounding_LLMs_with_online_RL. We perform functional grounding of LLMs' knowledge in BabyAI-Text
★ 276ReAct. [ICLR 2023] ReAct: Synergizing Reasoning and Acting in Language Models
★ 4.1klmdeploy. LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
★ 8kvllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88kMetaGPT. 🌟 The Multi-Agent Framework: First AI Software Company, Towards Natural Language Programming
★ 70kChatLaw. ChatLaw:A Powerful LLM Tailored for Chinese Legal. 中文法律大模型
★ 7.6kFirefly. Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型
★ 6.7kstable-diffusion. A latent text-to-image diffusion model
★ 73kLLM-Blender. [ACL2023] We introduce LLM-Blender, an innovative ensembling framework to attain consistently superior performance by leveraging the diverse strengths of multiple open-source LLMs. LLM-Blender cut the weaknesses through ranking and integrate the strengths through fusing generation to enhance the capability of LLMs.
★ 990FasterViT. [ICLR 2024] Official PyTorch implementation of FasterViT: Fast Vision Transformers with Hierarchical Attention
★ 922InternLM-techreport.
★ 895Youku-mPLUG. Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Pre-training Dataset and Benchmarks
★ 307YuLan-Chat. YuLan: An Open-Source Large Language Model
★ 634TigerBot. TigerBot: A multi-language multi-task LLM
★ 2.3kawesome-chinese-legal-resources. 📝 An Awesome Collection of Chinese Legal Dataset and Relevant Resources. 致力于收集全面的中文法律数据源
★ 993DragGAN. Official Code for DragGAN (SIGGRAPH 2023)
★ 36kcc_net. Tools to download and cleanup Common Crawl data
★ 1klm-evaluation-harness. A framework for few-shot evaluation of language models.
★ 13kcriu. Checkpoint/Restore tool
★ 3.9kCoT-Collection. [EMNLP 2023] The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning
★ 258lit-llama. Implementation of the LLaMA language model based on nanoGPT. Supports flash attention, Int8 and GPTQ 4bit quantization, LoRA and LLaMA-Adapter fine-tuning, pre-training. Apache 2.0-licensed.
★ 6.1kllm-foundry. LLM training code for Databricks foundation models
★ 4.4kPandaGPT. [TLLM'23] PandaGPT: One Model To Instruction-Follow Them All
★ 864pile_dedupe. Pile Deduplication Code
★ 18MaskDINO. [CVPR 2023] Official implementation of the paper "Mask DINO: Towards A Unified Transformer-based Framework for Object Detection and Segmentation"
★ 1.5kList-of-Dirty-Naughty-Obscene-and-Otherwise-Bad-Words. List of Dirty, Naughty, Obscene, and Otherwise Bad Words
★ 3.4kbigscience. Central place for the engineering/scaling WG: documentation, SLURM scripts and logs, compute environment and data.
★ 1kqlora. QLoRA: Efficient Finetuning of Quantized LLMs
★ 11ktree-of-thoughts. Plug in and Play Implementation of Tree of Thoughts: Deliberate Problem Solving with Large Language Models that Elevates Model Reasoning by atleast 70%
★ 4.6kLLMSurvey. The official GitHub page for the survey paper "A Survey of Large Language Models".
★ 12kLinly. Chinese-LLaMA 1&2、Chinese-Falcon 基础模型;ChatFlow中文对话模型;中文OpenLLaMA模型;NLP预训练/指令微调数据集
★ 3kRL4LMs. A modular RL library to fine-tune language models to human preferences
★ 2.4kWizardLM. LLMs build upon Evol Insturct: WizardLM, WizardCoder, WizardMath
★ 9.5kunlimiformer. Public repo for the NeurIPS 2023 paper "Unlimiformer: Long-Range Transformers with Unlimited Length Input"
★ 1.1kchain-of-thought-hub. Benchmarking large language models' complex reasoning ability with chain-of-thought prompting
★ 2.8kalpa. Training and serving large-scale neural networks with auto parallelization.
★ 3.2kInstructionZoo.
★ 283LLMZoo. ⚡LLM Zoo is a project that provides data, models, and evaluation benchmark for large language models.⚡
★ 2.9kchameleon-llm. Codes for "Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models".
★ 1.1kMOSS. An open-source tool-augmented conversational language model from Fudan University
★ 12kStableLM. StableLM: Stability AI Language Models
★ 16kMiniGPT-4. Open-sourced codes for MiniGPT-4 and MiniGPT-v2 (https://minigpt-4.github.io, https://minigpt-v2.github.io/)
★ 26khh-rlhf. Human preference data for "Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback"
★ 1.9krandomfun. Notebooks and various random fun
★ 1.2kdolly. Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform
★ 11kTencentPretrain. Tencent Pre-training framework in PyTorch & Pre-trained Model Zoo
★ 1.1khelm. Holistic Evaluation of Language Models (HELM) is an open source Python framework created by the Center for Research on Foundation Models (CRFM) at Stanford for holistic, reproducible and transparent evaluation of foundation models, including large language models (LLMs) and multimodal models.
★ 2.9kLlama-X. Open Academic Research on Improving LLaMA to SOTA LLM
★ 1.6kMarkTool. DoTAT 是一款基于web、面向领域的通用文本标注工具,支持大规模实体标注、关系标注、事件标注、文本分类、基于字典匹配和正则匹配的自动标注以及用于实现归一化的标准名标注,同时也支持迭代标注、嵌套实体标注和嵌套事件标注。标注规范可自定义且同类型任务中可“一次创建多次复用”。通过分级实体集合扩大了实体类型的规模,并设计了全新高效的标注方式,提升了用户体验和标注效率。此外,本工具增加了审核环节,可对多人的标注结果进行一致性检验、自动合并和手动调整,提高了标注结果的准确率。
★ 625LMFlow. An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.
★ 8.5kJARVIS. JARVIS, a system to connect LLMs with ML community. Paper: https://arxiv.org/pdf/2303.17580.pdf
★ 25kFasterTransformer. Transformer related optimization, including BERT, GPT
★ 6.4kthe-algorithm. Source code for the X Recommendation Algorithm
★ 74kFastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40kBELLE. BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)
★ 8.3kLoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
★ 14kLuotuo-Chinese-LLM. 骆驼(Luotuo): Open Sourced Chinese Language Models. Developed by 陈启源 @ 华中师范大学 & 李鲁鲁 @ 商汤科技 & 冷子昂 @ 商汤科技
★ 3.6kpytorch-vit. An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
★ 309ChatDoctor. Python
★ 3.6kChinese-LLaMA-Alpaca. 中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)
★ 19kchatgpt-retrieval-plugin. The ChatGPT Retrieval Plugin lets you easily find personal or work documents by asking questions in natural language.
★ 21kalpaca-lora. Instruct-tune LLaMA on consumer hardware
★ 19kdocument.ai. 基于向量数据库与GPT3.5的通用本地知识库方案(A universal local knowledge base solution based on vector database and GPT3.5)
★ 3.7kevals. Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
★ 19kgpt-neox. An implementation of model parallel autoregressive transformers on GPUs, based on the Megatron and DeepSpeed libraries
★ 7.4kFengshenbang-LM. Fengshenbang-LM(封神榜大模型)是IDEA研究院认知计算与自然语言研究中心主导的大模型开源体系,成为中文AIGC和认知智能的基础设施。
★ 4.1kChatGLM-6B. ChatGLM-6B: An Open Bilingual Dialogue Language Model | 开源双语对话语言模型
★ 41kstanford_alpaca. Code and documentation to train Stanford's Alpaca models, and generate the data.
★ 30ktaichi. Productive, portable, and performant GPU programming in Python.
★ 28kTaskMatrix. Python
★ 34koptimate. A collection of libraries to optimise AI model performances
★ 8.3kGLM-130B. GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)
★ 7.7kCVPR2026-Papers-with-Code. CVPR 2026 论文和开源项目合集
★ 23kOpenDelta. A plug-and-play library for parameter-efficient-tuning (Delta Tuning)
★ 1klangchain. The agent engineering platform.
★ 143kllama_index. LlamaIndex is the leading document agent and OCR platform
★ 51kpaper-qa. High accuracy RAG for answering questions from scientific documents with citations
★ 9kopen_clip. An open source implementation of CLIP.
★ 14ktrlx. A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
★ 4.8kPaLM-rlhf-pytorch. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
★ 7.9kpai. Resource scheduling and cluster management for AI
★ 2.7kDeepSpeed. DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
★ 43kMegatron-LM. Ongoing research training transformer models at scale
★ 17kgluten. Gluten is a middle layer responsible for offloading JVM-based SQL engines' execution to native engines.
★ 1.6klatent-diffusion. High-Resolution Image Synthesis with Latent Diffusion Models
★ 14kminiF2F. An updated version of miniF2F with lots of fixes and informal statements / solutions.
★ 104ControlNet. Let us control diffusion models!
★ 34k