This is your work, valued
CTO @ Subconscious Systems, Research Scientist @ MIT CSAIL
SAIL. SAIL: Search Augmented Instruction Learning
★ 160LangCode. LangCode - Improving alignment and reasoning of large language models (LLMs) with natural language embedded program (NLEP).
★ 51OIWE. Online Interpretable Word Embeddings
★ 37UniLC. Interpretable unified language safety checking with large language models
★ 32EntST. Entailment self-training
★ 27RGX. Synthetic QA generation for long documents.
★ 16PILM. Language model with phrase induction
★ 14bdfa-torch. Training neural networks with back-prop, feedback-alignment and direct feedback-alignment
★ 11ESP. Python
★ 5coatt-coref. Python
★ 4FG-GCN. GCN with feature graph
★ 3TIM-Multi-Page-Manga. Python
★ 3mooc_os. mooc_os
★ 2notebook. Jupyter Interactive Notebook
★ 1t-dialog. Topic driven dialog system
★ 1rex-encoder. REX encoder for dialog response selection
★ 1autodef. Lua
★ 1os_exercises. 清华大学OS公开课——课后练习
★ 1my-transformers. Python
★ 1Sipp. AI inference, packed simply. A blazing-fast, zero-dependency WebGPU runtime to run GGUF models directly in the browser. Features a symmetric API for seamless local execution and cloud provider routing. Built with Rust & C++.
★ 100rewardgen. Python
★ 6tapestry. Project Tapestry aims to give every nation and participant frontier AI they can call their own — uniting a global consortium to train a shared frontier model from which partners build and own sovereign models aligned to their national, socio-cultural, and industrial needs.
★ 230Evaluator. Open-source library for scalable, reproducible evaluation of AI models and benchmarks.
★ 319Nemotron. Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models
★ 1.8ktruss-examples. Examples of models deployable with Truss
★ 228llm-calculator. Interactive display to see if an LLM can fit on a device
★ 1atlas. Pure Rust Inference Engine
★ 620tokenspeed. TokenSpeed is a speed-of-light LLM inference engine.
★ 1.8kVeOmni. VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
★ 2.1kCodeWhale. Open-source, community-driven agent harness
★ 40kjinguyuan-dumpling-skill. 金谷园饺子馆.Skill - 北邮旁的饺子馆
★ 651pi. AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
★ 81kOpenSWE. Python
★ 201twitter-cli. A CLI for Twitter/X — feed, bookmarks, and user timeline in terminal
★ 2.8kdimos. Dimensional is the agentic operating system for physical space. Command humanoids, quadrupeds, drones, and other hardware platforms in natural language and build multi-agent systems that work seamlessly with physical input (cameras, lidar, actuators).
★ 3.8kinteractive_world_sim. [RSS 2026] Interactive World Simulator for Robot Policy Training and Evaluation
★ 281ms-swift. Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
★ 15kplay2prompt. Implementation for paper "PLAY2PROMPT: Zero-shot Tool Instruction Optimization for LLM Agents via Tool Play"
★ 7dataclaw. Agent harness to publish your agent chat history as Huggingface datasets.
★ 2.1ktau2-bench. τ-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
★ 1.7karchipelago. Harness for running and evaluating AI agents against RL environments
★ 227subconscious-python. Subconscious Python SDK
★ 2flash-attention. Fast and memory-efficient exact attention
★ 9ai-agent-degree-planner. An AI agent powered by the Subconscious SDK that searches official university course catalogs, extracts degree requirements, and generates a complete semester by semester plan of study using an agent runtime built for reliable multi step reasoning and verification.
★ 2agent-browser. Browser automation CLI for AI agents
★ 40kpixio. [CVPR 2026] Pixio: a capable vision encoder dedicated to dense prediction, simply by pixel reconstruction
★ 472NitroGen. A Foundation Model for Generalist Gaming Agents
★ 2.1kmini-sglang. A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.
★ 4.7kmarp. The entrance repository of Markdown presentation ecosystem
★ 12kbrowserbench. The First Benchmark for Browser Infrastructure Stealth
★ 20miles. Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.
★ 1.8kAVCD. [NeurIPS 2025] AVCD: Mitigating Hallucinations in Audio-Visual Large Language Models through Contrastive Decoding
★ 27Parallel-R1. The offical repo for "Parallel-R1: Towards Parallel Thinking via Reinforcement Learning"
★ 260EvaMAE. Python
★ 2unitree_sdk2. Unitree robot sdk version 2. https://support.unitree.com/home/zh/developer
★ 1.3kCookLikeHOC. 🥢像老乡鸡🐔那样做饭。已添加2026年发布的《老乡鸡菜品溯源报告 2.0中新出现的菜品。主要部分于2024年完工,非老乡鸡官方仓库。文字来自《老乡鸡菜品溯源报告》,并做归纳、编辑与整理。CookLikeHOC.
★ 24kunifolm-world-model-action. Python
★ 1.1kAgentLab. AgentLab: An open-source framework for developing, testing, and benchmarking web agents on diverse tasks, designed for scalability and reproducibility.
★ 610visualwebarena. VisualWebArena is a benchmark for multimodal agents.
★ 484LessIsMore. [ICML 2026] Less Is More: Training-Free Sparse Attention with Global Locality for Efficient Reasoning
★ 34system_prompts_leaks. Extracted system prompts from Anthropic - Claude Fable 5, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-5.6-Sol, Codex. Google - Gemini 3.5 Flash, 3.1 Pro, Antigravity. xAI - Grok, Cursor, Copilot, VS Code, Perplexity, and more. Updated regularly.
★ 61kdeep_research_bench. DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
★ 805mini-swe-agent. The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench verified!
★ 6.1kagent-lightning. The absolute trainer to light up AI agents.
★ 17kbrowser-use. 🌐 Make websites accessible for AI agents. Automate tasks online with ease.
★ 107kDeepResearch. Tongyi Deep Research, the Leading Open-source Deep Research Agent
★ 20kawesome-web-agents. 🔥 A list of tools, frameworks, and resources for building AI web agents
★ 1.5kMultiverse.
★ 88Multiverse-Engine. Customized Inference Engine for Multiverse Models
★ 25subconscious. JavaScript
★ 71learn-claude-code. Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
★ 73kRL-Factory. Train your Agent model via our easy and efficient framework
★ 1.8kcodex. Lightweight coding agent that runs in your terminal
★ 103kSelfCite. Code for the ICML 2025 paper "SelfCite Self-Supervised Alignment for Context Attribution in Large Language Models"
★ 30Awesome-ML-SYS-Tutorial. My learning notes for ML SYS.
★ 6.8kMuSR. Python
★ 57GPU-Puzzles. Solve puzzles. Learn CUDA.
★ 12kawesome-deepseek-integration. Integrate the DeepSeek API into popular software
★ 38kAIChip_Paper_List.
★ 672OpenManus. No fortress, purely open ground. OpenManus is Coming.
★ 58kllm_benchmarks. A collection of benchmarks and datasets for evaluating LLM.
★ 575UbiMoE. C++
★ 25transfomers-silicon-research. Research and Materials on Hardware implementation of Transformer Model
★ 309deep-research. An AI-powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models. The goal of this repo is to provide the simplest implementation of a deep research agent - e.g. an agent that can refine its research direction overtime and deep dive into a topic.
★ 19kPYNQ_tutorials. Jupyter Notebook
★ 108s1. s1: Simple test-time scaling
★ 6.7kVitis_Accel_Examples. Vitis_Accel_Examples
★ 603Transformer-Accelerator-Based-on-FPGA. You can run it on pynq z1. The repository contains the relevant Verilog code, Vivado configuration and C code for sdk testing. The size of the systolic array can be changed, now it is 16X16.
★ 264inference. Reference implementations of MLPerf® inference benchmarks
★ 1.6kvllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88kopen-r1. Fully open reproduction of DeepSeek-R1
★ 26kCNN-FPGA. 使用Verilog实现的CNN模块,可以方便的在FPGA项目中使用
★ 590AutoPresent. Code for the paper "AutoPresent: Designing Structured Visuals From Scratch" (CVPR 2025)
★ 175Vitis-AI. Vitis AI is Xilinx’s development stack for AI inference on Xilinx hardware platforms, including both edge devices and Alveo cards.
★ 1.8kPYNQ_Workshop. Jupyter Notebook
★ 479neuralNetwork. Verilog
★ 314the_well. A 15TB Collection of Physics Simulation Datasets
★ 4.3kLlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74kVivado-Design-Tutorials. Tcl
★ 270Step-DPO. Implementation for "Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs"
★ 398Thinking-Claude. Let your Claude able to think
★ 17kAwesome-Text2SQL. Curated tutorials and resources for Large Language Models, Text2SQL, Text2DSL、Text2API、Text2Vis and more.
★ 3.7kVitis-AI-Tutorials.
★ 512Vitis-Tutorials. Vitis In-Depth Tutorials
★ 1.6kdocling. Get your documents ready for gen AI
★ 64kFirefly. Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型
★ 6.7kflashinfer. FlashInfer: Kernel Library for LLM Serving
★ 6.1kAtom. [MLSys'24] Atom: Low-bit Quantization for Efficient and Accurate LLM Serving
★ 345BitNet. Official inference framework for 1-bit LLMs
★ 40kMinerU. Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
★ 76ktext-generation-inference. Large Language Model Text Generation Inference
★ 11kmem0. Universal memory layer for AI Agents
★ 62kReBase. ReBase: Training Task Experts through Retrieval Based Distillation
★ 28hilbertcurve. maps between 1-D space filling hilbert curve and N-D coordinates
★ 276omniplex. Open-Source Perplexity
★ 1kHalfrost-Field. ✍🏻 Source Code Deep Dives, System Design & Engineering Blogs | Halfrost-Field 冰霜之地:源码解析、系统设计与工程实践笔记
★ 13komniparse. Ingest, parse, and optimize any data format ➡️ from documents to multimedia ➡️ for enhanced compatibility with GenAI frameworks
★ 7.7kopen-webui. User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
★ 147kVane. Vane is an AI-powered answering engine.
★ 36ksearxng. SearXNG is a free internet metasearch engine which aggregates results from various search services and databases. Users are neither tracked nor profiled.
★ 35kthread. Python
★ 22HALC. [ICML 2024] Official implementation for "HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding"
★ 115llm_speech_emotion_challenge. Jupyter Notebook
★ 23fructose. Python
★ 747instructor. structured outputs for llms
★ 14kresearch-career-tools. Python
★ 164interactions.py. A highly extensible, easy to use, and feature complete bot framework for Discord
★ 874onelinerizer. Shamelessly convert any Python 2 script into a terrible single line of code
★ 1.5kMathPile. [NeurlPS D&B 2024] Generative AI for Math: MathPile
★ 418dinosr. DinoSR: Self-Distillation and Online Clustering for Self-supervised Speech Representation Learning
★ 53cav-mae. Code and Pretrained Models for ICLR 2023 Paper "Contrastive Audio-Visual Masked Autoencoder".
★ 292LangCode. LangCode - Improving alignment and reasoning of large language models (LLMs) with natural language embedded program (NLEP).
★ 51pytorch. Ascend PyTorch adapter (torch_npu). Mirror of https://gitcode.com/Ascend/pytorch
★ 558GraphUnsupASR. Shell
★ 10anchoring-ai. An open-source no-code tool for teams to collaborate on building, evaluating, and hosting applications leveraging GPT and other large language models. You could easily build and share LLM-powered apps, manage your budget and run batch jobs.
★ 156DoLa. Official implementation for the paper "DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models"
★ 557openinterpreter. A coding agent for open models like Kimi K3
★ 67kibm-generative-ai. IBM-Generative-AI is a Python library built on IBM's large language model REST interface to seamlessly integrate and extend this service in Python programs.
★ 261codellama. Inference code for CodeLlama models
★ 16kast. Code for the Interspeech 2021 paper "AST: Audio Spectrogram Transformer".
★ 1.5kTextBlob. Simple, Pythonic, text processing--Sentiment analysis, part-of-speech tagging, noun phrase extraction, translation, and more.
★ 9.5kwhisper-at. Code and Pretrained Models for Interspeech 2023 Paper "Whisper-AT: Noise-Robust Automatic Speech Recognizers are Also Strong Audio Event Taggers"
★ 422Awesome-Multimodal-Large-Language-Models. :sparkles::sparkles:Latest Advances on Multimodal Large Language Models
★ 18kleaked-system-prompts. Collection of leaked system prompts
★ 15kFakeNewsNet. This is a dataset for fake news detection research
★ 1.3kOpen-Instruction-Generalist. Open Instruction Generalist is an assistant trained on massive synthetic instructions to perform many millions of tasks
★ 210EAR. Code for the ACL 2023 long paper - Expand, Rerank, and Retrieve: Query Reranking for Open-Domain Question Answering
★ 38LogiCoT. the instructions and demonstrations for building a formal logical reasoning capable GLM
★ 54ltu. Code, Dataset, and Pretrained Models for Audio and Speech Large Language Model "Listen, Think, and Understand".
★ 478impact. ML has an impact on the climate. But not all models are born equal. Compute your model's emissions with our calculator and add the results to your paper with our generated latex template
★ 269Dromedary. Dromedary: towards helpful, ethical and reliable LLMs.
★ 1.1kssast. Code for the AAAI 2022 paper "SSAST: Self-Supervised Audio Spectrogram Transformer".
★ 430awesome-instruction-learning. Papers and Datasets on Instruction Tuning and Following. ✨✨✨
★ 512Awesome-Transformer-Attention. An ultimately comprehensive paper list of Vision Transformer/Attention, including papers, codes, and related websites
★ 5.1kLLaMA-Adapter. [ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parameters
★ 5.9ktrlx. A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
★ 4.8kGPT-4-LLM. Instruction Tuning with GPT-4
★ 4.3kddgs. A metasearch library that aggregates results from diverse web search services
★ 2.9kchatgpt-retrieval-plugin. The ChatGPT Retrieval Plugin lets you easily find personal or work documents by asking questions in natural language.
★ 21kAutomated-Fact-Checking-Resources. Links to conference/journal publications in automated fact-checking (resources for the TACL22/EMNLP23 paper).
★ 578llama. Inference code for Llama models
★ 60kFlexLLMGen. Running large language models on a single GPU for throughput-oriented scenarios.
★ 9.4kuavm. Code for the IEEE Signal Processing Letters 2022 paper "UAVM: Towards Unifying Audio and Visual Models".
★ 57awesome-fairness-papers. Papers on fairness in NLP
★ 452tuning_playbook. A playbook for systematically maximizing the performance of deep learning models.
★ 30kadapters. A Unified Library for Parameter-Efficient and Modular Transfer Learning
★ 2.8ktransformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163kclevr-dataset-gen. A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning
★ 653Imppres. This repository houses the IMPlicature and PRESupposition diagnostic dataset (IMPPRES), consisting of >25k semiautomatically generated sentence pairs illustrating well-studied pragmatic inference types. IMPPRES is an NLI dataset following the format of SNLI (Bowman et al., 2015), MultiNLI (Williams et al., 2018) and XNLI (Conneau et al., 2018), which was created to determine how well trained NLI models do on recognizing several kinds of presuppositions and scalar implicatures.
★ 19CausalMediationAnalysis. Code for the paper "Causal Mediation Analysis for Interpreting Neural NLP: The Case of Gender Bias"
★ 81aclpubcheck. Tools for checking ACL paper submissions
★ 1kStereoSet. StereoSet: Measuring stereotypical bias in pretrained language models
★ 204ColossalAI. Making large AI models cheaper, faster and more accessible
★ 41kpyprobml. Python code for "Probabilistic Machine learning" book by Kevin Murphy
★ 7.1kHowToCook. Programmer's guide about how to cook at home.
★ 101kPython-100-Days. Python - 100天从新手到大师
★ 185kllm-beginner. LLM、Agent上手教程
★ 6.6klanguage. Shared repository for open-sourced projects from the Google AI Language team.
★ 1.8ksearch-tweets-python. Python client for the Twitter 'search Tweets' and 'count Tweets' endpoints (v2/Labs/premium/enterprise). Now supports Twitter API v2 /recent and /all search endpoints.
★ 898TAADpapers. Must-read Papers on Textual Adversarial Attack and Defense
★ 1.6kMedical_DS. Python
★ 62jiant. jiant is an nlp toolkit
★ 1.7kstate-of-the-art-result-for-machine-learning-problems. This repository provides state of the art (SoTA) results for all machine learning problems. We do our best to keep this repository up to date. If you do find a problem's SoTA result is out of date or missing, please raise this as an issue or submit Google form (with this information: research paper name, dataset, metric, source code and year). We will fix it immediately.
★ 8.9kgraph-based-nn. Graph Convolutional Networks (GCNs)
★ 920MIT-6.824-2016. MIT 6.824 2016
★ 156ALSTM-cQA. Python
★ 1e2e-coref. End-to-end Neural Coreference Resolution
★ 525deep-coref. Python
★ 269NeuralDialogPapers. Summary of deep learning models for dialog systems (Tiancheng Zhao LTI, CMU)
★ 643guesswhat. GuessWhat?! Baselines
★ 75the-incredible-pytorch. The Incredible PyTorch: a curated list of tutorials, papers, projects, communities and more relating to PyTorch.
★ 13kdeeplearning-papernotes. Summaries and notes on Deep Learning research papers
★ 4.4kParlAI. A framework for training and evaluating AI models on a variety of openly available dialogue datasets.
★ 11kdeep-reinforcement-learning-papers. A list of recent papers regarding deep reinforcement learning
★ 2.2kawesome-public-datasets. A topic-centric list of HQ open datasets.
★ 78kDeepRL-InformationExtraction. Code for the paper "Improving Information Extraction by Acquiring External Evidence with Reinforcement Learning" http://arxiv.org/abs/1603.07954
★ 235chatbot-retrieval. Dual LSTM Encoder for Dialog Response Generation
★ 1.6kreinforcement-learning. Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course.
★ 22kDeepQA. My tensorflow implementation of "A neural conversational model", a Deep learning based chatbot
★ 2.9kdeep-residual-networks. Deep Residual Learning for Image Recognition
★ 6.7kkmeans. kmeans unsupervised pretraining from "Adam Coates and Andrew Y. Ng DEMYSTIFYING UNSUPERVISED FEATURE LEARNING"
★ 2TensorFlow-Examples. TensorFlow Tutorial and Examples for Beginners (support TF v1 & v2)
★ 44kdemos. Demos and tutorials around Torch7.
★ 355MITIE. MITIE: library and tools for information extraction
★ 3k