This is your work, valued
Agentic Coding & Recursive Self-Improvement & AGI
TokenFormer. [ICLR2025 Spotlight🔥] Official Implementation of TokenFormer: Rethinking Transformer Scaling with Tokenized Model Parameters
★ 593DSVT. [CVPR2023] Official Implementation of "DSVT: Dynamic Sparse Voxel Transformer with Rotated Sets"
★ 455GiT. [ECCV2024 Oral🔥] Official Implementation of "GiT: Towards Generalist Vision Transformer through Universal Language Interface"
★ 364UniTR. [ICCV2023] Official Implementation of "UniTR: A Unified and Efficient Multi-Modal Transformer for Bird’s-Eye-View Representation"
★ 359CAGroup3D. [NeurIPS2022] This is the official code of "CAGroup3D: Class-Aware Grouping for 3D Object Detection on Point Clouds".
★ 96RBGNet. [CVPR2022] This is the official code of "RBGNet: Ray-based Grouping for 3D Object Detection".
★ 40DTI-GAT. Drug-Target Interaction Prediction with GraphAttention networks
★ 20MAVN. C
★ 15tmax. Training terminal-agents
★ 254agents-last-exam. Agents' Last Exam
★ 912SWE-CI. SWE-CI: Evaluating Agent Capabilities in Maintaining Codebases via Continuous Integration
★ 173DeepSpec. DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms
★ 6.8kcc-download. claude code install and update on windows OS
★ 87Qwen-AgentWorld. Qwen-AgentWorld: Language World Models for General Agents
★ 910mcore-bridge. MCore-Bridge: Providing Megatron-Core model definitions for state-of-the-art large models and making Megatron training as simple as Transformers — with support for 300+ large language models (Qwen3-Next, GLM-5.2, Deepseek-V4, MiniMax-2.7, ...) and 200+ multimodal large models (Qwen3.5, Qwen3-Omni, Gemma4, ...).
★ 87frontier-swe. FrontierSWE is an ultra long-horizon coding agent benchmark that tests implementation, performance eng and ML research
★ 197dgm. Darwin Gödel Machine: Open-Ended Evolution of Self-Improving Agents
★ 2.2kSEAL. Self-Adapting Language Models
★ 1.8kskills. Skills for Real Engineers. Straight from my .agents directory.
★ 194kmobilegym. MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research · 浏览器里运行的安卓模拟器 · Browser-hosted Android Simulator · Verifiable Evaluation · Scalable Online RL Training
★ 740OpenHarnessExtended. Python
★ 2Claw-Anything. official implementation of "Claw-Anything: Benchmarking Always-On Personal Assistants with Broader Access to the User's Digital World"
★ 45Awesome-Issue-Resolution. Advances and Frontiers of LLM-based Issue Resolution in Software Engineering A Comprehensive Survey
★ 86dexjoco. Python
★ 169OpenHarness. "OpenHarness: Open Agent Harness with a Built-in Personal Agent--Ohmo!"
★ 15kClawEnvKit. Open-source Environment toolkit of claw-like agents, support task/harness generation and evaluation
★ 58ClawGUI. Build, Evaluate, and Deploy GUI Agents — online RL training, standardized benchmarks, and real-device deployment in one framework.
★ 1.3kcli. Google Workspace CLI — one command-line tool for Drive, Gmail, Calendar, Sheets, Docs, Chat, Admin, and more. Dynamically built from Google Discovery Service. Includes AI agent skills.
★ 30kNemoClaw. Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference
★ 22khermes-agent. The agent that grows with you
★ 222kclaw-eval. Claw-Eval is an evaluation harness for evaluating LLM as agents. All tasks verified by humans.
★ 737WildClawBench. An in-the-wild benchmark for AI agents in the OpenClaw Environment.
★ 498public-image-mirror. 很多镜像都在国外。比如 gcr 。国内下载很慢,需要加速。致力于提供连接全世界的稳定可靠安全的容器镜像服务。
★ 15kawesome-autoresearch. Curated list of AutoResearch use cases with optimization traces and open source implementations
★ 1kray. Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
★ 43kclaude-relay-service. CRS-自建Claude Code镜像,一站式开源中转服务,让 Claude、OpenAI、Gemini、Droid 订阅统一接入,支持拼车共享,更高效分摊成本,原生工具无缝使用。
★ 12kROLL. An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models
★ 3.3kclaude-code-design-guide. From Early Internet Design Patterns to AI Agent Implementation — A Deep Dive into Claude Code for Developers
★ 862CatchMe. "CatchMe: Make Your AI Agents Truly Personal"
★ 429claw-bench. The Definitive AI Agent Benchmark
★ 179CLI-Anything. "CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/
★ 46kClawTeam. "ClawTeam: Agent Swarm Intelligence" (One Command → Full Automation)
★ 5.5kmercury. 🪽 Mercury — There are many claws, but this one is mine.
★ 145agency-agents. A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables.
★ 137kskill. PinchBench is a benchmarking system for evaluating LLM models as OpenClaw coding agents. Made with 🦀 by the humans at https://kilo.ai
★ 1.3kopenfang. Open-source Agent Operating System
★ 18keigent. Eigent: The Open Source Cowork Desktop - Local and Free Alternative to Claude Cowork and Codex
★ 15ksymphony. Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work instead of supervising coding agents.
★ 26knanobot. Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-agent workflows, automation, and chat apps
★ 46kawesome-openclaw. Python
★ 558awesome-openclaw-usecases. A community collection of OpenClaw use cases for making life easier.
★ 32kawesome-openclaw-skills. The awesome collection of OpenClaw skills. 5,400+ skills filtered and categorized from the official OpenClaw Skills Registry.🦞
★ 52kv1. Open-source multi-agent platform that structures software engineering as a coordinated team of specialized AI agents. It enforces strict role separation, isolated execution sandboxes, structured communication, and secrets management.
★ 50AgentTeams. An open-source Collaborative Multi-Agent OS for transparent, human-in-the-loop task coordination via Matrix rooms.
★ 5.2kagents. Multi-harness agentic plugin marketplace for Claude Code, Codex CLI, Cursor, OpenCode, GitHub Copilot, and Gemini CLI
★ 38kScaleSWE.
★ 87slime. slime is an LLM post-training framework for RL Scaling.
★ 7.7knanoclaw. A lightweight alternative to OpenClaw that runs in containers for security. Connects to WhatsApp, Telegram, Slack, Discord, Gmail and other messaging apps,, has memory, scheduled jobs, and runs directly on Anthropic's Agents SDK
★ 30kopenclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 384kCLI-Gym. Official Implementation of "CLI-Gym: Scalable CLI Task Generation via Agentic Environment Inversion"
★ 140FeatureBench. [ICLR 2026] Official Implementation of "FeatureBench: Benchmarking Agentic Coding for Complex Feature Development"
★ 83Open-AgentRL. RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios
★ 596voltagent. AI Agent Engineering Platform built on an Open Source TypeScript AI Agent Framework
★ 10kswe-factory. [FSE'2026] SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarks
★ 183opencode. The open source coding agent.
★ 191kEngram. Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
★ 4.6kIQuest-Coder-V1. Python
★ 1.4kbenchmarks. Evaluation harness for OpenHands V1.
★ 109UniSplat. [ICLR 2026] Official Implementation of "UniSplat: Unified Spatio-Temporal Fusion via 3D Latent Scaffolds for Dynamic Driving Scene Reconstruction""
★ 85clash-for-linux-install. 😼 优雅地使用基于 clash/mihomo 的代理环境
★ 14kQwen3-VL. Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
★ 20kstatic-analysis. ⚙️ A curated list of static analysis (SAST) tools and linters for all programming languages, config files, build tools, and more. The focus is on tools which improve code quality.
★ 15kAwesome-Self-Evolving-Agents. [Survey] A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems
★ 2.4kDeepCode. "DeepCode: Open Agentic Coding (Paper2Code & Text2Web & Text2Backend)"
★ 16kcontext7. Context7 Platform -- Up-to-date code documentation for LLMs and AI code editors
★ 60kmini-swe-agent. The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench verified!
★ 6.1kpolars. Extremely fast Query Engine for DataFrames, written in Rust
★ 39kSWE-Swiss. SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution
★ 105CURE. [NeurIPS 2025 Spotlight] Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning
★ 167R-Zero. [ICLR2026] codes for R-Zero: Self-Evolving Reasoning LLM from Zero Data (https://www.arxiv.org/pdf/2508.05004)
★ 827deepwiki-open. Open Source DeepWiki: AI-Powered Wiki Generator for GitHub/Gitlab/Bitbucket Repositories. Join the discord: https://discord.gg/gMwThUMeme
★ 17kqwen-code. An open-source AI coding agent that lives in your terminal.
★ 26kgemini-cli. An open-source AI agent that brings the power of Gemini directly into your terminal.
★ 106klearn-claude-code. Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
★ 73kKode-CLI. Kode CLI — Design for post-human workflows. One unit agent for every human & computer task.
★ 5.2kml-diffucoder. DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation
★ 832AI-Coding-Style-Guides. A set of coding style guidelines for Vibe Coding or SWE-Agents that maximize efficiency and improve human readability.
★ 491trae-agent. Trae Agent is an LLM-based agent for general purpose software engineering tasks.
★ 12kKernelBench. KernelBench: Can LLMs Write GPU Kernels? - Benchmark + Toolkit with Torch -> CUDA (+ more DSLs)
★ 1.2kservers. Model Context Protocol Servers
★ 89kswe-rl. [NeurIPS'25] Official codebase for "SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution"
★ 712Awesome-Inference-Time-Scaling. Paper List of Inference/Test Time Scaling/Computing
★ 399LlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74kms-swift. Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
★ 15kdiscrete-diffusion-papers. A collection of papers on discrete diffusion models
★ 164openevolve. Open-source implementation of AlphaEvolve
★ 6.8kcodex. Lightweight coding agent that runs in your terminal
★ 102kawesome-hallucination-detection. List of papers on hallucination detection in LLMs.
★ 1.1khallucination-leaderboard. Leaderboard Comparing LLM Performance at Producing Hallucinations when Summarizing Short Documents
★ 3.3kCode-Agent-Survey. A survey of Code Agents / Foundation Models for improving development productivity. Become 10x SWE, MLE, etc.
★ 22code-act. Official Repo for ICML 2024 paper "Executable Code Actions Elicit Better LLM Agents" by Xingyao Wang, Yangyi Chen, Lifan Yuan, Yizhe Zhang, Yunzhu Li, Hao Peng, Heng Ji.
★ 1.7krllm. Democratizing Reinforcement Learning for LLMs
★ 5.7kSearch-R1. Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL
★ 5.2kDeepRetrieval. [COLM’25] DeepRetrieval — 🔥 Training Search Agent by RLVR with Retrieval Outcome
★ 715Search-o1. 🔍 Search-o1: Agentic Search-Enhanced Large Reasoning Models [EMNLP 2025]
★ 1.2kopen-deep-research. An open source deep research clone. AI Agent that reasons large amounts of web data extracted with Firecrawl
★ 6.3kaugment-swebench-agent. The #1 open-source SWE-bench Verified implementation
★ 877OpenHands. 🙌 OpenHands: AI-Driven Development
★ 83kawesome-language-agents. List of language agents based on paper "Cognitive Architectures for Language Agents"
★ 1.2kQwen3-Coder. Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team.
★ 17kSWE-agent. SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding challenges. [NeurIPS 2024]
★ 20kAgent4SE-Paper-List. Repository for the paper "Large Language Model-Based Agents for Software Engineering: A Survey". Keep updating.
★ 553Qwen2.5-Omni. Qwen2.5-Omni is an end-to-end multimodal model by Qwen team at Alibaba Cloud, capable of understanding text, audio, vision, video, and performing real-time speech generation.
★ 4.1kDeepSeek-Coder-V2. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
★ 6.9kawesome-deepseek-coder. A curated list of open-source projects related to DeepSeek Coder
★ 801bigcode-dataset. Jupyter Notebook
★ 497SWE-bench. SWE-bench: Can Language Models Resolve Real-world Github Issues?
★ 5.5kAwesome-Agent-Papers. [Up-to-date] Large Language Model Agent: A Survey on Methodology, Applications and Challenges
★ 2.8kAwesome-LLM. Awesome-LLM: a curated list of Large Language Model
★ 27kAwesome-Knowledge-Distillation-of-LLMs. This repository collects papers for "A Survey on Knowledge Distillation of Large Language Models". We break down KD into Knowledge Elicitation and Distillation Algorithms, and explore the Skill & Vertical Distillation of LLMs.
★ 1.3kself-correction-llm-papers. This is a collection of research papers for Self-Correcting Large Language Models with Automated Feedback.
★ 573fastText. Library for fast text representation and classification.
★ 27kcc_net. Tools to download and cleanup Common Crawl data
★ 1kfrontier-evals. OpenAI Frontier Evals
★ 1.3kopeninterpreter. A coding agent for open models like Kimi K3
★ 67kOpenManus. No fortress, purely open ground. OpenManus is Coming.
★ 58kowl. 🦉 OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation
★ 20kVICT. [CVPR 2025] Test-Time Visual In-Context Tuning
★ 30repo-level-codegen-papers. Repo-Level Code generation papers
★ 235Awesome-Code-LLM. [TMLR] A curated list of language modeling researches for code (and other software engineering activities), plus related datasets.
★ 3.4kAwesome-LLM-Reasoning. From Chain-of-Thought prompting to OpenAI o1 and DeepSeek-R1 🍓
★ 3.7kTinyZero. Minimal reproduction of DeepSeek R1-Zero
★ 13kSPIN. The official implementation of Self-Play Fine-Tuning (SPIN)
★ 1.2kAwesome-LLM-Post-training. Awesome Reasoning LLM Tutorial/Survey/Guide
★ 2.5kGoT. Official repository of "GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing"
★ 317Mani-GS. [CVPR' 2025'] Mani-GS: Gaussian Splatting Manipulation with Triangular Mesh
★ 197camel. 🐫 CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.org
★ 18kUFO. [NeurIPS2025 Spotlight 🔥 ] Official implementation of 🛸 "UFO: A Unified Approach to Fine-grained Visual Perception via Open-ended Language Interface"
★ 281Qwen-Agent. Agent framework and applications built upon Qwen>=3.0, featuring Function Calling, MCP, Code Interpreter, RAG, Chrome extension, etc.
★ 17knative-sparse-attention. 🐳 Efficient Triton implementations for "Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention"
★ 1kDeepEP. DeepEP: an efficient expert-parallel communication library
★ 9.9kopen-infra-index. Production-tested AI infrastructure tools for efficient AGI development and community-driven innovation
★ 8kFlashMLA. FlashMLA: Efficient Multi-head Latent Attention Kernels
★ 13kOpen-Reasoner-Zero. Official Repo for Open-Reasoner-Zero
★ 2.1kMoBA. MoBA: Mixture of Block Attention for Long-Context LLMs
★ 2.2kthinking-in-space. Official repo and evaluation implementation of VSI-Bench
★ 735leetcode-master. 《代码随想录》LeetCode 刷题攻略:200道经典题目刷题顺序,共60w字的详细图解,视频难点剖析,50余张思维导图,支持C++,Java,Python,Go,JavaScript等多语言版本,从此算法学习不再迷茫!🔥🔥 来看看,你会发现相见恨晚!🚀
★ 62kDeepSeek-Math. DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
★ 3.4kQwen. The official repo of Qwen (通义千问) chat & pretrained large language model proposed by Alibaba Cloud.
★ 22kopen-r1. Fully open reproduction of DeepSeek-R1
★ 26kDeepSeek-R1.
★ 92kOpenAgents. [COLM 2024] OpenAgents: An Open Platform for Language Agents in the Wild
★ 4.9kawesome-llm-powered-agent. Awesome things about LLM-powered agents. Papers / Repos / Blogs / ...
★ 2.3kgenesis-world. Simulation platform for general-purpose robotics & embodied AI learning.
★ 30kMoE-Jetpack. [NeurIPS 24] MoE Jetpack: From Dense Checkpoints to Adaptive Mixture of Experts for Vision Tasks
★ 137Pyramid-Flow. [ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
★ 3.2kTokenFormer. [ICLR2025 Spotlight🔥] Official Implementation of TokenFormer: Rethinking Transformer Scaling with Tokenized Model Parameters
★ 593Book-Mathematical-Foundation-of-Reinforcement-Learning. This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."
★ 17kToolBench. [ICLR'24 spotlight] An open platform for training, serving, and evaluating large language model for tool learning.
★ 5.7kToolLearningPapers.
★ 923quiet-star. Code for Quiet-STaR
★ 739Awesome-LLM-Strawberry. A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.
★ 6.9ktrl. Train transformer language models with reinforcement learning.
★ 19kdirect-preference-optimization. Reference implementation for DPO (Direct Preference Optimization)
★ 2.9kPaLM-rlhf-pytorch. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
★ 7.9kOpenRLHF. An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
★ 9.9kllama-moe. ⛷️ LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training (EMNLP 2024)
★ 1kRedPajama-Data. The RedPajama-Data repository contains code for preparing large datasets for training large language models.
★ 5kgemma. Gemma open-weight LLM library, from Google DeepMind
★ 5.6kLL3M. LL3M: Large Language and Multi-Modal Model in Jax
★ 74EasyLM. Large language models (LLMs) made easy, EasyLM is a one stop solution for pre-training, finetuning, evaluating and serving LLMs in JAX/Flax.
★ 2.5kmaxtext. A simple, performant, and scalable Jax LLM!
★ 2.4kgpt-jax. Jax/Flax rewrite of Karpathy's nanoGPT
★ 66MambaOut. MambaOut: Do We Really Need Mamba for Vision? (CVPR 2025)
★ 2.7kpykan. Kolmogorov Arnold Networks
★ 16kConTex-Human. [CVPR' 2024'] ConTex-Human: Free-View Rendering of Human from a Single Image with Texture-Consistent Synthesis
★ 48CoSSL. Official PyTorch Implementation of "CoSSL: Co-Learning of Representation and Classifier for Imbalanced Semi-Supervised Learning" (CVPR 2022)
★ 52latentsplat. [ECCV 2024] Implementation of latentSplat: Autoencoding Variational Gaussians for Fast Generalizable 3D Reconstruction
★ 239BLIP. PyTorch code for BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation
★ 5.7kdatacomp. DataComp: In search of the next generation of multimodal datasets
★ 787pythia. The hub for EleutherAI's work on interpretability and learning dynamics
★ 2.9kopen_clip. An open source implementation of CLIP.
★ 14kFouriScale. Official implementation of FouriScale (ECCV2024)
★ 161Graphormer-GD. [ICLR 2023 notable top-5%] Rethinking the Expressive Power of GNNs via Graph Biconnectivity (official implementation)
★ 104VAR. [NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". An *ultra-simple, user-friendly yet state-of-the-art* codebase for autoregressive image generation!
★ 8.7kgpt-neox. An implementation of model parallel autoregressive transformers on GPUs, based on the Megatron and DeepSpeed libraries
★ 7.4kFastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40kVIRL. (ECCV 2024) Code for V-IRL: Grounding Virtual Intelligence in Real Life
★ 367grok-1. Grok open release
★ 52kOpen-Sora. Open-Sora: Democratizing Efficient Video Production for All
★ 29kLLoVi. Official implementation for "A Simple LLM Framework for Long-Range Video Question-Answering"
★ 106MoE-LLaVA. 【TMM 2025🔥】 Mixture-of-Experts for Large Vision-Language Models
★ 2.3kSparK. [ICLR'23 Spotlight🔥] The first successful BERT/MAE-style pretraining on any convolutional network; Pytorch impl. of "Designing BERT for Convolutional Networks: Sparse and Hierarchical Masked Modeling"
★ 1.4kGiT. [ECCV2024 Oral🔥] Official Implementation of "GiT: Towards Generalist Vision Transformer through Universal Language Interface"
★ 364RWKV-LM. RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.
★ 15kgemma_pytorch. The official PyTorch implementation of Google's Gemma models
★ 5.7klm-evaluation-harness. A framework for few-shot evaluation of language models.
★ 13ks4. Structured state space sequence models
★ 2.9kmamba. Mamba SSM architecture
★ 19kLoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
★ 14kawesome-ChatGPT-repositories. A curated list of open source GitHub repositories related to ChatGPT, the OpenAI API, and Codex. Searchable via Claude Code skills.
★ 3.2klit-llama. Implementation of the LLaMA language model based on nanoGPT. Supports flash attention, Int8 and GPTQ 4bit quantization, LoRA and LLaMA-Adapter fine-tuning, pre-training. Apache 2.0-licensed.
★ 6.1kpicoGPT. An unnecessarily tiny implementation of GPT-2 in NumPy.
★ 3.5ktorchscale. Foundation Architecture for (M)LLMs
★ 3.1kScatterFormer. ScatterFormer: Efficient Voxel Transformer with Scattered Linear Attention (ECCV 2024)
★ 80bidiff. [CVPR'24] Text-to-3D Generation with Bidirectional Diffusion using both 2D and 3D priors
★ 168nanoGPT. The simplest, fastest repository for training/finetuning medium-sized GPTs.
★ 62kHEDNet. HEDNet (NeurIPS 2023) & SAFDNet (CVPR 2024 Oral)
★ 189