This is your work, valued
Neural-Style-Transfer-Papers. :pencil2: Neural Style Transfer: A Review
★ 1.6kAwesome-Model-Merging. :couple: A curated list of Model Merging methods.
★ 95AmalgamateGNN.PyTorch. PyTorch implementation of AmalgamateGNN (CVPR'21)
★ 21Character-Stylization.
★ 7FastMaskRCNN. Mask RCNN in TensorFlow
★ 1WorldFoundry. Unified World Model Inference & Evaluation Infrastructure
★ 275WorldModelLens. Python
★ 12GenDSA-V2. Python
★ 4GenDSA. [Nature Medicine, 2025 & Med (Cell Press), 2024] Large-scale Pretrained Frame Generative Model Enables Real-Time Low-Dose DSA Imaging: an AI System Development and Multicenter Validation Study
★ 20SQUALL-release. Integrating Histology with Spatial Molecular Programs Using a Multimodal Foundation Model
★ 12path2space-companion. Companion code for Shulman et al., Cell 2026: AI-predicted spatial transcriptomics unlocks breast cancer biomarkers from pathology.
★ 8path2space-extended. Extended Path2Space code: analysis and downstream methods.
★ 5awesome-autoresearch. A curated list of awesome autonomous researcher frameworks
★ 143MIRA. MIRA: Medical Time Series Foundation Model for Real-World Health Data
★ 414MediGuide. TypeScript
★ 1Kosmos. Kosmos: An AI Scientist for Autonomous Discovery - An implementation and adaptation to be driven by Claude Code or API - Based on the Kosmos AI Paper - https://arxiv.org/abs/2511.02824
★ 556andrej-karpathy-skills. A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.
★ 198kwriting-style-skill. Transform AI-generated text into your personal voice. Skill for Claude Code / OpenCode.
★ 8humanizer. Agent skill that removes signs of AI-generated writing from text
★ 32kclaude-scholar. Semi-automated research assistant for academic research and software development. Supports Claude Code, Codex CLI, Kimi Code CLI, and OpenCode across ideation, coding, experiments, writing, and publication.
★ 4.9kAuto-claude-code-research-in-sleep. ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works with Claude Code, Codex, OpenClaw, or any LLM agent.
★ 14kAwesome-AI-Scientist. This is a survey of research on AI scientists, AI researchers, AI engineers, and a series of AI-driven research studies
★ 301slime. slime is an LLM post-training framework for RL Scaling.
★ 7.7kTFB. [PVLDB 2024 Best Paper Nomination] TFB: Towards Comprehensive and Fair Benchmarking of Time Series Forecasting Methods
★ 1.7kautoresearch. AI agents running research on single-GPU nanochat training automatically
★ 92kTSFMs.
★ 1awesome-time-series. list of papers, code, and other resources
★ 1.1kHY-WorldPlay. HY-World 1.5: A Systematic Framework for Interactive World Modeling with Real-Time Latency and Geometric Consistency
★ 1.6kRFormer. Official Repository for the NeurIPS 2024 paper Rough Transformers: Lightweight Continuous-Time Sequence Modelling with Path Signatures
★ 27ToolOrchestra. ToolOrchestra is an end-to-end RL training framework for orchestrating tools and agentic workflows.
★ 753Awesome-Agent-Memory. Curated systems, benchmarks, and papers etc. on memory for LLMs/MLLMs --- long-term context, retrieval, and reasoning.
★ 564titans-pytorch. Unofficial implementation of Titans, SOTA memory for transformers, in Pytorch
★ 2kawesome-llm-time-series. tracking papers, datasets, and models of "large language model (LLM) for time series"
★ 518awesome-time-series-papers. An Awesome List of the latest time series papers and code from top AI venues.
★ 1.1kAwesome-time-series. A comprehensive survey on the time series domains
★ 550nested_learning. A Reproduction of GDM's Nested Learning Paper
★ 705giga-world-0. GigaWorld-0: World Models as Data Engine to Empower Embodied AI
★ 1.6kAwesome-Self-Evolving-Agents. [Survey] A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems
★ 2.4kllm-council. LLM Council works together to answer your hardest questions
★ 23kgraphiti. Build Real-Time Knowledge Graphs for AI Agents
★ 29kEventTriplesExtraction. An experiment and demo-level tool for text information extraction (event-triples extraction), which can be a route to the event chain and topic graph, 基于依存句法与语义角色标注的事件三元组抽取,可用于文本理解如文档主题链,事件线等应用。
★ 927awesome_Chinese_medical_NLP. 中文医学NLP公开资源整理:术语集/语料库/词向量/预训练模型/知识图谱/命名实体识别/QA/信息抽取/模型/论文/etc
★ 2.6kMedAgentBench. MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents
★ 308AgenticRAG-Survey. Agentic-RAG explores advanced Retrieval-Augmented Generation systems enhanced with AI LLM agents.
★ 1.7kdInfer. dInfer: An Efficient Inference Framework for Diffusion Language Models
★ 476Awesome-Interpretability-in-Large-Language-Models. This repository collects all relevant resources about interpretability in LLMs
★ 402DeepAgent. [WWW‘26 Oral🔥] DeepAgent: A General Reasoning Agent with Scalable Toolsets
★ 1.1kAwesome-LLM-Long-Context-Modeling. 📰 Must-read papers and blogs on LLM based Long Context Modeling 🔥
★ 2.1kA-mem. The code for NeurIPS 2025 paper "A-Mem: Agentic Memory for LLM Agents"
★ 932Awesome-Agentic-MLLMs. Agentic MLLMs
★ 216agent-distillation. Official Code Repository for the paper "Distilling LLM Agent into Small Models with Retrieval and Code Tools"
★ 250GTA1. Python
★ 130LifelongAgentBench. Code repo for "LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners"
★ 95MCPBench. The evaluation benchmark on MCP servers
★ 251mcp-bench. MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers
★ 495text_or_pixels. Codebase for EMNLP 2025 Findings paper "Text or Pixels? Evaluating Efficiency and Understanding of LLMs with Visual Text Inputs"
★ 19LightMem. [ICLR 2026] LightMem: Lightweight and Efficient Memory-Augmented Generation
★ 1.1kgem. A Gym for Agentic LLMs
★ 503DeepSeek-OCR. Contexts Optical Compression
★ 24kuni2ts. Unified Training of Universal Time Series Forecasting Transformers
★ 1.6kAwesome-AgenticLLM-RL-Papers.
★ 1.8kdeepconf. DeepConf: Deep Think with Confidence
★ 408Embodied-AI-Guide. [Lumina具身智能社区] 具身智能技术指南 Embodied-AI-Guide
★ 15kAwesome-GraphRAG. Awesome-GraphRAG: A curated list of resources (surveys, papers, benchmarks, and opensource projects) on graph-based retrieval-augmented generation.
★ 2.6klangchain. The agent engineering platform.
★ 143kGWM. [ICML 2025]"Graph World Model", Tao Feng, Yexin Wu, Guanyu Lin, Jiaxuan You
★ 44timesfm. TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting.
★ 27kMedical-Graph-RAG. A Graph RAG System for Evidenced-based Medical Information Retrieval [ACL 2025]
★ 814RD-Agent. Research and development (R&D) is crucial for the enhancement of industrial productivity, especially in the AI era, where the core aspects of R&D are mainly focused on data and models. We are committed to automating these high-value generic R&D processes through R&D-Agent, which lets AI drive data-driven AI. 🔗https://aka.ms/RD-Agent-Tech-Report
★ 14kinformation_flow. Jupyter Notebook
★ 43DLLM-Survey. [Arxiv] Discrete Diffusion in Large Language and Multimodal Models: A Survey
★ 387Reasoning-SFT. SFT of Reasoning LLMs with Megatron-LM
★ 23TxGNN. TxGNN: Zero-shot prediction of therapeutic use with geometric deep learning and clinician centered design
★ 279Isaac-GR00T. NVIDIA Isaac GR00T N1.7 - A Foundation Model for Generalist Robots.
★ 7.7klegged-loco. Low-level locomotion policy training in Isaac Lab
★ 447openpi. Python
★ 13kVeBrain. Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces
★ 86vjepa2. PyTorch code and models for VJEPA2 self-supervised learning from video.
★ 4.4kAwesome-Unified-Multimodal-Models. Awesome Unified Multimodal Models
★ 1.3kKG-LLM-Papers. [Paper List] Papers integrating knowledge graphs (KGs) and large language models (LLMs)
★ 2.2kAdversarial_Examples_Papers. A list of recent papers about adversarial learning
★ 372IEAP. [NeurIPS 2025] IEAP: Image Editing As Programs with Diffusion Models
★ 118One-Shot-RLVR. [NeurIPS 2025] Reinforcement Learning for Reasoning in Large Language Models with One Training Example
★ 444CrimeKgAssitant. Crime assistant including crime type prediction and crime consult service based on nlp methods and crime kg,罪名法务智能项目,内容包括856项罪名知识图谱, 基于280万罪名训练库的罪名预测,基于20W法务问答对的13类问题分类与法律资讯问答功能.
★ 1.6kDPTS. Official implementation of Dynamic Parallel Tree Search for accelerating LLM reasoning with test-time parallel search.
★ 4LLaDA. Official PyTorch implementation for "Large Language Diffusion Models"
★ 3.9kGraphRouter. [ICLR 2025] "GraphRouter: A Graph-based Router for LLM Selections", Tao Feng, Yanzhen Shen, Jiaxuan You
★ 75one-shot-em. One-shot Entropy Minimization
★ 190Paper2Poster. [NeurIPS 2025] Open-source Multi-agent Poster Generation from Papers
★ 3.9kMinitron. A family of compressed models obtained via pruning and knowledge distillation
★ 384MiniPLM. [ICLR 2025] MiniPLM: Knowledge Distillation for Pre-Training Language Models
★ 79Dimple. Dimple, the first Discrete Diffusion Multimodal Large Language Model
★ 117Medical_LVLM_Sycophancy. Python
★ 14ContexualDynamicMapping. Offical code for "Enhancing Cross-Tokenizer Knowledge Distillation with Contextual Dynamical Mapping"
★ 12O3-LLM-UNLEARNING. Python
★ 19DSKD. Repo for the EMNLP'24 Paper "Dual-Space Knowledge Distillation for Large Language Models". A general white-box KD framework for both same-tokenizer and cross-tokenizer LLM distillation.
★ 64merge-tokenizers. Package to align tokens from different tokenizations.
★ 16Multi-Level-OT. Pytorch Implementation of "Multi-Level Optimal Transport for Universal Cross-Tokenizer Knowledge Distillation on Language Models", AAAI 2025
★ 38FuseAI. FuseAI Project
★ 600FlanT5-CoT-Specialization. Implementation of ICML 23 Paper: Specializing Smaller Language Models towards Multi-Step Reasoning.
★ 131arxiv-sanity-lite. arxiv-sanity lite: tag arxiv papers of interest get recommendations of similar papers in a nice UI using SVMs over tfidf feature vectors based on paper abstracts.
★ 1.7kChineseMedicalQA. 基于中医药知识图谱智能问答
★ 197Search-R1. Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL
★ 5.2kAwesome-GenAI-Unlearning.
★ 188RAP. Code for Retrieval-Augmented Perception (ICML 2025)
★ 74awesome-llm-unlearning. A resource repository for machine unlearning in large language models
★ 615PERMU. Python
★ 34Awesome-Knowledge-Distillation-of-LLMs. This repository collects papers for "A Survey on Knowledge Distillation of Large Language Models". We break down KD into Knowledge Elicitation and Distillation Algorithms, and explore the Skill & Vertical Distillation of LLMs.
★ 1.3kGraphing-the-Face-of-3D-Beauty.
★ 1textgrad. TextGrad: Automatic ''Differentiation'' via Text -- using large language models to backpropagate textual gradients. Published in Nature.
★ 3.7kAwesome-Efficient-Reasoning-LLMs. [TMLR 2025] Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
★ 786Awesome-Efficient-Reasoning-Models. [TMLR 2025] Efficient Reasoning Models: A Survey
★ 318Awesome-Efficient-Reasoning. Paper list for Efficient Reasoning.
★ 898Awesome-Multimodal-Large-Language-Models. 🔥Awesome Multimodal Large Language Models Paper List
★ 154Open-R1-Video. ✨First Open-Source R1-like Video-LLM [2025/02/18]
★ 382lmm-r1. Extend OpenRLHF to support LMM RL training for reproduction of DeepSeek-R1 on multimodal tasks.
★ 848ProcessBench. Official repository for ACL 2025 paper "ProcessBench: Identifying Process Errors in Mathematical Reasoning"
★ 191awesome-deepseek. DeepSeek 资料大全🔥,DeepSeek 使用,指令指南,应用开发指南,精选资源清单,更好的使用 DeepSeek 让你的生产力 10倍提升! 🚀
★ 151LLM4Chemistry. [COLING 2025]A curated paper list about LLMs for chemistry
★ 155conformal-credal-sets. Conformalized Credal Set Predictors (NeurIPS 2024)
★ 7simpleRL-reason. Simple RL training for reasoning
★ 3.9kAIInfra. AIInfra(AI 基础设施)指AI系统从底层芯片等硬件,到上层软件栈支持AI大模型训练和推理。
★ 7.8kRAG-Survey. Collecting awesome papers of RAG for AIGC. We propose a taxonomy of RAG foundations, enhancements, and applications in paper "Retrieval-Augmented Generation for AI-Generated Content: A Survey".
★ 1.8kSkyThought. Sky-T1: Train your own O1 preview model within $450
★ 3.4kDeepSeek-V3. Python
★ 104kunsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kDRT. Deep Reasoning Translation (DRT) Project
★ 242CLEAR. [NeurIPS 2025] Official PyTorch implementation of paper "CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up".
★ 219Awesome-LLM-Self-Improvement. A curated list of awesome LLM Inference-Time Self-Improvement (ITSI, pronounced "itsy") papers from our recent survey: A Survey on Large Language Model Inference-Time Self-Improvement.
★ 109TinyFusion. [CVPR 2025 Highlight] TinyFusion: Diffusion Transformers Learned Shallow
★ 170VAR. [NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". An *ultra-simple, user-friendly yet state-of-the-art* codebase for autoregressive image generation!
★ 8.7kawesome-llm-attributions. A Survey of Attributions for Large Language Models
★ 229o1_Reasoning_Patterns_Study. Python
★ 105OOP-eval. The first Object-Oriented Programming (OOP) Evaluation Benchmark for LLMs
★ 27vllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88kreasoning-on-graphs. Official Implementation of ICLR 2024 paper: "Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning"
★ 527InternVL. [CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型
★ 10kQwen3-VL. Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
★ 20kllm-self-correction-papers. List of papers on Self-Correction of LLMs.
★ 82Awesome-LLM-Strawberry. A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.
★ 6.9kawesome-o1. A bibliography and survey of the papers surrounding o1
★ 1.2kLLaMA-O1. Large Reasoning Models
★ 803CPO. [NeurIPS 2024] The official implementation of paper: Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs.
★ 137HARDMath. A new dataset of difficult graduate-level applied mathematics problems; evaluations demonstrate that leading LLMs currently exhibit low accuracy in solving these problems.
★ 30JEEBench. Repository for the code and dataset for the paper: "Have LLMs Advanced enough? Towards Harder Problem Solving Benchmarks For Large Language Models"
★ 39Qwen3. Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.
★ 27kCoDe. [CVPR 2025] CoDe: Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient
★ 108Qwen-VL. The official repo of Qwen-VL (通义千问-VL) chat & pretrained large vision language model proposed by Alibaba Cloud.
★ 6.7kMarco-o1. An Open Large Reasoning Model for Real-World Solutions
★ 1.5kFailureLLMUnlearning. An official implementation of "Catastrophic Failure of LLM Unlearning via Quantization" (ICLR 2025)
★ 39smile. [EMNLP 2024] 中文领域心理健康对话大模型MeChat
★ 535LLaVA-CoT. [ICCV 2025] LLaVA-CoT, a visual language model capable of spontaneous, systematic reasoning
★ 2.1kqr-code-styling. Automaticly generate your styled QR code in your web app.
★ 2.9kAwesome-LLMs-in-Graph-tasks. A curated collection of research papers exploring the utilization of LLMs for graph-related tasks.
★ 656LLMs-from-scratch. Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
★ 100kO1-Journey. O1 Replication Journey
★ 2kg1. g1: Using Llama-3.1 70b on Groq to create o1-like reasoning chains
★ 4.2kDPO. dpo算法实现
★ 53Awesome-LLM-Uncertainty-Reliability-Robustness. Awesome-LLM-Robustness: a curated list of Uncertainty, Reliability and Robustness in Large Language Models
★ 831LLM_Tree_Search. (ICML 2024) Alphazero-like Tree-Search can guide large language model decoding and training
★ 287o1_inference_scaling_laws. Replicating O1 inference-time scaling laws
★ 94awesome-monte-carlo-tree-search-papers. A curated list of Monte Carlo tree search papers with implementations.
★ 713openr. OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
★ 1.8kAwesome-LLM-Inference. 📚A curated list of Awesome LLM/VLM Inference Papers with Codes: Flash-Attention, Paged-Attention, WINT8/4, Parallelism, etc.🎉
★ 5.4kgraphrag-accelerator. One-click deploy of a Knowledge Graph powered RAG (GraphRAG) in Azure
★ 2.4kChatTTS. A generative speech model for daily dialogue.
★ 40kefaqa-corpus-zh. ❤️Emotional First Aid Dataset, 心理咨询问答、聊天机器人语料库
★ 767knowledge_graph. Convert any text to a graph of knowledge. This can be used for Graph Augmented Generation or Knowledge Graph based QnA
★ 3.6kapiprompting. [ECCV 2024] API: Attention Prompting on Image for Large Vision-Language Models
★ 112Emu3. Next-Token Prediction is All You Need
★ 2.4kmini-omni. open-source multimodal large language model that can hear, talk while thinking. Featuring real-time end-to-end speech input and streaming audio output conversational capabilities.
★ 3.6kGraphRAG-Survey.
★ 285finetuned-qlora-falcon7b-medical. Finetuning of Falcon-7B LLM using QLoRA on Mental Health Conversational Dataset
★ 263BiasEval-LLM-MentalHealth. Unveiling and Mitigating Bias in Mental Health Analysis with Large Language Models
★ 12LLM-Dojo. 轻量级 LLM Post-training 框架,支持 SFT、RLVR、On-Policy KD、Guide KD 及混合训练;实现单轮/多轮 Guide 蒸馏、多教师蒸馏、Reward 混合训练与自动化数据分流👩🎓👨🎓
★ 938Hands-On-Large-Language-Models. Official code repo for the O'Reilly Book - "Hands-On Large Language Models"
★ 28kmaskllm-project-page. JavaScript
★ 3ReaLHF. Super-Efficient RLHF Training of LLMs with Parameter Reallocation
★ 336awesome-mechanistic-interpretability-lm-papers.
★ 260EfficientUsableVision. Repository for NSF Award: RI: Medium: Democratizing Visual AI: Enhancing Efficiency and Usability of Large Vision Models for Fostering Under-Resourced Access
★ 1ai-for-grant-writing. A curated list of resources for using LLMs to develop more competitive grant applications.
★ 4.2kkat. [ICLR2025] Kolmogorov-Arnold Transformer
★ 848graspologic. Python package for graph statistics
★ 1kMindGPT. A Mental Health conversational LLM
★ 44graphrag. A modular graph-based Retrieval-Augmented Generation (RAG) system
★ 35kbigcows. Scrapes citation statistics from Google Scholar
★ 61awesome-multimodal-in-medical-imaging. A collection of resources on applications of multi-modal learning in medical imaging.
★ 973SoulChat. 中文领域心理健康对话大模型SoulChat
★ 749EmoLLM. 心理健康大模型 (LLM x Mental Health), Pre & Post-training & Dataset & Evaluation & Depoly & RAG, with InternLM / Qwen / Baichuan / DeepSeek / Mixtral / LLama / GLM series models
★ 1.8kMentalLLaMA. This repository introduces MentaLLaMA, the first open-source instruction following large language model for interpretable mental health analysis.
★ 323Mental-LLM. The repo for paper "Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data"
★ 96ToG. This is the official github repo of Think-on-Graph (ICLR 2024). If you are interested in our work or willing to join our research team in Shenzhen, please feel free to contact us by email (xuchengjin@idea.edu.cn)
★ 655LightCompress. [EMNLP 2024 & AAAI 2026] A powerful toolkit for compressing large models including LLMs, VLMs, and video generative models.
★ 736lgl-feature-matching. [CVPR 2022 Oral] Lifelong Graph Learning (Feature Matching)
★ 31DeeperGCN-dgl. A DGL implementation of "DeeperGCN: All You Need to Train Deeper GCNs".
★ 9Deep_GCN_Benchmarking. [TPAMI 2022] "Bag of Tricks for Training Deeper Graph Neural Networks A Comprehensive Benchmark Study" by Tianlong Chen*, Kaixiong Zhou*, Keyu Duan, Wenqing Zheng, Peihao Wang, Xia Hu, Zhangyang Wang
★ 124Mamba_State_Space_Model_Paper_List. [Mamba-Survey-2024] Paper list for State-Space-Model/Mamba and it's Applications
★ 754MindBridge. [CVPR 2024 Highlight] Official PyTorch implementation of "MindBridge: A Cross-Subject Brain Decoding Framework"
★ 124Awesome-Text-to-3D. A growing curation of Text-to-3D, Diffusion-to-3D works.
★ 6hash3D. Hash3D: Training-free Acceleration for 3D Generation
★ 179LCAT. Python
★ 20Graphormer-GD. [ICLR 2023 notable top-5%] Rethinking the Expressive Power of GNNs via Graph Biconnectivity (official implementation)
★ 104Intra-Fusion. Towards Meta-Pruning via Optimal Transport, ICLR 2024 (Spotlight)
★ 18rebasin. Replicating the Git Re-Basin paper
★ 4dissect-git-re-basin. Replicating and dissecting the git-re-basin project in one-click-replication Colabs
★ 37