This is your work, valued
A PhD student in NLP at Nanjing University (Joint PhD Program in Linguistics and Translation with City University of Hong Kong).
CLiKA. Evaluation of the Cross-Lingual Knowledge Alignment in LLMs
★ 9scaling_finetuning. Python
★ 2ffn_composition_analysis. Python
★ 1GeoTranslate. Rule-based geographical names translation to Chinese
★ 1ExpLang. Python
★ 1TESSY. A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data
★ 35mempalace. The best-benchmarked open-source AI memory system. And it's free.
★ 58kreverse-engineering-gemma-3n. Reverse Engineering Gemma 3n: Google's New Edge-Optimized Language Model
★ 279CLI-Anything. "CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/
★ 46kBLEnD. BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages
★ 50spider. Streamline on-policy/off-policy distillation workflows in a few lines of code
★ 109trl. Train transformer language models with reinforcement learning.
★ 19kMath-Verify. Python
★ 1.2kDevelpoment-of-Upper-Limb-Exoskeleton. Upper Limb Exoskeleton development
★ 5LEAP-upper-limb. Upper limb powered exoskeleton (Stop updating)
★ 12de_vito. Kinematic control algorithms and software for 🦾 DE VITO 🦾: A dual-arm, high degree-of-freedom, lightweight, inexpensive, passive upper-limb exoskeleton for robot teleoperation.
★ 22dataset-ssvep-exoskeleton. SSVEP-based BCI recording of 12 subjects operating an upper limb exoskeleton during a shared control task. The exoskeleton is either controlled with a touchless interface detecting hand poses or with BCI.
★ 47VST. [ECCV2026] Visual Spatial Tuning
★ 200nanochat. The best ChatGPT that $100 can buy.
★ 57kSAELens. Training Sparse Autoencoders on Language Models
★ 1.5kJanusCoder. [ICLR 2026] JanusCoder: Towards a Foundational Visual-Programmatic Interface for Code Intelligence
★ 78Embodied-AI-Guide. [Lumina具身智能社区] 具身智能技术指南 Embodied-AI-Guide
★ 15kalgonauts-2025. Training and evaluating encoding models to predict fMRI brain responses to naturalistic video stimuli
★ 310unsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kms-swift. Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
★ 15kLoRA-Models-for-SAEs. Code for reproducing our paper "Low Rank Adapting Models for Sparse Autoencoder Features"
★ 17Gradient_Unified. How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients
★ 20circuit-tracer. Python
★ 2.9kmtob. Shell
★ 48MUR. Python
★ 49OpenAI_WebUI. WebUI to various OpenAI API-compatible GPTs and Image generation API (requires valid API keys)
★ 38open-thoughts. Fully open data curation for reasoning models
★ 2.3kuv. An extremely fast Python package and project manager, written in Rust.
★ 88kkodcode. ✨ A synthetic dataset generation framework that produces diverse coding questions and verifiable solutions - all in one framwork
★ 321CapArena. An Arena-style Automated Evaluation Benchmark for Detailed Captioning
★ 59DeepMath. A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
★ 297TLDR. Code for Research Project TLDR
★ 26s1. s1: Simple test-time scaling
★ 6.7kLogic-LLM. The project page for "LOGIC-LM: Empowering Large Language Models with Symbolic Solvers for Faithful Logical Reasoning"
★ 404ScienceBoard. [ICLR 2026] Code, benchmark and environment for "ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows"
★ 131rllm. Democratizing Reinforcement Learning for LLMs
★ 5.7kAbsolute-Zero-Reasoner. Official Repository of Absolute Zero Reasoner
★ 1.9kdingo. Dingo: A Comprehensive AI Data, Model and Application Quality Evaluation Tool
★ 730QuRating. [ICML 2024] Selecting High-Quality Data for Training Language Models
★ 204continuous-thought-machines. Continuous Thought Machines, because thought takes time and reasoning is a process.
★ 2katom. [NeurIPS 2025] Atom of Thoughts for Markov LLM Test-Time Scaling
★ 664TFB. [PVLDB 2024 Best Paper Nomination] TFB: Towards Comprehensive and Fair Benchmarking of Time Series Forecasting Methods
★ 1.7krbloom. A fast, simple and lightweight Bloom filter library for Python, implemented in Rust.
★ 318verl. verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
★ 23kSuperGPQA. Python
★ 191steplaw. Python
★ 235MAPO. The implement of ACL2024: "MAPO: Advancing Multilingual Reasoning through Multilingual Alignment-as-Preference Optimization"
★ 44near-duplicate-code-detector. A simple tool for detecting near-duplicate source code
★ 104BenchMAX. Python
★ 29open-r1. Fully open reproduction of DeepSeek-R1
★ 26kllm-reasoners. A library for advanced large language model reasoning
★ 2.3kSimPO. [NeurIPS 2024] SimPO: Simple Preference Optimization with a Reference-Free Reward
★ 956chat_templates. Chat Templates for 🤗 HuggingFace Large Language Models
★ 719pydec. Linear decomposition toolkit for neural network.
★ 5xtuner. A Next-Generation Training Engine Built for Ultra-Large MoE Models
★ 5.2kLRP-eXplains-Transformers. Layer-wise Relevance Propagation for Large Language Models and Vision Transformers [ICML 2024]
★ 242captum. Model interpretability and understanding for PyTorch
★ 5.7kOpen-LLaVA-NeXT. An open-source implementation for training LLaVA-NeXT.
★ 439sglang. SGLang is a high-performance serving framework for large language models and multimodal models.
★ 31khuman-eval. Code for the paper "Evaluating Large Language Models Trained on Code"
★ 3.3klingua-py. The most accurate natural language detection library for Python, suitable for short text and mixed-language text
★ 1.8klm-evaluation-harness. A framework for few-shot evaluation of language models.
★ 13kdedupe. :id: A python library for accurate and scalable fuzzy matching, record deduplication and entity-resolution.
★ 4.5kdeduplicate-text-datasets. Rust
★ 1.3kMMBench. Official Repo of "MMBench: Is Your Multi-modal Model an All-around Player?"
★ 307lmms-eval. One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
★ 4.3kLLaVA-NeXT. Python
★ 4.7kbelebele. Repo for the Belebele dataset, a massively multilingual reading comprehension dataset.
★ 341mt-bench-101. [ACL 2024] MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues
★ 153Deep-Fourier-based-Arbitrary-scale-Super-resolution-for-Real-time-Rendering. SIGGRAPH 2024 Conference Paper: Deep Fourier-based Arbitrary-scale Super-resolution for Real-time Rendering
★ 150InternVL. [CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型
★ 10kVLMEvalKit. Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
★ 4.3kawesome-brain-decoding. collection of awesome research in brain decoding, including interaction with multi-modalities, theories, and foundation models.
★ 109the-story-of-heads. This is a repository with the code for the ACL 2019 paper "Analyzing Multi-Head Self-Attention: Specialized Heads Do the Heavy Lifting, the Rest Can Be Pruned" and the ACL 2021 paper "Analyzing Source and Target Contributions to NMT Predictions".
★ 324pydec. Linear decomposition toolkit for neural network.
★ 8Classical-Modern. 非常全的文言文(古文)-现代文平行语料
★ 1.5kqlora. QLoRA: Efficient Finetuning of Quantized LLMs
★ 11kgfn-lm-tuning. Jupyter Notebook
★ 191LlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74kLWM. Large World Model -- Modeling Text and Video with Millions Context
★ 7.4kvllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88kEasyNLP. EasyNLP: A Comprehensive and Easy-to-use NLP Toolkit
★ 2.2kXGLUE. Cross-lingual GLUE
★ 49LLMs-are-parallel-multilingual-learners. The implementation of Revealing the Parallel Multilingual Learning within Large Language Models (EMNLP Main 2024).
★ 13MLQA. New dataset
★ 311mlmm-evaluation. Multilingual Large Language Models Evaluation Benchmark
★ 135grok-1. Grok open release
★ 52kCross-Lingual-MRC. Cross-Lingual Machine Reading Comprehension (EMNLP 2019)
★ 67NextChat. ✨ Light and Fast AI Assistant. Support: Web | iOS | MacOS | Android | Linux | Windows
★ 89kllm-latent-language. Repo accompanying our paper "Do Llamas Work in English? On the Latent Language of Multilingual Transformers".
★ 87Prompt-Engineering-Guide. 🐙 Guides, papers, lessons, notebooks and resources for prompt engineering, context engineering, RAG, and AI Agents.
★ 77kllm-course. Course to get into Large Language Models (LLMs) with roadmaps and Colab notebooks.
★ 81kbrain_syntactic_representations. Code to produce syntactic representations that can be used to study syntax processing in the human brain
★ 12QAlign. Ruby
★ 39mLoRA. An Efficient "Factory" to Build Multiple LoRA Adapters
★ 384Chinese-LLaMA-Alpaca-2. 中文LLaMA-2 & Alpaca-2大模型二期项目 + 64K超长上下文模型 (Chinese LLaMA-2 & Alpaca-2 LLMs with 64K long context models)
★ 7.1kTracIn. Implementation of Estimating Training Data Influence by Tracing Gradient Descent (NeurIPS 2020)
★ 243GAOKAO-Bench. GAOKAO-Bench is an evaluation framework that utilizes GAOKAO questions as a dataset to evaluate large language models.
★ 784XCSR. Code Repo for the ACL21 paper "Common Sense Beyond English: Evaluating and Improving Multilingual LMs for Commonsense Reasoning"
★ 23LAMA. LAnguage Model Analysis
★ 1.4kX-FACTR. Python
★ 24WhatICLLearns. [ACL 2023 Findings] What In-Context Learning “Learns” In-Context: Disentangling Task Recognition and Task Learning
★ 21Chinese-LLaMA-Alpaca. 中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)
★ 19kinstruct-eval. This repository contains code to quantitatively evaluate instruction-tuned models such as Alpaca and Flan-T5 on held-out tasks.
★ 552ff-layers. The accompanying code for "Transformer Feed-Forward Layers Are Key-Value Memories". Mor Geva, Roei Schuster, Jonathan Berant, and Omer Levy. EMNLP, 2021.
★ 103scibench. Python
★ 134relative_importance. Python
★ 17ValueZeroing. The official repo for the EACL 2023 paper "Quantifying Context Mixing in Transformers"
★ 12attention_flow. Jupyter Notebook
★ 274transformer-contributions. Measuring the Mixing of Contextual Information in the Transformer
★ 35logit-explanations. Jupyter Notebook
★ 18norm-analysis-of-transformer. Jupyter Notebook
★ 87cross_lingual_relative_importance. Relative word importance of different languages and multilingualism (new work)
★ 4modelblocks-release. A psycholinguistic modeling toolkit
★ 35lm-question-generation. Multilingual/multidomain question generation datasets, models, and python library for question generation.
★ 368CoGnition. Python
★ 17brain-inspired-replay. A brain-inspired version of generative replay for continual learning with deep neural networks (e.g., class-incremental learning on CIFAR-100; PyTorch code).
★ 252knn-box. an easy-to-use knn-mt toolkit
★ 104test. Measuring Massive Multitask Language Understanding | ICLR 2021
★ 1.6kRWKV-LM. RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.
★ 15kalpaca-lora. Instruct-tune LLaMA on consumer hardware
★ 19kLMFlow. An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.
★ 8.5klangchain. The agent engineering platform.
★ 143kMOSS. An open-source tool-augmented conversational language model from Fudan University
★ 12kstanford_alpaca. Code and documentation to train Stanford's Alpaca models, and generate the data.
★ 30kself-instruct. Aligning pretrained language models with instruction data generated by themselves.
★ 4.6kFastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40kLoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
★ 14kevals. Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
★ 19kopenai-cookbook. Examples and guides for using the OpenAI API
★ 75kIntroComputationModel. 南京大学宋方敏《计算模型导引》题解
★ 140NJU-TOC-Solutions. Solutions to "Introduction to Models of Computation" of Nanjing University
★ 61llama. Inference code for Llama models
★ 60kfairscale. PyTorch extensions for high performance and large scale training.
★ 3.4kDeepSpeed. DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
★ 43kautoml. Google Brain AutoML
★ 6.5kattention-analysis. Jupyter Notebook
★ 474structural-probes. Codebase for testing whether hidden states of neural networks encode discrete structures.
★ 403Open-Assistant. OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.
★ 37kStructAdapt. Structural Adapters in Pretrained Language Models for AMR-to-Text Generation (EMNLP 2021)
★ 29ATP. Source code for paper "ATP: AMRize Than Parse! Enhancing AMR Parsing with PseudoAMRs" @NAACL-2022
★ 9AMR-IE. The code repository for AMR guided joint information extraction model (NAACL-2021).
★ 37transition-amr-parser. SoTA Abstract Meaning Representation (AMR) parsing with word-node alignments in Pytorch. Includes checkpoints and other tools such as statistical significance Smatch.
★ 272amrlib. A python library that makes AMR parsing, generation and visualization simple.
★ 267BLINK. Entity Linker solution
★ 1.2kcdec. Decoder, aligner, and model optimizer for statistical machine translation and other structured prediction models based on (mostly) context-free formalisms
★ 185AMR-Process. Code for preprocessing AMR graphs.
★ 6AMRBART. Code for our paper "Graph Pre-training for AMR Parsing and Generation" in ACL2022
★ 106penman. PENMAN notation (e.g. AMR) in Python
★ 151NJUNMT-pytorch. Python
★ 93perin. PERIN is Permutation-Invariant Semantic Parser developed for MRP 2020
★ 45camr. The system of SUDA-HUAWEI submitted at CAMR2022.
★ 12Two-Stage-CAMRP. Source code for paper "A Two-Stage Method for Chinese AMR Parsing" @ CAMRP-2022 & CCL-2022
★ 24spring. SPRING is a seq2seq model for Text-to-AMR and AMR-to-Text (AAAI2021).
★ 136kit. 🧱 Describe your site, AI builds it, you own it as Markdown. Snap together Tailwind blocks like Lego — landing pages, blogs, portfolios, docs & more. No AI slop. Free to deploy anywhere 👇
★ 9.6ktext2digits. Converts text such as "twenty three" to number/digit "23" in any sentence
★ 67pycnnum. Converting Chinese number string <=> int/float/str
★ 20w2n. Convert number words (eg. twenty one) to numeric digits (21)
★ 179cybertron-ai. mindspore implementation of transformers
★ 69fenci. 中文分词模块:继承了jieba分词的基本算法逻辑,进行了全方位的代码优化,还额外提供了HMM算法的训练功能支持。
★ 39PTR. Prompt Tuning with Rules
★ 161transformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163knju-health-report. 南京大学APP每日健康打卡。自用
★ 9English-to-IPA. Converts English text to IPA notation
★ 408nlp-datasets. Alphabetical list of free/public domain datasets with text data for use in Natural Language Processing (NLP)
★ 6kDANN. pytorch implementation of Domain-Adversarial Training of Neural Networks
★ 951MyDeepLearning. A deep learning library to provide algs in pure Numpy or Tensorflow.
★ 290vaderSentiment. VADER Sentiment Analysis. VADER (Valence Aware Dictionary and sEntiment Reasoner) is a lexicon and rule-based sentiment analysis tool that is specifically attuned to sentiments expressed in social media, and works well on texts from other domains.
★ 5k