This is your work, valued
KESA. The source code of KESA
★ 31D-HAN. The source code of D-HAN
★ 26tracenet. This is the source code of TraceNet: Tracing and Locating the Key Elements in Sentiment Analysis
★ 5wsdm-digg-2020. Python
★ 5Ered. Python
★ 2KnowledgeCorpus. Some corpus used for pre-training entity- or description-related tasks
★ 1jepa-wms. Code, data and weights for the paper **What drives success in physical planning with Joint-Embedding Predictive World Models?**
★ 436openpi. Python
★ 13kLeafLink. LeafLink is a lightweight CLI tool for syncing local LaTeX projects with Overleaf, featuring pull/push workflows and pseudo real-time collaboration.
★ 11BridgeDepth. [ICCV 2025 Highlight] BridgeDepth: Bridging Monocular and Stereo Reasoning with Latent Alignment
★ 159MonSter. 【CVPR 2025 Highlight】MonSter: Marry Monodepth to Stereo Unleashes Power
★ 699Embodied_AI_Paper_List. [Embodied-AI-Survey-2025] Paper List and Resource Repository for Embodied AI
★ 2.1kYJYpaper. 一个用来记录武汉大学杨景媛论文问题的仓库
★ 3.6khurtlex. A multilingual lexicon of words to hurt.
★ 100hate-speech-lexicons. Useful resources for hate speech detection
★ 10Sentiment-Reasoning. [ACL 2025 Industry Track, Oral] Sentiment Reasoning for Healthcare
★ 164neuronpedia. open source interpretability platform 🧠
★ 1.1knnsight. The nnsight package enables interpreting and manipulating the internals of deep learned models.
★ 1kSAELens. Training Sparse Autoencoders on Language Models
★ 1.5kTransformerLens. A library for mechanistic interpretability of GPT-style language models
★ 3.7kcircuit-tracer. Python
★ 2.9kICLR2024-OpenReviewData. Crawl & Visualize ICLR 2024 Data from OpenReview
★ 8deciphering_cot. [EMNLP 2024 Findings] Code for deciphering CoT using shift ciphers
★ 12MemPi. localize a memorized sequence in LLMs (NAACL 2024)
★ 9dataset. Multi30k Dataset
★ 192semantic-memorization. Jupyter Notebook
★ 44NSFC-application-template-latex. 国家自然科学基金申请书正文(面上项目)LaTeX 模板(非官方)
★ 1.1knumeric-property-repr. Code for the paper: Monotonic Representation of Numeric Properties in Language Models (ACL 2024)
★ 6FreeCtrl. Python
★ 5SpinQuant. Code repo for the paper "SpinQuant LLM quantization with learned rotations"
★ 419CAA. Steering Llama 2 with Contrastive Activation Addition
★ 245sensitivity-hardness. Code for the paper
★ 7ML4RoadSafety. A dataset for traffic accident analysis in the US
★ 30Same-Task-More-Tokens. The code for the paper: "Same Task, More Tokens: the Impact of Input Length on the Reasoning Performance of Large Language Models"
★ 57semiring-backprop-exps. Jupyter Notebook
★ 16dolma. Data and tools for generating and inspecting OLMo pre-training data.
★ 1.5kmassive-activations. Code accompanying the paper "Massive Activations in Large Language Models"
★ 204awesome-mechanistic-interpretability-lm-papers.
★ 260ff-layers. The accompanying code for "Transformer Feed-Forward Layers Are Key-Value Memories". Mor Geva, Roei Schuster, Jonathan Berant, and Omer Levy. EMNLP, 2021.
★ 104platonic-rep. Python
★ 715AI-Scientist. The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery 🧑🔬
★ 14kneuron-analysis-cot-arithmetic-reasoning. Python
★ 14QSLAW. The official code for "Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation" | [MM2024]
★ 14diagNNose. diagNNose is a Python library that facilitates a broad set of tools for analysing hidden activations of neural models.
★ 81SCAPT-ABSA. Code for EMNLP 2021 paper: "Learning Implicit Sentiment in Aspect-based Sentiment Analysis with Supervised Contrastive Pre-Training"
★ 96XICL. HTML
★ 2LLM-Sentiment. [NAACL 2024] Data and code for our paper "Sentiment Analysis in the Era of Large Language Models: A Reality Check"
★ 116LLM101n. LLM101n: Let's build a Storyteller
★ 38kllm_interview_note. 主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题
★ 15kInterpret_Instruction_Tuning_LLMs. Understanding Why and How Instruction Tuning Changes Pre-trained Models
★ 25heldout-influence-estimation. Python
★ 62AQuA. A algebraic word problem dataset, with multiple choice questions annotated with rationales.
★ 338Artificial-Intelligence-Terminology-Database. A comprehensive mapping database of English to Chinese technical vocabulary in the artificial intelligence domain
★ 2kawesome_Chinese_medical_NLP. 中文医学NLP公开资源整理:术语集/语料库/词向量/预训练模型/知识图谱/命名实体识别/QA/信息抽取/模型/论文/etc
★ 2.6kKEPT. auto icd coding with prompt
★ 49MIMIC-IV-ICD-data-processing. Jupyter Notebook
★ 48medical-coding-reproducibility. Jupyter Notebook
★ 97BioT5. BioT5 (EMNLP 2023) and BioT5+ (ACL 2024 Findings)
★ 127OpenBioMed. Python
★ 1.1kScientific-LLM-Survey. Scientific Large Language Models: A Survey on Biological & Chemical Domains
★ 359chain-of-thought. Research papers about Chain of Thought (CoT)
★ 64beir. A Heterogeneous Benchmark for Information Retrieval. Easy to use, evaluate your models across 15+ diverse IR datasets.
★ 2.3kragas. Supercharge Your LLM Application Evaluations 🚀
★ 15kAutoGPT. AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
★ 186kChatGLM-Efficient-Tuning. Fine-tuning ChatGLM-6B with PEFT | 基于 PEFT 的高效 ChatGLM 微调
★ 3.7kLlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74kQwen. The official repo of Qwen (通义千问) chat & pretrained large language model proposed by Alibaba Cloud.
★ 22kmachine-learning-interview. 算法工程师-机器学习面试题总结
★ 1.7kLLMs_interview_notes. 该仓库主要记录 大模型(LLMs) 算法工程师相关的面试题
★ 2.6kebert. Python
★ 37idiomem. Python
★ 10lm-extraction-benchmark. Python
★ 308LM_Memorization. Training data extraction on GPT-2
★ 196reversal_curse. Python
★ 313BUAAThesis. 北航研究生学位论文模板(Word+LaTeX).
★ 608aigc. 《构筑大语言模型应用:应用开发与架构设计》一本关于 LLM 在真实世界应用的开源电子书,介绍了大语言模型的基础知识和应用,以及如何构建自己的模型。其中包括Prompt的编写、开发和管理,探索最好的大语言模型能带来什么,以及LLM应用开发的模式和架构设计。
★ 1.6kAGIEval. Python
★ 774meerkat. Explore and understand your training and validation data.
★ 852Code-LMs. Guide to using pre-trained large language models of source code
★ 1.8kmodelzoo. Python
★ 1.2kzeno-build. Build, evaluate, understand, and fix LLM-based apps
★ 491DPL. Python
★ 17SSEGCN-ABSA. SSEGCN: Syntactic and Semantic Enhanced Graph Convolutional Network for Aspect-based Sentiment Analysis
★ 46AspectBasedSentimentAnalysis. Aspect Based Sentiment Analysis is a special type of sentiment analysis. In an explicit aspect, opinion is expressed on a target(opinion target), this aspect-polarity extraction is known as ABSA.
★ 73Aspect-Based-Sentiment-Analysis. A paper list for aspect based sentiment analysis.
★ 141fine-grained-sentiment. A comparison and discussion of different NLP methods for 5-class sentiment classification on the SST-5 dataset.
★ 171prm800k. 800,000 step-level correctness labels on LLM solutions to MATH problems
★ 2.2kRWKV-LM. RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.
★ 15kChatRWKV. ChatRWKV is like ChatGPT but powered by RWKV (100% RNN) language model, and open source.
★ 9.5kFactualityPrompt. Python
★ 90language. Shared repository for open-sourced projects from the Google AI Language team.
★ 1.8kcrosslingual_winograd. "It's All in the Heads" (Findings of ACL 2021), official implementation and data
★ 10xtreme. XTREME is a benchmark for the evaluation of the cross-lingual generalization ability of pre-trained multilingual models that covers 40 typologically diverse languages and includes nine tasks.
★ 651XED. XED multilingual emotion datasets
★ 64ChatALL. Concurrently chat with ChatGPT, Bing Chat, Bard, Alpaca, Vicuna, Claude, ChatGLM, MOSS, 讯飞星火, 文心一言 and more, discover the best answers
★ 16klearning_research. 本人的科研经验
★ 14kLLaMA-Adapter. [ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parameters
★ 5.9kOpenAlpaca. OpenAlpaca: A Fully Open-Source Instruction-Following Model Based On OpenLLaMA
★ 303open_llama. OpenLLaMA, a permissively licensed open source reproduction of Meta AI’s LLaMA 7B trained on the RedPajama dataset
★ 7.5kawesome-instruction-dataset. A collection of open-source dataset to train instruction-following LLMs (ChatGPT,LLaMA,Alpaca)
★ 1.2kLLM-Zoo. LLM Zoo collects information of various open- and close-sourced LLMs
★ 270anthropic-sdk-python. Python
★ 3.8kSVAMP. NAACL 2021: Are NLP Models really able to Solve Simple Math Word Problems?
★ 143HumanPrompt. A framework for human-readable prompt-based method with large language models. Specially designed for researchers. (Deprecated, check out LangChain for better usage!)
★ 131ml-mkqa. We introduce MKQA, an open-domain question answering evaluation set comprising 10k question-answer pairs aligned across 26 typologically diverse languages (260k question-answer pairs in total). The goal of this dataset is to provide a challenging benchmark for question answering quality across a wide set of languages. Please refer to our paper for details, MKQA: A Linguistically Diverse Benchmark for Multilingual Open Domain Question Answering
★ 194llama-chat. Chat with Meta's LLaMA models at home made easy
★ 840evals. Evals is a framework for evaluating OpenAI models and an open-source registry of benchmarks.
★ 18pythia. The hub for EleutherAI's work on interpretability and learning dynamics
★ 2.9kdolly. Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform
★ 11kRedPajama-Data. The RedPajama-Data repository contains code for preparing large datasets for training large language models.
★ 5kCommitmentBank. Materials related to our Sinn und Bedeutung 23 paper
★ 40url-nlp. Python
★ 273realtimeqa_public. Python
★ 87exams-qa. A Multi-subject High School Examinations Dataset for Cross-lingual and Multilingual Question Answering
★ 49MLQA. New dataset
★ 311xquad.
★ 211nanoGPT. The simplest, fastest repository for training/finetuning medium-sized GPTs.
★ 62kPaLM-rlhf-pytorch. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
★ 7.9klmql. A language for constraint-guided and efficient LLM programming.
★ 4.2kLLMSurvey. The official GitHub page for the survey paper "A Survey of Large Language Models".
★ 12kSemantic-Segment-Anything. Automated dense category annotation engine that serves as the initial semantic labeling for the Segment Anything dataset (SA-1B).
★ 2.3kJARVIS. JARVIS, a system to connect LLMs with ML community. Paper: https://arxiv.org/pdf/2303.17580.pdf
★ 25kBIG-bench. Beyond the Imitation Game collaborative benchmark for measuring and extrapolating the capabilities of language models
★ 3.2klm-evaluation-harness. A framework for few-shot evaluation of language models.
★ 13kperspectiveapi. Perspective is an API that uses machine learning models to score the perceived impact a comment might have on a conversation. See https://developers.perspectiveapi.com for more information.
★ 924LMFlow. An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.
★ 8.5kBELLE. BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)
★ 8.3kgrade-school-math. Python
★ 1.5kreal-toxicity-prompts. Jupyter Notebook
★ 234TruthfulQA. TruthfulQA: Measuring How Models Imitate Human Falsehoods
★ 935alpaca-lora. Instruct-tune LLaMA on consumer hardware
★ 19kstanford_alpaca. Code and documentation to train Stanford's Alpaca models, and generate the data.
★ 30kpyllama. LLaMA: Open and Efficient Foundation Language Models
★ 2.8kevals. Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
★ 19kOpenChatKit. Python
★ 9koptimate. A collection of libraries to optimise AI model performances
★ 8.3kAwesome-ChatGPT. ChatGPT资料汇总学习,持续更新......
★ 4.2kminimal-llama. Python
★ 455ChatGPT-Official. ChatGPT Client using official OpenAI API
★ 99PyGPT. Python implementation of Unofficial ChatGPT Client
★ 47chatllama. ChatLLaMA 📢 Open source implementation for LLaMA-based ChatGPT runnable in a single GPU. 15x faster training process than ChatGPT
★ 1.2kGEMBA. GEMBA — GPT Estimation Metric Based Assessment
★ 153LLM-Augmenter.
★ 445TaskMatrix. Python
★ 34kmetaseq. Repo for external large-scale work
★ 6.6kCodeGen. CodeGen is a family of open-source model for program synthesis. Trained on TPU-v4. Competitive with OpenAI Codex.
★ 5.2kOpen-Assistant. OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.
★ 37kFlexLLMGen. Running large language models on a single GPU for throughput-oriented scenarios.
★ 9.4kllama. Inference code for Llama models
★ 60kMaxwelsDonc.
★ 7mozart. [COLING 2022] Data for our paper "Are Pretrained Multilingual Models Equally Fair Across Languages?"
★ 4following-instructions-human-feedback.
★ 1.3ksummarize-from-feedback. Code for "Learning to summarize from human feedback"
★ 1.1klm-human-preferences. Code for the paper Fine-Tuning Language Models from Human Preferences
★ 1.4kcrows-pairs. This repository contains the data and code introduced in the paper "CrowS-Pairs: A Challenge Dataset for Measuring Social Biases in Masked Language Models" (EMNLP 2020).
★ 138sentence-prediction. Python
★ 5PoLitBert. Polish RoBERTA model trained on Polish literature, Wikipedia, and Oscar. The major assumption is that quality text will give a good model.
★ 34transformer_grammars. Transformer Grammars: Augmenting Transformer Language Models with Syntactic Inductive Biases at Scale, TACL (2022)
★ 139detect-SCC. Detecting Strongly Connected Components for Scholarly Data
★ 2cramming. Cramming the training of a (BERT-type) language model into limited compute.
★ 1.4knotebooks. Interactive Jupyter Notebooks for learning materials
★ 56wiki_parse. Tools for multilingual constiuency / dependency parsing of wiki*.
★ 4stanford-openie-python. Stanford Open Information Extraction made simple!
★ 682enhanced-subject-verb-object-extraction. Enhanced Subject Word Object Extraction
★ 154fastText. Library for fast text representation and classification.
★ 27kresponsibleNLPresearch. templates and other documents regarding responsible NLP research
★ 72WAX. The respository describing a novel datasets for word association explanations
★ 13order_position. Python
★ 1sentence-transformers-for-analogies. Jupyter Notebook
★ 6jax. Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more
★ 36kdeepmind-research. This repository contains implementations and illustrative code to accompany DeepMind publications
★ 15kivy. Convert Machine Learning Code Between Frameworks
★ 14kaclpubcheck. Tools for checking ACL paper submissions
★ 1kResearch. novel deep learning research works with PaddlePaddle
★ 1.8kwordnet-randomwalk-python. Repository of code used for WordNet random walk embedding experiments
★ 9robustLM. Code for paper "Better Language Model with Hypernym Class Prediction"
★ 5sense-bert. This is the code for loading the SenseBERT model, described in our paper from ACL 2020.
★ 48LIBERT. Code from the paper "Specializing Unsupervised Pretraining Models for Word-Level Semantic Similarity"
★ 19KTNET. Python
★ 6nltk_data. NLTK Data
★ 1.8ksentime. Ensemble Approach to Extract Sentiments from Tweets or Product Reviews
★ 1entity2vec. Generates a set of property-specific entity embeddings from knowledge graphs using node2vec
★ 78ACL2022_KnowledgeNLP_Tutorial. Materials for ACL-2022 tutorial: Knowledge-Augmented Methods for Natural Language Processing
★ 286MRFN. Tensorflow implementation for MRFN in Retrieval-based Chatbots
★ 48datasets. 🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
★ 22kAllenNLP-Tutorials-Chinese. 中文AllenNLP教程(持续更新)
★ 14allennlp-tutorials. a detail tutorials of allennlp , which is based on my own view.
★ 10gensen. Learning General Purpose Distributed Sentence Representations via Large Scale Multi-task Learning
★ 311spaCy. 💫 Industrial-strength Natural Language Processing (NLP) in Python
★ 34kkb. KnowBert -- Knowledge Enhanced Contextual Word Representations
★ 374evaluate. 🤗 Evaluate: A library for easily evaluating machine learning models and datasets.
★ 2.5kKILT. Library for Knowledge Intensive Language Tasks
★ 978KESA. The source code of KESA
★ 31Paper-Writing-Tips. MLNLP社区用来帮助大家避免论文投稿小错误的整理仓库。 Paper Writing Tips
★ 4.6kprompt-vr. Jupyter Notebook
★ 1OTT-QA. Code and Data for ICLR2021 Paper "Open Question Answering over Tables and Text"
★ 164Wikipedia_annotator. Wikipedia plain text, anchors and their positions, associated entities extractor
★ 1alphafold. Open source code for AlphaFold 2.
★ 15kAnnotated-WikiExtractor. Simple Wikipedia plain text extractor with article link annotations and Hadoop support.
★ 103KGPT. Code and Data for EMNLP2020 Paper "KGPT: Knowledge-Grounded Pre-Training for Data-to-Text Generation"
★ 147Variational-Vocabulary-Selection. Code for NAACL19 Paper "How Large a Vocabulary Does Text Classification Need? A Variational Approach to Vocabulary Selection"
★ 42el-helpfulness-dataset. A dataset used in the paper "Evaluating the Helpfulness of Linked Entities to Readers"
★ 5textent. Representation Learning of Entities and Documents from Knowledge Base Descriptions
★ 18wikipedia-nlp. Sample code for natural language processing using Wikipedia
★ 19KnowledgeCorpus. Some corpus used for pre-training entity- or description-related tasks
★ 1DisExtract. The library that uses dependency parsing to preprocess text to train DisSent model
★ 31