This is your work, valued
mechanistic interpretability of LLM
awesome-llm-understanding-mechanism. awesome papers in LLM interpretability
625srnn. sliced-rnn
468awesome-SAE. awesome SAE papers
79sli_rec. Python
75neuron-attribution. code for EMNLP 2024 paper: Neuron-Level Knowledge Attribution in Large Language Models
52awesome-LLM-neuron.
36in-context-mechanism. code for EMNLP 2024 paper: How do Large Language Models Learn In-Context? Query and Key Matrices of In-Context Heads are Two Towers for Metric Learning
13arithmetic-mechanism. code for EMNLP 2024 paper: Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis
12llava-mechanism. code for arxiv paper: Understanding Multimodal LLMs: the Mechanistic Interpretability of Llava in Visual Question Answering
7