This is your work, valued
NJUNMT-pytorch. Python
★ 93synextractor. extractor chinese synonyms in large corpus
★ 11Iterative-Domain-Repaired-Back-Translation. Code for Paper 'Iterative Domain-Repaired Back-Translation' in EMNLP2020
★ 6ODC-NMT. Python
★ 5theano_cudnn_rnn_benchmark. Python
★ 2EXAMPLES-FOR-NMT-EXP.
★ 1picotron. Minimalistic 4D-parallelism distributed training framework for education purpose
★ 2.3kimmersive-translate. 沉浸式双语网页翻译扩展 , 支持输入框翻译, 鼠标悬停翻译, PDF, Epub, 字幕文件, TXT 文件翻译 - Immersive Dual Web Page Translation Extension
★ 18kTheAgentCompany. An agent benchmark with tasks in a simulated software company.
★ 752multilingual-needle-in-a-haystack. Jupyter Notebook
★ 6PRIME. Scalable RL solution for advanced reasoning of language models
★ 1.9kidioms-incontext-mt. idioms in context dataset
★ 5minimal-stable-PPO. A minimal and stable PPO.
★ 148EQ-Bench. A benchmark for emotional intelligence in large language models
★ 444mt-evaluation. A framework for evaluating Machine Translation models.
★ 13llm-datasets. A collection of datasets for language model pretraining including scripts for downloading, preprocesssing, and sampling.
★ 66prompt-cache. Modular and structured prompt caching for low-latency LLM inference
★ 115LanguageCodes. We present a list of languages with their codes, families, regions and etc. We also present a list of multi-lingual corpora (with urls).
★ 87GeniL. GeniL dataset is an effort for detecting various types of generalization in language. This multilingual dataset covers sentences in EN, FR, ES, PT, AR, HI, BN, MS, and ID and is annotated by native speakers of each language. Each sentence is collected from a public corpora of language and contains at least one identity group name and an attribute.
★ 3ethnologue. Data from Ethnologue (language code conversion & language family) and the web scraper.
★ 9ai-notes. notes for software engineers getting up to speed on new AI developments. Serves as datastore for https://latent.space writing, and product brainstorming, but has cleaned up canonical references under the /Resources folder.
★ 6.2kOmniSearch. Repo for Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent
★ 429lingua. Meta Lingua: a lean, efficient, and easy-to-hack codebase to research LLMs.
★ 4.8kEverything-of-Thoughts-XoT. An implemtation of Everyting of Thoughts (XoT).
★ 161Awesome-LLM-Strawberry. A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.
★ 6.9kO1-Journey. O1 Replication Journey
★ 2kurl-nlp. Python
★ 273lars_datasets. Lars's datasets
★ 12List-of-Dirty-Naughty-Obscene-and-Otherwise-Bad-Words. List of Dirty, Naughty, Obscene, and Otherwise Bad Words
★ 3.4ksacreCOMET. Python
★ 8Qwen3-VL. Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
★ 20kChatLearn. A flexible and efficient training framework for large-scale alignment tasks
★ 452Liger-Kernel. Efficient Triton Kernels for LLM Training
★ 6.5kcourses. Anthropic's educational courses
★ 22kglobal-exams. Multilingual Global Exams
★ 10fragments. Open-source Next.js template for building apps that are fully generated by AI. By E2B.
★ 6.4kMinitron. A family of compressed models obtained via pruning and knowledge distillation
★ 384cookbook. Deep learning for dummies. All the practical details and useful utilities that go into working with real models.
★ 845llm-datasets. Curated list of datasets and tools for post-training.
★ 4.7kToolLearningPapers.
★ 923deep-significance. Enabling easy statistical significance testing for deep neural networks.
★ 339opensubtitles-scraper. scrape subtitles from opensubtitles.org
★ 51llm-leaderboard. Project of llm evaluation to Japanese tasks
★ 94crewAI. Framework for orchestrating role-playing, autonomous AI agents. By fostering collaborative intelligence, CrewAI empowers agents to work together seamlessly, tackling complex tasks.
★ 56kMinerU. Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
★ 76kpypandoc. Thin wrapper for "pandoc" (MIT)
★ 1.1kpandoc. Universal markup converter
★ 46kmath-evaluation-harness. A simple toolkit for benchmarking LLMs on mathematical reasoning tasks. 🧮✨
★ 278ddgs. A metasearch library that aggregates results from diverse web search services
★ 2.9kmaps. Multicultural Proverbs and Sayings
★ 13LLM101n. LLM101n: Let's build a Storyteller
★ 38kMETAL-Towards-Multilingual-Meta-Evaluation. Official Codebase for "METAL: Towards Multilingual Meta-Evaluation" (published in Findings of ACL: NAACL 2024)
★ 6reddit-to-threads. Convert arctic_shift Reddit data dumps into thread-view documents
★ 11arctic_shift. Making Reddit data accessible to researchers, moderators and everyone else. Interact with the data through large dumps, an API or web interface.
★ 1.3kclemcore. A Framework for the Systematic Evaluation of Chat-Optimized Language Models as Conversational Agents and an Extensible Benchmark
★ 32segmoe. Python
★ 433paxqa. Code and Data for "PAXQA: Generating Cross-lingual Question Answering Examples at Training Scale" (EMNLP 2023)
★ 2translation-agent. Python
★ 5.8kNTREX. NTREX -- News Test References for MT Evaluation
★ 87xstreet.
★ 6Sophia. Effortless plugin and play Optimizer to cut model training costs by 50%. New optimizer that is 2x faster than Adam on LLMs.
★ 4online_merging_optimizers. Implementations of online merging optimizers proposed by Online Merging Optimizers for Boosting Rewards and Mitigating Tax in Alignment
★ 82torchtitan. A PyTorch native platform for training generative AI models
★ 5.6ktransagents. The official repository of the paper "(Perhaps) Beyond Human Translation: Harnessing Multi-Agent Collaboration for Translating Ultra-Long Literary Texts"
★ 604Collie. [ICLR 2024] COLLIE: Systematic Construction of Constrained Text Generation Tasks
★ 64JMMLU. 日本語マルチタスク言語理解ベンチマーク Japanese Massive Multitask Language Understanding Benchmark
★ 40GuoFeng-Webnovel. Multilingual Corpus of Web Fiction
★ 2033AM. Official code and data of "3AM: An Ambiguity-Aware Multi-Modal Machine Translation Dataset"
★ 12SeaEval. NAACL 2024: SeaEval for Multilingual Foundation Models: From Cross-Lingual Alignment to Cultural Reasoning
★ 26awesome-synthetic-datasets. awesome synthetic (text) datasets
★ 336UltraEval. [ACL 2024 Demo] Official GitHub repo for UltraEval: An open source framework for evaluating foundation models.
★ 257multilingual-register-labeling. Multilingual, multilabel modeling of registers
★ 3NeuScraper. [ACL 2024] This is the code repo for our ACL’24 paper "Cleaner Pretraining Corpus Curation with Neural Web Scraping".
★ 228QuRating. [ICML 2024] Selecting High-Quality Data for Training Language Models
★ 204LlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74kThunderKittens. Tile primitives for speedy kernels
★ 3.6kXpersona. XPersona: Evaluating Multilingual Personalized Chatbot
★ 71simple-evals. Python
★ 4.6kseahorse. Seahorse is a dataset for multilingual, multi-faceted summarization evaluation. It consists of 96K summaries with human ratings along 6 quality dimensions: comprehensibility, repetition, grammar, attribution, main idea(s), and conciseness, covering 6 languages, 9 systems and 4 datasets.
★ 90marker. Convert PDF to markdown + JSON quickly with high accuracy
★ 38ksailcraft. 🚢 Data Toolkit for Sailor Language Models
★ 94databonsai. clean & curate your data with LLMs.
★ 487RTP-LX. Repository for the paper "RTP-LX: Can LLMs Evaluate Toxicity in Multilingual Scenarios?"
★ 29megalodon. Reference implementation of Megalodon 7B model
★ 526CLIcK. CLIcK: A Benchmark Dataset of Cultural and Linguistic Intelligence in Korean
★ 48CLiKA. Evaluation of the Cross-Lingual Knowledge Alignment in LLMs
★ 9EasyContext. Memory optimization and training recipes to extrapolate language models' context length to 1 million tokens, with minimal hardware.
★ 761tasksource. Datasets collection and preprocessings framework for NLP extreme multitask learning
★ 196GlotLID. [EMNLP 2023] 💬 Language Identification with Support for More Than 2000 Labels
★ 214mc2_corpus. [ACL'24] MC^2: A Multilingual Corpus of Minority Languages in China (Tibetan, Uyghur, Kazakh, and Mongolian)
★ 37flores. The FLORES+ Machine Translation Benchmark
★ 112mlama. Python
★ 25DeepSeek-VL. DeepSeek-VL: Towards Real-World Vision-Language Understanding
★ 4.1kinfinity. The AI-native database built for LLM applications, providing incredibly fast hybrid search of dense vector, sparse vector, tensor (multi-vector), and full-text.
★ 4.7kaisys-building-blocks. Building blocks for foundation models.
★ 635OpenCodeInterpreter. OpenCodeInterpreter is a suite of open-source code generation systems aimed at bridging the gap between large language models and sophisticated proprietary systems like the GPT-4 Code Interpreter. It significantly enhances code generation capabilities by integrating execution and iterative refinement functionalities.
★ 1.7kLWM. Large World Model -- Modeling Text and Video with Millions Context
★ 7.4kunsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kGPTFast. Accelerate your Hugging Face Transformers 7.6-9x. Native to Hugging Face and PyTorch.
★ 684AceGPT. Python
★ 127nusa-crowd. A collaborative project to collect datasets in Indonesian languages.
★ 288wikipedia-utils. Utility scripts for preprocessing Wikipedia texts for NLP
★ 78Mr.Translate. ChatGPT 翻译助手 Prompt
★ 347Noi. 🚀 Less chaos. More flow.
★ 9kmlx. MLX: An array framework for Apple silicon
★ 28kSALMON. Self-Alignment with Principle-Following Reward Models
★ 170minbpe. Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization.
★ 11kmamba-scans. Blog post
★ 17JCoLA.
★ 19FollowBench. [ACL 2024] FollowBench: A Multi-level Fine-grained Constraints Following Benchmark for Large Language Models
★ 118Qwen3. Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.
★ 27kLong-Context-Data-Engineering. Implementation of paper Data Engineering for Scaling Language Models to 128K Context
★ 502alignment-handbook. Robust recipes to align language models with human and AI preferences
★ 5.7klm-human-preferences. Code for the paper Fine-Tuning Language Models from Human Preferences
★ 1.4kHaluEval. This is the repository of HaluEval, a large-scale hallucination evaluation benchmark for Large Language Models.
★ 593XGLUE. Cross-lingual GLUE
★ 49genshin-langdata. English-Chinese-Japanese translation dataset of the terms in Genshin Impact
★ 42nanotron. Minimalistic large language model 3D-parallelism training
★ 2.8kdeita. Deita: Data-Efficient Instruction Tuning for Alignment [ICLR2024]
★ 600self-rewarding-lm-pytorch. Implementation of the training framework proposed in Self-Rewarding Language Model, from MetaAI
★ 1.4kprimeqa. The prime repository for state-of-the-art Multilingual Question Answering research and development.
★ 739ToRA. ToRA is a series of Tool-integrated Reasoning LLM Agents designed to solve challenging mathematical reasoning problems by interacting with tools [ICLR'24].
★ 1.1kAwesome-LLM-Long-Context-Modeling. 📰 Must-read papers and blogs on LLM based Long Context Modeling 🔥
★ 2.1kMultilingual-Evaluation-of-Generative-AI-MEGA. Code for Multilingual Eval of Generative AI paper published at EMNLP 2023
★ 72globalbench. GlobalBench: A Benchmark for Global Progress in Language Technology
★ 7annotated-mamba. Annotated version of the Mamba paper
★ 502llama-moe. ⛷️ LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training (EMNLP 2024)
★ 1kmars. Mars is a tensor-based unified framework for large-scale data computation which scales numpy, pandas, scikit-learn and Python functions.
★ 2.7kPowerInfer. High-speed Large Language Model Serving for Local Deployment
★ 9.7kcatwalk. This project studies the performance and robustness of language models and task-adaptation methods.
★ 154agentlego. Enhance LLM agents with rich tool APIs
★ 413SafetyBench. Official github repo for SafetyBench, a comprehensive benchmark to evaluate LLMs' safety. [ACL 2024]
★ 296InfiniteBench. Codes for the paper "∞Bench: Extending Long Context Evaluation Beyond 100K Tokens": https://arxiv.org/abs/2402.13718
★ 387weak-to-strong. Python
★ 2.6kpromptbase. All things prompt engineering
★ 5.8kSafety-Prompts. Chinese safety prompts for evaluating and improving the safety of LLMs. 中文安全prompts,用于评估和提升大模型的安全性。
★ 1.2kOMGEval. OMGEval😮: An Open Multilingual Generative Evaluation Benchmark for Foundation Models
★ 36mamba. Mamba SSM architecture
★ 19kseamless_communication. Foundational Models for State-of-the-Art Speech and Text Translation
★ 12ksoftware-documentation-data-set-for-machine-translation. A parallel evaluation data set of SAP software documentation with document structure annotation
★ 15RefGPT. Python
★ 98gpt-fast. Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.
★ 6.2kDeepSeek-LLM. DeepSeek LLM: Let there be answers
★ 7.2kchatgpt_system_prompt. A collection of GPT system prompts and various prompt injection/leaking knowledge.
★ 11kMELA. TeX
★ 7XLingEval. Code and Resources for the paper, "Better to Ask in English: Cross-Lingual Evaluation of Large Language Models for Healthcare Queries"
★ 25Yuan-2.0. Yuan 2.0 Large Language Model
★ 687bff. Rust
★ 39open-instruct. AllenAI's post-training codebase
★ 3.8kwork_space. NLP 项目记录档案
★ 66InstructionWild.
★ 462LookaheadDecoding. [ICML 2024] Break the Sequential Dependency of LLM Inference Using Lookahead Decoding
★ 1.3kAwesome-Embodied-Robotics-and-Agent. This is a curated list of "Embodied AI or robot with Large Language Models" research. Watch this repository for the latest updates! 🔥
★ 1.8kActive-IT. Code for our EMNLP-2023 paper: "Active Instruction Tuning: Improving Cross-Task Generalization by Training on Prompt Sensitive Tasks"
★ 26UniversalNER.
★ 28lm-polygraph. Python
★ 498PiPPy. Pipeline Parallelism for PyTorch
★ 786XVERSE-65B. XVERSE-65B: A multilingual large language model developed by XVERSE Technology Inc.
★ 139fondant. Production-ready data processing made easy and shareable
★ 358Yi. A series of large language models trained from scratch by developers @01-ai
★ 7.8kDeepSeek-Coder. DeepSeek Coder: Let the Code Write Itself
★ 24kSWE-bench. SWE-bench: Can Language Models Resolve Real-world Github Issues?
★ 5.5kFastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40kdata-juicer. Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷
★ 6.8kScienceWorld. ScienceWorld is a text-based virtual environment centered around accomplishing tasks from the standardized elementary science curriculum.
★ 370LLMDataHub. A quick guide (especially) for trending instruction finetuning datasets
★ 3.4kTensorRT-LLM. TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
★ 14kmaxtext. A simple, performant, and scalable Jax LLM!
★ 2.4kautogen. A programming framework for agentic AI
★ 60kCloudflareSpeedTest. 🌩「自选优选 IP」测试 Cloudflare CDN 延迟和速度,获取最快 IP !当然也支持其他 CDN / 多个解析 IP 的网站 ~
★ 28kLightSeq. Official repository for DistFlashAttn: Distributed Memory-efficient Attention for Long-context LLMs Training
★ 223chathub. TypeScript
★ 11kDeep-reinforcement-learning-with-pytorch. PyTorch implementation of DQN, AC, ACER, A2C, A3C, PG, DDPG, TRPO, PPO, SAC, TD3 and ....
★ 4.6kstreaming-llm. [ICLR 2024] Efficient Streaming Language Models with Attention Sinks
★ 7.3kSmartPlay. SmartPlay is a benchmark for Large Language Models (LLMs). Uses a variety of games to test various important LLM capabilities as agents. SmartPlay is designed to be easy to use, and to support future development of LLMs.
★ 146NGT. Nearest Neighbor Search with Neighborhood Graph and Tree for High-dimensional Data
★ 1.4kbelebele. Repo for the Belebele dataset, a massively multilingual reading comprehension dataset.
★ 341ebooklib. A versatile Python library for EPUB2/EPUB3 manipulation and processing.
★ 1.8kBMPrinciples. A collection of phenomenons observed during the scaling of big foundation models, which may be developed into consensus, principles, or laws in the future
★ 284Xwin-LM. Xwin-LM: Powerful, Stable, and Reproducible LLM Alignment
★ 1kOpenRLHF. An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
★ 9.9kLLM-Agent-Survey.
★ 2.9kcog. Containers for machine learning
★ 9.4kLLM-Agent-Paper-List. The paper list of the 86-page SCIS cover paper "The Rise and Potential of Large Language Model Based Agents: A Survey" by Zhiheng Xi et al.
★ 8.2kPai-Megatron-Patch. The official repo of Pai-Megatron-Patch for LLM & VLM large scale training developed by Alibaba Cloud.
★ 1.6kguidance. A guidance language for controlling large language models.
★ 22kagents. An Open-source Framework for Data-centric, Self-evolving Autonomous Language Agents
★ 6kreflexion. [NeurIPS 2023] Reflexion: Language Agents with Verbal Reinforcement Learning
★ 3.2kmlc-llm. Universal LLM Deployment Engine with ML Compilation
★ 23kOmniQuant. [ICLR2024 spotlight] OmniQuant is a simple and powerful quantization technique for LLMs.
★ 905tm2tb. Bilingual term extractor
★ 60linguee-api. Proxy to convert HTML responses from linguee.com to JSON format
★ 209localization-xml-mt. A High-Quality Multilingual Dataset for Structured Documentation Translation
★ 39LLMAgentPapers. Must-read Papers on LLM Agents.
★ 3.1kakshare. AKShare is an elegant and simple financial data interface library for Python, built for human beings! 开源财经数据接口库
★ 22kms-agent. MS-Agent: a lightweight framework to empower agentic execution of complex tasks
★ 4.3kToolQA. ToolQA, a new dataset to evaluate the capabilities of LLMs in answering challenging questions with external tools. It offers two levels (easy/hard) across eight real-life scenarios.
★ 286openinterpreter. A coding agent for open models like Kimi K3
★ 67kBaichuan2. A series of large language models developed by Baichuan Intelligent Technology
★ 4.1kDiffusionLLM. Code for paper "Diffusion Language Models Can Perform Many Tasks with Scaling and Instruction-Finetuning"
★ 84Awesome-Code-LLM. 👨💻 An awesome and curated list of best code-LLM for research.
★ 1.3kGPTQ-triton. GPTQ inference Triton kernel
★ 322Mol-Instructions. [ICLR 2024] Mol-Instructions: A Large-Scale Biomolecular Instruction Dataset for Large Language Models
★ 293xtuner. A Next-Generation Training Engine Built for Ultra-Large MoE Models
★ 5.2kFinGPT. FinGPT: Open-Source Financial Large Language Models! Revolutionize 🔥 We release the trained model on HuggingFace.
★ 21kllama_index. LlamaIndex is the leading document agent and OCR platform
★ 51k