RAG-Survey. Collecting awesome papers of RAG for AIGC. We propose a taxonomy of RAG foundations, enhancements, and applications in paper "Retrieval-Augmented Generation for AI-Generated Content: A Survey".
1.8kvllm. A high-throughput and memory-efficient inference and serving engine for LLMs
2EMTI_journey. Efficient Multi-Turn Inference official code for NeurIPS 2025
1hymiezhao.github.io. Github Pages template for academic personal websites, forked from mmistakes/minimal-mistakes
1DistServe. Disaggregated serving system for Large Language Models (LLMs).
1