Awesome-VLM-Streaming-Video. 📚 A curated collection of papers and open-source code repositories dedicated to the application of Vision-Language Models (VLMs) for streaming video.
190TailorKV. Official implementation of "TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization" (Findings of ACL 2025).
21VecInfer. [ACL 2026 Main] Official implementation of "VecInfer: Efficient LLM Inference with Low-Bit KV Cache via Outlier-Suppressed Vector Quantization" .
10course. HNU_COURSE
2ydyhello.github.io. Stylus
1studydemo. 一些å¦ä¹ æ—¶çš„demo
1