Rare find

RocketKV. [ICML 2025] RocketKV: Accelerating Long-Context LLM Inference via Two-Stage KV Cache Compression

github.com/NVlabs/RocketKV

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.