llm-inference-simulator. 🚀 LLM inference optimization simulator, modeling compute-bound prefill and memory-bound decode phases.

github.com/Muhtasham/llm-inference-simulator

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.