lm-evaluation-harness-fast. speedup for lm-evaluation-harness; support tensor-parallel inference and data-parallel inference; support gptq, bitsandbytes, peft and exllamav2.

github.com/LZY-the-boys/lm-evaluation-harness-fast

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.