Rare find

large-model-proxy. Run multiple resource-heavy Large Models (LM) on the same machine with limited amount of VRAM/other resources by exposing them on different ports and loading/unloading them on demand

github.com/perk11/large-model-proxy

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.