oxillama. Pure Rust LLM Inference Engine — The Sovereign Alternative to llama.cpp License Rust Complete GGUF model loading, multi-format quantized inference, and an OpenAI-compatible API server — all without a single line of C, C++, or Fortran code.

github.com/cool-japan/oxillama

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.