lit-llama. Implementation of the LLaMA language model based on nanoGPT. Supports flash attention, quantization, LoRA fine-tuning, pre-training. Apache 2.0-licensed.

github.com/t-vi/lit-llama

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.