Rare find

flash-attention-minimal. Flash Attention in ~100 lines of CUDA (forward pass only)

github.com/tspeterkim/flash-attention-minimal

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.