flash-linear-attention. Fast implementations of causal linear attention for autogressive language modeling (Pytorch)

github.com/berlino/flash-linear-attention

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.