Rare find

cuLA. CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.

github.com/inclusionAI/cuLA

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.