memory_efficient_attention.pytorch. A human-readable PyTorch implementation of "Self-attention Does Not Need O(n^2) Memory" (Rabe&Staats'21).

github.com/moskomule/memory_efficient_attention.pytorch

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.