inclusionAI/cuLA
CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.
github.com/inclusionAI/cuLA ↗
Outside contributors get real replies here, and their work gets merged.
10 of 26outside PRs merged (38%)
checked 1 day ago
Starter issues
README
Where newcomer work lands
The evidence