Search

NVIDIA AI Platform · fla-org/flash-linear-attention

2 skills found.
Category:
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Profile and optimize FLA kernels on NVIDIA GPUs, with same-hardware benchmarks and targeted Nsight Compute analysis.

fla-org/flash-linear-attention5.8k—~818Automated safety check: PassMITtoday
2

Port an FLA Triton kernel to Gluon when explicit layouts, asynchronous transfers, or scheduling can address a measured bottleneck.

fla-org/flash-linear-attention5.8k—~1.6kAutomated safety check: PassMITtoday