xlite-dev / ffpa-attnView on GitHub
Fast and Memory-Efficient Exact Attention (BF16/FP16/FP8/FP4) for Large Headdim, 1.5x~15x speedup over PyTorch SDPA.
329Aug 31, 2026Updated this week

Alternatives and similar repositories for ffpa-attn

Users that are interested in ffpa-attn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

Are these results useful?