Package profile
flash-attn-4
- Summary: Flash Attention CUTE (CUDA Template Engine) implementation
- Author: Tri Dao
- License: BSD License
- Homepage: https://github.com/Dao-AILab/flash-attention
- Source: https://github.com/Dao-AILab/flash-attention
- Number of releases: 31
- First release: 0.0.1 on 2026-02-09
- Latest release: 4.0.0b33 on 2026-09-30
- Latest release size: 509.9 KB (pure Python wheel)