| --- |
| language: |
| - en |
| --- |
| |
| | Prebuilt Wheels | Python Versions | PyTorch Versions | CUDA Versions | Source | |
| |------------------------------------------------|-----------------|------------------|----------------|---------------------------------------------------------------------------| |
| | [Flash-Attention 2.7.4.post1](https://huggingface.co/lym00/win_amd64_prebuilt_wheels/blob/main/flash_attn-2.7.4.post1-cp312-cp312-win_amd64.whl) | 3.12 | 2.8.0.dev | 12.8.1 | [Dao-AILab/flash-attention](https://github.com/Dao-AILab/flash-attention) | |
| | SageAttention2++ ([pending official approval](https://huggingface.co/jt-zhang/SageAttention2_plus/discussions/4)) | 3.12 | 2.8.0.dev | 12.8.1 | [jt-zhang/SageAttention2_plus](https://huggingface.co/jt-zhang/SageAttention2_plus) | |
| | SageAttention3 (pending official release) | 3.12 | 2.8.0.dev | 12.8.1 | INSERT | |
| | INSERT | INSERT | INSERT | INSERT | INSERT | |
|
|