sntdismas/Sageattention-3-torch2.11.0-cu130-py312-Blackwell
SageAttention 3 Wheel for Blackwell / CUDA 13.0 / PyTorch 2.11
This repository contains a locally compiled Python wheel for SageAttention 3 built for NVIDIA Blackwell GPUs.
The wheel was compiled for a RunPod ComfyUI environment using:
- Python: 3.12
- PyTorch: 2.11.0+cu130
- CUDA: 13.0
- GPU architecture: Blackwell
sm_120 - Target platform: Linux x86_64
- Package:
sageattn3
About SageAttention 3
SageAttention 3 is a Blackwell-oriented attention implementation designed to take advantage of NVIDIA RTX 50-series / Blackwell capabilities, including NVFP4-related acceleration paths.
It is intended for experimental acceleration of heavy image/video generation workloads such as ComfyUI pipelines with models like Flux or Wan, where attention performance can be a major bottleneck.
Source and credits
This wheel was built using the official SageAttention source repository:
- https://github.com/thu-ml/SageAttention
The build process was based on the RunPod compilation guide published by Seryoger on HuggingFace:
- https://huggingface.co/Seryoger/Sageattention-3-cu130-5090-endpoint
This wheel was rebuilt locally for a newer ComfyUI environment using:
- PyTorch 2.11.0+cu130
- CUDA 13.0
- Blackwell `sm_120`
Build notes
The wheel was compiled inside a RunPod Blackwell pod.
The important build parameters were:
export TORCH_CUDA_ARCH_LIST="12.0"