Latest wheel from woct0rdho includes the torch.compile fix: https://github.com/woct0rdho/SageAttention/releases Based on my quick testing this reduces peak VRAM usage a bit when running sageattn + torch.compile
Latest wheel from woct0rdho includes the torch.compile fix: https://github.com/woct0rdho/SageAttention/releases Based on my quick testing this reduces peak VRAM usage a bit when running sageattn + torch.compile