mirror of
https://github.com/jamiepine/voicebox.git
synced 2026-09-15 04:40:40 -07:00
feat: add Blackwell GPU (sm_120) CUDA support (#401)
Set TORCH_CUDA_ARCH_LIST in the CUDA build step to include 12.0+PTX for forward compatibility with Blackwell GPUs (RTX 5070 Ti, 5080, etc). Pre-built PyTorch cu128 wheels only ship native kernels for sm_80/86/89/90. Without this, Blackwell GPU users get "no kernel image is available for execution on the device" at runtime. Fixes #386 Related: #395, #396, #399, #400 Co-authored-by: Matt Van Horn <[email protected]>
This commit is contained in:
co-authored by
Matt Van Horn
parent
13ba5f1aa6
commit
c9d8142a78
@@ -203,6 +203,12 @@ jobs:
|
||||
- name: Build CUDA server binary (onedir)
|
||||
shell: bash
|
||||
working-directory: backend
|
||||
env:
|
||||
# Include Blackwell (sm_120) via PTX forward compatibility.
|
||||
# Pre-built PyTorch cu128 wheels ship native kernels for sm_80/86/89/90
|
||||
# but not sm_120. Setting this env var causes torch.utils.cpp_extension
|
||||
# (and any JIT-compiled kernels) to target Blackwell GPUs as well.
|
||||
TORCH_CUDA_ARCH_LIST: "8.0;8.6;8.9;9.0;12.0+PTX"
|
||||
run: python build_binary.py --cuda
|
||||
|
||||
- name: Package into server core + CUDA libs archives
|
||||
|
||||
Reference in New Issue
Block a user