Details
cuda : enable CUDA graphs for MMID 1 <= BS <= 4 (#19645)
-
cuda : enable CUDA graphs for MMID BS <= 4
-
cont : add stream capture check
Co-authored-by: Oliver Simons osimons@nvidia.com
- cont : add MMVQ_MMID_MAX_BATCH_SIZE
Co-authored-by: Oliver Simons osimons@nvidia.com
macOS/iOS:
Linux:
Windows:
- Windows x64 (CPU)
- Windows arm64 (CPU)
- Windows x64 (CUDA 12) - CUDA 12.4 DLLs
- Windows x64 (CUDA 13) - CUDA 13.1 DLLs
- Windows x64 (Vulkan)
- Windows x64 (SYCL)
- Windows x64 (HIP)
openEuler: