Details
ggml : optimize cuda ssm_scan using warp-level reduction (#18505)
-
ggml : optimize cuda ssm_scan using warp-level reduction
-
ggml : apply code review suggestions (style, const, constexpr)
-
ggml : add TODO regarding stride consistency
macOS/iOS:
Linux:
Windows:
- Windows x64 (CPU)
- Windows arm64 (CPU)
- Windows x64 (CUDA 12) - CUDA 12.4 DLLs
- Windows x64 (CUDA 13) - CUDA 13.1 DLLs
- Windows x64 (Vulkan)
- Windows x64 (SYCL)
- Windows x64 (HIP)
openEuler: