github ggml-org/llama.cpp b10359

one hour ago
Details

ggml-webgpu: fix CI errors from #25025 and #25262 (#26566)

  • test new flash_attn test

  • rebase and fix to disable subgrou matrices when max_kv_tile == 0

  • delete log output

  • Add i32 support to cpy and enables the all ops test

  • restore the non target ci tests

  • comment out of TODO of build-cpu.yml

  • fix format

Website:

macOS/iOS:

Linux:

Android:

Windows:

openEuler:

  • DISABLED
  • openEuler x86 (310p)
  • openEuler x86 (910b, ACL Graph)
  • openEuler aarch64 (310p)
  • openEuler aarch64 (910b, ACL Graph)

UI:

Don't miss a new llama.cpp release

NewReleases is sending notifications on new releases.