github ggml-org/llama.cpp b11236

pre-releaseone hour ago
Details

batch: migrate speculative, mtmd and server to batch_ext (#29385)

  • adapt common

  • add common_batch

  • wip

  • wip: spec

  • cont

  • common_speculative_process

  • server_batch to use common_batch

  • rm some stale calls

Assisted-by: Claude Fable 5.1

  • migrate mtmd

  • handle imrope, handle return val of add()/add_embd()

  • add spec zeros vector

  • add warning on zero fill path

Website:

Attestations:

macOS/iOS:

Linux:

Android:

Windows:

openEuler:

  • DISABLED
  • openEuler x86 (310p)
  • openEuler x86 (910b, ACL Graph)
  • openEuler aarch64 (310p)
  • openEuler aarch64 (910b, ACL Graph)

UI:

Don't miss a new llama.cpp release

NewReleases is sending notifications on new releases.