github ggml-org/llama.cpp b9320

latest releases: b11549, b11548, b11547...
4 months ago
Details

TP: fix ggml context size calculation (#22616)

  • TP: fix ggml context size calculation, memory leak

  • move split state cache back into the context

  • revert to constant ggml context size for cgraphs

  • increase headroom for statically allocated tensors

  • remove obsolete include

macOS/iOS:

Linux:

Android:

Windows:

openEuler:

UI:

Don't miss a new llama.cpp release

NewReleases is sending notifications on new releases.