llama.cpp releases · · 1 min read

b10359

Mirrored from llama.cpp releases for archival readability. Support the source by reading on the original site.

ggml-webgpu: fix CI errors from #25025 and #25262 (#26566)

  • test new flash_attn test

  • rebase and fix to disable subgrou matrices when max_kv_tile == 0

  • delete log output

  • Add i32 support to cpy and enables the all ops test

  • restore the non target ci tests

  • comment out of TODO of build-cpu.yml

  • fix format

Website:

macOS/iOS:

Linux:

Android:

Windows:

openEuler:

  • DISABLED
  • openEuler x86 (310p)
  • openEuler x86 (910b, ACL Graph)
  • openEuler aarch64 (310p)
  • openEuler aarch64 (910b, ACL Graph)

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from llama.cpp releases