llama.cpp releases · · 1 min read

b11126

Mirrored from llama.cpp releases for archival readability. Support the source by reading on the original site.

vulkan: add IQ4_XS MMQ/MMV matmul kernels (#28415)

  • vulkan: optimize IQ4_XS matmul kernels

Assisted-by: OpenAI Codex

  • vulkan: address IQ4_XS review nits
  • drop the dead LOAD_VEC_A != 8 branch in the IQ4_XS shmem load; iq4_xs is
    in lut_load_vec_a()'s "8" list, so that path is never generated
  • disable MMVQ for IQ4_XS on Intel (27.3% tg regression on A770)
  • remove a stray empty line in types.glsl

Assisted-By: Claude Opus 5 [email protected]

Website:

Attestations:

macOS/iOS:

Linux:

Android:

Windows:

openEuler:

  • DISABLED
  • openEuler x86 (310p)
  • openEuler x86 (910b, ACL Graph)
  • openEuler aarch64 (310p)
  • openEuler aarch64 (910b, ACL Graph)

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from llama.cpp releases