llama.cpp releases · · 1 min read

b9939

Mirrored from llama.cpp releases for archival readability. Support the source by reading on the original site.

metal : add CONV_2D_DW (depthwise convolution) support (#21565)

  • metal : add CONV_2D_DW (depthwise 2D convolution) support

  • test : add perf cases for CONV_2D_DW

  • metal : use 3D dispatch for CONV_2D_DW kernel

  • metal : add channel-tiled CONV_2D_DW kernel for non-contiguous layouts

  • metal : simplify CONV_2D_DW dispatch and trim comments

  • metal : merge duplicate CONV_2D_DW pipeline getters

  • tests : add F16 CONV2D_DW tests

  • cpu : fix F16 kernel support for CONV_2D_DW

  • tests : remove commented-out CONV_2D_DW test block


Co-authored-by: Georgi Gerganov [email protected]

macOS/iOS:

Linux:

Android:

Windows:

openEuler:

  • DISABLED
  • openEuler x86 (310p)
  • openEuler x86 (910b, ACL Graph)
  • openEuler aarch64 (310p)
  • openEuler aarch64 (910b, ACL Graph)

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from llama.cpp releases