github
GitHub APIExtracts: repo, release, stars delta · metadata or bounded excerpt
Retention: D1 event history with canonical source link and deduplication metadata.
Access basis: Public API; token recommended for production rate limits.
Last ingested: 2026-08-25 · run status: success with data
- 2026-07-17ggerganov/llama.cpp b10066: b10066
<details open> opencl: load and use `kernel_gemm_moe_q6_k_f32_ns` from bin kernel lib (#25797) </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b10066/llama-b10066-bin-macos-ar
github:ggerganov/llama.cpp - 2026-07-17ggerganov/llama.cpp b10064: b10064
<details open> opencl: transpose q4_K noshuffle scales for coalesced reads (#25805) </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b10064/llama-b10064-bin-macos-arm64.tar.gz)
github:ggerganov/llama.cpp - 2026-07-17ggerganov/llama.cpp b10063: b10063
<details open> sync : ggml </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b10063/llama-b10063-bin-macos-arm64.tar.gz) - macOS Apple Silicon (arm64, KleidiAI enabled) [DISABLE
github:ggerganov/llama.cpp - 2026-07-17ggerganov/llama.cpp b10061: b10061
<details open> tests : initialize all tensors in test_dsv4_hc to avoid NaNs in sentinel tensors (#25822) Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com> </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/gg
github:ggerganov/llama.cpp - 2026-07-17ggerganov/llama.cpp b10059: b10059
<details open> ggml-blas: default hadamard mul_mat to cpu routine (#25710) Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b1005
github:ggerganov/llama.cpp - 2026-07-17ggerganov/llama.cpp b10058: b10058
<details open> vulkan: Support Q2_0 (#25430) * vulkan: Support Q2_0 The backend perf tests for mat-vec-mul weren't very good at first (worse than q2_k), doubling the rows per workgroup made a big difference. * reorder * resolve merge conflict, adjust err threshold for f16->q
github:ggerganov/llama.cpp - 2026-07-17ggerganov/llama.cpp b10057: b10057
<details open> sycl: fix row calculation when K_QUANTS_PER_ITERATION is 1 (#25690) * sycl: fix incorrect row calculation when K_QUANTS_PER_ITERATION=1 Signed-off-by: Todd Malsbary <todd.malsbary@intel.com> * sycl: use K_QUANTS_PER_ITERATION for non-reordered Q5_K kernel This
github:ggerganov/llama.cpp - 2026-07-17ggerganov/llama.cpp b10056: b10056
<details open> opencl: add ABS op (#25115) </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b10056/llama-b10056-bin-macos-arm64.tar.gz) - macOS Apple Silicon (arm64, KleidiAI e
github:ggerganov/llama.cpp - 2026-07-17ggerganov/llama.cpp b10054: b10054
<details open> docs: added a note about using OpenCl with Adreno 810 (#25786) </details> **Website:** - <https://llama.app> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b10054/llama-b10054-bin-macos-arm64.tar.gz) - mac
github:ggerganov/llama.cpp