github
GitHub APIExtracts: repo, release, stars delta · metadata or bounded excerpt
Retention: D1 event history with canonical source link and deduplication metadata.
Access basis: Public API; token recommended for production rate limits.
Last ingested: 2026-08-26 · run status: success with data
- 2026-07-06ggerganov/llama.cpp b9888: b9888
<details open> CUDA: extend K-type validation to V-types for flash attention (#24403) * CUDA: extend K-type validation to V-types for flash attention * reorder </details> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b
github:ggerganov/llama.cpp - 2026-07-06ggerganov/llama.cpp b9886: b9886
<details open> ggml-cpu: use UE4M3 LUT in ARM NVFP4 dot product (#25331) </details> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b9886/llama-b9886-bin-macos-arm64.tar.gz) - macOS Apple Silicon (arm64, KleidiAI enabled)
github:ggerganov/llama.cpp - 2026-07-06ggerganov/llama.cpp b9885: b9885
<details open> ggml-cpu: Enable tiled matmul on AIX (#25199) The matmul_tiled path uses large local stack buffers for A_pack and B_pack. On AIX this can trigger a segmentation fault, so reduce the buffer footprint there to keep the tiled path usable. Performance Impact: ~
github:ggerganov/llama.cpp - 2026-07-06ggerganov/llama.cpp b9884: b9884
<details open> vulkan: fix 32-bit integer overflow in CEIL_DIV (#25245) </details> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b9884/llama-b9884-bin-macos-arm64.tar.gz) - macOS Apple Silicon (arm64, KleidiAI enabled) [
github:ggerganov/llama.cpp - 2026-07-06ggerganov/llama.cpp b9882: b9882
<details open> scripts : use HF_TOKEN when downloading UI assets (#25280) Signed-off-by: Adrien Gallouët <angt@huggingface.co> </details> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b9882/llama-b9882-bin-macos-arm64.t
github:ggerganov/llama.cpp - 2026-07-06ggerganov/llama.cpp b9881: b9881
<details open> ggml-hip: enable -ffast-math for HIP builds (#23862) </details> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b9881/llama-b9881-bin-macos-arm64.tar.gz) - macOS Apple Silicon (arm64, KleidiAI enabled) [DISA
github:ggerganov/llama.cpp - 2026-07-05ggerganov/llama.cpp b9878: b9878
<details open> Fix stale tensor-split params for draft models (#24814) * meta: fix tensor split metadata for GQA attention * Tidied the code a bit to match existing style * Revert "Tidied the code a bit to match existing style" This reverts commit b90c6c6300091fe09e2350a3d4e
github:ggerganov/llama.cpp - 2026-07-05ggerganov/llama.cpp b9877: b9877
<details open> abort if we see a multi buffer (#25276) </details> **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b9877/llama-b9877-bin-macos-arm64.tar.gz) - macOS Apple Silicon (arm64, KleidiAI enabled) [DISABLED](https:/
github:ggerganov/llama.cpp