ggml · Issues· 363 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #1630
Why doesn't ggml distrbute libggml-cpu-{arch}.so got all CPU architectures?
Updated Sep 16, 2026 - #1498
ggml-cuda: missing __syncthreads() in ssm-scan.cu
Updated Sep 15, 2026 - #1628
[Feature request] Use a stable GGML CPU thread pool instead of re-creating it for each request
Updated Sep 14, 2026 - #1223
Upstream the mmap and splits loader from llama.cpp
Updated Sep 14, 2026 - #1624
ggml-cuda: missing <cuda/iterator>/<cuda/execution>/<cuda/stream_ref> includes break build with CCCL 3.1+ (argsort.cu, top-k.cu)
Updated Sep 11, 2026 - #1621
MUSA: flash-attn TILE GQA path (ncols2>1) emits all-NaN logits when attention window length is a multiple of 256
Updated Sep 7, 2026 - #1615
vulkan: vkGetDeviceQueue2 returns VK_NULL_HANDLE on older AMD drivers, first vkQueueSubmit aborts
Updated Sep 4, 2026 - #1609
Convert ggml to Rust
Updated Aug 27, 2026 - #1608
cpu: support GGML_CPU_ALL_VARIANTS with static linking
Updated Aug 27, 2026 - #1545
ggml-metal: a transient command-buffer failure (e.g. iOS background-GPU rejection) permanently latches the whole backend into an error state
Updated Aug 11, 2026 - #1554
libggml 0.13.1 fails in MUL_MAT_ID() on i915 (Intel comet lake) on freebsd-16
Updated Jul 15, 2026 - #1565
libggml isn't throwing an error if intel gpu fence timeout happens but DOES if GGML_VK_PERF_LOGGER=1
Updated Jul 15, 2026 - #1557
CUDA backend crashes with "an illegal memory access" on SM 8.9 (RTX 4060 Laptop) when processing ViT-style small matmul shapes
Updated Jul 9, 2026 - #1555
Question about requirements.txt dependencies
Updated Jul 6, 2026 - #1486
Windows build fails due to MLX backend (undefined symbols, no Windows build tags)
Updated Jun 22, 2026