ik_llama_opt/ggml
Joel Farthing 08b500b958
ggml: fix HC_POST single-token CPU chunk count (#2357)
2026-08-25 16:29:06 +02:00
..
cmake Merge mainline llama.cpp (#3) 2024-07-27 07:55:01 +02:00
include Fix KQ mask padding for the Vulkan back-end (#2350) 2026-08-24 18:31:17 +02:00
src ggml: fix HC_POST single-token CPU chunk count (#2357) 2026-08-25 16:29:06 +02:00
.gitignore Merge mainline llama.cpp (#3) 2024-07-27 07:55:01 +02:00
CMakeLists.txt Chunked experts (CPU) (#2202) 2026-07-30 13:16:02 +03:00