ik_llama_opt/docs
Samuel Oliveira Alves 2f068b5d87
dflash: use draft context as capacity contract (#2341)
dflash: account selector drafts in telemetry
2026-08-26 18:26:50 +02:00
..
backend Merge mainline - Aug 12 2024 (#17) 2024-08-12 15:14:32 +02:00
development hotswap: keep load-time-derived and transformed tensors coherent after reloads (follow-up to #2131) (#2163) 2026-07-22 17:34:43 +03:00
android.md Merge mainline llama.cpp (#3) 2024-07-27 07:55:01 +02:00
autoparser.md common: handle Laguna chat delimiters (#1943) 2026-06-10 07:46:19 +02:00
build.md Update repository clone instructions in build.md (#1753) 2026-05-07 12:57:06 +03:00
docker.md Update Docker documentation with important notice 2026-03-15 12:35:04 +01:00
function-calling.md common : introduce composable PEG parser combinators for chat parsing and new jinja template engine (#1369) 2026-03-09 11:03:33 +01:00
install.md Merge mainline llama.cpp (#3) 2024-07-27 07:55:01 +02:00
llguidance.md Tool calls support from mainline (#723) 2025-09-01 08:38:49 +03:00
parameters.md dflash: use draft context as capacity contract (#2341) 2026-08-26 18:26:50 +02:00
speculative.md model: add openPangu-2.0-Flash (92B-A6B) with MLA-latent cache, DSA/SWA, mHC, and multi-head MTP (#2065) 2026-07-11 12:29:20 +03:00