ik_llama_opt/common
dungquixote42 a903409a5e
fix adaptive p sampler rewinding too far back (#1359)
* fix adaptive p sampler rewinding too far back

* update comments

* correct default value for total_weight, more comments

* new variables/names

* update comment for n_rewind

* move null pointer check back to common_sampler_review()

* refactor weighted_sum and total_weight to vector<pair>, better boundary check in llama_review_adaptive_p_impl()
2026-03-04 13:26:25 +01:00
..
cmake Merge mainline llama.cpp (#3) 2024-07-27 07:55:01 +02:00
CMakeLists.txt spec : add self speculative decoding, ngram and refactor (#1261) 2026-02-13 19:04:55 +01:00
base64.hpp llava : expose as a shared library for downstream projects (#3613) 2023-11-07 00:36:23 +03:00
build-info.cpp.in build : link against build info instead of compiling against it (#3879) 2023-11-02 08:50:16 +02:00
chat-parser-xml-toolcall.cpp Allow arbitrary arguments order for Q3C, Q3CN, and Qwen3.5 (#1352) 2026-03-03 15:39:16 +01:00
chat-parser-xml-toolcall.h Allow arbitrary arguments order for Q3C, Q3CN, and Qwen3.5 (#1352) 2026-03-03 15:39:16 +01:00
chat-parser.cpp Add chat parser for MiroThinker (#1138) 2026-01-13 08:07:12 +02:00
chat-parser.h Refactor chat and server file (#1062) 2025-12-15 08:27:20 +01:00
chat.cpp Allow arbitrary arguments order for Q3C, Q3CN, and Qwen3.5 (#1352) 2026-03-03 15:39:16 +01:00
chat.h Add chat parser for MiroThinker (#1138) 2026-01-13 08:07:12 +02:00
common.cpp Fix clang warnings on macOS (#1354) 2026-03-03 16:27:16 +01:00
common.h server: add checkpoint tolerance and fix grammar_trigger init (#1346) 2026-03-02 07:45:32 +01:00
console.cpp check C++ code with -Wmissing-declarations (#3184) 2023-09-15 15:38:27 -04:00
console.h gguf : new file format with flexible meta data (beta) (#2398) 2023-08-21 23:07:43 +03:00
json-partial.cpp common: Generalized XML-style tool-call parsing with streaming support (#958) 2025-11-18 15:29:58 +01:00
json-partial.h Move minja and nlohmann/json to vendor (#802) 2025-09-27 09:12:35 +02:00
json-schema-to-grammar.cpp Update grammar (#1023) 2025-11-30 18:45:38 +01:00
json-schema-to-grammar.h common: Generalized XML-style tool-call parsing with streaming support (#958) 2025-11-18 15:29:58 +01:00
llguidance.cpp Tool calls support from mainline (#723) 2025-09-01 08:38:49 +03:00
log.cpp Refactor chat and server file (#1062) 2025-12-15 08:27:20 +01:00
log.h Server: refactor and rename functions (#1151) 2026-01-18 08:16:57 +02:00
ngram-cache.cpp spec : add self speculative decoding, ngram and refactor (#1261) 2026-02-13 19:04:55 +01:00
ngram-cache.h spec : add self speculative decoding, ngram and refactor (#1261) 2026-02-13 19:04:55 +01:00
ngram-map.cpp spec : add self speculative decoding, ngram and refactor (#1261) 2026-02-13 19:04:55 +01:00
ngram-map.h spec : add self speculative decoding, ngram and refactor (#1261) 2026-02-13 19:04:55 +01:00
ngram-mod.cpp spec : add self speculative decoding, ngram and refactor (#1261) 2026-02-13 19:04:55 +01:00
ngram-mod.h spec : add self speculative decoding, ngram and refactor (#1261) 2026-02-13 19:04:55 +01:00
regex-partial.cpp llama : add token matching support to llama-grammar (#1220) 2026-02-03 07:57:17 +02:00
regex-partial.h Tool calls support from mainline (#723) 2025-09-01 08:38:49 +03:00
sampling.cpp fix adaptive p sampler rewinding too far back (#1359) 2026-03-04 13:26:25 +01:00
sampling.h fix adaptive p sampler rewinding too far back (#1359) 2026-03-04 13:26:25 +01:00
speculative.cpp Add MTP decoding support for GLM-4.x MoE (#1270) 2026-02-22 18:14:39 +01:00
speculative.h Add MTP decoding support for GLM-4.x MoE (#1270) 2026-02-22 18:14:39 +01:00
train.cpp Server: refactor and rename functions (#1151) 2026-01-18 08:16:57 +02:00
train.h sync : ggml (backend v2) (#3912) 2023-11-13 14:16:23 +02:00