Logo
Explore Help
Sign In
Superminaren/llama.cpp
Watch 1
Star 0
Fork 0
mirror of https://github.com/ggml-org/llama.cpp.git synced 2026-09-04 10:47:38 +02:00
Code Issues Packages Projects Releases Wiki Activity
Files
b8552
llama.cpp/tools
T
History
mtmcp 37f230dd7c completion : session_tokens insert range in completion tool (no-op → correct) (#20917)
The embd.begin(), embd.begin() range is empty and inserts nothing, so session_tokens never gets updated after
  decoding. Should be embd.begin(), embd.end(). Introduced in commit 2b6dfe8.
2026-03-27 09:25:58 +01:00
..
batched-bench
…
cli
docs : rerun llama-gen-docs to include new CLI args (#20892)
2026-03-23 12:33:38 +01:00
completion
completion : session_tokens insert range in completion tool (no-op → correct) (#20917)
2026-03-27 09:25:58 +01:00
cvector-generator
…
export-lora
…
fit-params
…
gguf-split
gguf-split : clarify operation of gguf-split (#19749)
2026-03-25 13:12:50 +02:00
imatrix
imatrix : fix crash when using --show-statistics with zero counts (#19532)
2026-03-26 08:14:36 +01:00
llama-bench
llama-bench: print -n-cpu-moe when offloaded layers > 1 (#20984)
2026-03-25 21:17:27 +08:00
mtmd
mtmd: refactor image preprocessing (#21031)
2026-03-26 19:49:20 +01:00
parser
common/parser: add proper reasoning tag prefill reading (#20424)
2026-03-19 16:58:21 +01:00
perplexity
tools : enable kvu in perplexity for hellaswag, winogrande, multiple-choice (#19954)
2026-03-13 21:25:57 +01:00
quantize
…
results
…
rpc
…
server
Send reasoning content back to the model across turns via the reasoning_content API field (#21036)
2026-03-27 08:17:35 +01:00
tokenize
…
tts
…
CMakeLists.txt
…
Powered by Gitea Version: 1.27.2 Page: 2164ms Template: 43ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API