This website requires JavaScript.
Explore
Help
Sign In
Superminaren
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-04 10:47:38 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
b6305
llama.cpp
/
common
T
History
Georgi Gerganov
da54f9f1a2
presets : add qwen3-30B-a3b FIM (
#15616
)
2025-08-27 15:48:07 +03:00
..
arg.cpp
presets : add qwen3-30B-a3b FIM (
#15616
)
2025-08-27 15:48:07 +03:00
arg.h
…
base64.hpp
…
build-info.cpp.in
…
chat-parser.cpp
chat : support Granite model reasoning and tool call (
#14864
)
2025-08-06 20:27:30 +02:00
chat-parser.h
…
chat.cpp
model : gpt-oss add response_format support (
#15494
)
2025-08-22 11:04:08 -05:00
chat.h
chat : include kwargs in template example (
#15309
)
2025-08-14 10:28:29 -07:00
CMakeLists.txt
…
common.cpp
llama : remove KV cache defragmentation logic (
#15473
)
2025-08-22 12:22:13 +03:00
common.h
llama : remove KV cache defragmentation logic (
#15473
)
2025-08-22 12:22:13 +03:00
console.cpp
…
console.h
…
json-partial.cpp
…
json-partial.h
…
json-schema-to-grammar.cpp
…
json-schema-to-grammar.h
…
llguidance.cpp
…
log.cpp
…
log.h
…
ngram-cache.cpp
…
ngram-cache.h
…
regex-partial.cpp
…
regex-partial.h
…
sampling.cpp
…
sampling.h
…
speculative.cpp
server : implement universal assisted decoding (
#12635
)
2025-07-31 14:25:23 +02:00
speculative.h
server : implement universal assisted decoding (
#12635
)
2025-07-31 14:25:23 +02:00