This website requires JavaScript.
Explore
Help
Sign In
Superminaren
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-04 10:47:38 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
b5354
llama.cpp
/
tools
T
History
City
c104023994
mtmd : Use RMS norm for InternVL 3 38B and 78B mmproj (
#13459
)
2025-05-12 00:39:06 +02:00
..
batched-bench
…
cvector-generator
…
export-lora
…
gguf-split
…
imatrix
imatrix : Add --parse-special for enabling parsing of special tokens in imatrix calculation (
#13389
)
2025-05-09 11:53:58 +02:00
llama-bench
Add
--no-op-offload
to improve
-ot
pp perf in MoE models like llama4 400B (
#13386
)
2025-05-11 14:18:39 +02:00
main
llama : do not crash if there is no CPU backend (
#13395
)
2025-05-09 13:02:07 +02:00
mtmd
mtmd : Use RMS norm for InternVL 3 38B and 78B mmproj (
#13459
)
2025-05-12 00:39:06 +02:00
perplexity
…
quantize
…
rpc
llama : do not crash if there is no CPU backend (
#13395
)
2025-05-09 13:02:07 +02:00
run
llama-run: add support for downloading models from ModelScope (
#13370
)
2025-05-09 10:25:50 +01:00
server
tools : fix uninitialized llama_batch in server (
#13436
)
2025-05-11 17:08:26 +02:00
tokenize
…
tts
…
CMakeLists.txt
…