This website requires JavaScript.
Explore
Help
Sign In
Superminaren
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-07 20:47:30 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
c0c7fa930d67c5ca63c22449b4790dcd15ae62dd
llama.cpp
/
include
T
History
Xuan Son Nguyen
c0c7fa930d
quantize: cap working memory size to avoid loading big tensors onto RAM
2026-08-27 13:48:24 +02:00
..
llama-cpp.h
llama : re-enable manual LoRA adapter free (
#19983
)
2026-03-18 12:03:26 +02:00
llama.h
quantize: cap working memory size to avoid loading big tensors onto RAM
2026-08-27 13:48:24 +02:00