This website requires JavaScript.
Explore
Help
Sign In
Superminaren
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-24 13:37:01 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
27aef3dd91e7cde049e7c242dbf6c8fe86574d01
llama.cpp
/
ggml
T
History
Rithik Sharma
45155597aa
add fast matmul iquants (
#22504
)
2026-04-29 22:58:32 -07:00
..
cmake
ggml: backend-agnostic tensor parallelism (experimental) (
#19378
)
2026-04-09 16:42:19 +02:00
include
CUDA: manage NCCL communicators in context (
#21891
)
2026-04-15 15:58:40 +02:00
src
add fast matmul iquants (
#22504
)
2026-04-29 22:58:32 -07:00
.gitignore
…
CMakeLists.txt
ggml : bump version to 0.10.1 (ggml/1469)
2026-04-29 16:43:47 +03:00