This website requires JavaScript.
Explore
Help
Sign In
Superminaren
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-04 10:47:38 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
b9733
llama.cpp
/
ggml
T
History
Masashi Yoshimura
f449e05537
ggml-webgpu: add adapter toggles for F16 on Vulkan + NVIDIA
2026-06-20 08:12:32 +09:00
..
cmake
ggml : Parallelize quant LUT init (
#23595
)
2026-05-25 10:15:46 +03:00
include
Remove padding and multiple D2D copies for MTP (
#24086
)
2026-06-10 23:21:16 +05:30
src
ggml-webgpu: add adapter toggles for F16 on Vulkan + NVIDIA
2026-06-20 08:12:32 +09:00
.gitignore
…
CMakeLists.txt
ggml : bump version to 0.15.2 (ggml/1548)
2026-06-19 10:19:14 +03:00