Logo
Explore Help
Sign In
Superminaren/llama.cpp
Watch 1
Star 0
Fork 0
mirror of https://github.com/ggml-org/llama.cpp.git synced 2026-09-04 10:47:38 +02:00
Code Issues Packages Projects Releases Wiki Activity
Files
b9549
llama.cpp/ggml/include
T
History
Johannes Gäßler 8e6fff84de TP: quantized KV cache support (#23792)
* TP: quantized KV cache support

* fix partial view

* remove overly strict assert
2026-06-01 12:30:10 +02:00
..
ggml-alloc.h
TP: fix entirely zero-sized slices per device (#23525)
2026-05-24 08:19:33 +02:00
ggml-backend.h
TP: quantized KV cache support (#23792)
2026-06-01 12:30:10 +02:00
ggml-blas.h
…
ggml-cann.h
…
ggml-cpp.h
…
ggml-cpu.h
…
ggml-cuda.h
ggml: backend-agnostic tensor parallelism (experimental) (#19378)
2026-04-09 16:42:19 +02:00
ggml-hexagon.h
…
ggml-metal.h
…
ggml-opencl.h
…
ggml-openvino.h
ggml : add OpenVINO backend (#15307)
2026-03-14 07:56:55 +02:00
ggml-opt.h
chore : correct typos [no ci] (#20041)
2026-03-05 08:50:21 +01:00
ggml-rpc.h
rpc : add native RDMA transport for RPC backend (RoCEv2) (#20590)
2026-04-15 16:44:02 +03:00
ggml-sycl.h
…
ggml-virtgpu.h
ggml-virtgpu: make the code thread safe (#19204)
2026-02-04 10:46:18 +08:00
ggml-vulkan.h
…
ggml-webgpu.h
…
ggml-zdnn.h
…
ggml-zendnn.h
…
ggml.h
ggml.h: correct ggml_silu_back arg docstring (a=dy, b=x) (ggml/1500)
2026-05-25 12:38:01 +03:00
gguf.h
ggml: gguf_init_from_callback and gguf_init_from_buffer (#22341)
2026-05-25 11:33:29 +02:00
Powered by Gitea Version: 1.27.2 Page: 1649ms Template: 16ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API