Logo
Explore Help
Sign In
Superminaren/llama.cpp
Watch 1
Star 0
Fork 0
mirror of https://github.com/ggml-org/llama.cpp.git synced 2026-09-07 20:47:30 +02:00
Code Issues Packages Projects Releases Wiki Activity
5,771 Commits 780 Branches 7,352 Tags
118d52fefcd677d3e8bc94dc11735ce2990ad7be
Commit Graph
7 Commits
Author SHA1 Message Date
Francis Couture-Harpin 118d52fefc Merge branch 'master' into compilade/imatrix-batched-chunks 2025-06-23 12:54:56 -04:00
Francis Couture-Harpin 0e79355075 quantize : fix dataset name loading from gguf imatrix 2025-06-23 12:43:25 -04:00
Francis Couture-Harpin 43cd2b3eb5 imatrix : support 3d tensors with MUL_MAT 2025-06-23 12:20:55 -04:00
Ed Addario fa4a9f2a1c quantize : handle user-defined pruning of whole layers (blocks) (#13037) 2025-06-22 23:16:26 +02:00
Francis Couture-Harpin 2c0945027a Merge branch 'master' into compilade/imatrix-batched-chunks 2025-06-18 16:32:35 -04:00
Ed Addario e5c834f718 quantize : improve tensor-type pattern matching (#13033) 2025-05-13 19:12:31 +02:00
Diego DevesaandXuan Son Nguyen 1d36b3670b llama : move end-user examples to tools directory (#13249)
* llama : move end-user examples to tools directory

---------

Co-authored-by: Xuan Son Nguyen <[email protected]>
2025-05-02 20:27:13 +02:00
Powered by Gitea Version: 1.27.2 Page: 470ms Template: 14ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API