Files
llama.cpp/src
Ruixiang WangandGeorgi Gerganov 571d0d540d model: rotate injected K/V cache for DFlash (#25823)
* dflash: rotate injected K/V cache when using K/V quantization

* Update src/models/dflash.cpp

Co-authored-by: Georgi Gerganov <[email protected]>

* clearer format

* remove trailing whitespace

---------

Co-authored-by: Georgi Gerganov <[email protected]>
2026-07-18 15:02:18 +02:00
..
2026-06-29 16:58:51 +08:00
2026-06-29 16:58:51 +08:00
2026-06-29 16:58:51 +08:00
2026-06-29 16:58:51 +08:00
2026-06-07 20:50:54 +08:00
2026-04-03 10:33:03 +02:00