mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-09-27 21:46:57 +02:00
Recreated from #24546 --------- Co-authored-by: Carl Philipp Klemm <[email protected]> * CUDA: pick MMQ tile size against ncols_opt set on the host side Assisted-by: Claude Fable 5.1 Claude-Session: https://claude.ai/code/session_011SYPfRhKoUpU3gMsGxq6go --------- Co-authored-by: ravel7524 <[email protected]> Co-authored-by: Carl Philipp Klemm <[email protected]>