mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-08-29 08:31:18 +02:00
39eab74a05
* add a direct size condition for `large` weights; the original dimension condition is insufficient -- q6_K lm_head for gemma-4 E2B has [1536, 262144], which is big enough to slowdown gemv_noshuffle but does not satisfy the dimension condition (ne0 >= 2048)