Files
llama.cpp/ggml/src
Colin Kealty 212152923a Batched gemm for grid IQ quants
Style updates and a bit more performance

Clean up comments

Move code around

Vectorize IQ panel decode, lower threshold for speedup

IQ panel: single-source gather layout, gate bias, vectorize interleave

Add ggml_gemm_iqp_8x8_q8_K_p4 kernel, remove gather buffer

Move IQ panel code out of repack into iqp.cpp, clean up comments

Another comment sweep
2026-08-30 08:10:13 -04:00
..
2026-08-30 08:10:13 -04:00
2026-08-24 10:43:04 +03:00
2026-08-30 08:10:13 -04:00
2026-04-16 17:21:28 +08:00
2026-08-24 10:43:04 +03:00