mirror of
https://github.com/LostRuins/koboldcpp.git
synced 2026-09-10 06:49:12 +02:00
116efee0ee
llama: enable K-shift for quantized KV cache It will fail on unsupported backends or quant types.