mirror of
https://github.com/LostRuins/koboldcpp.git
synced 2026-09-02 02:51:17 +02:00
3aec5ed0fd
revert https://github.com/ggml-org/llama.cpp/pull/16715 (+2 squashed commit) Squashed commit: [289af2ee2] Revert "Hide latency of bias and gate-loading (#16847)" This reverts commit8b11deea46. [a3e5c1e95] Revert "CUDA: add unused vars to mmvf and mmvq (#16807)" This reverts commit463bbf20bf.