Files
llama.cpp/src
Aman Gupta b0539c43ed DeepseekV4: fix rollback with multi-seq (#26756)
* DeepseekV4: fix rollback with multi-seq

* fix model loading

* make pending rollback single use

* only clear cache for seq_id for full load

* add assert for compress ratio

* make graph topology static

* pass true instead of flags in clear_compressed

* cont : clean-up + TODOs

---------

Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
2026-08-23 13:57:49 +03:00
..
2026-08-21 19:52:34 +02:00
2026-08-21 19:52:34 +02:00
2026-08-21 19:52:34 +02:00
2026-08-21 19:52:34 +02:00
2026-08-21 19:52:34 +02:00
2026-06-29 16:58:51 +08:00
2026-06-07 20:50:54 +08:00
2026-08-21 19:52:34 +02:00
2026-04-03 10:33:03 +02:00