mirror of
https://github.com/LostRuins/koboldcpp.git
synced 2026-08-27 17:29:52 +00:00
* DeepseekV4: fix rollback with multi-seq * fix model loading * make pending rollback single use * only clear cache for seq_id for full load * add assert for compress ratio * make graph topology static * pass true instead of flags in clear_compressed * cont : clean-up + TODOs --------- Co-authored-by: Georgi Gerganov <ggerganov@gmail.com> |
||
|---|---|---|
| .. | ||
| llama-cpp.h | ||
| llama.h | ||