mirror of
https://github.com/LostRuins/koboldcpp.git
synced 2026-08-07 15:35:42 +00:00
* llama : load MTP tensors only if they are really used * llama : skip loading MTP (if not used) in remaining models that support MTP --------- Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com> |
||
|---|---|---|
| .. | ||
| llama-cpp.h | ||
| llama.h | ||