koboldcpp/tools
Xuan-Son Nguyen 56db501e73
mtmd: use pillow-accurate algo, correct resize_algo for all models (#27594)
* mtmd: use pillow-accurate resize algo, correct resize_algo for all models

* speed optimization
2026-08-23 18:35:41 +02:00
..
batched-bench cmake : add install() for impl libraries + fix apple builds (#23511) 2026-05-22 11:46:26 +03:00
cli common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
completion common: migrate the deprecated --mmap/--no-mmap to --load-mode (#26934) 2026-08-15 16:35:53 +08:00
cvector-generator cmake : introduce semantic versioning (#26839) 2026-08-12 14:15:03 +02:00
export-lora docs: fix export-lora --lora-scaled syntax [no release] (#24703) 2026-06-18 16:46:17 +02:00
fit-params fit: also take into account n_streams (#27496) 2026-08-22 16:16:06 +02:00
gguf-split cmake : introduce semantic versioning (#26839) 2026-08-12 14:15:03 +02:00
imatrix imatrix.cpp: Move finite check and only check touched experts (#26861) 2026-08-11 11:18:19 -04:00
llama-bench fit: also take into account n_streams (#27496) 2026-08-22 16:16:06 +02:00
mtmd mtmd: use pillow-accurate algo, correct resize_algo for all models (#27594) 2026-08-23 18:35:41 +02:00
parser common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
perplexity quant : Optimise memory usage by evicting weights after processing each layer (#22877) 2026-08-18 16:22:32 +02:00
quantize cmake : introduce semantic versioning (#26839) 2026-08-12 14:15:03 +02:00
results
rpc binaries : Improve rpc-server and export-graph-ops names. (#25045) 2026-06-27 10:31:29 +03:00
server server : add LLAMA_SERVER_SLOTS_N_DIFF (#27600) 2026-08-23 15:55:51 +03:00
tokenize tokenize : drop --stdin mutual-exclusion check (#25672) 2026-07-15 18:41:51 +02:00
tts mtmd: add --mmproj-device argument (#23255) 2026-08-20 18:45:37 +02:00
ui ui: Chat Conversation Tabbed navigation (#27263) 2026-08-23 10:46:49 +02:00
CMakeLists.txt cmake: skip cvector-generator and export-lora when CPU backend is disabled (#24053) 2026-06-04 13:13:19 +03:00