..
batched-bench
cmake : add install() for impl libraries + fix apple builds ( #23511 )
2026-05-22 11:46:26 +03:00
cli
cli: fix model params not propagated ( #23893 )
2026-06-05 17:29:41 +02:00
completion
completion : remove useless statics ( #24226 )
2026-06-06 12:16:16 +02:00
cvector-generator
libs : rename libcommon -> libllama-common ( #21936 )
2026-04-17 11:11:46 +03:00
export-lora
libs : rename libcommon -> libllama-common ( #21936 )
2026-04-17 11:11:46 +03:00
fit-params
cmake : add install() for impl libraries + fix apple builds ( #23511 )
2026-05-22 11:46:26 +03:00
gguf-split
libs : rename libcommon -> libllama-common ( #21936 )
2026-04-17 11:11:46 +03:00
imatrix
Move duplicated imatrix code into single common imatrix-loader.cpp ( #22445 )
2026-06-04 17:45:40 +02:00
llama-bench
Support -fa auto in llama-bench ( #23714 )
2026-05-31 02:03:57 +05:30
mtmd
mtmd: support "frame merge" for qwen-vl-based models ( #21858 )
2026-06-06 21:17:25 +02:00
parser
libs : rename libcommon -> libllama-common ( #21936 )
2026-04-17 11:11:46 +03:00
perplexity
perplexity : fix format specifier in LOG_ERR ( #23788 )
2026-05-28 10:34:58 +03:00
quantize
docs: Update quantization readme ( #24133 )
2026-06-05 12:21:26 +02:00
results
libs : rename libcommon -> libllama-common ( #21936 )
2026-04-17 11:11:46 +03:00
rpc
fix: rpc-server cache may not work in Windows environments ( #22394 )
2026-04-27 17:25:09 +03:00
server
mtmd, server: add "placeholder bitmap" for counting tokens , add */input_tokens API ( #23913 )
2026-06-06 11:06:51 +02:00
tokenize
libs : rename libcommon -> libllama-common ( #21936 )
2026-04-17 11:11:46 +03:00
tts
logs : reduce ( #23021 )
2026-05-14 13:05:52 +03:00
ui
ui: add ignore-scripts=true to npmrc ( #24149 )
2026-06-05 14:31:03 +02:00
CMakeLists.txt
cmake: skip cvector-generator and export-lora when CPU backend is disabled ( #24053 )
2026-06-04 13:13:19 +03:00