..
batched-bench
cmake : add install() for impl libraries + fix apple builds ( #23511 )
2026-05-22 11:46:26 +03:00
cli
args: add env vars for temperature, top-p, min-p and penalties ( #27380 )
2026-09-21 12:47:38 +02:00
completion
args: add env vars for temperature, top-p, min-p and penalties ( #27380 )
2026-09-21 12:47:38 +02:00
cvector-generator
cmake : introduce semantic versioning ( #26839 )
2026-08-12 14:15:03 +02:00
export-lora
docs: fix export-lora --lora-scaled syntax [no release] ( #24703 )
2026-06-18 16:46:17 +02:00
fit-params
fit: also take into account n_streams ( #27496 )
2026-08-22 16:16:06 +02:00
gguf-split
cmake : introduce semantic versioning ( #26839 )
2026-08-12 14:15:03 +02:00
imatrix
imatrix.cpp: Move finite check and only check touched experts ( #26861 )
2026-08-11 11:18:19 -04:00
llama-bench
llama-bench: support --version to print build info ( #28971 )
2026-09-16 13:39:43 +08:00
mtmd
model : add Ling 3.0 VL support ( #29151 )
2026-09-24 08:57:31 +02:00
perplexity
quant : Optimise memory usage by evicting weights after processing each layer ( #22877 )
2026-08-18 16:22:32 +02:00
quantize
quantize: cap working memory size to avoid loading big tensors onto RAM ( #27795 )
2026-08-27 18:31:13 +02:00
results
libs : rename libcommon -> libllama-common ( #21936 )
2026-04-17 11:11:46 +03:00
rpc
rpc : skip ACCEL devices ( #29020 )
2026-09-17 18:36:53 +03:00
server
server,common : fix the GCC 12 stringop-overread false positive (again) ( #29325 )
2026-09-24 08:40:06 +02:00
tokenize
tokenize : drop --stdin mutual-exclusion check ( #25672 )
2026-07-15 18:41:51 +02:00
tts
args: add --video-* CLI arguments ( #24318 )
2026-08-27 12:11:12 +02:00
tuning
metal : key the fa-vec tuned table by family instead of SKU ( #29075 )
2026-09-23 19:23:15 +08:00
ui
ui : Accept WEBM video files ( #28622 )
2026-09-22 15:40:24 +02:00
CMakeLists.txt
metal : per-device tuned (Q, NE) for flash-attn vec ( #26570 )
2026-08-24 19:22:27 +03:00