| .. |
|
batched-bench
|
cmake : add install() for impl libraries + fix apple builds (#23511)
|
2026-05-22 11:46:26 +03:00 |
|
cli
|
common: add json.h abstraction (#27511)
|
2026-08-22 16:28:28 +02:00 |
|
completion
|
common: migrate the deprecated --mmap/--no-mmap to --load-mode (#26934)
|
2026-08-15 16:35:53 +08:00 |
|
cvector-generator
|
cmake : introduce semantic versioning (#26839)
|
2026-08-12 14:15:03 +02:00 |
|
export-lora
|
docs: fix export-lora --lora-scaled syntax [no release] (#24703)
|
2026-06-18 16:46:17 +02:00 |
|
fit-params
|
fit: also take into account n_streams (#27496)
|
2026-08-22 16:16:06 +02:00 |
|
gguf-split
|
cmake : introduce semantic versioning (#26839)
|
2026-08-12 14:15:03 +02:00 |
|
imatrix
|
imatrix.cpp: Move finite check and only check touched experts (#26861)
|
2026-08-11 11:18:19 -04:00 |
|
llama-bench
|
fit: also take into account n_streams (#27496)
|
2026-08-22 16:16:06 +02:00 |
|
mtmd
|
mtmd: use pillow-accurate algo, correct resize_algo for all models (#27594)
|
2026-08-23 18:35:41 +02:00 |
|
parser
|
common: add json.h abstraction (#27511)
|
2026-08-22 16:28:28 +02:00 |
|
perplexity
|
quant : Optimise memory usage by evicting weights after processing each layer (#22877)
|
2026-08-18 16:22:32 +02:00 |
|
quantize
|
cmake : introduce semantic versioning (#26839)
|
2026-08-12 14:15:03 +02:00 |
|
results
|
|
|
|
rpc
|
binaries : Improve rpc-server and export-graph-ops names. (#25045)
|
2026-06-27 10:31:29 +03:00 |
|
server
|
server : add LLAMA_SERVER_SLOTS_N_DIFF (#27600)
|
2026-08-23 15:55:51 +03:00 |
|
tokenize
|
tokenize : drop --stdin mutual-exclusion check (#25672)
|
2026-07-15 18:41:51 +02:00 |
|
tts
|
mtmd: add --mmproj-device argument (#23255)
|
2026-08-20 18:45:37 +02:00 |
|
ui
|
ui: Chat Conversation Tabbed navigation (#27263)
|
2026-08-23 10:46:49 +02:00 |
|
CMakeLists.txt
|
cmake: skip cvector-generator and export-lora when CPU backend is disabled (#24053)
|
2026-06-04 13:13:19 +03:00 |