| .. |
|
bench
|
Merge branch 'upstream' into concedo_experimental
|
2026-06-16 17:55:04 +08:00 |
|
tests
|
Merge commit 'ae9afff8d2' into concedo_experimental
|
2026-09-13 23:53:41 +08:00 |
|
main.cpp
|
app : introduce the llama unified executable (#23296)
|
2026-05-20 13:22:22 +02:00 |
|
README-dev.md
|
server: refactor sleep handling, allow access /metrics during sleep (#27376)
|
2026-08-19 20:48:09 +02:00 |
|
server-chat.cpp
|
common: add json.h abstraction (#27511)
|
2026-08-22 16:28:28 +02:00 |
|
server-chat.h
|
common: add json.h abstraction (#27511)
|
2026-08-22 16:28:28 +02:00 |
|
server-common.cpp
|
common : implement common_schema internal representation for JSON schemas (#28736)
|
2026-09-12 16:14:50 -05:00 |
|
server-common.h
|
server: refactor subproc handling (#28555)
|
2026-09-12 00:53:07 +02:00 |
|
server-context.cpp
|
server: fix speculation after an image (#28715)
|
2026-09-11 11:33:26 +03:00 |
|
server-context.h
|
args: add --video-* CLI arguments (#24318)
|
2026-08-27 12:11:12 +02:00 |
|
server-cors-proxy.h
|
common,server: handle bracketed IPv6 literals in URL authority (#25140)
|
2026-06-30 16:16:44 +02:00 |
|
server-http.cpp
|
server : make models endpoints private when authentication is enabled (#26347)
|
2026-08-19 20:44:42 +02:00 |
|
server-http.h
|
server: refactor server_stream (#25541)
|
2026-07-11 12:41:47 +02:00 |
|
server-mcp.cpp
|
common: add subproc.h wrapper, disabled on android/ios (#26102)
|
2026-07-26 20:54:25 +02:00 |
|
server-mcp.h
|
server: support MCP stdio (#26062)
|
2026-07-26 01:08:49 +02:00 |
|
server-models.cpp
|
server : allow model downloads at model limit fix issue #26809 (#28530)
|
2026-09-12 11:50:35 +02:00 |
|
server-models.h
|
server: refactor subproc handling (#28555)
|
2026-09-12 00:53:07 +02:00 |
|
server-queue.cpp
|
server: refactor sleep handling, allow access /metrics during sleep (#27376)
|
2026-08-19 20:48:09 +02:00 |
|
server-queue.h
|
server: refactor sleep handling, allow access /metrics during sleep (#27376)
|
2026-08-19 20:48:09 +02:00 |
|
server-schema.cpp
|
common : implement common_schema internal representation for JSON schemas (#28736)
|
2026-09-12 16:14:50 -05:00 |
|
server-schema.h
|
sampler : remove "full-context windows" from history-based samplers (#26524)
|
2026-08-04 21:28:55 +03:00 |
|
server-stream.cpp
|
server + ui: fix stream routes for model names containing a slash (#26137)
|
2026-07-27 07:34:47 +02:00 |
|
server-stream.h
|
server + ui: fix stream routes for model names containing a slash (#26137)
|
2026-07-27 07:34:47 +02:00 |
|
server-task.cpp
|
common: add json.h abstraction (#27511)
|
2026-08-22 16:28:28 +02:00 |
|
server-task.h
|
common: add json.h abstraction (#27511)
|
2026-08-22 16:28:28 +02:00 |
|
server-tools.cpp
|
common: add json.h abstraction (#27511)
|
2026-08-22 16:28:28 +02:00 |
|
server-tools.h
|
ui: Refactor Built-In Tools naming (Server/Browser) (#27271)
|
2026-08-17 22:23:22 +02:00 |
|
server.cpp
|
server: add ctx-per-slot (--kv-unified-per-slot) (#24124)
|
2026-08-27 22:39:14 +02:00 |
|
ui.cpp
|
can build llama server now
|
2026-06-14 15:06:37 +08:00 |
|
ui.h
|
can build llama server now
|
2026-06-14 15:06:37 +08:00 |