koboldcpp/tools/server
Concedo 4447017602 Merge branch 'upstream' into concedo_experimental
# Conflicts:
#	.github/actions/ccache-clear/action.yml
#	.github/workflows/build-apple.yml
#	.github/workflows/build-cpu.yml
#	.github/workflows/build-cuda-ubuntu.yml
#	.github/workflows/build-opencl.yml
#	.github/workflows/build-openvino.yml
#	.github/workflows/build-sycl.yml
#	.github/workflows/build-vulkan.yml
#	.github/workflows/build-wasm.yml
#	.github/workflows/build-webgpu.yml
#	.github/workflows/hip-quality-check.yml
#	.github/workflows/server.yml
#	CONTRIBUTING.md
#	README.md
#	ci/run.sh
#	common/CMakeLists.txt
#	common/chat.cpp
#	docs/autoparser.md
#	ggml/src/ggml-webgpu/wgsl-shaders/flash_attn.wgsl
#	ggml/src/ggml-webgpu/wgsl-shaders/flash_attn_tile.wgsl
#	ggml/src/ggml-webgpu/wgsl-shaders/flash_attn_vec_split.wgsl
#	scripts/sync_vendor.py
#	tests/CMakeLists.txt
#	tests/peg-parser/test-json-serialization.cpp
#	tests/peg-parser/tests.h
#	tests/test-chat-auto-parser.cpp
#	tests/test-chat-peg-parser.cpp
#	tests/test-chat-template.cpp
#	tests/test-chat.cpp
#	tests/test-grammar-integration.cpp
#	tests/test-jinja.cpp
#	tests/test-json-schema-to-grammar.cpp
#	tests/test-llama-archs.cpp
#	tests/test-model-resolution.cpp
#	tests/test-recurrent-state-rollback.cpp
#	tools/CMakeLists.txt
2026-08-25 20:44:34 +08:00
..
bench Merge branch 'upstream' into concedo_experimental 2026-06-16 17:55:04 +08:00
tests Merge branch 'upstream' into concedo_experimental 2026-08-25 20:44:34 +08:00
main.cpp app : introduce the llama unified executable (#23296) 2026-05-20 13:22:22 +02:00
README-dev.md server: refactor sleep handling, allow access /metrics during sleep (#27376) 2026-08-19 20:48:09 +02:00
server-chat.cpp common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-chat.h common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-common.cpp common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-common.h common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-context.cpp server : add LLAMA_SERVER_SLOTS_N_DIFF (#27600) 2026-08-23 15:55:51 +03:00
server-context.h common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-cors-proxy.h common,server: handle bracketed IPv6 literals in URL authority (#25140) 2026-06-30 16:16:44 +02:00
server-http.cpp server : make models endpoints private when authentication is enabled (#26347) 2026-08-19 20:44:42 +02:00
server-http.h server: refactor server_stream (#25541) 2026-07-11 12:41:47 +02:00
server-mcp.cpp common: add subproc.h wrapper, disabled on android/ios (#26102) 2026-07-26 20:54:25 +02:00
server-mcp.h server: support MCP stdio (#26062) 2026-07-26 01:08:49 +02:00
server-models.cpp common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-models.h server: (router) lazy-load startup_models after main setup (#27424) 2026-08-20 15:22:16 +02:00
server-queue.cpp server: refactor sleep handling, allow access /metrics during sleep (#27376) 2026-08-19 20:48:09 +02:00
server-queue.h server: refactor sleep handling, allow access /metrics during sleep (#27376) 2026-08-19 20:48:09 +02:00
server-schema.cpp common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-schema.h sampler : remove "full-context windows" from history-based samplers (#26524) 2026-08-04 21:28:55 +03:00
server-stream.cpp server + ui: fix stream routes for model names containing a slash (#26137) 2026-07-27 07:34:47 +02:00
server-stream.h server + ui: fix stream routes for model names containing a slash (#26137) 2026-07-27 07:34:47 +02:00
server-task.cpp common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-task.h common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-tools.cpp common: add json.h abstraction (#27511) 2026-08-22 16:28:28 +02:00
server-tools.h ui: Refactor Built-In Tools naming (Server/Browser) (#27271) 2026-08-17 22:23:22 +02:00
server.cpp server: (router) lazy-load startup_models after main setup (#27424) 2026-08-20 15:22:16 +02:00
ui.cpp can build llama server now 2026-06-14 15:06:37 +08:00
ui.h can build llama server now 2026-06-14 15:06:37 +08:00