koboldcpp/tools
Concedo ccaed41be0 Merge branch 'upstream' into concedo_experimental
# Conflicts:
#	.github/workflows/build-openvino.yml
#	.github/workflows/server-sanitize.yml
#	AUTHORS
#	README.md
#	app/llama.cpp
#	ci/run.sh
#	common/build-info.h
#	docs/backend/SYCL.md
#	docs/build.md
#	docs/ops.md
#	ggml/CMakeLists.txt
#	ggml/src/ggml-cuda/CMakeLists.txt
#	ggml/src/ggml-musa/CMakeLists.txt
#	ggml/src/ggml-opencl/CMakeLists.txt
#	ggml/src/ggml-opencl/ggml-opencl.cpp
#	ggml/src/ggml-opencl/kernels/cvt.cl
#	ggml/src/ggml-opencl/kernels/gemm_noshuffle_q4_k_f32.cl
#	ggml/src/ggml-opencl/kernels/gemm_noshuffle_q6_k_f32.cl
#	ggml/src/ggml-opencl/kernels/gemv_noshuffle_q4_0_f32.cl
#	ggml/src/ggml-opencl/kernels/gemv_noshuffle_q4_1_f32.cl
#	ggml/src/ggml-opencl/kernels/gemv_noshuffle_q4_k_f32.cl
#	ggml/src/ggml-opencl/kernels/gemv_noshuffle_q5_k_f32.cl
#	ggml/src/ggml-opencl/kernels/gemv_noshuffle_q6_k_f32.cl
#	ggml/src/ggml-opencl/kernels/gemv_noshuffle_q8_0_f32.cl
#	ggml/src/ggml-opencl/kernels/mul_mm_f32_f32_l4_lm.cl
#	ggml/src/ggml-opencl/kernels/rms_norm.cl
#	ggml/src/ggml-sycl/binbcast.cpp
#	ggml/src/ggml-sycl/binbcast.hpp
#	ggml/src/ggml-sycl/common.hpp
#	ggml/src/ggml-sycl/fattn.cpp
#	ggml/src/ggml-sycl/fusion.cpp
#	ggml/src/ggml-sycl/ggml-sycl.cpp
#	ggml/src/ggml-sycl/norm.cpp
#	ggml/src/ggml-sycl/norm.hpp
#	scripts/snapdragon/build.py
#	scripts/snapdragon/qdc/run_qdc_jobs.py
#	scripts/snapdragon/qdc/tests/linux/run_linux.sh
#	scripts/snapdragon/qdc/tests/run_backend_ops_posix.py
#	scripts/snapdragon/qdc/tests/run_bench_tests_posix.py
#	scripts/snapdragon/qdc/tests/utils.py
#	scripts/snapdragon/run.py
#	src/CMakeLists.txt
#	src/llama.cpp
#	tests/test-backend-ops.cpp
#	tests/test-json-schema-to-grammar.cpp
2026-09-04 16:59:17 +08:00
..
batched-bench
cli common: rename --tensor-read-lazy to --lazy-mode, add -lzm shorthand (#27969) 2026-08-30 09:18:10 +03:00
completion common: rename --tensor-read-lazy to --lazy-mode, add -lzm shorthand (#27969) 2026-08-30 09:18:10 +03:00
fit-params fit: also take into account n_streams (#27496) 2026-08-22 16:16:06 +02:00
gguf-split Merge branch 'upstream' into concedo_experimental 2026-08-14 18:10:18 +08:00
llama-bench common: rename --tensor-read-lazy to --lazy-mode, add -lzm shorthand (#27969) 2026-08-30 09:18:10 +03:00
mtmd Merge branch 'upstream' into concedo_experimental 2026-09-04 16:59:17 +08:00
perplexity quant : Optimise memory usage by evicting weights after processing each layer (#22877) 2026-08-18 16:22:32 +02:00
quantize Merge branch 'upstream' into concedo_experimental 2026-08-28 22:29:03 +08:00
rpc rpc: avoid serializing buffers from other servers (#26500) 2026-08-30 20:26:16 +03:00
server Merge branch 'upstream' into concedo_experimental 2026-09-04 16:59:17 +08:00
tts Merge commit 'deae5ee133' into concedo_experimental 2026-08-28 19:31:29 +08:00
tuning Merge commit '3737e41370' into concedo_experimental 2026-08-28 17:23:25 +08:00
ui Merge commit '2a74817f93' into concedo_experimental 2026-09-02 23:13:14 +08:00
kcpplauncherhook.py