koboldcpp/docs
Gaurav Garg 2bb9bddafa
spec: Add benchmark-only synthetic speculative acceptance options (#27711)
* Add benchmark-only synthetic speculative acceptance to llama-server and llama-cli

* Address review comments

* Address review comments

* Add some comments in the code
2026-08-27 13:53:42 +03:00
..
android
backend hexagon: support for multi-NPU devices (IQ9, IQ10) and fully asynchronous backend (#26501) 2026-08-26 18:46:50 -07:00
development model: use ggml_rope_set_offset() (#27382) 2026-08-21 18:54:29 +02:00
multimodal
ops Implemented vulkan cross_entropy_loss and cross_entropy_loss_back (#27216) 2026-08-26 16:49:32 +02:00
android.md docs/android.md: Add dependency libandroid-spawn for building in termux (#21812) 2026-06-22 05:48:31 +02:00
autoparser.md test: move tools/parser to tests (#27548) 2026-08-23 18:38:51 +02:00
build-riscv64-spacemit.md
build-s390x.md
build.md docs: improve Windows build instructions (#27381) 2026-08-21 21:49:27 +03:00
completions.md readme : refresh (#26280) 2026-07-30 16:14:37 +03:00
docker.md sycl : Improve SYCL doc (#23025) 2026-06-04 08:02:54 +03:00
function-calling.md docs: Update documentation with Granite 4.0/4.1 (#23404) 2026-05-22 20:35:46 +08:00
install.md docs: Adapt conda-forge package name (#26229) 2026-07-28 16:51:20 +02:00
llguidance.md
models.md readme : refresh (#26280) 2026-07-30 16:14:37 +03:00
multi-gpu.md Write a readme on Multi-GPU usage in llama.cpp (#22729) 2026-05-07 17:48:40 +02:00
multimodal.md
ops.md Implemented vulkan cross_entropy_loss and cross_entropy_loss_back (#27216) 2026-08-26 16:49:32 +02:00
preset.md common: add system-level config file (#26118) 2026-08-13 00:02:27 +02:00
release.md ci : allow make-release to target a specific commit (#27234) 2026-08-17 11:52:46 +03:00
speculative.md spec: Add benchmark-only synthetic speculative acceptance options (#27711) 2026-08-27 13:53:42 +03:00
xcframework.md readme : refresh (#26280) 2026-07-30 16:14:37 +03:00