koboldcpp/tests
Yash Raj Pandey f8def7fe16
ggml : require contiguous src for ROLL on CUDA and Metal (#25928)
ggml_roll only asserts nb[0] == ggml_type_size, so a permuted src is a
valid input, but the CUDA and Metal roll kernels index by ne alone and
never read the nb strides. A non-contiguous src therefore produced
silently wrong results. Neither backend declared a contiguity
requirement in supports_op, so the scheduler did not fall back to the
CPU implementation, which does handle strides correctly.

Add the requirement to both backends, matching the existing
GGML_OP_ROPE guard, and add a permuted test_roll case.
2026-08-10 15:01:44 +03:00
..
peg-parser
snapshots tests : avoid building get-model.cpp many times (#26317) 2026-07-30 19:34:04 +03:00
.gitignore
CMakeLists.txt tests: add model resolution test on synthetic repo listings (#26172) 2026-08-03 18:58:15 +02:00
gguf-model-data.cpp
gguf-model-data.h
test-alloc.cpp
test-arg-parser.cpp sampler : remove "full-context windows" from history-based samplers (#26524) 2026-08-04 21:28:55 +03:00
test-autorelease.cpp tests : avoid building get-model.cpp many times (#26317) 2026-07-30 19:34:04 +03:00
test-backend-ops.cpp ggml : require contiguous src for ROLL on CUDA and Metal (#25928) 2026-08-10 15:01:44 +03:00
test-backend-sampler.cpp ci : onboard AMD ROCm CI with gfx1151 fixes (#26544) 2026-08-06 10:43:26 +02:00
test-barrier.cpp
test-batch-alloc.cpp
test-c.c
test-chat-auto-parser.cpp chat : Align Laguna-S-2.1 chat template to huggingface (#26232) 2026-08-10 05:20:59 -05:00
test-chat-peg-parser.cpp chat : add qwen3 specialized parser (#26252) 2026-08-02 04:13:20 -05:00
test-chat-template.cpp
test-chat.cpp chat : add new template for DeepSeek V4 Flash 0731 (#26398) 2026-08-03 17:59:11 -05:00
test-col2im-1d.cpp
test-double-float.cpp
test-export-graph-ops.cpp
test-gbnf-validator.cpp
test-gguf-model-data.cpp
test-gguf.cpp
test-grammar-integration.cpp
test-grammar-llguidance.cpp
test-grammar-parser.cpp grammar : degrade max repetition >= 2000 to unbounded (#26613) 2026-08-05 07:39:10 -05:00
test-jinja.cpp
test-json-schema-to-grammar.cpp
test-llama-archs.cpp model: Muse Glimmer Support (#26841) 2026-08-10 13:07:27 +02:00
test-llama-grammar.cpp
test-log.cpp
test-lora-conversion-inference.sh
test-model-load-cancel.cpp tests : avoid building get-model.cpp many times (#26317) 2026-07-30 19:34:04 +03:00
test-model-resolution.cpp Mitigate crashing issue on Windows MSYS2 UCRT64 environment (GCC 16.1.0) (#26555) 2026-08-07 11:17:16 +02:00
test-mtmd-c-api.c mtmd: add chunk save/load function (#26645) 2026-08-06 19:46:40 +02:00
test-opt.cpp
test-peg-parser.cpp
test-quant-type-selection.cpp tests : avoid building get-model.cpp many times (#26317) 2026-07-30 19:34:04 +03:00
test-quantize-fns.cpp
test-quantize-perf.cpp
test-quantize-stats.cpp
test-reasoning-budget.cpp
test-recurrent-state-rollback.cpp DeepseekV4 MTP + DSpark (#25784) 2026-08-02 20:55:34 +08:00
test-rope.cpp
test-rset-release.cpp tests : avoid building get-model.cpp many times (#26317) 2026-07-30 19:34:04 +03:00
test-sampling.cpp sampler : remove "full-context windows" from history-based samplers (#26524) 2026-08-04 21:28:55 +03:00
test-save-load-state.cpp tests : remove unnecessary sync in test-save-load-state (#26166) 2026-07-27 13:11:20 +03:00
test-state-restore-fragmented.cpp
test-thread-safety.cpp
test-tokenizer-0.cpp
test-tokenizer-0.py
test-tokenizer-0.sh
test-tokenizer-1-bpe.cpp
test-tokenizer-1-spm.cpp
test-tokenizer-random.py
test-tokenizers-repo.sh
testing.h