koboldcpp/ggml
Michael de Gans 0f8a414b75
metal : gate mul_mm_id src1 rescale behind ggml_prec (#29029)
* metal : gate mul_mm_id src1 rescale behind ggml_prec

Assisted-by: Claude Fable 5.1

* ggml-webgpu: reject MUL_MAT_ID when src1 precision is F32

* cuda/vulkan: reject MUL_MAT_ID in supports_op when src1 prec is F32

fix `supports_op` to return false for failing backends when the specified src1 precision is f32

Assisted-by: Claude Fable 5.1

---------

Co-authored-by: yomaytk <yoshimura.masashi.frbs@gmail.com>
2026-09-22 18:32:28 +03:00
..
cmake CUDA: replace GGML_FA_ALL_QUANTS with GGML_FA_QUANTS, more control over what is compiled (#28079) 2026-09-09 12:50:08 +02:00
include sycl : pinned memory use right device context instead of 0 (#28895) 2026-09-21 13:58:59 +03:00
src metal : gate mul_mm_id src1 rescale behind ggml_prec (#29029) 2026-09-22 18:32:28 +03:00
.gitignore
CMakeLists.txt ggml : bump version to 0.24.0 (ggml/1627) 2026-09-14 16:45:33 +03:00