koboldcpp/ggml
Anant Shrivastava 1aa2954bde
sycl : coalesce MKL-FA softmax loads instead of one work-item per row (#28918)
* sycl : coalesce MKL-FA softmax loads instead of one work-item per row

* better human readable variable name
2026-09-21 11:07:04 +03:00
..
cmake CUDA: replace GGML_FA_ALL_QUANTS with GGML_FA_QUANTS, more control over what is compiled (#28079) 2026-09-09 12:50:08 +02:00
include [SYCL] Fix function signature for ggml_backend_sycl_split_buffer_type (#28981) 2026-09-16 22:15:47 -04:00
src sycl : coalesce MKL-FA softmax loads instead of one work-item per row (#28918) 2026-09-21 11:07:04 +03:00
.gitignore
CMakeLists.txt ggml : bump version to 0.24.0 (ggml/1627) 2026-09-14 16:45:33 +03:00