mirror of
https://github.com/LostRuins/koboldcpp.git
synced 2026-10-07 23:30:15 +00:00
ggml-webgpu: fix flash_attn supports_op check for overlapping KV (#28205)
update-ops-docs.yml #482 -Commit
36a73916ee
pushed by
vrr
ggml-webgpu: fix flash_attn supports_op check for overlapping KV (#28205)
python-type-check.yml #481 -Commit
36a73916ee
pushed by
vrr
ggml-webgpu: fix flash_attn supports_op check for overlapping KV (#28205)
python-check-requirements.yml #480 -Commit
36a73916ee
pushed by
vrr
ggml-webgpu: fix flash_attn supports_op check for overlapping KV (#28205)
pre-tokenizer-hashes.yml #479 -Commit
36a73916ee
pushed by
vrr
common : fix HF cache paths on Windows (#29475)
update-ops-docs.yml #478 -Commit
6c7a87f7e5
pushed by
vrr
common : fix HF cache paths on Windows (#29475)
python-type-check.yml #477 -Commit
6c7a87f7e5
pushed by
vrr
common : fix HF cache paths on Windows (#29475)
pre-tokenizer-hashes.yml #476 -Commit
6c7a87f7e5
pushed by
vrr
cuda : add conv3d with implicit GEMM (#29137)
update-ops-docs.yml #475 -Commit
53ed051ce5
pushed by
vrr
cuda : add conv3d with implicit GEMM (#29137)
python-type-check.yml #474 -Commit
53ed051ce5
pushed by
vrr
cuda : add conv3d with implicit GEMM (#29137)
pre-tokenizer-hashes.yml #473 -Commit
53ed051ce5
pushed by
vrr
cuda : tune MMVQ to MMQ crossover for SM70 (Volta) (#28912)
update-ops-docs.yml #472 -Commit
68d9053afd
pushed by
vrr
cuda : tune MMVQ to MMQ crossover for SM70 (Volta) (#28912)
python-type-check.yml #471 -Commit
68d9053afd
pushed by
vrr
ggml-cpu: add F16 input to the FWHT (#27779)
update-ops-docs.yml #470 -Commit
4fea119de3
pushed by
vrr
ggml-cpu: add F16 input to the FWHT (#27779)
python-type-check.yml #469 -Commit
4fea119de3
pushed by
vrr
ggml-cpu: add F16 input to the FWHT (#27779)
python-check-requirements.yml #468 -Commit
4fea119de3
pushed by
vrr
ggml-cpu: add F16 input to the FWHT (#27779)
pre-tokenizer-hashes.yml #467 -Commit
4fea119de3
pushed by
vrr
ggml-cpu: add F16 input to the FWHT (#27779)
copilot-setup-steps.yml #466 -Commit
4fea119de3
pushed by
vrr
hexagon: add back missing contiguous fast-path and hvx_copy_uu for each run (#28886)
update-ops-docs.yml #465 -Commit
930e2fa599
pushed by
vrr
hexagon: add back missing contiguous fast-path and hvx_copy_uu for each run (#28886)
python-type-check.yml #464 -Commit
930e2fa599
pushed by
vrr
qwen4exp: enable rms_norm + mul fusion (#28896)
python-type-check.yml #463 -Commit
41abbfd599
pushed by
vrr
tests : reduce FA test sizes (#28842)
python-type-check.yml #462 -Commit
4a89937354
pushed by
vrr
tests : reduce FA test sizes (#28842)
python-check-requirements.yml #461 -Commit
4a89937354
pushed by
vrr
Add IQ type handling for MoE (#28476)
python-type-check.yml #460 -Commit
304665fe7a
pushed by
vrr
Revert "CUDA: size routed MoE MMQ N-tiles from typical expert width on RDNA3 (#24546)" (#28551)
python-type-check.yml #459 -Commit
e71b80510c
pushed by
vrr
Revert "CUDA: size routed MoE MMQ N-tiles from typical expert width on RDNA3 (#24546)" (#28551)
python-check-requirements.yml #458 -Commit
e71b80510c
pushed by
vrr
Revert "CUDA: size routed MoE MMQ N-tiles from typical expert width on RDNA3 (#24546)" (#28551)
pre-tokenizer-hashes.yml #457 -Commit
e71b80510c
pushed by
vrr
sycl : fix test-backend-ops CI break && restore Kronecker product FWHT support (#28016) (#28254)
python-type-check.yml #456 -Commit
4d9176092d
pushed by
vrr
sycl : fix test-backend-ops CI break && restore Kronecker product FWHT support (#28016) (#28254)
python-check-requirements.yml #455 -Commit
4d9176092d
pushed by
vrr
sycl : fix test-backend-ops CI break && restore Kronecker product FWHT support (#28016) (#28254)
pre-tokenizer-hashes.yml #454 -Commit
4d9176092d
pushed by
vrr
ggml : don't crash when backend search path can't be read (#28271)
update-ops-docs.yml #453 -Commit
4cbe8b070b
pushed by
vrr