Commit graph

783 commits

Author SHA1 Message Date
Wagner Bruna
0bd509c596
sd: sync with master-866-42d6c0a (#2458)
* sd: sync with master-856-e06b205

* sd: sync with master-857-7f986a9

* sd: sync with master-859-7f410a3

* sd: sync with master-866-42d6c0a

* sd: update generate_video call
2026-09-24 16:21:20 +08:00
Concedo
60acd9b92f metadata forcing 2026-09-24 00:20:36 +08:00
Concedo
444b9e4419 acestep qwen3 lm caption comes later after generating lyrics. 2026-09-24 00:04:57 +08:00
Wagner Bruna
b3887208fc
sd: round info floating point parameters (#2484) 2026-09-21 21:31:49 +08:00
Concedo
c49fbd7003 cleanup and remove sdcpp_logger_adapter 2026-09-21 21:00:24 +08:00
Wagner Bruna
3f490d1133
sd: fix: inhibit ggml log changes from sdcpp code (#2483)
* sd: apply logging changes from master-885-b8248a8

* sd: fix: inhibit ggml log changes from sdcpp code

* sd: do not set ggml logging callback, leave the default

* sd: do not set ggml logging callback

* sd: fix kcpp log level test
2026-09-21 20:56:15 +08:00
Concedo
e5e51240b5 limit max preview size 2026-09-20 20:57:30 +08:00
Concedo
cc9fa7d5e4 fix sd logger move out of makefiles 2026-09-19 18:04:08 +08:00
Wagner Bruna
469f456003
sd: sync with master-852-14eddb3 (#2457)
* bump GGML_MAX_NAME to 160

https://github.com/leejet/stable-diffusion.cpp/pull/1950

* sd: sync with master-852-14eddb3

* fix int8 and fp8 tensor size validation

* temporarily disable LoRA caching
2026-09-19 17:20:14 +08:00
Concedo
9b7397f140 fix sd logger (codex assisted) 2026-09-17 00:02:23 +08:00
Wagner Bruna
27786379e5
sd: sync with master-849-d04e895 (#2453) 2026-09-16 23:13:56 +08:00
Wagner Bruna
3f0a418624
sd: fix video LoRAs (#2460) 2026-09-14 22:21:11 +08:00
Wagner Bruna
da1a9d4f4c
sd: merge upstream int8 convrot and fp8 scaled at master-841-6b3edaa (#2452)
Gated behind disabled macros KCPP_MAINLINE_INT8_CONVROT and
KCPP_MAINLINE_FP8_SCALED .
2026-09-14 22:11:37 +08:00
Concedo
de926911e1 fixed a arm compile issue 2026-09-06 18:18:24 +08:00
Wagner Bruna
5ac1407e87
sd: free audio VAE weights after encoding and decoding (#2430) 2026-09-04 10:15:22 +08:00
Wagner Bruna
b83cdc6f42
sd: cherry-pick changes up to master-841-6b3edaa (#2426)
fix: match exact weights in LLM config detection
feat: add LTX-2.5 support
fix: correct MiniMax H3 reference audio encoding
fix: correct MiniMax H3 audio Euler steps
feat: additional `--preview-interval` values
feat: support numbering for preview images
fix: use carrier sampling for MiniMax H3 audio
feat: generalize temporal tiling across video VAEs
2026-09-03 16:32:59 +08:00
Concedo
5f9ecb13a6 allow removing the clamp limits for flux1 too 2026-09-01 16:39:07 +08:00
Concedo
e3dd9daf97 fixed ltx audio to video conditioning 2026-08-30 17:49:15 +08:00
Concedo
cec13b80a4 wip on true media references for h3 video gen 2026-08-30 14:49:24 +08:00
Wagner Bruna
4ac5721b5e
sd: cherry-pick changes up to master-827-97d2990 (#2408)
feat: add taeh3 support
fix: prevent gallocr hash overflow in tiny graph-cut segments
fix: re-clamp streaming VRAM budget to currently free memory
fix: mark graph cuts with both a prefix and a suffix
fix: make max_order of lms sampler configurable
fix: guard against missing sampler/scheduler names
chore: format code
2026-08-28 22:55:53 +08:00
Concedo
dd4fa6718a dont set_preview_images within abort itself 2026-08-26 12:49:10 +08:00
Concedo
eb6c7376a1 clear old preview when aborted 2026-08-25 20:56:42 +08:00
Wagner Bruna
ad2b22a41d
sd: set step as 0 for the first noisy preview (#2416)
The initial noisy preview is reported before denoising, so it shows the
ongoing step number; but we immediately switch to denoised previews or
progress callback info, which instead show concluded steps. So we
currently report step '1' during the first two inference steps.
2026-08-25 20:53:15 +08:00
Concedo
b6a253517e fix whisper device selection 2026-08-22 18:16:59 +08:00
Concedo
2b78904b4a bug fixes for backend code 2026-08-22 17:28:48 +08:00
Wagner Bruna
b9b3bbcec1
refactor: add new module for backend-specific code (#2308)
* refactor: add new module for backend-specific code

* replace ifdefs on rwkv_v3

* replace ifdefs on llama_v2

* replace ifdefs on llama_v3

* replace ifdefs on gpt2_v3

* replace ifdefs on gptj_v3

* replace ifdefs on mpt_v3

* replace ifdefs on neox_v3

* replace ifdefs on whisper

* adjust build for kcpp_backend

Mostly a:
sed 's/ gpttype_adapter\([^ ]*\)\.o / gpttype_adapter_default.o kcpp_backend\1.o /g' Makefile

and include variant objects for kcpp_backend.
2026-08-22 10:00:00 +08:00
Concedo
5b80635beb fix for https://github.com/LostRuins/koboldcpp/security/advisories/GHSA-qhvp-gj7g-rw26 2026-08-18 18:20:05 +08:00
Wagner Bruna
493f6e6d9c
sd: fix progress report for second order samplers (#2403)
Each first half step is reported as negative.
2026-08-18 18:18:58 +08:00
Concedo
d81ae290d2 move a print to debug 2026-08-14 21:22:42 +08:00
Wagner Bruna
7ec78d48e4
sd: sync with master-816-487de75 (#2395) 2026-08-12 21:35:09 +08:00
Concedo
75c8184210 hold the mutex for the whole step_callback function 2026-08-12 18:05:05 +08:00
Concedo
533aa18897 make preview sticky for sdcpp 2026-08-12 17:13:20 +08:00
Concedo
5925082d19 fix noisy callback no longer encodes a preview, and preview generation triggers only one-shot 2026-08-11 21:21:50 +08:00
Wagner Bruna
c98c00f9ba
sd: generation progress fixes (#2391)
* sd: generation progress fixes

The preview callback is not called if preview images are not enabled,
so when a preview image wasn't requested, the step count wouldn't be
updated. So move the update to the progress callback. Additionally,
adjust the total step count when the progress call reports a lower
total (e.g. for img2img).

Also remove the preview reset from inside the callback, since it
often caused a preview miss, depending on when the next preview
request arrived.

* sd: fix image preview behavior for VAE encoding / tiling

The progress callback is also called for VAE encoding and decoding,
receiving the number of tiles as step count, so there is no simple
way to detect the diffusion beginning. So we set up the first preview
callback to detect it, and transition to the decoding phase when
we reach the last step.
2026-08-11 21:03:11 +08:00
Concedo
16d9abb6b3 vae free buf fix sdcpp memory leak 2026-08-08 18:44:59 +08:00
Concedo
56368c0fd4 newline in debug (+1 squashed commits)
Squashed commits:

[dbc58568d] newline in debug
2026-08-07 18:42:09 +08:00
Concedo
569532f7e4 improve image gen display 2026-08-07 15:21:42 +08:00
Concedo
d49b7a62e8 fixed compile order 2026-08-07 14:44:54 +08:00
Concedo
28e98198db handle race condition for image preview 2026-08-07 10:21:20 +08:00
Concedo
61bfce83de fix some issues with the preview image: Preview generation is disabled by default and only done when requested
Cleared stale generation state at job start/end.
Fixed the animated preview GIF buffer leak.
2026-08-07 00:19:39 +08:00
Wagner Bruna
e3cb5e9e44
sd: support for the /sdapi/v1/progress endpoint (#2316)
Co-authored-by: LostRuins Concedo <39025047+LostRuins@users.noreply.github.com>
2026-08-06 23:52:31 +08:00
Concedo
910ab962eb load lora at runtime, to solve eager/lazy loading bug 2026-08-06 23:43:49 +08:00
Wagner Bruna
0ddb9190c8
sd: sync with master-812-ea7f0c8 (#2371)
* sd: sync with master-801-9cfe2af

* sd: sync with master-802-e92e86f

* sd: sync with master-805-e31a86c

* sd: sync with master-810-db99efd

* sd: sync with master-812-ea7f0c8

* sd: minimax-h3 support
2026-08-06 23:42:23 +08:00
Concedo
7a69646196 qwen3tts support languages 2026-07-27 22:45:04 +08:00
Concedo
133411c3fd fixed incorrect line removal 2026-07-25 19:06:56 +08:00
Wagner Bruna
fd44ba2c61
sd: sync with master-795-87a0177 (#2338)
* sd: sync with master-782-b290693

* sd: sync with master-788-8a51eb9

* sd: sync with master-789-5114672

* sd: sync with master-795-87a0177

* sd: expose ref_image_args and make it trigger edit mode
2026-07-25 19:04:50 +08:00
Concedo
49dbdaaab5 Merge branch 'upstream' into concedo_experimental
# Conflicts:
#	AGENTS.md
#	CODEOWNERS
#	CONTRIBUTING.md
#	docs/backend/OPENCL.md
#	docs/development/HOWTO-add-model.md
#	examples/training/finetune.cpp
#	ggml/src/ggml-hexagon/ggml-hexagon.cpp
#	ggml/src/ggml-hexagon/htp-drv.cpp
#	ggml/src/ggml-hexagon/htp/act-ops.c
#	ggml/src/ggml-hexagon/htp/dma-queue.c
#	ggml/src/ggml-hexagon/htp/dma-queue.h
#	ggml/src/ggml-hexagon/htp/flash-attn-ops.c
#	ggml/src/ggml-hexagon/htp/flash-attn-ops.h
#	ggml/src/ggml-hexagon/htp/hmx-mm-kernels-tiled.h
#	ggml/src/ggml-hexagon/htp/htp-ctx.h
#	ggml/src/ggml-hexagon/htp/htp-ops.h
#	ggml/src/ggml-hexagon/htp/htp-tensor.c
#	ggml/src/ggml-hexagon/htp/htp-tensor.h
#	ggml/src/ggml-hexagon/htp/hvx-fa-kernels.h
#	ggml/src/ggml-hexagon/htp/hvx-reduce.h
#	ggml/src/ggml-hexagon/htp/main.c
#	ggml/src/ggml-hexagon/htp/matmul-ops.c
#	ggml/src/ggml-hexagon/htp/matmul-ops.h
#	ggml/src/ggml-hexagon/htp/unary-ops.c
#	ggml/src/ggml-hexagon/htp/unary-ops.h
#	ggml/src/ggml-opencl/CMakeLists.txt
#	ggml/src/ggml-opencl/ggml-opencl.cpp
#	scripts/compare-llama-bench.py
#	scripts/snapdragon/ggml-hexagon-profile.py
#	scripts/snapdragon/ggml-hexagon-trace.py
#	scripts/sync_vendor.py
#	tests/test-arg-parser.cpp
#	tests/test-chat.cpp
#	tests/test-model-load-cancel.cpp
#	tests/test-quantize-stats.cpp
#	tools/cli/README.md
#	tools/completion/README.md
#	tools/llama-bench/llama-bench.cpp
#	tools/server/README.md
#	tools/ui/src/lib/constants/settings-registry.ts
2026-07-25 12:20:51 +08:00
Concedo
68db0dda84 fixed a qwen image edit regression from https://github.com/LostRuins/koboldcpp/pull/2251 2026-07-21 22:20:48 +08:00
Concedo
d7ca5a4d01 use u8 helper 2026-07-19 17:32:43 +08:00
Wagner Bruna
1f9dd9c398
sd: sync with master-775-b5d8120 (#2321)
* sd: expose extra_sample_args parameter

* sd: sync with master-773-1b04283

* sd: sync with master-775-b5d8120
2026-07-13 20:42:22 +08:00