Wagner Bruna
f273fd35b9
sd: sync to master-601-eeac950 ( #2206 )
...
* sd: sync to master-601-eeac950
* sd: add mmap support
2026-05-16 11:23:10 +08:00
Wagner Bruna
bfe9548fd5
sd: sync to master-596-90e87bc ( #2204 )
...
* sd: reuse source lists between make and cmake
* sd: sync to master-596-90e87bc
* Update source file path for sdtype_adapter.cpp
---------
Co-authored-by: LostRuins Concedo <39025047+LostRuins@users.noreply.github.com >
2026-05-14 23:14:33 +08:00
Wagner Bruna
e2bdd6d7aa
sd: sync to master-591-331cfa5 ( #2155 )
...
* sd: sync to master-585-44cca3d
* sd: sync to master-587-b8bdffc
* sd: sync to master-591-331cfa5
2026-05-01 16:33:28 +08:00
Wagner Bruna
bad9b61064
sd: sync to master-582-7023fc4 ( #2150 )
...
* sd: remove sampler alias handling from the C++ layer
It's already handled at the Python layer.
* sd: sync to master-580-7d33d4b
* sd: sync to master-582-7023fc4
2026-04-21 23:01:33 +08:00
Concedo
5529748a01
Merge commit 'de1aa6fa73e135839109e09fec1a0997f4207b2a' into concedo_experimental
...
# Conflicts:
# docs/build.md
# docs/ops.md
# docs/ops/WebGPU.csv
# ggml/src/ggml-sycl/dequantize.hpp
# ggml/src/ggml-sycl/dmmv.cpp
# ggml/src/ggml-sycl/ggml-sycl.cpp
# ggml/src/ggml-sycl/mmvq.cpp
# ggml/src/ggml-sycl/quants.hpp
# ggml/src/ggml-sycl/vecdotq.hpp
# ggml/src/ggml-webgpu/ggml-webgpu-shader-lib.hpp
# ggml/src/ggml-webgpu/ggml-webgpu.cpp
# ggml/src/ggml-webgpu/wgsl-shaders/mul_mat_decls.tmpl
# tests/test-backend-ops.cpp
# tests/test-quantize-fns.cpp
2026-04-09 17:16:33 +08:00
Wagner Bruna
f371bb14d4
sd: sync to master-560-e8323ca ( #2082 )
...
* sd: sync to master-540-f16a110
* tae post-merge fixes
* build fixes
* restore image mask for non-inpainting models
* sd: sync to master-551-99c1de3
* avoid nlohmann/json.hpp include diffs
* Euler A now works on Flux
* sd: sync to master-555-7397dda
avi_writer.h got removed upstream, but I've simply kept the local
copy for now.
* sd: sync to master-558-8afbeb6
* sd: sync to master-560-e8323ca
2026-04-09 14:44:59 +08:00
Wagner Bruna
9223f41320
sd: call SetCircularAxesAll directly ( #2078 )
2026-03-29 01:17:48 +08:00
Wagner Bruna
b437d18319
add support for cache modes to accelerate image generation ( #2021 )
...
* sd: sync to master-525-d6dd6d7
* sd: add support for cache modes for inference acceleration
* keep gendefaults as a JSON object inside the config file
* covered more invalid cases on gendefaults parsing
2026-03-15 15:27:14 +08:00
Concedo
b1c500ae2b
Merge commit '2948e6049a4ad0f96a4ab15246db2d2086b80703' into concedo_experimental
...
# Conflicts:
# .github/workflows/build.yml
# CONTRIBUTING.md
# docs/backend/VirtGPU/development.md
# docs/ops.md
# docs/ops/WebGPU.csv
# embd_res/templates/GigaChat3-10B-A1.8B.jinja
# embd_res/templates/GigaChat3.1-10B-A1.8B.jinja
# ggml/src/ggml-hip/CMakeLists.txt
# ggml/src/ggml-opencl/CMakeLists.txt
# ggml/src/ggml-opencl/ggml-opencl.cpp
# ggml/src/ggml-webgpu/ggml-webgpu-shader-lib.hpp
# ggml/src/ggml-webgpu/ggml-webgpu.cpp
# scripts/sync_vendor.py
# tests/CMakeLists.txt
# tests/test-backend-ops.cpp
# tests/test-chat.cpp
# tests/test-grammar-integration.cpp
# tests/test-quantize-fns.cpp
2026-03-15 11:21:24 +08:00
Wagner Bruna
5c40f07d4a
sd: sync to 0752cc9 (master-507-b314d80 +1) ( #1999 )
...
* sd: sync to 0752cc9 (master-507-b314d80 +1)
* sd: add flow-shift support to gendefaults
2026-02-28 12:22:32 +08:00
Wagner Bruna
d9ac52a01a
sd: sync to master-492-f957fa3 ( #1957 )
...
* sd: sync to master-492-f957fa3
* add Res Multistep and Res 2s samplers
* make sdflashattention control flash_attn too
2026-02-04 16:12:39 +08:00
Wagner Bruna
c91fc850c1
sd: sync to master-467-0e52afc ( #1916 )
2026-01-15 23:06:51 +08:00
Wagner Bruna
0ef55844d3
sd: sync to master-453-4ff2c8c ( #1907 )
2026-01-03 15:28:27 +08:00
Concedo
442fa7cd7c
support for circular textures in sdcpp
2026-01-01 16:34:09 +08:00
Wagner Bruna
84765f5967
sd: sync to master-447-ccb6b0a ( #1898 )
...
* sd: sync to master-438-298b110
* sd: sync to master-440-3e81246
* sd: sync to master-444-a0adcfb
* sd: sync to master-447-ccb6b0a
2025-12-27 16:30:52 +08:00
Wagner Bruna
44ce1a80b3
sd: sync to master-431-23fce0b ( #1893 )
...
* sd: sync to master-427-78e15bd
* add kl_optimal to the available schedulers list
* more robust workaround to avoid stb linkage issues
* sd: sync to master-431-23fce0b
* add TAEHV support and disable TAE if the model isn't found
2025-12-22 15:07:09 +08:00
Wagner Bruna
78bbe89956
sd: sync to master-417-43a70e8 ( #1889 )
...
* sd: sync to master-417-43a70e8
* fix sdmain build
* switch to upstream apply_loras()
* refactor u8 path conversions and add it to the gguf reader
2025-12-16 16:16:48 +08:00
Wagner Bruna
801840d3bd
sd: sync to master-391-5865b5e ( #1878 )
2025-12-08 19:53:52 +08:00
Wagner Bruna
3db48a1536
sd: sync to master-383-20eb674
2025-12-01 17:04:56 -03:00
Wagner Bruna
717d9c6401
sd: sync to master-377-2034588
2025-11-23 19:29:01 -03:00
Wagner Bruna
703bbf67d8
sd: sync to master-371-5498cc0
2025-11-23 19:29:01 -03:00
Wagner Bruna
8ef66e90c1
sd: sync to master-366-f532972
2025-11-23 19:29:01 -03:00
Wagner Bruna
3a7dd1a97f
sd: sync to master-358-347710f
...
Also adapt Koboldcpp LoRA loading function, and add
backend support for lora_apply_mode.
2025-11-23 19:28:54 -03:00
Wagner Bruna
3318b73c94
sd: sync to master-355-694f0d9
2025-11-23 19:28:34 -03:00
Wagner Bruna
fef73919ea
sd: clean up changes against stable-diffusion.cpp 90ef5f8 ( #1804 )
...
* sd: clean up changes against stable-diffusion.cpp 90ef5f8
Clean up the diff, and include a few missing changes, mainly from
the upscaler and model weight type statistics.
* added line clear again
* remove excess spaces
---------
Co-authored-by: LostRuins Concedo <39025047+LostRuins@users.noreply.github.com >
2025-10-23 22:00:33 +08:00
Concedo
86b94456de
Sync sd.cpp to 90ef5f8
2025-10-20 11:07:42 +08:00
Concedo
fab2ff0687
sync sd.cpp to e370258
2025-10-20 10:45:34 +08:00
Concedo
80f88eb703
wip qwen image edit. not working yet
2025-10-11 11:24:17 +08:00
Concedo
f282362414
added qwen image support (+1 squashed commits)
...
Squashed commits:
[92df28061] added qwen image support (+1 squashed commits)
Squashed commits:
[1485c71ed] wip adding qwen image
2025-10-03 18:58:48 +08:00
Wagner Bruna
42087c3622
update stable-diffusion.cpp to master-306-2abe945 ( #1732 )
...
* update stable-diffusion.cpp to master-52a97b3
* update stable-diffusion.cpp to master-0ebe6fe
* update stable-diffusion.cpp to master-301-fd693ac
* update stable-diffusion.cpp to master-306-2abe945
* fix taesd file selection
2025-09-27 16:52:58 +08:00
Wagner Bruna
5de7ed3d56
WIP: update stable-diffusion.cpp to 5900ef6605c6 (new API) ( #1669 )
...
* Update stable-diffusion.cpp to 5900ef6605c6 (new API)
* Clean up pending LoRA code and simplify LoRA changes to upstream
* Move VAE tiling disabling for TAESD to sdtype_adapter.cpp
* Move auxiliary ctx functions to sdtype_adapter.cpp
* Use ref_images parameter for Kontext images
* Drop clip skip workaround (fixed upstream)
* Workaround for flash attention with img2img
leejet/stable-diffusion.cpp#756
* Workaround for Chroma with flash attention, debug prints
* Disable forcing CLIP weights to F32 for reduced memory usage
2025-08-12 23:25:02 +08:00
Wagner Bruna
eed5577aaa
fix unintended sd model quantization ( #1672 )
...
The recent ggml update added another quant type, GGML_TYPE_MXFP4,
which got the same value as SD_TYPE_COUNT. That made the embedded
sd.cpp quantize to GGML_TYPE_MXFP4 by default.
Photomaker in particular ends up crashing due to
"Missing CPY op for types: f32 mxfp4".
2025-08-08 10:19:58 +08:00
Concedo
186227fc26
sync with sd.cpp
2025-06-30 00:10:51 +08:00
Concedo
4ec0e0fd21
now accept multiple images for reference images
2025-06-28 17:30:28 +08:00
Concedo
ed289227e5
added support for flux kontext
2025-06-28 11:37:19 +08:00
Concedo
ce58d1253f
fixed build and workflow
2025-06-21 00:56:27 +08:00
Concedo
2d4c1aa5a0
chroma support is now usable
2025-06-08 18:53:59 +08:00
Concedo
fea3b2bd4a
updated sdcpp prepare for inpaint
...
fixed img2img (+1 squashed commits)
Squashed commits:
[42c48f14] try update sdcpp, feels kind of buggy
2025-04-09 20:26:10 +08:00
Concedo
de64b9198c
merge checkpoint 2 - functional merge without q4_0_4_4 (need regen shaders)
2024-12-13 17:04:19 +08:00
Concedo
2ba5949054
updated sdcpp, also set euler as default sampler
2024-12-01 17:00:20 +08:00
Concedo
409e393d10
fixed critical bug in image model loader
2024-11-30 23:28:24 +08:00
Concedo
3cfc4dc581
avoid euler a for flux (+4 squashed commit)
...
Squashed commit:
[5a4b72385] fix cuda build
[5f969a645] add vulkan information
[6849e7398] fixed flux
[740e80419] update readme
2024-11-05 22:50:14 +08:00
Concedo
5b90eeaf17
fixed sd to work on larger images by adding tiling, also limit res for sd1.5
2024-11-04 23:26:15 +08:00
Concedo
f32a874966
resync and updated sdcpp for flux and sd3 support
2024-11-03 22:03:16 +08:00
Concedo
f75e479db0
WIP on sdcpp integration
2024-02-29 00:40:07 +08:00
Concedo
26696970ce
initial files from sdcpp (not working)
2024-02-28 15:45:13 +08:00