Qi Yuhang
|
dcfb92ddbc
|
[diffusion] fix: fix using the sub-optimal fa kernel (#16382)
|
2026-01-05 16:25:41 +08:00 |
|
Xiaoyu Zhang
|
520c048d55
|
[diffusion] CI: add script for automatically generation ci perf baseline (#16389)
|
2026-01-05 13:18:35 +08:00 |
|
Kangyan-Zhou
|
ca80c19b55
|
Revert "[diffusion] feat: support warmup with resolutions" (#16433)
|
2026-01-04 18:44:05 -08:00 |
|
CHEN Xi
|
e267ca0beb
|
[diffusion] doc: document LoRA support in CLI (#16375)
|
2026-01-05 10:20:30 +08:00 |
|
Mick
|
9a8ba3c189
|
[diffusion] feat: support warmup with resolutions (#16330)
|
2026-01-05 10:16:26 +08:00 |
|
Alison Shao
|
52c604342c
|
chore: print test list at beginning and end of run_suite.py (#16334)
Co-authored-by: Mick <mickjagger19@icloud.com>
|
2026-01-04 11:52:52 -08:00 |
|
Chi McIsaac
|
dcacc492d0
|
[diffusion] fix: fix RuntimeError in SageAttention3 on Blackwell with Qwen-Image (#16335)
Co-authored-by: qimcis <qimcis@users.noreply.github.com>
Co-authored-by: Mick <mickjagger19@icloud.com>
|
2026-01-03 22:12:53 +08:00 |
|
Mick
|
2c09de343e
|
[diffusion] improve: skip loading vision module for text encoders (#16304)
|
2026-01-03 19:30:45 +08:00 |
|
Mick
|
38d48de93d
|
[diffusion] CI: simplify warmup (#16303)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2026-01-03 18:09:38 +08:00 |
|
Xiaoyu Zhang
|
d0fb24ee7b
|
[Diffusion] Flux2 tp support (#16219)
Co-authored-by: Mick <mickjagger19@icloud.com>
|
2026-01-03 17:02:27 +08:00 |
|
Alison Shao
|
5b4f790200
|
[diffusion] CI: add CI validation for diffusion model downloads (#16311)
|
2026-01-03 14:12:45 +08:00 |
|
triple-mu
|
888e126ac9
|
[diffusion] comment: fix typo (#16257)
Co-authored-by: Mick <mickjagger19@icloud.com>
|
2026-01-03 11:42:04 +08:00 |
|
sunxxuns
|
30cfb687fa
|
fixed amd multimodal CI failures caused by refactor in #15812 #15813 (#16287)
|
2026-01-02 14:55:42 -08:00 |
|
Xiaoyu Zhang
|
1cfd2b2ded
|
[diffusion] chore: remove redundant ulysses nccl warmup (#16301)
|
2026-01-03 00:19:35 +08:00 |
|
Changyi Yang
|
0eae831797
|
[diffusion] fix: Increase text length from 256 to 1024 in Qwen-Image (#16248)
Co-authored-by: Mick <mickjagger19@icloud.com>
|
2026-01-02 23:25:31 +08:00 |
|
Mick
|
f26f6c2c99
|
[diffusion] fix: make lora compatible with layerwise-offload (#16298)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2026-01-02 22:33:26 +08:00 |
|
Mick
|
5062537b67
|
[diffusion] feat: support lightweight e2e warmup for benchmarking (#16213)
|
2026-01-02 20:10:27 +08:00 |
|
Yuhao Yang
|
6c8587b5db
|
[diffusion] fix: align negative prompt with official readme for new model (#16222)
|
2026-01-02 15:20:56 +08:00 |
|
Xiaoyu Zhang
|
bd48ad5e6b
|
[Diffusion] Fix broken ring_attention when use upstream fa3 (#16270)
|
2026-01-02 14:33:12 +08:00 |
|
Mick
|
21de3e1406
|
[diffusion] webui: tiny fix loading output image (#16251)
|
2026-01-01 14:34:49 +08:00 |
|
Yuhao Yang
|
4280a18a13
|
[diffusion] CI: add test for cache-dit (#16204)
Co-authored-by: Mick <mickjagger19@icloud.com>
|
2025-12-31 19:58:10 +08:00 |
|
chhnb
|
3619ec61b4
|
[diffusion] feat: support multi-frame image output (#15878)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2025-12-31 13:50:02 +08:00 |
|
Mick
|
5bf0d862dd
|
[diffusion] CI: fix generate mode and add cli test (#16174)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2025-12-31 09:51:48 +08:00 |
|
Xiaoyu Zhang
|
733a0c1a37
|
[Diffusion] Zimage opt with qknorm and flashinfer rope (#16161)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2025-12-30 23:39:32 +08:00 |
|
Xiaoyu Zhang
|
b369aaa23f
|
[Diffusion] Refine diffusion profling doc (#16163)
Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>
|
2025-12-30 23:38:46 +08:00 |
|
Yuhao Yang
|
39ca57cd28
|
[diffusion] chore: tiny fix model config (#16159)
Co-authored-by: Mick <mickjagger19@icloud.com>
|
2025-12-30 22:11:38 +08:00 |
|
Mick
|
3449806727
|
[diffusion] feat: generalize layer-wise-offload to all supported models (#16150)
|
2025-12-30 22:06:57 +08:00 |
|
jiapingW
|
5d200dd8d9
|
[diffusion] bench: distinguish between video generation and image generation in the bench_serving (#16149)
|
2025-12-30 21:29:18 +08:00 |
|
yh0903
|
49adb37e37
|
[diffusion] chore: fix ZMQ binding and model loading for FastWan compatibility (#13978)
Co-authored-by: Han Yu <hyu5@dt-login01.delta.ncsa.illinois.edu>
Co-authored-by: Mick <mickjagger19@icloud.com>
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2025-12-30 18:09:38 +08:00 |
|
HuangJi
|
8a84b1e7e0
|
[diffusion] model: support TurboWan2.1-T2V-1.3B/14B SLA (#15888)
|
2025-12-30 14:30:22 +08:00 |
|
Prozac614
|
f253f43c9d
|
[diffusion] CI: fix LoRA downloading issues and respect offline flag (#15813)
|
2025-12-30 11:39:27 +08:00 |
|
Mick
|
1e45320198
|
[diffusion] improve: tiny improve layerwise offload manager by consolidating weights per layer (#16081)
|
2025-12-30 11:31:00 +08:00 |
|
Mick
|
26e17f9076
|
[diffusion] improve: tiny speedup qwen-image-edit-2511 by avoiding unnecessary calculation (#15896)
|
2025-12-30 10:10:32 +08:00 |
|
Mick
|
8e08207c18
|
[diffusion] fix: fix serving with dit-layerwise-offload enabled (#16066)
Co-authored-by: Xiaoyu Zhang <35585791+BBuf@users.noreply.github.com>
|
2025-12-30 00:21:42 +08:00 |
|
Mick
|
b5d9fc873b
|
[diffusion] chore: minor refactor by streamlining the VAE class hierarchy (#16069)
|
2025-12-29 23:37:59 +08:00 |
|
Xiaoyu Zhang
|
24616c5234
|
[Diffusion] Qwen image edit support qknorm optimization (#16062)
|
2025-12-29 21:37:24 +08:00 |
|
Xiaoyu Zhang
|
f3d73b0199
|
[Diffusion] Refactor qwen_image's rope in a single helper func (#16047)
|
2025-12-29 17:24:26 +08:00 |
|
Xiaoyu Zhang
|
8305dc1718
|
[Diffusion] Disable packed QKV for FLUX & Z-Image (#16038)
|
2025-12-29 14:33:05 +08:00 |
|
Liangsheng Yin
|
a435f55d18
|
Tiny print launch command with shlex (#16010)
|
2025-12-29 11:26:46 +08:00 |
|
Li Jinliang
|
b840d6aaeb
|
[diffusion] webui: reference to content task and better visualization capabilities (#16017)
|
2025-12-29 10:22:21 +08:00 |
|
Mick
|
d7a3336ebe
|
[diffusion] fix: fix stages not logged when perf_dump_path is provided (#16016)
|
2025-12-28 23:17:43 +08:00 |
|
Mick
|
9f8e23071a
|
[diffusion] chore: fix default offload setting for image generation model (#15928)
|
2025-12-28 20:45:33 +08:00 |
|
Mick
|
3881bc8d0b
|
[diffusion] CI: relax threshold by supporting different profiles (#16002)
|
2025-12-28 20:05:07 +08:00 |
|
Mick
|
b4a00ed2d9
|
[diffusion] chore: clean ComposedPipelineBase (#15937)
|
2025-12-28 11:43:25 +08:00 |
|
Yuhao Yang
|
0cd2b719a5
|
[diffusion] chore: remove useless params (#15925)
|
2025-12-28 01:01:08 +08:00 |
|
Mick
|
39d56196a0
|
[diffusion] logging: log available gpu mem while loading and generating (#15936)
|
2025-12-28 00:34:58 +08:00 |
|
Mick
|
aa89c6a7e2
|
[diffusion] refactor: unify model loading and offloading behavior (#15923)
|
2025-12-27 16:18:24 +08:00 |
|
Yuhao Yang
|
29ce7b3612
|
[diffusion] chore: remove stepvideo code (#15918)
|
2025-12-27 13:25:05 +08:00 |
|
Lianmin Zheng
|
43e1bbc0d5
|
Revert "[feat] Init support for webui-I2I" (#15906)
|
2025-12-26 09:53:03 -08:00 |
|
Li Jinliang
|
59b12996fb
|
[diffusion] apps: support I2I tasks in webui (#15778)
|
2025-12-26 23:45:11 +08:00 |
|