Commit Graph
363 Commits
Author SHA1 Message Date
kijai b06c7d2d6d Add UltraViCo -sage attention mode and refactor some attention code
https://github.com/thu-ml/DiT-Extrapolation/
2025-12-03 13:32:45 +02:00
kijai 014e711972 cleanup 2025-12-03 11:09:09 +02:00
kijai b9f6c9aa50 Update __init__.py 2025-12-02 00:45:22 +02:00
kijai c1fbc93521 Add ViBTScheduler to use ViBT models
https://github.com/Yuanshi9815/ViBT/tree/main
2025-12-02 00:26:50 +02:00
kijai c4db00609a Use torch.chunk for chunking 2025-12-01 17:39:50 +02:00
kijai 5a52d6b92f Add VAE feat_cache offloading, VAE tqdm progress bar and memory usage report 2025-12-01 15:05:20 +02:00
kijai 196d39695f Fix sageattn_varlen 2025-12-01 01:11:06 +02:00
kijai e5be3e5263 Use torch custom_ops to avoid graph breaks with torch.compile
Hopefully finally fixes the torch.compile VRAM issues...
2025-12-01 00:29:27 +02:00
kijai a6071c7be5 Merge branch 'main' into steadydancer 2025-11-30 17:56:50 +02:00
kijai a9cd073f29 Remove unnecessary recompile when using cfg 2025-11-30 17:52:56 +02:00
kijai 1e9e2be622 Avoid recompile here 2025-11-30 17:32:24 +02:00
kijai 99c3978da4 Reduce peak VRAM usage when not using torch.compile (and some even with it)
Found some intermediates that weren't freed which should reduce VRAM usage overall, and modified RoPE application outside torch compile for similar gains than when using torch.compile.
2025-11-30 17:14:53 +02:00
kijai 66d44ec8db This doesn't really do anything useful 2025-11-28 21:15:25 +02:00
kijai 394c7c13d2 Add strength controls 2025-11-28 21:11:20 +02:00
kijai e54fa5d059 Init 2025-11-28 20:32:16 +02:00
kijai fa7a967ee7 Fix stand-in 2025-11-17 11:56:02 +02:00
kijai f872460285 Fix TTM for dual sampler setups 2025-11-16 19:23:19 +02:00
kijai e3c2a1431b Squashed commit of the following:
commit f685ee33ac
Merge: bb5707f 4e31081
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Nov 13 16:37:38 2025 +0200

    Merge branch 'main' into bindweave

commit bb5707f601
Merge: acb662b ff26836
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Tue Nov 11 18:53:19 2025 +0200

    Merge branch 'main' into bindweave

commit acb662b5af
Merge: 907c9e1 e926f7a
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Tue Nov 11 11:44:26 2025 +0200

    Merge branch 'main' into bindweave

commit 907c9e1cdd
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Mon Nov 10 21:02:58 2025 +0200

    Update nodes.py

commit e4a4d22537
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Sat Nov 8 16:04:55 2025 +0200

    Update nodes_sampler.py

commit a3b2f67337
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Sat Nov 8 16:03:00 2025 +0200

    Pad clip vision embeds like in original code

commit 1e00c8fb28
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Sat Nov 8 12:21:11 2025 +0200

    Update nodes.py

commit ff16dce5c0
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Fri Nov 7 01:15:13 2025 +0200

    Update nodes_sampler.py

commit f972b31bf2
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Fri Nov 7 01:09:22 2025 +0200

    Update nodes.py

commit 3dacd6a719
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Fri Nov 7 00:45:06 2025 +0200

    Update nodes_sampler.py

commit 7bf99791ad
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Fri Nov 7 00:44:13 2025 +0200

    Update nodes.py

commit 7a5587b5af
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Fri Nov 7 00:39:10 2025 +0200

    Let the user resize for QwenVL

    Seems to need smaller resolutions

commit d6cf172846
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Nov 6 23:41:24 2025 +0200

    Update nodes.py

commit cf86f4f0a4
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Nov 6 21:01:03 2025 +0200

    Update model.py

commit b1f8309a20
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Nov 6 19:55:54 2025 +0200

    Update nodes_model_loading.py

commit 8992c6af64
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Nov 6 19:22:56 2025 +0200

    Don't include padding for scheduler

commit e4084a961b
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Nov 6 18:32:53 2025 +0200

    Update nodes.py

commit 3ec1edefbe
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Nov 6 17:35:48 2025 +0200

    init

    For testing, no idea if it works yet
2025-11-13 17:10:57 +02:00
kijai 1d0516a2a9 Avoid graph break for LongCat 2025-11-04 10:26:37 +02:00
kijai 8ce6916d72 Fix for some cases of using comfy_chunked rope 2025-11-03 10:36:55 +02:00
kijai 0d0d28569a Fix cases where text encoder isn't used (eg. Minimax remover) 2025-11-03 10:29:10 +02:00
kijai 75109fdb79 Fix custom sigmas with euler 2025-11-03 10:12:05 +02:00
kijai 5eae7087fa Fix for S2V 2025-11-02 01:24:50 +02:00
kijai 393fe78ec2 Update model.py 2025-10-31 23:39:55 +02:00
kijai cc9bf1e4f5 Store lora diffs in buffers for GGUF as well 2025-10-30 16:44:03 +02:00
kijai d2614a9a49 Merge branch 'main' into longcat 2025-10-29 02:33:37 +02:00
kijai 9d45b9f0de Use comfy core Conv3D workaround for VAE rather than the fp32 cast 2025-10-29 02:23:49 +02:00
chengzeyiandClaude d15cf3001f Fix dtype mismatch in ref_conv forward pass
This commit fixes a RuntimeError that occurs when using Fun-Control
reference images: "Input type (float) and bias type (c10::Half)
should be the same"

Root cause:
- Commit 1ba1a16 changed the dtype handling strategy to convert
  the main latent `x` to `base_dtype` instead of converting
  embeddings to match `x.dtype`
- This caused `fun_ref` input to be in a different dtype than
  the `ref_conv` layer's weights and bias
- Line 2324 already handles this correctly for `attn_cond` by
  converting to `self.attn_conv_in.weight.dtype`

Solution:
- Convert `fun_ref` to match `self.ref_conv.weight.dtype` before
  passing through the convolution layer
- This follows the same pattern used for `attn_cond` on line 2324

🤖 Generated with [Claude Code](https://claude.com/claude-code)

Co-Authored-By: Claude <noreply@anthropic.com>
2025-10-28 12:08:23 +00:00
kijai eebbcd5ee0 Update model.py 2025-10-28 02:04:50 +02:00
kijai 90908df260 Update model.py 2025-10-28 01:54:42 +02:00
kijai c80a488f70 Use fp32 norms for other models too and other fixes 2025-10-28 01:52:48 +02:00
kijai e560366600 Update model.py 2025-10-27 18:55:30 +02:00
kijai c59e52ca44 Precision adjustments 2025-10-27 00:23:32 +02:00
kijai a0bdf20817 Some cleanup and allow full block swap 2025-10-26 23:05:17 +02:00
kijai 8ad7e50f33 Fix cross attention split point 2025-10-26 22:07:19 +02:00
kijai d504c96174 Separate attention for input images like in original 2025-10-26 19:14:47 +02:00
kijai 43acf83adb Update model.py 2025-10-26 16:57:25 +02:00
kijai fb00932cad Init
https://huggingface.co/Kijai/LongCat-Video_comfy/tree/main
2025-10-26 16:36:20 +02:00
kijai 41168b1e82 Support light VAE
https://huggingface.co/lightx2v/Autoencoders/tree/main
2025-10-23 12:31:55 +03:00
kijai 67fcf0ba52 Reduce needless torch.compile recompiles 2025-10-22 13:29:13 +03:00
kijai 1f0861b649 MoCha: modify RoPE function to be more torch.compile friendly 2025-10-21 17:54:25 +03:00
unrealMJ 88defbfdd1 add MoCha 2025-10-21 09:40:54 +08:00
kijai 200f6943e3 Add sageattn mode that allows torch.compile
Latest wheel from woct0rdho includes the torch.compile fix:

https://github.com/woct0rdho/SageAttention/releases

Based on my quick testing this reduces peak VRAM usage a bit when running sageattn + torch.compile
2025-10-20 15:16:43 +03:00
kijai 9cd79d3d4a Add experimental rCM scheduler
Based on the original code, works but doesn't feel better than dpm++_sde so far
2025-10-19 19:34:17 +03:00
kijai cc06d71ca0 Workaround for bug in pytorch 2.9.0 that makes the VAE use crazy amounts of VRAM
This is probably caused by:

https://github.com/pytorch/pytorch/pull/164027/files

That disables cudnn when using half precision VAE, this workaround simply uses fp32 for the Conv3D operations.
2025-10-16 17:10:05 +03:00
kijai 9a88b9e40a FlashVSR: Add strength setting 2025-10-15 19:11:41 +03:00
kijai bb75cddd60 Add minimal FlashVSR upscale support
https://zhuang2002.github.io/FlashVSR/

This only implements the projection model and the VAE, which seems to be enough for upscaling. This does NOT implement any of the streaming and sparse attention code.
2025-10-15 18:24:28 +03:00
kijai 139bdf827f Squashed commit of the following:
commit 73dd1a06d33953912f5dd684f168028b14e42a36
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Mon Oct 13 19:47:38 2025 +0300

    cleanup

commit 39bc2cecf493e2eb176b55e8841d933f0da1ec39
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Mon Oct 13 19:24:20 2025 +0300

    Allow scheduling ovi cfg

commit 2c153c5f324dbd59670ad9c51a7995459504a3cd
Merge: dba7667 32eb6b4
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Mon Oct 13 17:48:20 2025 +0300

    Merge branch 'main' into ovi

commit dba76674c71af7bf94c82834a0b0e40d94043c99
Merge: 0f11a43 5a0456e
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Sun Oct 12 22:45:43 2025 +0300

    Merge branch 'main' into ovi

commit 0f11a439622799ad8070f8a2b8cc8e6a041b761d
Merge: 0999f50 e2d8c9b
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Sat Oct 11 07:48:06 2025 +0300

    Merge branch 'main' into ovi

commit 0999f50cfe025290cd7ce88a8dd1acff0b38d9bd
Merge: d45df1f f1d1c83
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Fri Oct 10 22:16:09 2025 +0300

    Merge branch 'main' into ovi

commit d45df1fb5b7c629b15eabc197357d62bdc232aaf
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Oct 9 20:21:37 2025 +0300

    Remove dependency for librosa

commit d8e7533fdf7eab1d2489c3e025a908c02d997444
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Oct 9 19:57:28 2025 +0300

    Remove omegaconf dependency

commit f4e27ff018e98cb5b09655dceda399baea36b240
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Oct 9 19:31:06 2025 +0300

    Fix VACE

commit 35d3df39294831e5e7568b6f7e16d2ecf2d790a0
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Thu Oct 9 00:26:40 2025 +0300

    small update

commit 96f8ea1d26869ab7e49e12a07f19d5d5a2023253
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 22:32:57 2025 +0300

    Create wanvideo_2_2_5B_ovi_testing.json

commit a2511be73b9da7019fd21aeb0b521af941c09150
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 22:32:54 2025 +0300

    Update nodes_sampler.py

commit d3688b8db71452ea1f7c9a2bc0216441d524e56c
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 21:43:02 2025 +0300

    Allow EasyCache to work with ovi

commit 586d9148a0306ef5d30e9a971a9c3be4cd3ecc97
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 19:09:06 2025 +0300

    Update model.py

commit 61eedd2839decdb7d4c2ddd5f1310fdaf49d36ad
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 19:09:02 2025 +0300

    I2V fix

commit a97fcb1b9ae9fb7bbfdf668c24816e014a1b58d1
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 17:57:28 2025 +0300

    Add nodes to set audio latent size

commit d41e42a697f3d561dabbc22566f633b5f1bbd952
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 16:42:04 2025 +0300

    Support loading mmaudio vae from .safetensors

commit 1b0e28ec41e3c97fe1f2f057fef9b9bbcb87bca7
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 16:19:53 2025 +0300

    Update nodes_sampler.py

commit fbd18f45fe85ede8edcb5aebaea7ceb5b6eab5a2
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 10:16:44 2025 +0300

    Fixes for other workflows

commit b06993b637198f7fad92208f3b3dc9a7d7f57c7f
Author: kijai <40791699+kijai@users.noreply.github.com>
Date:   Wed Oct 8 09:46:27 2025 +0300

    initial commit

    T2V works
2025-10-13 20:16:53 +03:00
kijai 32eb6b480d Allow canceling mid step 2025-10-13 17:47:28 +03:00
kijai 6d2ff33466 Fix double encode 2025-10-10 09:36:37 +03:00