EasyAI代码托管平台

mirror of https://github.com/comfyanonymous/ComfyUI.git synced 2026-04-13 20:12:30 +08:00

Author	SHA1	Message	Date
comfyanonymous	27d1bd8829	Fix rope scaling. (#10560 )	2025-10-30 22:51:58 -04:00
comfyanonymous	614cf9805e	Add a ScaleROPE node. Currently only works on WAN models. (#10559 )	2025-10-30 22:11:38 -04:00
comfyanonymous	0cf33953a7	Fix batch size above 1 giving bad output in chroma radiance. (#10394 )	2025-10-18 23:15:34 -04:00
rattus128	95ca2e56c8	WAN2.2: Fix cache VRAM leak on error (#10308 ) Same change pattern as `7e8dd275c2` applied to WAN2.2 If this suffers an exception (such as a VRAM oom) it will leave the encode() and decode() methods which skips the cleanup of the WAN feature cache. The comfy node cache then ultimately keeps a reference this object which is in turn reffing large tensors from the failed execution. The feature cache is currently setup at a class variable on the encoder/decoder however, the encode and decode functions always clear it on both entry and exit of normal execution. Its likely the design intent is this is usable as a streaming encoder where the input comes in batches, however the functions as they are today don't support that. So simplify by bringing the cache back to local variable, so that if it does VRAM OOM the cache itself is properly garbage when the encode()/decode() functions dissappear from the stack.	2025-10-13 15:23:11 -04:00
comfyanonymous	84e9ce32c6	Implement the mmaudio VAE. (#10300 )	2025-10-11 22:57:23 -04:00
comfyanonymous	195e0b0639	Remove useless code. (#10223 )	2025-10-05 15:41:19 -04:00
Finn-Hecker	93d859cfaa	Fix type annotation syntax in MotionEncoder_tc __init__ (#10186 ) ## Summary Fixed incorrect type hint syntax in `MotionEncoder_tc.__init__()` parameter list. ## Changes - Line 647: Changed `num_heads=int` to `num_heads: int` - This corrects the parameter annotation from a default value assignment to proper type hint syntax ## Details The parameter was using assignment syntax (`=`) instead of type annotation syntax (`:`), which would incorrectly set the default value to the `int` class itself rather than annotating the expected type.	2025-10-03 14:32:19 -07:00
rattus128	4965c0e2ac	WAN: Fix cache VRAM leak on error (#10141 ) If this suffers an exception (such as a VRAM oom) it will leave the encode() and decode() methods which skips the cleanup of the WAN feature cache. The comfy node cache then ultimately keeps a reference this object which is in turn reffing large tensors from the failed execution. The feature cache is currently setup at a class variable on the encoder/decoder however, the encode and decode functions always clear it on both entry and exit of normal execution. Its likely the design intent is this is usable as a streaming encoder where the input comes in batches, however the functions as they are today don't support that. So simplify by bringing the cache back to local variable, so that if it does VRAM OOM the cache itself is properly garbage when the encode()/decode() functions dissappear from the stack.	2025-10-01 18:42:16 -04:00
comfyanonymous	a6f83a4a1a	Support the new hunyuan vae. (#10150 )	2025-10-01 17:19:13 -04:00
rattus128	653ceab414	Reduce Peak WAN inference VRAM usage - part II (#10062 ) * flux: math: Use _addcmul to avoid expensive VRAM intermediate The rope process can be the VRAM peak and this intermediate for the addition result before releasing the original can OOM. addcmul_ it. * wan: Delete the self attention before cross attention This saves VRAM when the cross attention and FFN are in play as the VRAM peak.	2025-09-27 18:14:16 -04:00
comfyanonymous	fccab99ec0	Fix issue with .view() in HuMo. (#10014 )	2025-09-24 20:09:42 -04:00
comfyanonymous	e8df53b764	Update WanAnimateToVideo to more easily extend videos. (#9959 )	2025-09-19 18:48:56 -04:00
comfyanonymous	dc95b6acc0	Basic WIP support for the wan animate model. (#9939 )	2025-09-19 03:07:17 -04:00
comfyanonymous	24b0fce099	Do padding of audio embed in model for humo for more flexibility. (#9935 )	2025-09-18 19:54:16 -04:00
comfyanonymous	dd611a7700	Support the HuMo 17B model. (#9912 )	2025-09-17 18:39:24 -04:00
comfyanonymous	9288c78fc5	Support the HuMo model. (#9903 )	2025-09-17 00:12:48 -04:00
rattus128	e42682b24e	Reduce Peak WAN inference VRAM usage (#9898 ) * flux: Do the xq and xk ropes one at a time This was doing independendent interleaved tensor math on the q and k tensors, leading to the holding of more than the minimum intermediates in VRAM. On a bad day, it would VRAM OOM on xk intermediates. Do everything q and then everything k, so torch can garbage collect all of qs intermediates before k allocates its intermediates. This reduces peak VRAM usage for some WAN2.2 inferences (at least). * wan: Optimize qkv intermediates on attention As commented. The former logic computed independent pieces of QKV in parallel which help more inference intermediates in VRAM spiking VRAM usage. Fully roping Q and garbage collecting the intermediates before touching K reduces the peak inference VRAM usage.	2025-09-16 19:21:14 -04:00
blepping	1a85483da1	Fix depending on asserts to raise an exception in BatchedBrownianTree and Flash attn module (#9884 ) Correctly handle the case where w0 is passed by kwargs in BatchedBrownianTree	2025-09-15 20:05:03 -04:00
Jedrzej Kosinski	f228367c5e	Make ModuleNotFoundError ImportError instead (#9850 )	2025-09-13 21:34:21 -04:00
comfyanonymous	80b7c9455b	Changes to the previous radiance commit. (#9851 )	2025-09-13 18:03:34 -04:00
blepping	c1297f4eb3	Add support for Chroma Radiance (#9682 ) * Initial Chroma Radiance support * Minor Chroma Radiance cleanups * Update Radiance nodes to ensure latents/images are on the intermediate device * Fix Chroma Radiance memory estimation. * Increase Chroma Radiance memory usage factor * Increase Chroma Radiance memory usage factor once again * Ensure images are multiples of 16 for Chroma Radiance Add batch dimension and fix channels when necessary in ChromaRadianceImageToLatent node * Tile Chroma Radiance NeRF to reduce memory consumption, update memory usage factor * Update Radiance to support conv nerf final head type. * Allow setting NeRF embedder dtype for Radiance Bump Radiance nerf tile size to 32 Support EasyCache/LazyCache on Radiance (maybe) * Add ChromaRadianceStubVAE node * Crop Radiance image inputs to multiples of 16 instead of erroring to be in line with existing VAE behavior * Convert Chroma Radiance nodes to V3 schema. * Add ChromaRadianceOptions node and backend support. Cleanups/refactoring to reduce code duplication with Chroma. * Fix overriding the NeRF embedder dtype for Chroma Radiance * Minor Chroma Radiance cleanups * Move Chroma Radiance to its own directory in ldm Minor code cleanups and tooltip improvements * Fix Chroma Radiance embedder dtype overriding * Remove Radiance dynamic nerf_embedder dtype override feature * Unbork Radiance NeRF embedder init * Remove Chroma Radiance image conversion and stub VAE nodes Add a chroma_radiance option to the VAELoader builtin node which uses comfy.sd.PixelspaceConversionVAE Add a PixelspaceConversionVAE to comfy.sd for converting BHWC 0..1 <-> BCHW -1..1	2025-09-13 17:58:43 -04:00
comfyanonymous	a3b04de700	Hunyuan refiner vae now works with tiled. (#9836 )	2025-09-12 19:46:46 -04:00
Jedrzej Kosinski	d7f40442f9	Enable Runtime Selection of Attention Functions (#9639 ) * Looking into a @wrap_attn decorator to look for 'optimized_attention_override' entry in transformer_options * Created logging code for this branch so that it can be used to track down all the code paths where transformer_options would need to be added * Fix memory usage issue with inspect * Made WAN attention receive transformer_options, test node added to wan to test out attention override later * Added *kwargs to all attention functions so transformer_options could potentially be passed through Make sure wrap_attn doesn't make itself recurse infinitely, attempt to load SageAttention and FlashAttention if not enabled so that they can be marked as available or not, create registry for available attention * Turn off attention logging for now, make AttentionOverrideTestNode have a dropdown with available attention (this is a test node only) * Make flux work with optimized_attention_override * Add logs to verify optimized_attention_override is passed all the way into attention function * Make Qwen work with optimized_attention_override * Made hidream work with optimized_attention_override * Made wan patches_replace work with optimized_attention_override * Made SD3 work with optimized_attention_override * Made HunyuanVideo work with optimized_attention_override * Made Mochi work with optimized_attention_override * Made LTX work with optimized_attention_override * Made StableAudio work with optimized_attention_override * Made optimized_attention_override work with ACE Step * Made Hunyuan3D work with optimized_attention_override * Make CosmosPredict2 work with optimized_attention_override * Made CosmosVideo work with optimized_attention_override * Made Omnigen 2 work with optimized_attention_override * Made StableCascade work with optimized_attention_override * Made AuraFlow work with optimized_attention_override * Made Lumina work with optimized_attention_override * Made Chroma work with optimized_attention_override * Made SVD work with optimized_attention_override * Fix WanI2VCrossAttention so that it expects to receive transformer_options * Fixed Wan2.1 Fun Camera transformer_options passthrough * Fixed WAN 2.1 VACE transformer_options passthrough * Add optimized to get_attention_function * Disable attention logs for now * Remove attention logging code * Remove _register_core_attention_functions, as we wouldn't want someone to call that, just in case * Satisfy ruff * Remove AttentionOverrideTest node, that's something to cook up for later	2025-09-12 18:07:38 -04:00
comfyanonymous	33bd9ed9cb	Implement hunyuan image refiner model. (#9817 )	2025-09-12 00:43:20 -04:00
comfyanonymous	e01e99d075	Support hunyuan image distilled model. (#9807 )	2025-09-10 23:17:34 -04:00
comfyanonymous	543888d3d8	Fix lowvram issue with hunyuan image vae. (#9794 )	2025-09-10 02:15:34 -04:00
comfyanonymous	85e34643f8	Support hunyuan image 2.1 regular model. (#9792 )	2025-09-10 02:05:07 -04:00
comfyanonymous	5c33872e2f	Fix issue on old torch. (#9791 )	2025-09-10 00:23:47 -04:00
comfyanonymous	b288fb0db8	Small refactor of some vae code. (#9787 )	2025-09-09 18:09:56 -04:00
comfyanonymous	2ee7879a0b	Fix lowvram issues with hunyuan3d 2.1 (#9735 )	2025-09-05 14:57:35 -04:00
Yousef R. Gamaleldin	261421e218	Add Hunyuan 3D 2.1 Support (#8714 )	2025-09-04 20:36:20 -04:00
comfyanonymous	72855db715	Fix potential rope issue. (#9710 )	2025-09-03 22:20:13 -04:00
comfyanonymous	e3018c2a5a	uso -> uxo/uno as requested. (#9688 )	2025-09-02 16:12:07 -04:00
comfyanonymous	3412d53b1d	USO style reference. (#9677 ) Load the projector.safetensors file with the ModelPatchLoader node and use the siglip_vision_patch14_384.safetensors "clip vision" model and the USOStyleReferenceNode.	2025-09-02 15:36:22 -04:00
comfyanonymous	27e067ce50	Implement the USO subject identity lora. (#9674 ) Use the lora with FluxContextMultiReferenceLatentMethod node set to "uso" and a ReferenceLatent node with the reference image.	2025-09-01 18:54:02 -04:00
comfyanonymous	491755325c	Better s2v memory estimation. (#9584 )	2025-08-27 19:02:42 -04:00
comfyanonymous	496888fd68	Improve s2v performance when generating videos longer than 120 frames. (#9582 )	2025-08-27 16:06:40 -04:00
comfyanonymous	b5ac6ed7ce	Fixes to make controlnet type models work on qwen edit and kontext. (#9581 )	2025-08-27 15:26:28 -04:00
comfyanonymous	88aee596a3	WIP Wan 2.2 S2V model. (#9568 )	2025-08-27 01:10:34 -04:00
Jedrzej Kosinski	fc247150fe	Implement EasyCache and Invent LazyCache (#9496 ) * Attempting a universal implementation of EasyCache, starting with flux as test; I screwed up the math a bit, but when I set it just right it works. * Fixed math to make threshold work as expected, refactored code to use EasyCacheHolder instead of a dict wrapped by object * Use sigmas from transformer_options instead of timesteps to be compatible with a greater amount of models, make end_percent work * Make log statement when not skipping useful, preparing for per-cond caching * Added DIFFUSION_MODEL wrapper around forward function for wan model * Add subsampling for heuristic inputs * Add subsampling to output_prev (output_prev_subsampled now) * Properly consider conds in EasyCache logic * Created SuperEasyCache to test what happens if caching and reuse is moved outside the scope of conds, added PREDICT_NOISE wrapper to facilitate this test * Change max reuse_threshold to 3.0 * Mark EasyCache/SuperEasyCache as experimental (beta) * Make Lumina2 compatible with EasyCache * Add EasyCache support for Qwen Image * Fix missing comma, curse you Cursor * Add EasyCache support to AceStep * Add EasyCache support to Chroma * Added EasyCache support to Cosmos Predict t2i * Make EasyCache not crash with Cosmos Predict ImagToVideo latents, but does not work well at all * Add EasyCache support to hidream * Added EasyCache support to hunyuan video * Added EasyCache support to hunyuan3d * Added EasyCache support to LTXV (not very good, but does not crash) * Implemented EasyCache for aura_flow * Renamed SuperEasyCache to LazyCache, hardcoded subsample_factor to 8 on nodes * Eatra logging when verbose is true for EasyCache	2025-08-22 22:41:08 -04:00
contentis	fe31ad0276	Add elementwise fusions (#9495 ) * Add elementwise fusions * Add addcmul pattern to Qwen	2025-08-22 19:39:15 -04:00
comfyanonymous	ff57793659	Support InstantX Qwen controlnet. (#9488 )	2025-08-22 00:53:11 -04:00
comfyanonymous	f7bd5e58dd	Make it easier to implement future qwen controlnets. (#9485 )	2025-08-21 23:18:04 -04:00
comfyanonymous	0963493a9c	Support for Qwen Diffsynth Controlnets canny and depth. (#9465 ) These are not real controlnets but actually a patch on the model so they will be treated as such. Put them in the models/model_patches/ folder. Use the new ModelPatchLoader and QwenImageDiffsynthControlnet nodes.	2025-08-20 22:26:37 -04:00
comfyanonymous	8d38ea3bbf	Fix bf16 precision issue with qwen image embeddings. (#9441 )	2025-08-20 02:58:54 -04:00
comfyanonymous	7cd2c4bd6a	Qwen rotary embeddings should now match reference code. (#9437 )	2025-08-20 00:45:27 -04:00
comfyanonymous	ed43784b0d	WIP Qwen edit model: The diffusion model part. (#9383 )	2025-08-17 16:45:39 -04:00
comfyanonymous	0f2b8525bc	Qwen image model refactor. (#9375 )	2025-08-16 17:51:28 -04:00
comfyanonymous	1702e6df16	Implement wan2.2 camera model. (#9357 ) Use the old WanCameraImageToVideo node.	2025-08-15 17:29:58 -04:00
comfyanonymous	c308a8840a	Add FluxKontextMultiReferenceLatentMethod node. (#9356 ) This node is only useful if someone trains the kontext model to properly use multiple reference images via the index method. The default is the offset method which feeds the multiple images like if they were stitched together as one. This method works with the current flux kontext model.	2025-08-15 15:50:39 -04:00
comfyanonymous	ad19a069f6	Make SLG nodes work on Qwen Image model. (#9345 )	2025-08-14 23:16:01 -04:00
comfyanonymous	9df8792d4b	Make last PR not crash comfy on old pytorch. (#9324 )	2025-08-13 15:12:41 -04:00
contentis	3da5a07510	SDPA backend priority (#9299 )	2025-08-13 14:53:27 -04:00
comfyanonymous	560d38f34c	Wan2.2 fun control support. (#9292 )	2025-08-12 23:26:33 -04:00
comfyanonymous	d044a24398	Fix default shift and any latent size for qwen image model. (#9186 )	2025-08-05 06:12:27 -04:00
comfyanonymous	c012400240	Initial support for qwen image model. (#9179 )	2025-08-04 22:53:25 -04:00
comfyanonymous	1e638a140b	Tiny wan vae optimizations. (#9136 )	2025-08-01 05:25:38 -04:00
chaObserv	61b08d4ba6	Replace manual x * sigmoid(x) with torch silu in VAE nonlinearity (#9057 )	2025-07-30 19:25:56 -04:00
comfyanonymous	da9dab7edd	Small wan camera memory optimization. (#9111 )	2025-07-30 05:55:26 -04:00
comfyanonymous	dca6bdd4fa	Make wan2.2 5B i2v take a lot less memory. (#9102 )	2025-07-29 19:44:18 -04:00
comfyanonymous	c60dc4177c	Remove unecessary clones in the wan2.2 VAE. (#9083 )	2025-07-28 14:48:19 -04:00
comfyanonymous	a88788dce6	Wan 2.2 support. (#9080 )	2025-07-28 08:00:23 -04:00
comfyanonymous	0621d73a9c	Remove useless code. (#9059 )	2025-07-26 04:44:19 -04:00
comfyanonymous	e6e5d33b35	Remove useless code. (#9041 ) This is only needed on old pytorch 2.0 and older.	2025-07-25 04:58:28 -04:00
Harel Cain	9bc2798f72	LTXV VAE decoder: switch default padding mode (#8930 )	2025-07-16 13:54:38 -04:00
comfyanonymous	938d3e8216	Remove windows line endings. (#8866 )	2025-07-11 02:37:51 -04:00
josephrocca	974254218a	Un-hardcode chroma patch_size (#8840 )	2025-07-08 15:56:59 -04:00
comfyanonymous	9093301a49	Don't add tiny bit of random noise when VAE encoding. (#8705 ) Shouldn't change outputs but might make things a tiny bit more deterministic.	2025-06-27 14:14:56 -04:00
comfyanonymous	ef5266b1c1	Support Flux Kontext Dev model. (#8679 )	2025-06-26 11:28:41 -04:00
comfyanonymous	ec70ed6aea	Omnigen2 model implementation. (#8669 )	2025-06-25 19:35:57 -04:00
comfyanonymous	f7fb193712	Small flux optimization. (#8611 )	2025-06-20 05:37:32 -04:00
comfyanonymous	7e9267fa77	Make flux controlnet work with sd3 text enc. (#8599 )	2025-06-19 18:50:05 -04:00
comfyanonymous	91d40086db	Fix pytorch warning. (#8593 )	2025-06-19 11:04:52 -04:00
comfyanonymous	7ea79ebb9d	Add correct eps to ltxv rmsnorm. (#8542 )	2025-06-15 12:21:25 -04:00
comfyanonymous	29596bd53f	Small cosmos attention code refactor. (#8530 )	2025-06-14 05:02:05 -04:00
Kohaku-Blueleaf	520eb77b72	LoRA Trainer: LoRA training node in weight adapter scheme (#8446 )	2025-06-13 19:25:59 -04:00
comfyanonymous	c69af655aa	Uncap cosmos predict2 res and fix mem estimation. (#8518 )	2025-06-13 07:30:18 -04:00
comfyanonymous	251f54a2ad	Basic initial support for cosmos predict2 text to image 2B and 14B models. (#8517 )	2025-06-13 07:05:23 -04:00
comfyanonymous	8a4ff747bd	Fix mistake in last commit. (#8496 ) * Move to right place.	2025-06-11 15:13:29 -04:00
comfyanonymous	af1eb58be8	Fix black images on some flux models in fp16. (#8495 )	2025-06-11 15:09:11 -04:00
comfyanonymous	4248b1618f	Let chroma TE work on regular flux. (#8429 )	2025-06-05 10:07:17 -04:00
drhead	08b7cc7506	use fused multiply-add pointwise ops in chroma (#8279 )	2025-05-30 18:09:54 -04:00
comfyanonymous	5e5e46d40c	Not really tested WAN Phantom Support. (#8321 )	2025-05-28 23:46:15 -04:00
comfyanonymous	5a87757ef9	Better error if sageattention is installed but a dependency is missing. (#8264 )	2025-05-24 06:43:12 -04:00
drhead	30b2eb8a93	create arange on-device (#8255 )	2025-05-23 16:15:06 -04:00
comfyanonymous	f85c08df06	Make VACE conditionings stackable. (#8240 )	2025-05-22 19:22:26 -04:00
comfyanonymous	87f9130778	Revert "This doesn't seem to be needed on chroma. (#8209 )" (#8210 ) This reverts commit `7e84bf5373`.	2025-05-20 05:39:55 -04:00
comfyanonymous	7e84bf5373	This doesn't seem to be needed on chroma. (#8209 )	2025-05-20 05:29:23 -04:00
George0726	c820ef950d	Add Wan-FUN Camera Control models and Add WanCameraImageToVideo node (#8013 ) * support wan camera models * fix by ruff check * change camera_condition type; make camera_condition optional * support camera trajectory nodes * fix camera direction --------- Co-authored-by: Qirui Sun <sunqr0667@126.com>	2025-05-15 19:00:43 -04:00
comfyanonymous	4a9014e201	Hunyuan Custom initial untested implementation. (#8101 )	2025-05-13 15:53:47 -04:00
comfyanonymous	640c47e7de	Fix torch warning about deprecated function. (#8075 ) Drop support for torch versions below 2.2 on the audio VAEs.	2025-05-12 14:32:01 -04:00
blepping	42da274717	Use normal ComfyUI attention in ACE-Steps model (#8023 ) * Use normal ComfyUI attention in ACE-Steps model * Let optimized_attention handle output reshape for ACE	2025-05-09 13:51:02 -04:00
comfyanonymous	fd08e39588	Make torchaudio not a hard requirement. (#7987 ) Some platforms can't install it apparently so if it's not there it should only break models that actually use it.	2025-05-07 21:37:12 -04:00
comfyanonymous	cc33cd3422	Experimental lyrics strength for ACE. (#7984 )	2025-05-07 19:22:07 -04:00
comfyanonymous	16417b40d9	Initial ACE-Step model implementation. (#7972 )	2025-05-07 08:33:34 -04:00
comfyanonymous	80a44b97f5	Change lumina to native RMSNorm. (#7935 )	2025-05-04 06:39:23 -04:00
comfyanonymous	9187a09483	Change cosmos and hydit models to use the native RMSNorm. (#7934 )	2025-05-04 06:26:20 -04:00
comfyanonymous	3041e5c354	Switch mochi and wan modes to use pytorch RMSNorm. (#7925 ) * Switch genmo model to native RMSNorm. * Switch WAN to native RMSNorm.	2025-05-03 19:07:55 -04:00
comfyanonymous	aa9d759df3	Switch ltxv to use the pytorch RMSNorm. (#7897 )	2025-05-01 06:33:42 -04:00
comfyanonymous	08ff5fa08a	Cleanup chroma PR.	2025-04-30 20:57:30 -04:00
Silver	4ca3d84277	Support for Chroma - Flux1 Schnell distilled with CFG (#7355 ) * Upload files for Chroma Implementation * Remove trailing whitespace * trim more trailing whitespace..oops * remove unused imports * Add supported_inference_dtypes * Set min_length to 0 and remove attention_mask=True * Set min_length to 1 * get_mdulations added from blepping and minor changes * Add lora conversion if statement in lora.py * Update supported_models.py * update model_base.py * add uptream commits * set modelType.FLOW, will cause beta scheduler to work properly * Adjust memory usage factor and remove unnecessary code * fix mistake * reduce code duplication * remove unused imports * refactor for upstream sync * sync chroma-support with upstream via syncbranch patch * Update sd.py * Add Chroma as option for the OptimalStepsScheduler node	2025-04-30 20:57:00 -04:00
comfyanonymous	dbc726f80c	Better vace memory estimation. (#7875 )	2025-04-29 20:42:00 -04:00
comfyanonymous	83d04717b6	Support HiDream E1 model. (#7857 )	2025-04-28 15:01:15 -04:00
comfyanonymous	3ab231f01f	Fix issue with WAN VACE implementation. (#7724 )	2025-04-21 23:36:12 -04:00
comfyanonymous	5d0d4ee98a	Add strength control for vace. (#7717 )	2025-04-21 19:36:20 -04:00
comfyanonymous	ce22f687cc	Support for WAN VACE preview model. (#7711 ) * Support for WAN VACE preview model. * Remove print.	2025-04-21 14:40:29 -04:00
comfyanonymous	dbcfd092a2	Set default context_img_len to 257	2025-04-17 12:42:34 -04:00
comfyanonymous	c14429940f	Support loading WAN FLF model.	2025-04-17 12:04:48 -04:00
comfyanonymous	0d720e4367	Don't hardcode length of context_img in wan code.	2025-04-17 06:25:39 -04:00
comfyanonymous	1fc00ba4b6	Make hidream work with any latent resolution.	2025-04-16 18:34:14 -04:00
comfyanonymous	f00f340a56	Reuse code from flux model.	2025-04-16 17:43:55 -04:00
comfyanonymous	9ad792f927	Basic support for hidream i1 model.	2025-04-15 17:35:05 -04:00
comfyanonymous	8a438115fb	add RMSNorm to comfy.ops	2025-04-14 18:00:33 -04:00
Raphael Walker	89e4ea0175	Add activations_shape info in UNet models (#7482 ) * Add activations_shape info in UNet models * activations_shape should be a list	2025-04-04 21:27:54 -04:00
comfyanonymous	e471c726e5	Fallback to pytorch attention if sage attention fails.	2025-03-22 15:45:56 -04:00
comfyanonymous	32ca0805b7	Fix orientation of hunyuan 3d model.	2025-03-19 19:55:24 -04:00
comfyanonymous	11f1b41bab	Initial Hunyuan3Dv2 implementation. Supports the multiview, mini, turbo models and VAEs.	2025-03-19 16:52:58 -04:00
comfyanonymous	3b19fc76e3	Allow disabling pe in flux code for some other models.	2025-03-18 05:09:25 -04:00
comfyanonymous	e8e990d6b8	Cleanup code.	2025-03-16 06:29:12 -04:00
comfyanonymous	a2448fc527	Remove useless code.	2025-03-14 18:10:37 -04:00
comfyanonymous	6a0daa79b6	Make the SkipLayerGuidanceDIT node work on WAN.	2025-03-14 10:55:19 -04:00
FeepingCreature	9c98c6358b	Tolerate missing `@torch.library.custom_op` (#7234 ) This can happen on Pytorch versions older than 2.4.	2025-03-14 09:51:26 -04:00
FeepingCreature	7aceb9f91c	Add --use-flash-attention flag. (#7223 ) * Add --use-flash-attention flag. This is useful on AMD systems, as FA builds are still 10% faster than Pytorch cross-attention.	2025-03-14 03:22:41 -04:00
comfyanonymous	9aac21f894	Fix issues with new hunyuan img2vid model and bumb version to v0.3.26	2025-03-09 05:07:22 -04:00
comfyanonymous	7395b0c0d1	Support new hunyuan video i2v model. Use the new "v2 (replace)" guidance type in HunyuanImageToVideo and set image_interleave to 4 on the "Text Encode Hunyuan Video" node.	2025-03-08 20:34:47 -05:00
comfyanonymous	0952569493	Fix stable cascade VAE on some lowvram machines.	2025-03-08 20:24:04 -05:00
comfyanonymous	369b079ff6	Fix lowvram issue with ltxv vae.	2025-03-05 05:26:08 -05:00
comfyanonymous	93fedd92fe	Support LTXV 0.9.5. Credits: Lightricks team.	2025-03-05 00:13:49 -05:00
comfyanonymous	f4dac8ab6f	Wan code small cleanup.	2025-02-27 07:22:42 -05:00
comfyanonymous	3ea3bc8546	Fix wan issues when prompt length is long.	2025-02-26 20:34:02 -05:00
comfyanonymous	0270a0b41c	Reduce artifacts on Wan by doing the patch embedding in fp32.	2025-02-26 16:59:26 -05:00
comfyanonymous	4bca7367f3	Don't try to use clip_fea on t2v model.	2025-02-26 08:38:09 -05:00
comfyanonymous	fa62287f1f	More code reuse in wan. Fix bug when changing the compute dtype on wan.	2025-02-26 05:22:29 -05:00
comfyanonymous	4ced06b879	WIP support for Wan I2V model.	2025-02-26 01:49:43 -05:00
comfyanonymous	9a66bb972d	Make wan work with all latent resolutions. Cleanup some code.	2025-02-25 19:56:04 -05:00
comfyanonymous	ea0f939df3	Fix issue with wan and other attention implementations.	2025-02-25 19:13:39 -05:00
comfyanonymous	f37551c1d2	Change wan rope implementation to the flux one. Should be more compatible.	2025-02-25 19:11:14 -05:00
comfyanonymous	63023011b9	WIP support for Wan t2v model.	2025-02-25 17:20:35 -05:00
comfyanonymous	96d891cb94	Speedup on some models by not upcasting bfloat16 to float32 on mac.	2025-02-24 05:41:32 -05:00
comfyanonymous	aff16532d4	Remove some useless code.	2025-02-22 04:45:14 -05:00
Jukka Seppänen	acc152b674	Support loading and using SkyReels-V1-Hunyuan-I2V (#6862 ) * Support SkyReels-V1-Hunyuan-I2V * VAE scaling * Fix T2V oops * Proper latent scaling	2025-02-18 17:06:54 -05:00
comfyanonymous	1cd6cd6080	Disable pytorch attention in VAE for AMD.	2025-02-14 05:42:14 -05:00
HishamC	b124256817	Fix for running via DirectML (#6542 ) * Fix for running via DirectML Fix DirectML empty image generation issue with Flux1. add CPU fallback for unsupported path. Verified the model works on AMD GPUs * fix formating * update casual mask calculation	2025-02-11 17:11:32 -05:00
comfyanonymous	4027466c80	Make lumina model work with any latent resolution.	2025-02-10 00:24:20 -05:00
comfyanonymous	14880e6dba	Remove some useless code.	2025-02-06 05:00:37 -05:00
comfyanonymous	94f21f9301	Upcasting rope to fp32 seems to make no difference in this model.	2025-02-05 04:32:47 -05:00
comfyanonymous	60653004e5	Use regular numbers for rope in lumina model.	2025-02-05 04:17:25 -05:00
comfyanonymous	a57d635c5f	Fix lumina 2 batches.	2025-02-04 21:48:11 -05:00
comfyanonymous	3e880ac709	Fix on python 3.9	2025-02-04 04:20:56 -05:00
comfyanonymous	e5ea112a90	Support Lumina 2 model.	2025-02-04 04:16:30 -05:00
Dr.Lt.Data	0a0df5f136	better guide message for sageattention (#6634 )	2025-02-02 09:26:47 -05:00
comfyanonymous	96e2a45193	Remove useless code.	2025-01-23 05:56:23 -05:00
comfyanonymous	fb2ad645a3	Add FluxDisableGuidance node to disable using the guidance embed.	2025-01-20 14:50:24 -05:00
comfyanonymous	0aa2368e46	Fix some cosmos fp8 issues.	2025-01-16 17:45:37 -05:00
comfyanonymous	cca96a85ae	Fix cosmos VAE failing with videos longer than 121 frames.	2025-01-16 16:30:06 -05:00
comfyanonymous	31831e6ef1	Code refactor.	2025-01-16 07:23:54 -05:00
comfyanonymous	23289a6a5c	Clean up some debug lines.	2025-01-16 04:24:39 -05:00
comfyanonymous	6320d05696	Slightly lower hunyuan video memory usage.	2025-01-16 00:23:01 -05:00
comfyanonymous	25683b5b02	Lower cosmos diffusion model memory usage.	2025-01-15 23:46:42 -05:00
comfyanonymous	4758fb64b9	Lower cosmos VAE memory usage by a bit.	2025-01-15 22:57:52 -05:00
comfyanonymous	008761166f	Optimize first attention block in cosmos VAE.	2025-01-15 21:48:46 -05:00
comfyanonymous	1f1c7b7b56	Remove useless code.	2025-01-13 03:52:37 -05:00
comfyanonymous	2ff3104f70	WIP support for Nvidia Cosmos 7B and 14B text to world (video) models.	2025-01-10 09:14:16 -05:00
comfyanonymous	129d8908f7	Add argument to skip the output reshaping in the attention functions.	2025-01-10 06:27:37 -05:00
comfyanonymous	ff838657fa	Cleaner handling of attention mask in ltxv model code.	2025-01-09 07:12:03 -05:00
comfyanonymous	79eea51a1d	Fix and enforce all ruff W rules.	2025-01-01 03:08:33 -05:00
comfyanonymous	b7572b2f87	Fix and enforce no trailing whitespace.	2024-12-31 03:16:37 -05:00
comfyanonymous	d9b7cfac7e	Fix and enforce new lines at the end of files.	2024-12-30 04:14:59 -05:00
comfyanonymous	d170292594	Remove some trailing white space.	2024-12-27 18:02:30 -05:00
comfyanonymous	e44d0ac7f7	Make --novram completely offload weights. This flag is mainly used for testing the weight offloading, it shouldn't actually be used in practice. Remove useless import.	2024-12-23 01:51:08 -05:00
comfyanonymous	56bc64f351	Comment out some useless code.	2024-12-22 23:51:14 -05:00
zhangp365	f7d83b72e0	fixed a bug in ldm/pixart/blocks.py (#6158 )	2024-12-22 23:44:20 -05:00
comfyanonymous	80f07952d2	Fix lowvram issue with ltxv vae.	2024-12-22 23:20:17 -05:00
comfyanonymous	da13b6b827	Get rid of meshgrid warning.	2024-12-20 18:02:12 -05:00
comfyanonymous	c86cd58573	Remove useless code.	2024-12-20 17:50:03 -05:00
comfyanonymous	b5fe39211a	Remove some useless code.	2024-12-20 17:43:50 -05:00
comfyanonymous	e946667216	Some fixes/cleanups to pixart code. Commented out the masking related code because it is never used in this implementation.	2024-12-20 17:10:52 -05:00
Chenlei Hu	d7969cb070	Replace print with logging (#6138 ) * Replace print with logging * nit * nit * nit * nit * nit * nit	2024-12-20 16:24:55 -05:00
City	bddb02660c	Add PixArt model support (#6055 ) * PixArt initial version * PixArt Diffusers convert logic * pos_emb and interpolation logic * Reduce duplicate code * Formatting * Use optimized attention * Edit empty token logic * Basic PixArt LoRA support * Fix aspect ratio logic * PixArtAlpha text encode with conds * Use same detection key logic for PixArt diffusers	2024-12-20 15:25:00 -05:00
comfyanonymous	418eb7062d	Support new LTXV VAE.	2024-12-20 04:38:29 -05:00
comfyanonymous	cbbf077593	Small optimizations.	2024-12-18 18:23:28 -05:00
comfyanonymous	4c5c4ddeda	Fix regression in VAE code on old pytorch versions.	2024-12-18 03:08:28 -05:00
comfyanonymous	37e5390f5f	Add: --use-sage-attention to enable SageAttention. You need to have the library installed first.	2024-12-18 01:56:10 -05:00
comfyanonymous	bda1482a27	Basic Hunyuan Video model support.	2024-12-16 19:35:40 -05:00
comfyanonymous	19ee5d9d8b	Don't expand mask when not necessary. Expanding seems to slow down inference.	2024-12-16 18:22:50 -05:00
Raphael Walker	61b50720d0	Add support for attention masking in Flux (#5942 ) * fix attention OOM in xformers * allow passing attention mask in flux attention * allow an attn_mask in flux * attn masks can be done using replace patches instead of a separate dict * fix return types * fix return order * enumerate * patch the right keys * arg names * fix a silly bug * fix xformers masks * replace match with if, elif, else * mask with image_ref_size * remove unused import * remove unused import 2 * fix pytorch/xformers attention This corrects a weird inconsistency with skip_reshape. It also allows masks of various shapes to be passed, which will be automtically expanded (in a memory-efficient way) to a size that is compatible with xformers or pytorch sdpa respectively. * fix mask shapes	2024-12-16 18:21:17 -05:00
comfyanonymous	e83063bf24	Support conv3d in PatchEmbed.	2024-12-14 05:46:04 -05:00
comfyanonymous	4e14032c02	Make pad_to_patch_size function work on multi dim.	2024-12-13 07:22:05 -05:00
Chenlei Hu	563291ee51	Enforce all pyflake lint rules (#6033 ) * Enforce F821 undefined-name * Enforce all pyflake lint rules	2024-12-12 19:29:37 -05:00
Chenlei Hu	2cddbf0821	Lint and fix undefined names (1/N) (#6028 )	2024-12-12 18:55:26 -05:00
Chenlei Hu	60749f345d	Lint and fix undefined names (3/N) (#6030 )	2024-12-12 18:49:40 -05:00
Chenlei Hu	d9d7f3c619	Lint all unused variables (#5989 ) * Enable F841 * Autofix * Remove all unused variable assignment	2024-12-12 17:59:16 -05:00
Chenlei Hu	0fd4e6c778	Lint unused import (#5973 ) * Lint unused import * nit * Remove unused imports * revert fix_torch import * nit	2024-12-09 15:24:39 -05:00
comfyanonymous	8af9a91e0c	A few improvements to #5937 .	2024-12-06 05:49:15 -05:00
Michael Kupchick	005d2d3a13	ltxv: add noise to guidance image to ensure generated motion. (#5937 )	2024-12-06 05:46:08 -05:00
Jedrzej Kosinski	0ee322ec5f	ModelPatcher Overhaul and Hook Support (#5583 ) * Added hook_patches to ModelPatcher for weights (model) * Initial changes to calc_cond_batch to eventually support hook_patches * Added current_patcher property to BaseModel * Consolidated add_hook_patches_as_diffs into add_hook_patches func, fixed fp8 support for model-as-lora feature * Added call to initialize_timesteps on hooks in process_conds func, and added call prepare current keyframe on hooks in calc_cond_batch * Added default_conds support in calc_cond_batch func * Added initial set of hook-related nodes, added code to register hooks for loras/model-as-loras, small renaming/refactoring * Made CLIP work with hook patches * Added initial hook scheduling nodes, small renaming/refactoring * Fixed MaxSpeed and default conds implementations * Added support for adding weight hooks that aren't registered on the ModelPatcher at sampling time * Made Set Clip Hooks node work with hooks from Create Hook nodes, began work on better Create Hook Model As LoRA node * Initial work on adding 'model_as_lora' lora type to calculate_weight * Continued work on simpler Create Hook Model As LoRA node, started to implement ModelPatcher callbacks, attachments, and additional_models * Fix incorrect ref to create_hook_patches_clone after moving function * Added injections support to ModelPatcher + necessary bookkeeping, added additional_models support in ModelPatcher, conds, and hooks * Added wrappers to ModelPatcher to facilitate standardized function wrapping * Started scaffolding for other hook types, refactored get_hooks_from_cond to organize hooks by type * Fix skip_until_exit logic bug breaking injection after first run of model * Updated clone_has_same_weights function to account for new ModelPatcher properties, improved AutoPatcherEjector usage in partially_load * Added WrapperExecutor for non-classbound functions, added calc_cond_batch wrappers * Refactored callbacks+wrappers to allow storing lists by id * Added forward_timestep_embed_patch type, added helper functions on ModelPatcher for emb_patch and forward_timestep_embed_patch, added helper functions for removing callbacks/wrappers/additional_models by key, added custom_should_register prop to hooks * Added get_attachment func on ModelPatcher * Implement basic MemoryCounter system for determing with cached weights due to hooks should be offloaded in hooks_backup * Modified ControlNet/T2IAdapter get_control function to receive transformer_options as additional parameter, made the model_options stored in extra_args in inner_sample be a clone of the original model_options instead of same ref * Added create_model_options_clone func, modified type annotations to use __future__ so that I can use the better type annotations * Refactored WrapperExecutor code to remove need for WrapperClassExecutor (now gone), added sampler.sample wrapper (pending review, will likely keep but will see what hacks this could currently let me get rid of in ACN/ADE) * Added Combine versions of Cond/Cond Pair Set Props nodes, renamed Pair Cond to Cond Pair, fixed default conds never applying hooks (due to hooks key typo) * Renamed Create Hook Model As LoRA nodes to make the test node the main one (more changes pending) * Added uuid to conds in CFGGuider and uuids to transformer_options to allow uniquely identifying conds in batches during sampling * Fixed models not being unloaded properly due to current_patcher reference; the current ComfyUI model cleanup code requires that nothing else has a reference to the ModelPatcher instances * Fixed default conds not respecting hook keyframes, made keyframes not reset cache when strength is unchanged, fixed Cond Set Default Combine throwing error, fixed model-as-lora throwing error during calculate_weight after a recent ComfyUI update, small refactoring/scaffolding changes for hooks * Changed CreateHookModelAsLoraTest to be the new CreateHookModelAsLora, rename old ones as 'direct' and will be removed prior to merge * Added initial support within CLIP Text Encode (Prompt) node for scheduling weight hook CLIP strength via clip_start_percent/clip_end_percent on conds, added schedule_clip toggle to Set CLIP Hooks node, small cleanup/fixes * Fix range check in get_hooks_for_clip_schedule so that proper keyframes get assigned to corresponding ranges * Optimized CLIP hook scheduling to treat same strength as same keyframe * Less fragile memory management. * Make encode_from_tokens_scheduled call cleaner, rollback change in model_patcher.py for hook_patches_backup dict * Fix issue. * Remove useless function. * Prevent and detect some types of memory leaks. * Run garbage collector when switching workflow if needed. * Moved WrappersMP/CallbacksMP/WrapperExecutor to patcher_extension.py * Refactored code to store wrappers and callbacks in transformer_options, added apply_model and diffusion_model.forward wrappers * Fix issue. * Refactored hooks in calc_cond_batch to be part of get_area_and_mult tuple, added extra_hooks to ControlBase to allow custom controlnets w/ hooks, small cleanup and renaming * Fixed inconsistency of results when schedule_clip is set to False, small renaming/typo fixing, added initial support for ControlNet extra_hooks to work in tandem with normal cond hooks, initial work on calc_cond_batch merging all subdicts in returned transformer_options * Modified callbacks and wrappers so that unregistered types can be used, allowing custom_nodes to have their own unique callbacks/wrappers if desired * Updated different hook types to reflect actual progress of implementation, initial scaffolding for working WrapperHook functionality * Fixed existing weight hook_patches (pre-registered) not working properly for CLIP * Removed Register/Direct hook nodes since they were present only for testing, removed diff-related weight hook calculation as improved_memory removes unload_model_clones and using sample time registered hooks is less hacky * Added clip scheduling support to all other native ComfyUI text encoding nodes (sdxl, flux, hunyuan, sd3) * Made WrapperHook functional, added another wrapper/callback getter, added ON_DETACH callback to ModelPatcher * Made opt_hooks append by default instead of replace, renamed comfy.hooks set functions to be more accurate * Added apply_to_conds to Set CLIP Hooks, modified relevant code to allow text encoding to automatically apply hooks to output conds when apply_to_conds is set to True * Fix cached_hook_patches not respecting target_device/memory_counter results * Fixed issue with setting weights from hooks instead of copying them, added additional memory_counter check when caching hook patches * Remove unnecessary torch.no_grad calls for hook patches * Increased MemoryCounter minimum memory to leave free by 2 until a better way to get inference memory estimate of currently loaded models exists For encode_from_tokens_scheduled, allow start_percent and end_percent in add_dict to limit which scheduled conds get encoded for optimization purposes * Removed a .to call on results of calculate_weight in patch_hook_weight_to_device that was screwing up the intermediate results for fp8 prior to being passed into stochastic_rounding call * Made encode_from_tokens_scheduled work when no hooks are set on patcher * Small cleanup of comments * Turn off hook patch caching when only 1 hook present in sampling, replace some current_hook = None with calls to self.patch_hooks(None) instead to avoid a potential edge case * On Cond/Cond Pair nodes, removed opt_ prefix from optional inputs * Allow both FLOATS and FLOAT for floats_strength input * Revert change, does not work * Made patch_hook_weight_to_device respect set_func and convert_func * Make discard_model_sampling True by default * Add changes manually from 'master' so merge conflict resolution goes more smoothly * Cleaned up text encode nodes with just a single clip.encode_from_tokens_scheduled call * Make sure encode_from_tokens_scheduled will respect use_clip_schedule on clip * Made nodes in nodes_hooks be marked as experimental (beta) * Add get_nested_additional_models for cases where additional_models could have their own additional_models, and add robustness for circular additional_models references * Made finalize_default_conds area math consistent with other sampling code * Changed 'opt_hooks' input of Cond/Cond Pair Set Default Combine nodes to 'hooks' * Remove a couple old TODO's and a no longer necessary workaround	2024-12-02 14:51:02 -05:00
comfyanonymous	95d8713482	Missing parentheses.	2024-11-27 13:45:32 -05:00
comfyanonymous	b4526d3fc3	Skip layer guidance now works on hydit model.	2024-11-24 05:54:30 -05:00
comfyanonymous	ab885b33ba	Skip layer guidance node now works on LTX-Video.	2024-11-23 10:33:05 -05:00
comfyanonymous	e5c3f4b87f	LTXV lowvram fixes.	2024-11-22 17:17:11 -05:00

... 2 3 4 5 6 ...

576 Commits