EasyAI代码托管平台

mirror of https://github.com/comfyanonymous/ComfyUI.git synced 2026-03-09 11:17:42 +08:00

History

rattus 0fd1b78736 Reduce LTX2 VAE VRAM consumption (#12028 ) * causal_video_ae: Remove attention ResNet This attention_head_dim argument does not exist on this constructor so this is dead code. Remove as generic attention mid VAE conflicts with temporal roll. * ltx-vae: consoldate causal/non-causal code paths * ltx-vae: add cache rolling adder * ltx-vae: use cached adder for resnet * ltx-vae: Implement rolling VAE Implement a temporal rolling VAE for the LTX2 VAE. Usually when doing temporal rolling VAEs you can just chunk on time relying on causality and cache behind you as you go. The LTX VAE is however non-causal. So go whole hog and implement per layer run ahead and backpressure between the decoder layers using recursive state beween the layers. Operations are ammended with temporal_cache_state{} which they can use to hold any state then need for partial execution. Convolutions cache their inputs behind the up to N-1 frames, and skip connections need to cache the mismatch between convolution input and output that happens due to missing future (non-causal) input. Each call to run_up() processes a layer accross a range on input that may or may not be complete. It goes depth first to process as much as possible to try and digest frames to the final output ASAP. If layers run out of input due to convolution losses, they simply return without action effectively applying back-pressure to the earlier layers. As the earlier layers do more work and caller deeper, the partial states are reconciled and output continues to digest depth first as much as possible. Chunking is done using a size quota rather than a fixed frame length and any layer can initiate chunking, and multiple layers can chunk at different granulatiries. This remove the old limitation of always having to process 1 latent frame to entirety and having to hold 8 full decoded frames as the VRAM peak.		2026-01-22 16:54:18 -05:00
..
vae	Reduce LTX2 VAE VRAM consumption (#12028 )	2026-01-22 16:54:18 -05:00
vocoders	Support the LTXV 2 model. (#11632 )	2026-01-05 01:58:59 -05:00
av_model.py	Reduce LTX2 VRAM use by more efficient timestep embed handling (#11829 )	2026-01-12 17:28:59 -05:00
embeddings_connector.py	Fix lowvram issue with ltxv2 text encoder. (#11675 )	2026-01-06 17:33:03 -05:00
latent_upsampler.py	Support the LTXV 2 model. (#11632 )	2026-01-05 01:58:59 -05:00
model.py	Support the LTXV 2 model. (#11632 )	2026-01-05 01:58:59 -05:00
symmetric_patchifier.py	Support the LTXV 2 model. (#11632 )	2026-01-05 01:58:59 -05:00