IF YOU ARE WONDERING WHAT V2 IS OR YOU'RE HERE BECAUSE OF MISMATCHES AFTER UPDATING COMFY THE FIX IS IN!
Perfil de ejecución
Descripción de la fuente
IF YOU ARE WONDERING WHAT V2 IS OR YOU'RE HERE BECAUSE OF MISMATCHES AFTER UPDATING COMFY THE FIX IS IN!
UPDATE KJNODES AND COMFY!!!!!
WE NOW PUT THE EMBEDDINGS INTO THE MODEL LOADER AND THE CLIP BUT IT IS ONLY LOADED ONCE! DOES NOT USE MORE MEMORY!
UPDATE KJNODES AND COMFY!!!!!
THIS WAS A NECESSARY CHANGE AFTER COMFY REORGANIZED WHERE THE EMBEDDING MODEL WOULD LOAD INTO. REQUIRED UPDATE!
PLEASE TAKE NOTE OF NEW DEV+LORA COMBO! WE NOW USE FP4 GEMMA TEXT ENCODER!!!! CHECK MODELS!!! WE NOW HAVE PREVIEWS USING TINY VAE!!! CHECK MODELS!!!! CHECK MODELS!!! DID I MENTION TO CHECK ALL YOUR MODELS!!! DO EEEEEEET!
Comentario generado por IA
Explicación generada por IA basada en los detalles de la fuente y la configuración. Las sugerencias están claramente identificadas.
This ComfyUI workflow is for generating LTX-2 video from a text prompt, image, short video clip, or audio paired with text or an image; it supports text-to-video, image-to-video, video extension, text-plus-audio video, and image-plus-audio lip-sync, with video outputs.
The five described paths are t2v, i2v, v2v extend, ta2v, and ia2v.
Use v2v extend with a few seconds of video and a prompt to continue it; ta2v uses text and audio, while ia2v uses an image and audio.
Text-to-video uses a text prompt; image-to-video uses an image; video extension uses a few seconds of video and a prompt; audio paths use audio with text or an image.
The described outputs are generated videos, including audio-driven video and image-plus-audio lip-sync results.
The description lists at least 12 GB of VRAM and 48 GB of system RAM; compare those figures with your setup before use.
Sugerencia · sin verificar
Named non-optional files include: gemma_3_12B_it_fp4_mixed.safetensors; ltx-2-19b-dev-Q4_K_M.gguf; ltx-2-19b-distilled-lora-384.safetensors; and ltx-2-19b-embeddings_connector_dev_bf16.safetensors.
Other named non-optional files are: ltx-2-spatial-upscaler-x2-1.0.safetensors; LTX2_audio_vae_bf16.safetensors; LTX2_video_vae_bf16.safetensors; and taeltx_2.safetensors.
The named node pack is comfyui-kjnodes.
The V2 notes instruct you to update KJNodes and Comfy and to load embeddings through the model loader and CLIP once.
The description calls out an FP4 Gemma text encoder, a dev-plus-LoRA combination, Tiny VAE previews, and two audio-enhancement nodes near the end.
Before setup, verify the intended source and version for ltx-2-19b-distilled-lora-384.safetensors and ltx-2-spatial-upscaler-x2-1.0.safetensors; no checksum is listed for ltx-2-19b-embeddings_connector_dev_bf16.safetensors, LTX2_video_vae_bf16.safetensors, or taeltx_2.safetensors
Sugerencia · sin verificar
No license is specified on the source page; check the applicable terms before use.
Sugerencia · sin verificar
¿Es necesario editarlo?
Inicia sesión para solicitar una edición.
Fuentes
1 fuenteExtractos de fuentes
1 extractoSource context: 20067 downloads · Type Workflows · Base model LTXV2
WE ALSO HAVE THE LORAS SETUP CORRECTLY AND THERE ARE SOME FUN ONES OUT ALREADY! NODE IS READY TO GO FOR YOU!
5 TOTAL GGUF 12GB WORKFLOWS!
t2v, i2v, v2v extend, ta2v, ia2v!
Hello everyone! This workflow has come a long way since 1.0 actually. It doesn't seem like it when you first look but, boy this has been a project for me!
Here we have quite a few workflows for LTX-2 using GGUF and running on at least 12GB VRAM and 48GB system ram.
First we have your typical t2v and i2v workflows.
Second we now have two new audio driven workflows! ta2v which is supply a text prompt ONLY and some audio get a neat generated video with your audio! The other is an ia2v where you supply an image and an audio file and it lip-syncs up nicely. I tried to keep everything as simple as possible.
Then the one I like the most v2v extend. Feed ltx2 a few seconds of video, create a prompt to continue the video and watch the magic happen!!
I got done with the workflows, I now need to get all the info out there but I wanted to get these into the wild so everyone can start having fun with them!
I HAVE CREATED TWO ENHANCEMENT NODES FOR THE AUDIO!!
YOU WILL NOTICE 2 NEW NODES TOWARD THE END OF THE WORKFLOW FOR AUDIO ENHANCEMENT. CLICK THE BLUE LINK BELOW FOR MY GITHUB PAGE, INSTALLATION INSTRUCTIONS, AND USEAGE NOTES!
URABEWE-COMFYUI-AUDIOTOOLS
Requisito de VRAM estimado
Estimación no disponible
20,0 GB en 3 de 8 archivos de modelos. Total de archivos de modelos + 25 % de sobrecarga de carga + 2 GB de margen de ejecución, redondeado hacia arriba.
Requisitos
9 requisitosgemma_3_12B_it_fp4_mixed.safetensors
Archivo sin verificarText encoder · Civitai · 2817877
Checkpoint · 12.0 GB · Hugging Face · unsloth/LTX-2-GGUF
ltx-2-19b-distilled-lora-384.safetensors
EncontradoLoRA · 7.15 GB · Hugging Face · Lightricks/LTX-2
ltx-2-19b-embeddings_connector_dev_bf16.safetensors
No resueltoText encoder · Unknown
Upscaler · 950 MB · Hugging Face · Lightricks/LTX-2
LTX2_audio_vae_bf16.safetensors
Archivo sin verificarVAE · Hugging Face · novoluz/ltx2_audio_vae_bf16 · Ltx2 Audio VAE BF16
LTX2_video_vae_bf16.safetensors
No resueltoVAE · Unknown
taeltx_2.safetensors
No resueltoVAE · Unknown
Paquete de nodos · Registry