This is a simple multi model workflow I use to generate short videos.
Runtime profile
Source description
This is a simple multi model workflow I use to generate short videos.
Text2Images2Videos2Video.
1 - FLUX to create 3 similar images from similar prompts
2 - CogVideoX 1.5 5B I2V to create two "interpolations" between those images
3 - Hunyuan V2V to solve inconsistencies and sudden changes in the video, specially in the background.
https://github.com/chengzeyi/Comfy-WaveSpeed
This version uses WaveSpeed to accelerate both FLUX and Hunyuan. Comfy probably won't find it as a Missing Node, make sure to follow install instructions from the link above.
AI-generated commentary
AI-generated explanation based on source and configuration details. Suggestions are clearly labeled.
Use this workflow to turn similar text prompts into short videos: FLUX creates three similar images, CogVideoX 1.5 5B I2V creates two interpolations between them, and Hunyuan V2V addresses inconsistencies and sudden background changes.
The described sequence goes from prompts to three images, then to two interpolations, followed by Hunyuan V2V processing.
Input similar text prompts for FLUX.
Intermediate results are three similar images and two interpolations between them.
The final output is a short video processed with Hunyuan V2V.
Follow the Comfy-WaveSpeed installation instructions linked in the description; Comfy may not identify WaveSpeed as a missing node.
Suggestion · not verified
Listed model files: ae.safetensors, clip_l.safetensors, CogVideoX_5b_1_5_I2V_GGUF_Q4_0.safetensors, hunyuan_video_t2v_720p_bf16.safetensors, hunyuan_video_vae_bf16.safetensors, jibMixFlux_v72PixelHeaven.safetensors, llava_llama3_fp8_scaled.safetensors, and t5xxl_fp8_e4m3fn.safet
Node names include UNETLoader, VAEDecode, ModelSamplingSD3, VAEEncodeTiled, VAELoader, RandomNoise, GetNode, Seed (rgthree), Fast Groups Bypasser (rgthree), EmptySD3LatentImage, CLIPTextEncode, SetNode, ApplyFBCacheOnModel, ModelSamplingFlux, DownloadAndLoadCogVideoGGUFModel, and
Additional node names include FluxGuidance, ImageBatchMulti, VHS_SelectImages, DualCLIPLoader, KSamplerSelect, VAEDecodeTiled, CLIPLoader, CogVideoImageEncode, PurgeVRAMNode, CogVideoTextEncode, SamplerCustomAdvanced, CogVideoSampler, BasicGuider, VHS_VideoCombine, INTConstant, a
The base model is Hunyuan Video, and the workflow combines FLUX, CogVideoX 1.5 5B I2V, and Hunyuan V2V stages.
This version uses WaveSpeed with both FLUX and Hunyuan.
This item is listed as a workflow, and no Pipe version is provided.
Before use, check that the listed model files and node names are available.
Suggestion · not verified
Before use, account for WaveSpeed’s increased VRAM use; the description reports only a 24GB run.
Suggestion · not verified
Does this need to be edited?
Sign in to send an edit request.
Sources
1 sourceSource excerpts
1 excerptSource context: 1012 downloads · Type Workflows · Base model Hunyuan Video
WaveSpeed comes with the cost of increased VRAM usage.
So far I have only been able to run this workflow with 24GB.
Estimated VRAM requirement
Estimate unavailable
4.79 GB across 2 of 14 model files. Model file total + 25% loading overhead + 2 GB execution buffer, rounded up.
Requirements
43 requirementsVAE · Unknown
Text encoder · 235 MB · SAFETENSORS · Unknown
CLIPLoader
Not resolvedText encoder · Unknown
CogVideoX_5b_1_5_I2V_GGUF_Q4_0.safetensors
Not resolvedCheckpoint · Unknown
DualCLIPLoader
Not resolvedText encoder · Unknown
hunyuan_video_t2v_720p_bf16.safetensors
File unverifiedCheckpoint · SAFETENSORS · Unknown
hunyuan_video_vae_bf16.safetensors
PossibleVAE · Unknown
jibMixFlux_v72PixelHeaven.safetensors
Not resolvedCheckpoint · Unknown
llava_llama3_fp8_scaled.safetensors
ConflictText encoder · Unknown
Text encoder · 4.56 GB · SAFETENSOR · fp8 · Unknown
VAEDecode
Not resolvedVAE · Unknown
VAEDecodeTiled
Not resolvedVAE · Unknown
VAEEncodeTiled
Not resolvedVAE · Unknown
VAELoader
Not resolvedVAE · Unknown
ApplyFBCacheOnModel
Not resolvedNode pack · Unknown
BasicGuider
Not resolvedNode pack · Unknown
BasicScheduler
Not resolvedNode pack · Unknown
CLIPTextEncode
Node pack · Unknown
CogVideoDecode
Not resolvedNode pack · Unknown
CogVideoImageEncode
Not resolvedNode pack · Unknown
CogVideoSampler
Not resolvedNode pack · Unknown
CogVideoTextEncode
Not resolvedNode pack · Unknown
DownloadAndLoadCogVideoGGUFModel
Not resolvedNode pack · Unknown
EmptySD3LatentImage
Not resolvedNode pack · Unknown
Fast Groups Bypasser (rgthree)
Not resolvedNode pack · Unknown
FILM VFI
Not resolvedNode pack · Unknown
FluxGuidance
Not resolvedNode pack · Unknown
GetNode
Not resolvedNode pack · Unknown
ImageBatch
Not resolvedNode pack · Unknown
ImageBatchMulti
Not resolvedNode pack · Unknown
INTConstant
Not resolvedNode pack · Unknown
KSamplerSelect
Not resolvedNode pack · Unknown
ModelSamplingFlux
Not resolvedNode pack · Unknown
ModelSamplingSD3
Not resolvedNode pack · Unknown
PreviewImage
Not resolvedNode pack · Unknown
PurgeVRAMNode
Not resolvedNode pack · Unknown
RandomNoise
Not resolvedNode pack · Unknown
SamplerCustomAdvanced
Not resolvedNode pack · Unknown
Seed (rgthree)
Not resolvedNode pack · Unknown
SetNode
Not resolvedNode pack · Unknown
UNETLoader
Not resolvedNode pack · Unknown
VHS_SelectImages
Not resolvedNode pack · Unknown
VHS_VideoCombine
Not resolvedNode pack · Unknown