With the native comfy implementation of Hunyuan I have tweaked the workflow to work for 12GB VRAM cards. It looks like you can get at least to 73 frames and probably a bit more. It takes about 8 minutes for a 4070Ti...
Profil d'exécution
Description de la source
With the native comfy implementation of Hunyuan I have tweaked the workflow to work for 12GB VRAM cards. It looks like you can get at least to 73 frames and probably a bit more. It takes about 8 minutes for a 4070Ti to run 20 steps.
Do make sure to update your comfy and to get the exact result above the guidance I put up to 10 (I reduced it for the base workflow as it caused some burning for some prompts).
As its not as complete as the wrapper node there is a few less features than that for now. But there is also no crazy special installations that you need to do.
Links to model downloads here:
Commentaire généré par l’IA
Explication générée par l’IA à partir des détails de la source et de la configuration. Les suggestions sont clairement signalées.
This ComfyUI workflow is for generating Hunyuan Video clips from a text prompt: it encodes the text and outputs a combined video, and it is described as tweaked for cards with 12GB of VRAM.
The graph encodes text, creates an empty Hunyuan latent video, samples it, decodes it with tiled VAE processing, and combines the result into a video.
The description says the output can reach at least 73 frames, possibly more.
The listing reports about 8 minutes for a 4070Ti to run 20 steps.
Before using it, update ComfyUI, as the description instructs.
Suggestion · non vérifié
The native ComfyUI implementation is described as not requiring special installations.
Named model files are clip_l.safetensors, hunyuan_video_t2v_720p_bf16.safetensors, hunyuan_video_vae_bf16.safetensors, and llava_llama3_fp8_scaled.safetensors.
Named nodes include BasicGuider, BasicScheduler, CLIPTextEncode, DualCLIPLoader, EmptyHunyuanLatentVideo, FluxGuidance, KSamplerSelect, and ModelSamplingSD3.
It also names RandomNoise, SamplerCustomAdvanced, UNETLoader, VAEDecodeTiled, VAELoader, and VHS_VideoCombine.
If you are following the described setup, set guidance up to 10; the base workflow uses lower guidance because some prompts caused burning.
Suggestion · non vérifié
Before use, verify that the listed model files and nodes are available in your ComfyUI setup.
Suggestion · non vérifié
The description notes that this implementation has fewer features than the wrapper node.
Faut-il modifier cela ?
Connectez-vous pour envoyer une demande de modification.
Sources
1 sourceExtraits de sources
1 extraitSource context: 16736 downloads · Type Workflows · Base model Hunyuan Video
Estimation des besoins VRAM
Estimation indisponible
235 MB sur 1 de 7 fichiers de modèle. Total des fichiers du modèle + 25 % de surcharge de chargement + 2 Go de tampon d'exécution, arrondi à l'unité supérieure.
Exigences
Exigences 18Text encoder · 235 MB · SAFETENSORS · Unknown
DualCLIPLoader
Non résoluText encoder · Unknown
hunyuan_video_t2v_720p_bf16.safetensors
Fichier non vérifiéCheckpoint · SAFETENSORS · Unknown
hunyuan_video_vae_bf16.safetensors
PossibleVAE · Unknown
llava_llama3_fp8_scaled.safetensors
ConflitText encoder · Unknown
VAEDecodeTiled
Non résoluVAE · Unknown
VAELoader
Non résoluVAE · Unknown
BasicGuider
Non résoluPack de nœud · Unknown
BasicScheduler
Non résoluPack de nœud · Unknown
CLIPTextEncode
Non résoluPack de nœud · Unknown
EmptyHunyuanLatentVideo
Pack de nœud · Unknown
FluxGuidance
Non résoluPack de nœud · Unknown
KSamplerSelect
Non résoluPack de nœud · Unknown
ModelSamplingSD3
Non résoluPack de nœud · Unknown
RandomNoise
Non résoluPack de nœud · Unknown
SamplerCustomAdvanced
Non résoluPack de nœud · Unknown
UNETLoader
Non résoluPack de nœud · Unknown
VHS_VideoCombine
Non résoluPack de nœud · Unknown