With the native comfy implementation of Hunyuan I have tweaked the workflow to work for 12GB VRAM cards. It looks like you can get at least to 73 frames and probably a bit more. It takes about 8 minutes for a 4070Ti...
Laufzeitprofil
Quellenbeschreibung
With the native comfy implementation of Hunyuan I have tweaked the workflow to work for 12GB VRAM cards. It looks like you can get at least to 73 frames and probably a bit more. It takes about 8 minutes for a 4070Ti to run 20 steps.
Do make sure to update your comfy and to get the exact result above the guidance I put up to 10 (I reduced it for the base workflow as it caused some burning for some prompts).
As its not as complete as the wrapper node there is a few less features than that for now. But there is also no crazy special installations that you need to do.
Links to model downloads here:
KI-generierter Kommentar
KI-generierte Erklärung auf Grundlage von Quellen- und Konfigurationsdetails. Vorschläge sind klar gekennzeichnet.
This ComfyUI workflow is for generating Hunyuan Video clips from a text prompt: it encodes the text and outputs a combined video, and it is described as tweaked for cards with 12GB of VRAM.
The graph encodes text, creates an empty Hunyuan latent video, samples it, decodes it with tiled VAE processing, and combines the result into a video.
The description says the output can reach at least 73 frames, possibly more.
The listing reports about 8 minutes for a 4070Ti to run 20 steps.
Before using it, update ComfyUI, as the description instructs.
Vorschlag · nicht geprüft
The native ComfyUI implementation is described as not requiring special installations.
Named model files are clip_l.safetensors, hunyuan_video_t2v_720p_bf16.safetensors, hunyuan_video_vae_bf16.safetensors, and llava_llama3_fp8_scaled.safetensors.
Named nodes include BasicGuider, BasicScheduler, CLIPTextEncode, DualCLIPLoader, EmptyHunyuanLatentVideo, FluxGuidance, KSamplerSelect, and ModelSamplingSD3.
It also names RandomNoise, SamplerCustomAdvanced, UNETLoader, VAEDecodeTiled, VAELoader, and VHS_VideoCombine.
If you are following the described setup, set guidance up to 10; the base workflow uses lower guidance because some prompts caused burning.
Vorschlag · nicht geprüft
Before use, verify that the listed model files and nodes are available in your ComfyUI setup.
Vorschlag · nicht geprüft
The description notes that this implementation has fewer features than the wrapper node.
Muss das bearbeitet werden?
Melden Sie sich an, um eine Änderungsanfrage zu senden.
Quellen
1 QuelleQuellenauszüge
1 AuszugSource context: 16736 downloads · Type Workflows · Base model Hunyuan Video
Geschätzter VRAM Bedarf
Schätzung nicht verfügbar
235 MB über 1 von 7 Modell-Dateien. Gesamtmodell-Dateien + 25% Ladeaufwand + 2 GB Ausführungs-Puffer, aufgerundet.
Anforderungen
18 AnforderungenText encoder · 235 MB · SAFETENSORS · Unknown
DualCLIPLoader
Nicht aufgelöstText encoder · Unknown
hunyuan_video_t2v_720p_bf16.safetensors
Datei ungeprüftCheckpoint · SAFETENSORS · Unknown
hunyuan_video_vae_bf16.safetensors
MöglichVAE · Unknown
llava_llama3_fp8_scaled.safetensors
KonfliktText encoder · Unknown
VAEDecodeTiled
Nicht aufgelöstVAE · Unknown
VAELoader
Nicht aufgelöstVAE · Unknown
BasicGuider
Nicht aufgelöstKnotenpaket · Unknown
BasicScheduler
Nicht aufgelöstKnotenpaket · Unknown
CLIPTextEncode
Nicht aufgelöstKnotenpaket · Unknown
Knotenpaket · Unknown
FluxGuidance
Nicht aufgelöstKnotenpaket · Unknown
KSamplerSelect
Nicht aufgelöstKnotenpaket · Unknown
ModelSamplingSD3
Nicht aufgelöstKnotenpaket · Unknown
RandomNoise
Nicht aufgelöstKnotenpaket · Unknown
SamplerCustomAdvanced
Nicht aufgelöstKnotenpaket · Unknown
UNETLoader
Nicht aufgelöstKnotenpaket · Unknown
VHS_VideoCombine
Nicht aufgelöstKnotenpaket · Unknown