With the native comfy implementation of Hunyuan I have tweaked the workflow to work for 12GB VRAM cards. It looks like you can get at least to 73 frames and probably a bit more. It takes about 8 minutes for a 4070Ti...
Perfil de execução
Descrição da fonte
With the native comfy implementation of Hunyuan I have tweaked the workflow to work for 12GB VRAM cards. It looks like you can get at least to 73 frames and probably a bit more. It takes about 8 minutes for a 4070Ti to run 20 steps.
Do make sure to update your comfy and to get the exact result above the guidance I put up to 10 (I reduced it for the base workflow as it caused some burning for some prompts).
As its not as complete as the wrapper node there is a few less features than that for now. But there is also no crazy special installations that you need to do.
Links to model downloads here:
Comentário gerado por IA
Explicação gerada por IA com base nos detalhes da fonte e da configuração. As sugestões são claramente identificadas.
This ComfyUI workflow is for generating Hunyuan Video clips from a text prompt: it encodes the text and outputs a combined video, and it is described as tweaked for cards with 12GB of VRAM.
The graph encodes text, creates an empty Hunyuan latent video, samples it, decodes it with tiled VAE processing, and combines the result into a video.
The description says the output can reach at least 73 frames, possibly more.
The listing reports about 8 minutes for a 4070Ti to run 20 steps.
Before using it, update ComfyUI, as the description instructs.
Sugestão · não verificado
The native ComfyUI implementation is described as not requiring special installations.
Named model files are clip_l.safetensors, hunyuan_video_t2v_720p_bf16.safetensors, hunyuan_video_vae_bf16.safetensors, and llava_llama3_fp8_scaled.safetensors.
Named nodes include BasicGuider, BasicScheduler, CLIPTextEncode, DualCLIPLoader, EmptyHunyuanLatentVideo, FluxGuidance, KSamplerSelect, and ModelSamplingSD3.
It also names RandomNoise, SamplerCustomAdvanced, UNETLoader, VAEDecodeTiled, VAELoader, and VHS_VideoCombine.
If you are following the described setup, set guidance up to 10; the base workflow uses lower guidance because some prompts caused burning.
Sugestão · não verificado
Before use, verify that the listed model files and nodes are available in your ComfyUI setup.
Sugestão · não verificado
The description notes that this implementation has fewer features than the wrapper node.
Isso precisa ser editado?
Faça login para enviar uma solicitação de edição.
Fontes
1 fonteTrechos de fonte
1 trechoSource context: 16736 downloads · Type Workflows · Base model Hunyuan Video
Estimativa de requisito VRAM
Estimativa indisponível
235 MB distribuídos em 1 de 7 arquivos de modelo. Total de arquivos do modelo + 25% de overhead de carregamento + 2 GB de buffer de execução, arredondado para cima.
Requisitos
Requisitos 18Text encoder · 235 MB · SAFETENSORS · Unknown
DualCLIPLoader
Não resolvidoText encoder · Unknown
hunyuan_video_t2v_720p_bf16.safetensors
Arquivo não verificadoCheckpoint · SAFETENSORS · Unknown
hunyuan_video_vae_bf16.safetensors
PossívelVAE · Unknown
llava_llama3_fp8_scaled.safetensors
ConflitoText encoder · Unknown
VAEDecodeTiled
Não resolvidoVAE · Unknown
VAELoader
Não resolvidoVAE · Unknown
BasicGuider
Não resolvidoPacote de nós · Unknown
BasicScheduler
Não resolvidoPacote de nós · Unknown
CLIPTextEncode
Não resolvidoPacote de nós · Unknown
Pacote de nós · Unknown
FluxGuidance
Não resolvidoPacote de nós · Unknown
KSamplerSelect
Não resolvidoPacote de nós · Unknown
ModelSamplingSD3
Não resolvidoPacote de nós · Unknown
RandomNoise
Não resolvidoPacote de nós · Unknown
SamplerCustomAdvanced
Não resolvidoPacote de nós · Unknown
UNETLoader
Não resolvidoPacote de nós · Unknown
VHS_VideoCombine
Não resolvidoPacote de nós · Unknown