This lets you bring an image, a couple lines of dialog, and a short voice sample and get speaking character videos out of Wan 2.1. Thanks to Kijai for adding multitalk capability to the WanWrapper series of nodes.
Perfil de execução
Descrição da fonte
This lets you bring an image, a couple lines of dialog, and a short voice sample and get speaking character videos out of Wan 2.1. Thanks to Kijai for adding multitalk capability to the WanWrapper series of nodes.
Helpful links:
Github of the chatterbox node implementation I'm using:
https://github.com/diodiogod/ComfyUI_ChatterBox_SRT_Voice
Lots of demo voices you can download here:
Comentário gerado por IA
Explicação gerada por IA com base nos detalhes da fonte e da configuração. As sugestões são claramente identificadas.
Use this workflow to turn an image, a couple lines of dialog, and a short voice sample into speaking character videos with Wan 2.1 multitalk.
The output is a speaking character video generated from the supplied image, dialog, and voice sample.
The description includes links to the Chatterbox node implementation, downloadable demo voices, and Kijai’s WanVideoWrapper nodes.
The workflow lists these node packs: ComfyUI-WanVideoWrapper, ComfyUI-KJNodes, ComfyUI-VideoHelperSuite, rgthree-comfy, comfyui-ollama, and audio-separation-nodes-comfyui.
It also lists chatterbox_srt_voice and was-node-suite-comfyui.
Listed model files and checkpoints include CLIP-ViT-H-14-laion2B-s32B-b79K.safetensors, TencentGameMate/chinese-wav2vec2-base, Wan2_1_VAE_bf16.safetensors, umt5-xxl-enc-bf16.safetensors, Wan21_I2V_14B_lightx2v_cfg_step_distill_lora_rank64.safetensors, WanVideo_2_1_Multitalk_14B\n
The listed base model is Wan Video 14B t2v.
The workflow version is v1.0.
Before use, verify that every listed node pack and model file is available under the exact name shown.
Sugestão · não verificado
Usage terms are not provided here, so verify them before use.
Sugestão · não verificado
Isso precisa ser editado?
Faça login para enviar uma solicitação de edição.
Fontes
1 fonteTrechos de fonte
2 trechosSource context: 517 downloads · Type Workflows · Base model Wan Video 14B t2v
Kijai's nodes:
Estimativa de requisito VRAM
Estimativa indisponível
Os tamanhos de arquivos do modelo ainda não estão completos o suficiente para calcular este requisito.
Requisitos
Requisitos 15CLIP-ViT-H-14-laion2B-s32B-b79K.safetensors
Não resolvidoVision model · Hugging Face · h94/IP-Adapter
TencentGameMate/chinese-wav2vec2-base
Não resolvidoCheckpoint · Hugging Face · TencentGameMate/chinese-wav2vec2-base
Text encoder · Hugging Face · wanabmeya/WanVideo_comfy_umt5-xxl-enc-bf16
VAE · SAFETENSORS · Unknown
Wan2_1-I2V-14B-480P_fp8_e4m3fn.safetensors
Não resolvidoCheckpoint · Unknown
Wan21_I2V_14B_lightx2v_cfg_step_distill_lora_rank64.safetensors
Arquivo não verificadoLoRA · Hugging Face · lgylgy/Wan21_I2V_14B_lightx2v_cfg_step_distill_lora_rank64
WanVideo_2_1_Multitalk_14B_fp8_e4m3fn.safetensors
Não resolvidoCheckpoint · Unknown
audio-separation-nodes-comfyui
EncontradoPacote de nós · Registry
chatterbox_srt_voice
Não resolvidoPacote de nós · Registry
Source context: 502 downloads · Type Workflows · Base model Wan Video 14B t2v
Pacote de nós · Registry
Pacote de nós · Registry
Pacote de nós · Registry
Pacote de nós · Registry
Pacote de nós · Registry
was-node-suite-comfyui
PossívelPacote de nós · Registry