This lets you bring an image, a couple lines of dialog, and a short voice sample and get speaking character videos out of Wan 2.1. Thanks to Kijai for adding multitalk capability to the WanWrapper series of nodes.
Profil d'exécution
Description de la source
This lets you bring an image, a couple lines of dialog, and a short voice sample and get speaking character videos out of Wan 2.1. Thanks to Kijai for adding multitalk capability to the WanWrapper series of nodes.
Helpful links:
Github of the chatterbox node implementation I'm using:
https://github.com/diodiogod/ComfyUI_ChatterBox_SRT_Voice
Lots of demo voices you can download here:
Commentaire généré par l’IA
Explication générée par l’IA à partir des détails de la source et de la configuration. Les suggestions sont clairement signalées.
Use this workflow to turn an image, a couple lines of dialog, and a short voice sample into speaking character videos with Wan 2.1 multitalk.
The output is a speaking character video generated from the supplied image, dialog, and voice sample.
The description includes links to the Chatterbox node implementation, downloadable demo voices, and Kijai’s WanVideoWrapper nodes.
The workflow lists these node packs: ComfyUI-WanVideoWrapper, ComfyUI-KJNodes, ComfyUI-VideoHelperSuite, rgthree-comfy, comfyui-ollama, and audio-separation-nodes-comfyui.
It also lists chatterbox_srt_voice and was-node-suite-comfyui.
Listed model files and checkpoints include CLIP-ViT-H-14-laion2B-s32B-b79K.safetensors, TencentGameMate/chinese-wav2vec2-base, Wan2_1_VAE_bf16.safetensors, umt5-xxl-enc-bf16.safetensors, Wan21_I2V_14B_lightx2v_cfg_step_distill_lora_rank64.safetensors, WanVideo_2_1_Multitalk_14B\n
The listed base model is Wan Video 14B t2v.
The workflow version is v1.0.
Before use, verify that every listed node pack and model file is available under the exact name shown.
Suggestion · non vérifié
Usage terms are not provided here, so verify them before use.
Suggestion · non vérifié
Faut-il modifier cela ?
Connectez-vous pour envoyer une demande de modification.
Sources
1 sourceExtraits de sources
2 extraitsSource context: 517 downloads · Type Workflows · Base model Wan Video 14B t2v
Kijai's nodes:
Estimation des besoins VRAM
Estimation indisponible
Les tailles des fichiers du modèle ne sont pas encore suffisamment complètes pour calculer cette exigence.
Exigences
Exigences 15Vision model · Hugging Face · h94/IP-Adapter
TencentGameMate/chinese-wav2vec2-base
Non résoluCheckpoint · Hugging Face · TencentGameMate/chinese-wav2vec2-base
Text encoder · Hugging Face · wanabmeya/WanVideo_comfy_umt5-xxl-enc-bf16
VAE · SAFETENSORS · Unknown
Wan2_1-I2V-14B-480P_fp8_e4m3fn.safetensors
Non résoluCheckpoint · Unknown
Wan21_I2V_14B_lightx2v_cfg_step_distill_lora_rank64.safetensors
Fichier non vérifiéLoRA · Hugging Face · lgylgy/Wan21_I2V_14B_lightx2v_cfg_step_distill_lora_rank64
WanVideo_2_1_Multitalk_14B_fp8_e4m3fn.safetensors
Non résoluCheckpoint · Unknown
Pack de nœud · Registry
chatterbox_srt_voice
Non résoluPack de nœud · Registry
Source context: 502 downloads · Type Workflows · Base model Wan Video 14B t2v
Pack de nœud · Registry
Pack de nœud · Registry
Pack de nœud · Registry
Pack de nœud · Registry
Pack de nœud · Registry
was-node-suite-comfyui
PossiblePack de nœud · Registry