This lets you bring an image, a couple lines of dialog, and a short voice sample and get speaking character videos out of Wan 2.1. Thanks to Kijai for adding multitalk capability to the WanWrapper series of nodes.
Laufzeitprofil
Quellenbeschreibung
This lets you bring an image, a couple lines of dialog, and a short voice sample and get speaking character videos out of Wan 2.1. Thanks to Kijai for adding multitalk capability to the WanWrapper series of nodes.
Helpful links:
Github of the chatterbox node implementation I'm using:
https://github.com/diodiogod/ComfyUI_ChatterBox_SRT_Voice
Lots of demo voices you can download here:
KI-generierter Kommentar
KI-generierte Erklärung auf Grundlage von Quellen- und Konfigurationsdetails. Vorschläge sind klar gekennzeichnet.
Use this workflow to turn an image, a couple lines of dialog, and a short voice sample into speaking character videos with Wan 2.1 multitalk.
The output is a speaking character video generated from the supplied image, dialog, and voice sample.
The description includes links to the Chatterbox node implementation, downloadable demo voices, and Kijai’s WanVideoWrapper nodes.
The workflow lists these node packs: ComfyUI-WanVideoWrapper, ComfyUI-KJNodes, ComfyUI-VideoHelperSuite, rgthree-comfy, comfyui-ollama, and audio-separation-nodes-comfyui.
It also lists chatterbox_srt_voice and was-node-suite-comfyui.
Listed model files and checkpoints include CLIP-ViT-H-14-laion2B-s32B-b79K.safetensors, TencentGameMate/chinese-wav2vec2-base, Wan2_1_VAE_bf16.safetensors, umt5-xxl-enc-bf16.safetensors, Wan21_I2V_14B_lightx2v_cfg_step_distill_lora_rank64.safetensors, WanVideo_2_1_Multitalk_14B\n
The listed base model is Wan Video 14B t2v.
The workflow version is v1.0.
Before use, verify that every listed node pack and model file is available under the exact name shown.
Vorschlag · nicht geprüft
Usage terms are not provided here, so verify them before use.
Vorschlag · nicht geprüft
Muss das bearbeitet werden?
Melden Sie sich an, um eine Änderungsanfrage zu senden.
Quellen
1 QuelleQuellenauszüge
2 AuszügeSource context: 517 downloads · Type Workflows · Base model Wan Video 14B t2v
Kijai's nodes:
Geschätzter VRAM Bedarf
Schätzung nicht verfügbar
Die Modell-Dateigrößen sind noch nicht vollständig genug, um diese Anforderung zu berechnen.
Anforderungen
15 AnforderungenCLIP-ViT-H-14-laion2B-s32B-b79K.safetensors
Nicht aufgelöstVision model · Hugging Face · h94/IP-Adapter
TencentGameMate/chinese-wav2vec2-base
Nicht aufgelöstCheckpoint · Hugging Face · TencentGameMate/chinese-wav2vec2-base
Text encoder · Hugging Face · wanabmeya/WanVideo_comfy_umt5-xxl-enc-bf16
VAE · SAFETENSORS · Unknown
Wan2_1-I2V-14B-480P_fp8_e4m3fn.safetensors
Nicht aufgelöstCheckpoint · Unknown
LoRA · Hugging Face · lgylgy/Wan21_I2V_14B_lightx2v_cfg_step_distill_lora_rank64
WanVideo_2_1_Multitalk_14B_fp8_e4m3fn.safetensors
Nicht aufgelöstCheckpoint · Unknown
audio-separation-nodes-comfyui
GefundenKnotenpaket · Registry
chatterbox_srt_voice
Nicht aufgelöstKnotenpaket · Registry
Source context: 502 downloads · Type Workflows · Base model Wan Video 14B t2v
Knotenpaket · Registry
Knotenpaket · Registry
Knotenpaket · Registry
Knotenpaket · Registry
Knotenpaket · Registry
was-node-suite-comfyui
MöglichKnotenpaket · Registry