New BlockSwap for low-VRAM users,
Runtime profile
Source description
✨ WAN2.1 — Image to video — Simple Workflow
A clean, all-in-one WAN image-to-video workflow built entirely with the UmeAiRT Toolkit for ComfyUI. Only 12 nodes . No spaghetti wires. Just load your model, write your prompt, and hit generate.
⚠️ IMPORTANT — Nodes 2.0 Required
This workflow is built for the Nodes 2.0 (Vue) interface of ComfyUI. If you don't enable it, the workflow may have display problems.
How to activate Nodes 2.0:
Open ComfyUI
Go to Settings (⚙️ icon, bottom-left)
AI-generated commentary
AI-generated explanation based on source and configuration details. Suggestions are clearly labeled.
This ComfyUI workflow is for image-to-video and text-to-video generation: it uses an input image or text prompt to produce WAN video, can save generated images with their generation parameters, and uses the Wan Video 14B i2v 480p model.
It includes two workflow files: IMG to VIDEO 2.4 BASE SIMPLE.json and IMG to VIDEO 2.4 GGUF SIMPLE.json.
The release summary highlights a new BlockSwap for low-VRAM users.
Saved images include all generation parameters for online publishing and remixing.
In ComfyUI, turn on “Use Nodes V2 (Vue)” in Settings, refresh, then load the workflow; leaving it off may cause display problems.
Install ComfyUI-UmeAiRT-Toolkit through ComfyUI Manager by searching “UmeAiRT,” or use the UmeAiRT Auto-Installer.
The BASE variant lists these model files: clip_vision_h.safetensors; MiaoshouAI/Florence-2-base-PromptGen-v2.0; rife47.pth; umt5-xxl-encoder-fp8-e4m3fn-scaled.safetensors; wan_2.1_vae.safetensors; and wan2.1_i2v_480p_14B_fp8_e4m3fn.safetensors.
The GGUF variant lists these model files: clip_vision_h.safetensors; MiaoshouAI/Florence-2-base-PromptGen-v2.0; rife47.pth; umt5-xxl-encoder-Q8_0.gguf; wan_2.1_vae.safetensors; and wan2.1-i2v-14b-480p-Q8_0.gguf.
The listed node packs are comfyui-custom-scripts, comfyui-easy-use, comfyui-florence2, comfyui-frame-interpolation, comfyui-kjnodes, comfyui-lora-manager, comfyui-mxtoolkit, comfyui-videohelpersuite, wanblockswap, and was-node-suite-comfyui; the GGUF variant also lists ComfyUI-GG
The built-in SeedVR2 tiled upscaler can be toggled on or off.
Three LoRA slots have individual on/off toggles and strength controls, and additional LoRA modules can be connected.
The auto version downloads models automatically; the manual version uses model files placed in ComfyUI model folders.
The source page does not state a license, so check the terms before publishing or remixing.
Suggestion · not verified
Before downloading, verify that the page identifies publisher UmeAiRT and release 🟡 v2.4 (simple).
Suggestion · not verified
Does this need to be edited?
Sign in to send an edit request.
Sources
1 sourceSource excerpts
1 excerptSource context: 100975 downloads · Type Workflows · Base model Wan Video 14B i2v 480p
Find "Use Nodes V2 (Vue)" and toggle it ON
Refresh the page
Load the workflow
If you prefer the classic interface, check out my Legacy version of this workflow instead ( link ).
🎯 Features
Text-to-video generation
Automatic download of models in auto version
Built-in SeedVR2 upscaler — high-quality tiled upscaling (toggleable on/off) Slower than a classic upscaler, but significantly better quality
Full metadata embedding — your images are saved with all generation parameters, ready for online publishing and remixing
3 LoRA slots — with individual on/off toggles and strength control and you can connect as many other lora modules to each other for as many LoRA as you want.
📦 Custom Node Required
Only one custom node to install:
👉 ComfyUI-UmeAiRT-Toolkit
Install via ComfyUI Manager (search "UmeAiRT") or use the UmeAiRT Auto-Installer . The Toolkit packages everything internally — upscaler, face detailer, metadata saver. No other custom nodes needed.
📂 Files you need (in manual version)
For base version I2V Model : 480p or 720p In models/diffusion_models
For GGUF version I2V Quant Model :
Common files : CLIP: umt5_xxl_fp8_e4m3fn_scaled.safetensors in models/clip CLIP-VISION: clip_vision_h.safetensors in models/clip_vision VAE: wan_2.1_vae.safetensors in models/vae Speed LoRA: 480p , 720p in models/loras
ANY upscale model:
Realistic : RealESRGAN_x4plus.pth
Anime : RealESRGAN_x4plus_anime_6B.pth
in models/upscale_models
Estimated VRAM requirement
Estimate unavailable
242 MB across 3 of 6 model files. Model file total + 25% loading overhead + 2 GB execution buffer, rounded up.
Requirements
16 requirementsVision model · 135 B · Hugging Face · Comfy-Org/Wan_2.1_ComfyUI_repackaged
MiaoshouAI/Florence-2-base-PromptGen-v2.0
File unverifiedVision model · Hugging Face · MiaoshouAI/Florence-2-base-PromptGen-v2.0 · Florence 2 Base PromptGen v2.0
Checkpoint · Hugging Face · dci05049/rife47.pth
umt5-xxl-encoder-fp8-e4m3fn-scaled.safetensors
File unverifiedText encoder · Unknown · umt5_xxl_fp8_e4m3fn_scaled.safetensors
VAE · 242 MB · Hugging Face · Comfy-Org/Wan_2.1_ComfyUI_repackaged
Checkpoint · 136 B · Hugging Face · UmeAiRT/ComfyUI-Auto_installer
Node pack · Registry
Node pack · Registry
comfyui-florence2
Node pack · Registry
Node pack · Registry
Node pack · Registry
Node pack · Registry
Node pack · Registry
Node pack · Registry
Node pack · Registry
was-node-suite-comfyui
PossibleNode pack · Registry