ComfyUI nodes for vision-language models: Qwen3-VL, Moondream 3, Florence-2, SmolVLM2, InternVL, Gemma 3, MiniCPM-V. Plus open-vocabulary detection, SAM2/SAM3 segmentation, video temporal reasoning, GGUF via...
Perfil del paquete
Léeme
ComfyUI nodes for vision-language models: Qwen3-VL, Moondream 3, Florence-2, SmolVLM2, InternVL, Gemma 3, MiniCPM-V. Plus open-vocabulary detection, SAM2/SAM3 segmentation, video temporal reasoning, GGUF via llama.cpp, and hosted LLM/VLM APIs.
Nodos de este paquete
34 nodosFuentes
1 fuenteExtractos de fuentes
1 extractoSource context: Repo gokayfem/ComfyUI_VLM_nodes
AudioLDM2Node
No hay una descripción breve disponible para este nodo.
VLM Nodes/Audio · 2 entradas · 6 parámetros