ComfyUI nodes for vision-language models: Qwen3-VL, Moondream 3, Florence-2, SmolVLM2, InternVL, Gemma 3, MiniCPM-V. Plus open-vocabulary detection, SAM2/SAM3 segmentation, video temporal reasoning, GGUF via...
Paketprofil
README
ComfyUI nodes for vision-language models: Qwen3-VL, Moondream 3, Florence-2, SmolVLM2, InternVL, Gemma 3, MiniCPM-V. Plus open-vocabulary detection, SAM2/SAM3 segmentation, video temporal reasoning, GGUF via llama.cpp, and hosted LLM/VLM APIs.
Knoten in diesem Paket
34 Knoten (Plural)Quellen
1 QuelleQuellenauszüge
1 AuszugSource context: Repo gokayfem/ComfyUI_VLM_nodes
AudioLDM2Node
Kompakte Beschreibung für diesen Knoten nicht verfügbar.
VLM Nodes/Audio · 2 Eingaben · 6 Parameter