QwenVL-Mod: Enhanced Vision-Language

By huchukatoView on GitHub →

Enhanced QwenVL node with Smart Prompt Caching, multilingual WAN 2.2 presets, comprehensive visual style detection, and NSFW support. Latest v2.0.8 with bug fixes and stability improvements for professional multimodal workflows.

VRAM Requirements

Direct Answer: The ComfyUI node QwenVL-Mod: Enhanced Vision-Language requires a minimum base VRAM of 256MB and is optimized for GPUs with at least 4GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.

High (2-4GB)
Base VRAM:
256MB (0.3GB)
Recommended GPU:
4GB+ VRAM
Low VRAM Mode:
✓ Supported
Estimation Confidence:
HIGH

🧮 Interactive VRAM Compatibility Estimator

Estimated Total VRAM: 3.00 GBTarget: 8 GB
✅ Comfortable Fit

Your GPU has plenty of headroom. You can run this node safely with your active configurations!

Deploy on High-Performance GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.

🚀 Deploy on Vast.ai

Deploy on Cloud GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.

🚀 Deploy on RunPod

Python Dependencies

Direct Answer: Running QwenVL-Mod: Enhanced Vision-Language requires installing the following Python package dependencies: accelerate, bitsandbytes, hf_xet, huggingface-hub, numpy, opencv-python, pillow, psutil, torch, transformers. Ensure your ComfyUI environment has these packages active before launching.

requirements.txt
accelerate
bitsandbytes
hf_xet
huggingface-hub
numpy
opencv-python
pillow
psutil
torch
transformers

🛠️ Interactive Setup & Dependency Resolver

Operating System:
Environment Type:
Run this terminal command in your ComfyUI root folder:
# Loading command...

Compatible Foundations

This node is verified to support or optimize workflows for the following foundation model families: