Comfyui_Qwen3-VL-Instruct

By IuvenisSapiensView on GitHub →

This is an implementation of [Qwen3-VL-Instruct](https://github.com/QwenLM/Qwen3-VL) by [ComfyUI](https://github.com/comfyanonymous/ComfyUI), which includes, but is not limited to, support for text-based queries, video queries, single-image queries, and multi-image queries to generate captions or responses.

How much VRAM does Comfyui_Qwen3-VL-Instruct require?

Direct Answer: The ComfyUI node Comfyui_Qwen3-VL-Instruct requires a minimum base VRAM of 4096MB and is optimized for GPUs with at least 8GB of VRAM. Low VRAM mode is not supported for this node.

Very High (4-8GB)
Base VRAM:
4096MB (4.0GB)
Recommended GPU:
8GB+ VRAM
Low VRAM Mode:
✗ Not supported
Estimation Confidence:
MEDIUM

Interactive VRAM Compatibility Estimator

Estimated Total VRAM: 3.00 GBTarget: 8 GB
✅ Comfortable Fit

Your GPU has plenty of headroom. You can run this node safely with your active configurations!

Deploy on High-Performance GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.

🚀 Deploy on Vast.ai

Deploy on Cloud GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.

🚀 Deploy on RunPod

What Python packages are required for Comfyui_Qwen3-VL-Instruct?

Direct Answer: Running Comfyui_Qwen3-VL-Instruct requires installing the following Python package dependencies: accelerate, av, bitsandbytes, huggingface_hub, numpy, opencv-python, optimum, pillow, qwen-vl-utils, torch, torchaudio, torchvision, transformers>=4.57.1, triton; sys_platform == 'linux', triton-windows; sys_platform == 'win32'. Ensure your ComfyUI environment has these packages active before launching.

requirements.txt
accelerate
av
bitsandbytes
huggingface_hub
numpy
opencv-python
optimum
pillow
qwen-vl-utils
torch
torchaudio
torchvision
transformers>=4.57.1
triton; sys_platform == 'linux'
triton-windows; sys_platform == 'win32'

Interactive Setup & Dependency Resolver

Operating System:
Environment Type:
Run this terminal command in your ComfyUI root folder:
# Loading command...

Frequently Asked Questions

How much VRAM does Comfyui_Qwen3-VL-Instruct require?

Comfyui_Qwen3-VL-Instruct requires a minimum of 4096MB (4.0GB) of VRAM for base operation. For optimal performance, a GPU with at least 8GB of VRAM is recommended. Low VRAM mode is not supported for this node.

Can I run Comfyui_Qwen3-VL-Instruct on an RTX 3060, RTX 4070, or RTX 4090?

✅ RTX 3060 (12GB): Yes, fully compatible with 6.8GB headroom. ✅ RTX 4070 (12GB): Yes, fully compatible with 6.8GB headroom. ✅ RTX 4070 Ti (16GB): Yes, fully compatible with 10.4GB headroom. ✅ RTX 4090 (24GB): Yes, fully compatible with 17.6GB headroom

What PyTorch version does Comfyui_Qwen3-VL-Instruct need?

Comfyui_Qwen3-VL-Instruct requires the following PyTorch-related packages: torch, torchaudio, torchvision. Ensure your ComfyUI environment has these installed. Ensure your PyTorch installation matches your CUDA version (use torch.version.cuda to check).

What Python packages are required for Comfyui_Qwen3-VL-Instruct?

To run Comfyui_Qwen3-VL-Instruct, you need to install: accelerate, av, bitsandbytes, huggingface_hub, numpy, opencv-python, optimum, pillow, qwen-vl-utils, torch, torchaudio, torchvision, transformers>=4.57.1, triton; sys_platform == 'linux', triton-windows; sys_platform == 'win32'. You can install these using pip or add them to your requirements.txt file.

How do I install Comfyui_Qwen3-VL-Instruct in ComfyUI?

To install Comfyui_Qwen3-VL-Instruct: (1) Navigate to your ComfyUI/custom_nodes directory, (2) Clone the repository: git clone https://github.com/IuvenisSapiens/ComfyUI_Qwen3-VL-Instruct, (3) Install dependencies: pip install -r requirements.txt (if present), (4) Restart ComfyUI. Alternatively, use ComfyUI Manager for one-click installation.