VLM_nodes

Custom Nodes for Vision Language Models (VLM) , Large Language Models (LLM), Image Captioning, Automatic Prompt Generation, Creative and Consistent Prompt Suggestion, Keyword Extraction

How much VRAM does VLM_nodes require?

Direct Answer: The ComfyUI node VLM_nodes requires a minimum base VRAM of 256MB and is optimized for GPUs with at least 4GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.

High (2-4GB)
Base VRAM:
256MB (0.3GB)
Recommended GPU:
4GB+ VRAM
Low VRAM Mode:
✓ Supported
Estimation Confidence:
HIGH

Interactive VRAM Compatibility Estimator

Estimated Total VRAM: 3.00 GBTarget: 8 GB
✅ Comfortable Fit

Your GPU has plenty of headroom. You can run this node safely with your active configurations!

Deploy on High-Performance GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.

🚀 Deploy on Vast.ai

Deploy on Cloud GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.

🚀 Deploy on RunPod

What Python packages are required for VLM_nodes?

Direct Answer: Running VLM_nodes requires installing the following Python package dependencies: accelerate>=1.0, bitsandbytes, cffi, decord, diffusers>=0.31.0, diskcache, einops>=0.7.0, gitpython, huggingface-hub>=0.26.2, matplotlib, moviepy, numpy>=1.26.4,<2.0.0, openai>=0.27.8, opencv-python, optimum>=1.17.0, pillow>=9.4.0, py-cpuinfo>=3.3.0, python-dateutil>=2.7.0, pytz, qwen-vl-utils, safetensors>=0.4.1, scikit-build, six, soundfile, symusic, torch>=2.0.1, torchvision>=0.15.2, transformers>=4.46. This node specifically requires PyTorch version 2.0.1 or newer.

requirements.txt
accelerate>=1.0
bitsandbytes
cffi
decord
diffusers>=0.31.0
diskcache
einops>=0.7.0
gitpython
huggingface-hub>=0.26.2
matplotlib
moviepy
numpy>=1.26.4,<2.0.0
openai>=0.27.8
opencv-python
optimum>=1.17.0
pillow>=9.4.0
py-cpuinfo>=3.3.0
python-dateutil>=2.7.0
pytz
qwen-vl-utils
safetensors>=0.4.1
scikit-build
six
soundfile
symusic
torch>=2.0.1
torchvision>=0.15.2
transformers>=4.46

Interactive Setup & Dependency Resolver

Operating System:
Environment Type:
Run this terminal command in your ComfyUI root folder:
# Loading command...

Special Environment Requirements:

Requires PyTorch version 2.0.1+.

To update PyTorch for your selected setup, run:
# Loading PyTorch command...

Frequently Asked Questions

How much VRAM does VLM_nodes require?

VLM_nodes requires a minimum of 256MB (0.3GB) of VRAM for base operation. For optimal performance, a GPU with at least 4GB of VRAM is recommended. This node supports low VRAM mode for resource-constrained setups.

Can I run VLM_nodes on an RTX 3060, RTX 4070, or RTX 4090?

✅ RTX 3060 (12GB): Yes, fully compatible with 10.6GB headroom. ✅ RTX 4070 (12GB): Yes, fully compatible with 10.6GB headroom. ✅ RTX 4070 Ti (16GB): Yes, fully compatible with 14.2GB headroom. ✅ RTX 4090 (24GB): Yes, fully compatible with 21.4GB headroom

What PyTorch version does VLM_nodes need?

VLM_nodes requires the following PyTorch-related packages: torch>=2.0.1, torchvision>=0.15.2. Ensure your ComfyUI environment has these installed. Ensure your PyTorch installation matches your CUDA version (use torch.version.cuda to check).

What Python packages are required for VLM_nodes?

To run VLM_nodes, you need to install: accelerate>=1.0, bitsandbytes, cffi, decord, diffusers>=0.31.0, diskcache, einops>=0.7.0, gitpython, huggingface-hub>=0.26.2, matplotlib, moviepy, numpy>=1.26.4,<2.0.0, openai>=0.27.8, opencv-python, optimum>=1.17.0, pillow>=9.4.0, py-cpuinfo>=3.3.0, python-dateutil>=2.7.0, pytz, qwen-vl-utils, safetensors>=0.4.1, scikit-build, six, soundfile, symusic, torch>=2.0.1, torchvision>=0.15.2, transformers>=4.46. You can install these using pip or add them to your requirements.txt file.

How can I reduce VRAM usage when running VLM_nodes?

VLM_nodes supports low VRAM mode. To reduce memory usage: (1) Enable --lowvram or --medvram flags in ComfyUI, (2) Reduce batch size to 1, (3) Use fp16 or fp8 precision if supported, (4) Close other GPU applications.

How do I install VLM_nodes in ComfyUI?

To install VLM_nodes: (1) Navigate to your ComfyUI/custom_nodes directory, (2) Clone the repository: git clone https://github.com/gokayfem/ComfyUI_VLM_nodes, (3) Install dependencies: pip install -r requirements.txt (if present), (4) Restart ComfyUI. Alternatively, use ComfyUI Manager for one-click installation.