ComfyUI-MultiModal-Prompt-Nodes

By kantan-kantoView on GitHub →

Advanced multimodal prompt generation nodes for ComfyUI with local GGUF models (Qwen-VL) and cloud API support for vision-based prompt enhancement.

VRAM Requirements

Direct Answer: The ComfyUI node ComfyUI-MultiModal-Prompt-Nodes requires a minimum base VRAM of 256MB and is optimized for GPUs with at least 4GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.

High (2-4GB)
Base VRAM:
256MB (0.3GB)
Recommended GPU:
4GB+ VRAM
Low VRAM Mode:
✓ Supported
Estimation Confidence:
HIGH

🧮 Interactive VRAM Compatibility Estimator

Estimated Total VRAM: 3.00 GBTarget: 8 GB
✅ Comfortable Fit

Your GPU has plenty of headroom. You can run this node safely with your active configurations!

Deploy on High-Performance GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.

🚀 Deploy on Vast.ai

Deploy on Cloud GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.

🚀 Deploy on RunPod

Python Dependencies

Direct Answer: Running ComfyUI-MultiModal-Prompt-Nodes requires installing the following Python package dependencies: dashscope>=1.20.0, llama-cpp-python>=0.3.16 # Recent JamePeng fork recommended for Qwen3-VL/Qwen3.5/Qwen3.6, numpy>=1.24.0, pillow>=10.0.0. Ensure your ComfyUI environment has these packages active before launching.

requirements.txt
dashscope>=1.20.0
llama-cpp-python>=0.3.16  # Recent JamePeng fork recommended for Qwen3-VL/Qwen3.5/Qwen3.6
numpy>=1.24.0
pillow>=10.0.0

🛠️ Interactive Setup & Dependency Resolver

Operating System:
Environment Type:
Run this terminal command in your ComfyUI root folder:
# Loading command...