ComfyUI-Qwen-Omni
ComfyUI-Qwen-Omni is the first ComfyUI plugin that supports end-to-end multimodal interaction, enabling seamless joint generation and editing of text, images, and audio. Without intermediate steps, with just one operation, the model can simultaneously understand and process multiple input modalities, generating coherent text descriptions and voice outputs, providing an unprecedentedly smooth experience for AI creation.
How much VRAM does ComfyUI-Qwen-Omni require?
Direct Answer: The ComfyUI node ComfyUI-Qwen-Omni requires a minimum base VRAM of 1536MB and is optimized for GPUs with at least 6GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.
- Base VRAM:
- 1536MB (1.5GB)
- Recommended GPU:
- 6GB+ VRAM
- Low VRAM Mode:
- ✓ Supported
- Estimation Confidence:
- MEDIUM
Interactive VRAM Compatibility Estimator
Your GPU has plenty of headroom. You can run this node safely with your active configurations!
Deploy on High-Performance GPUs
Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.
Deploy on Cloud GPUs
Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.
What Python packages are required for ComfyUI-Qwen-Omni?
Direct Answer: Running ComfyUI-Qwen-Omni requires installing the following Python package dependencies: accelerate, bitsandbytes, modelscope, numpy, numpy, pillow, qwen_omni_utils, requests, soundfile, triton-windows. Ensure your ComfyUI environment has these packages active before launching.
accelerate
bitsandbytes
modelscope
numpy
numpy
pillow
qwen_omni_utils
requests
soundfile
triton-windowsInteractive Setup & Dependency Resolver
# Loading command...Frequently Asked Questions
How much VRAM does ComfyUI-Qwen-Omni require?
ComfyUI-Qwen-Omni requires a minimum of 1536MB (1.5GB) of VRAM for base operation. For optimal performance, a GPU with at least 6GB of VRAM is recommended. This node supports low VRAM mode for resource-constrained setups.
Can I run ComfyUI-Qwen-Omni on an RTX 3060, RTX 4070, or RTX 4090?
✅ RTX 3060 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 Ti (16GB): Yes, fully compatible with 12.9GB headroom. ✅ RTX 4090 (24GB): Yes, fully compatible with 20.1GB headroom
What Python packages are required for ComfyUI-Qwen-Omni?
To run ComfyUI-Qwen-Omni, you need to install: accelerate, bitsandbytes, modelscope, numpy, numpy, pillow, qwen_omni_utils, requests, soundfile, triton-windows. You can install these using pip or add them to your requirements.txt file.
How can I reduce VRAM usage when running ComfyUI-Qwen-Omni?
ComfyUI-Qwen-Omni supports low VRAM mode. To reduce memory usage: (1) Enable --lowvram or --medvram flags in ComfyUI, (2) Reduce batch size to 1, (3) Use fp16 or fp8 precision if supported, (4) Close other GPU applications.
How do I install ComfyUI-Qwen-Omni in ComfyUI?
To install ComfyUI-Qwen-Omni: (1) Navigate to your ComfyUI/custom_nodes directory, (2) Clone the repository: git clone https://github.com/SXQBW/ComfyUI-Qwen-Omni, (3) Install dependencies: pip install -r requirements.txt (if present), (4) Restart ComfyUI. Alternatively, use ComfyUI Manager for one-click installation.