ComfyUI-JoyCaption
Joy Caption is a ComfyUI custom node powered by the LLaVA model for efficient, stylized image captioning. Caption Tools nodes handle batch image processing and automatic separation of caption text.
VRAM Requirements
Direct Answer: The ComfyUI node ComfyUI-JoyCaption requires a minimum base VRAM of 256MB and is optimized for GPUs with at least 4GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.
- Base VRAM:
- 256MB (0.3GB)
- Recommended GPU:
- 4GB+ VRAM
- Low VRAM Mode:
- ✓ Supported
- Estimation Confidence:
- HIGH
🧮 Interactive VRAM Compatibility Estimator
Your GPU has plenty of headroom. You can run this node safely with your active configurations!
Deploy on High-Performance GPUs
Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.
Deploy on Cloud GPUs
Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.
Python Dependencies
Direct Answer: Running ComfyUI-JoyCaption requires installing the following Python package dependencies: accelerate, bitsandbytes>=0.42.0, compressed-tensors>=0.6.0, huggingface_hub>=0.19.0, pillow>=10.0.0, torch>=2.0.0, transformers>=4.36.0. This node specifically requires PyTorch version 2.0.0 or newer.
accelerate
bitsandbytes>=0.42.0
compressed-tensors>=0.6.0
huggingface_hub>=0.19.0
pillow>=10.0.0
torch>=2.0.0
transformers>=4.36.0🛠️ Interactive Setup & Dependency Resolver
# Loading command...⚠️ Special Environment Requirements:
Requires PyTorch version 2.0.0+.
# Loading PyTorch command...