VocalSeparation-ComfyUI
a custom node for separation vocals from music based on [a/ZFTurbo/Music-Source-Separation-Training](https://github.com/ZFTurbo/Music-Source-Separation-Training)
How much VRAM does VocalSeparation-ComfyUI require?
Direct Answer: The ComfyUI node VocalSeparation-ComfyUI requires a minimum base VRAM of 1536MB and is optimized for GPUs with at least 6GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.
- Base VRAM:
- 1536MB (1.5GB)
- Recommended GPU:
- 6GB+ VRAM
- Low VRAM Mode:
- ✓ Supported
- Estimation Confidence:
- MEDIUM
Interactive VRAM Compatibility Estimator
Your GPU has plenty of headroom. You can run this node safely with your active configurations!
Deploy on High-Performance GPUs
Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.
Deploy on Cloud GPUs
Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.
What Python packages are required for VocalSeparation-ComfyUI?
Direct Answer: Running VocalSeparation-ComfyUI requires installing the following Python package dependencies: asteroid, audiomentations, auraloss, beartype, demucs, einops, ml_collections, numpy, omegaconf, pandas, pedalboard, protobuf, rotary_embedding_torch, scipy, segmentation_models_pytorch, spafe, timm, torch_audiomentations, torchmetrics, torchseg, tqdm, transformers. Ensure your ComfyUI environment has these packages active before launching.
asteroid
audiomentations
auraloss
beartype
demucs
einops
ml_collections
numpy
omegaconf
pandas
pedalboard
protobuf
rotary_embedding_torch
scipy
segmentation_models_pytorch
spafe
timm
torch_audiomentations
torchmetrics
torchseg
tqdm
transformersInteractive Setup & Dependency Resolver
# Loading command...Frequently Asked Questions
How much VRAM does VocalSeparation-ComfyUI require?
VocalSeparation-ComfyUI requires a minimum of 1536MB (1.5GB) of VRAM for base operation. For optimal performance, a GPU with at least 6GB of VRAM is recommended. This node supports low VRAM mode for resource-constrained setups.
Can I run VocalSeparation-ComfyUI on an RTX 3060, RTX 4070, or RTX 4090?
✅ RTX 3060 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 Ti (16GB): Yes, fully compatible with 12.9GB headroom. ✅ RTX 4090 (24GB): Yes, fully compatible with 20.1GB headroom
What PyTorch version does VocalSeparation-ComfyUI need?
VocalSeparation-ComfyUI requires the following PyTorch-related packages: rotary_embedding_torch, segmentation_models_pytorch, torch_audiomentations, torchmetrics, torchseg. Ensure your ComfyUI environment has these installed. Ensure your PyTorch installation matches your CUDA version (use torch.version.cuda to check).
What Python packages are required for VocalSeparation-ComfyUI?
To run VocalSeparation-ComfyUI, you need to install: asteroid, audiomentations, auraloss, beartype, demucs, einops, ml_collections, numpy, omegaconf, pandas, pedalboard, protobuf, rotary_embedding_torch, scipy, segmentation_models_pytorch, spafe, timm, torch_audiomentations, torchmetrics, torchseg, tqdm, transformers. You can install these using pip or add them to your requirements.txt file.
How can I reduce VRAM usage when running VocalSeparation-ComfyUI?
VocalSeparation-ComfyUI supports low VRAM mode. To reduce memory usage: (1) Enable --lowvram or --medvram flags in ComfyUI, (2) Reduce batch size to 1, (3) Use fp16 or fp8 precision if supported, (4) Close other GPU applications.
How do I install VocalSeparation-ComfyUI in ComfyUI?
To install VocalSeparation-ComfyUI: (1) Navigate to your ComfyUI/custom_nodes directory, (2) Clone the repository: git clone https://github.com/AIFSH/VocalSeparation-ComfyUI, (3) Install dependencies: pip install -r requirements.txt (if present), (4) Restart ComfyUI. Alternatively, use ComfyUI Manager for one-click installation.