Yuan-ManX/ComfyUI-AudioX
Make AudioX avialbe in ComfyUI.
Quick Technical Summary: Yuan-ManX/ComfyUI-AudioX
- Base VRAM Footprint:
- 1536 MB (6 GB Tier)
- Primary Dependencies:
- aeiou, alias-free-torch==0.0.6, auraloss==0.4.0, decord==0.6.0, descript-audio-codec==1.0.0, einops, einops_exts, ema-pytorch==0.2.3, encodec==0.1.1, huggingface_hub, importlib-resources==5.12.0, k-diffusion==0.1.1, laion-clap==1.1.6, local-attention==1.8.6, pandas==2.0.2, pedalboard==0.9.14, prefigure==0.0.9, pytorch_lightning==2.4.0, pywavelets==1.4.1, safetensors, sentencepiece==0.1.99, torch>=2.0.1, torchaudio>=2.0.2, torchmetrics==0.11.4, tqdm, transformers, v-diffusion-pytorch==0.0.2, vector-quantize-pytorch==1.9.14, wandb, webdataset==0.2.48, x-transformers<1.27.0
- Min PyTorch / CUDA:
- PyTorch 2.0.1 | CUDA 12.1+
- GitHub Repository:
- https://github.com/Yuan-ManX/ComfyUI-AudioX
Citation Note: Data sourced from VRAM DB. For complete workflow OOM estimations, use the VRAM DB Workflow Analyzer.
How much VRAM does Yuan-ManX/ComfyUI-AudioX require?
Direct Answer: The ComfyUI node Yuan-ManX/ComfyUI-AudioX requires a minimum base VRAM of 1536MB and is optimized for GPUs with at least 6GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.
- Base VRAM:
- 1536MB (1.5GB)
- Recommended GPU:
- 6GB+ VRAM
- Low VRAM Mode:
- ✓ Supported
- Estimation Confidence:
- MEDIUM
Cheapest VRAM Upgrade Paths (Live Market Prices):
- GeForce RTX 3060 12GB (Ultimate Budget VRAM King)──► Used: $209.62View eBay ↗
- GeForce RTX 4060 8GB (Modern Entry-Level)──► New: $303.50View Amazon ↗
Interactive VRAM Compatibility Estimator
Your GPU has plenty of headroom. You can run this node safely with your active configurations!
Verify Compatibility for Your Specific GPU VRAM
Select your graphics card's VRAM capacity to view optimized batch sizes, suggested resolutions, and custom performance tips for Yuan-ManX/ComfyUI-AudioX:
Buy NVIDIA GeForce RTX 3060 (12GB VRAM)
Tired of renting cloud rigs? Run ComfyUI locally with absolute zero latency. Best entry-level ComfyUI experience. Avoids immediate VRAM limitations on basic LoRA training.
Are you the author of this node?
Help your users avoid out-of-memory errors by displaying this professional, dynamic VRAM badge on your GitHub README. Copy the markdown below to embed it with a backlink directly to this hardware specification profile.
What Python packages are required for Yuan-ManX/ComfyUI-AudioX?
Direct Answer: Running Yuan-ManX/ComfyUI-AudioX requires installing the following Python package dependencies: aeiou, alias-free-torch==0.0.6, auraloss==0.4.0, decord==0.6.0, descript-audio-codec==1.0.0, einops, einops_exts, ema-pytorch==0.2.3, encodec==0.1.1, huggingface_hub, importlib-resources==5.12.0, k-diffusion==0.1.1, laion-clap==1.1.6, local-attention==1.8.6, pandas==2.0.2, pedalboard==0.9.14, prefigure==0.0.9, pytorch_lightning==2.4.0, pywavelets==1.4.1, safetensors, sentencepiece==0.1.99, torch>=2.0.1, torchaudio>=2.0.2, torchmetrics==0.11.4, tqdm, transformers, v-diffusion-pytorch==0.0.2, vector-quantize-pytorch==1.9.14, wandb, webdataset==0.2.48, x-transformers<1.27.0. This node specifically requires PyTorch version 2.0.1 or newer.
aeiou
alias-free-torch==0.0.6
auraloss==0.4.0
decord==0.6.0
descript-audio-codec==1.0.0
einops
einops_exts
ema-pytorch==0.2.3
encodec==0.1.1
huggingface_hub
importlib-resources==5.12.0
k-diffusion==0.1.1
laion-clap==1.1.6
local-attention==1.8.6
pandas==2.0.2
pedalboard==0.9.14
prefigure==0.0.9
pytorch_lightning==2.4.0
pywavelets==1.4.1
safetensors
sentencepiece==0.1.99
torch>=2.0.1
torchaudio>=2.0.2
torchmetrics==0.11.4
tqdm
transformers
v-diffusion-pytorch==0.0.2
vector-quantize-pytorch==1.9.14
wandb
webdataset==0.2.48
x-transformers<1.27.0Interactive Setup & Dependency Resolver
# Loading command...Special Environment Requirements:
Requires PyTorch version 2.0.1+.
# Loading PyTorch command...Frequently Asked Questions
How much VRAM does Yuan-ManX/ComfyUI-AudioX require?
Yuan-ManX/ComfyUI-AudioX requires a minimum of 1536MB (1.5GB) of VRAM for base operation. For optimal performance, a GPU with at least 6GB of VRAM is recommended. This node supports low VRAM mode for resource-constrained setups.
Can I run Yuan-ManX/ComfyUI-AudioX on an RTX 3060, RTX 4070, or RTX 4090?
✅ RTX 3060 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 Ti (16GB): Yes, fully compatible with 12.9GB headroom. ✅ RTX 4090 (24GB): Yes, fully compatible with 20.1GB headroom
How much VRAM does Yuan-ManX/ComfyUI-AudioX take on an RTX 3060 vs RTX 4090?
On an RTX 3060 (12GB VRAM), Yuan-ManX/ComfyUI-AudioX runs smoothly on an RTX 3060 (12GB) with 9.3GB of headroom. This is sufficient to run the node alongside standard SD 1.5 and SDXL workflows in full precision. On an RTX 4090 (24GB VRAM), the node runs with extreme headroom on an RTX 4090 (24GB) with 20.1GB of dedicated headroom. This allows you to combine the node with massive models (like FLUX.1 Dev, Schnell, or Hunyuan Video) in full precision (FP16) without any offload flags.
What PyTorch version does Yuan-ManX/ComfyUI-AudioX need?
Yuan-ManX/ComfyUI-AudioX requires the following PyTorch-related packages: alias-free-torch==0.0.6, ema-pytorch==0.2.3, pytorch_lightning==2.4.0, torch>=2.0.1, torchaudio>=2.0.2, torchmetrics==0.11.4, v-diffusion-pytorch==0.0.2, vector-quantize-pytorch==1.9.14. Ensure your ComfyUI environment has these installed. Ensure your PyTorch installation matches your CUDA version (use torch.version.cuda to check).
What Python packages are required for Yuan-ManX/ComfyUI-AudioX?
To run Yuan-ManX/ComfyUI-AudioX, you need to install: aeiou, alias-free-torch==0.0.6, auraloss==0.4.0, decord==0.6.0, descript-audio-codec==1.0.0, einops, einops_exts, ema-pytorch==0.2.3, encodec==0.1.1, huggingface_hub, importlib-resources==5.12.0, k-diffusion==0.1.1, laion-clap==1.1.6, local-attention==1.8.6, pandas==2.0.2, pedalboard==0.9.14, prefigure==0.0.9, pytorch_lightning==2.4.0, pywavelets==1.4.1, safetensors, sentencepiece==0.1.99, torch>=2.0.1, torchaudio>=2.0.2, torchmetrics==0.11.4, tqdm, transformers, v-diffusion-pytorch==0.0.2, vector-quantize-pytorch==1.9.14, wandb, webdataset==0.2.48, x-transformers<1.27.0. You can install these using pip or add them to your requirements.txt file.
How can I reduce VRAM usage when running Yuan-ManX/ComfyUI-AudioX?
Yuan-ManX/ComfyUI-AudioX supports low VRAM mode. To reduce memory usage: (1) Enable --lowvram or --medvram flags in ComfyUI, (2) Reduce batch size to 1, (3) Use fp16 or fp8 precision if supported, (4) Close other GPU applications.
How do I install Yuan-ManX/ComfyUI-AudioX in ComfyUI?
To install Yuan-ManX/ComfyUI-AudioX: (1) Navigate to your ComfyUI/custom_nodes directory, (2) Clone the repository: git clone https://github.com/Yuan-ManX/ComfyUI-AudioX, (3) Install dependencies: pip install -r requirements.txt (if present), (4) Restart ComfyUI. Alternatively, use ComfyUI Manager for one-click installation.