Yuan-ManX/ComfyUI-AudioX

By Yuan-ManXView on GitHub →

Make AudioX avialbe in ComfyUI.

How much VRAM does Yuan-ManX/ComfyUI-AudioX require?

Direct Answer: The ComfyUI node Yuan-ManX/ComfyUI-AudioX requires a minimum base VRAM of 1536MB and is optimized for GPUs with at least 6GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.

High (2-4GB)
Base VRAM:
1536MB (1.5GB)
Recommended GPU:
6GB+ VRAM
Low VRAM Mode:
✓ Supported
Estimation Confidence:
MEDIUM

Interactive VRAM Compatibility Estimator

Estimated Total VRAM: 3.00 GBTarget: 8 GB
✅ Comfortable Fit

Your GPU has plenty of headroom. You can run this node safely with your active configurations!

Deploy on High-Performance GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.

🚀 Deploy on Vast.ai

Deploy on Cloud GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.

🚀 Deploy on RunPod

What Python packages are required for Yuan-ManX/ComfyUI-AudioX?

Direct Answer: Running Yuan-ManX/ComfyUI-AudioX requires installing the following Python package dependencies: aeiou, alias-free-torch==0.0.6, auraloss==0.4.0, decord==0.6.0, descript-audio-codec==1.0.0, einops, einops_exts, ema-pytorch==0.2.3, encodec==0.1.1, huggingface_hub, importlib-resources==5.12.0, k-diffusion==0.1.1, laion-clap==1.1.6, local-attention==1.8.6, pandas==2.0.2, pedalboard==0.9.14, prefigure==0.0.9, pytorch_lightning==2.4.0, pywavelets==1.4.1, safetensors, sentencepiece==0.1.99, torch>=2.0.1, torchaudio>=2.0.2, torchmetrics==0.11.4, tqdm, transformers, v-diffusion-pytorch==0.0.2, vector-quantize-pytorch==1.9.14, wandb, webdataset==0.2.48, x-transformers<1.27.0. This node specifically requires PyTorch version 2.0.1 or newer.

requirements.txt
aeiou
alias-free-torch==0.0.6
auraloss==0.4.0
decord==0.6.0
descript-audio-codec==1.0.0
einops
einops_exts
ema-pytorch==0.2.3
encodec==0.1.1
huggingface_hub
importlib-resources==5.12.0
k-diffusion==0.1.1
laion-clap==1.1.6
local-attention==1.8.6
pandas==2.0.2
pedalboard==0.9.14
prefigure==0.0.9
pytorch_lightning==2.4.0
pywavelets==1.4.1
safetensors
sentencepiece==0.1.99
torch>=2.0.1
torchaudio>=2.0.2
torchmetrics==0.11.4
tqdm
transformers
v-diffusion-pytorch==0.0.2
vector-quantize-pytorch==1.9.14
wandb
webdataset==0.2.48
x-transformers<1.27.0

Interactive Setup & Dependency Resolver

Operating System:
Environment Type:
Run this terminal command in your ComfyUI root folder:
# Loading command...

Special Environment Requirements:

Requires PyTorch version 2.0.1+.

To update PyTorch for your selected setup, run:
# Loading PyTorch command...

Frequently Asked Questions

How much VRAM does Yuan-ManX/ComfyUI-AudioX require?

Yuan-ManX/ComfyUI-AudioX requires a minimum of 1536MB (1.5GB) of VRAM for base operation. For optimal performance, a GPU with at least 6GB of VRAM is recommended. This node supports low VRAM mode for resource-constrained setups.

Can I run Yuan-ManX/ComfyUI-AudioX on an RTX 3060, RTX 4070, or RTX 4090?

✅ RTX 3060 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 Ti (16GB): Yes, fully compatible with 12.9GB headroom. ✅ RTX 4090 (24GB): Yes, fully compatible with 20.1GB headroom

What PyTorch version does Yuan-ManX/ComfyUI-AudioX need?

Yuan-ManX/ComfyUI-AudioX requires the following PyTorch-related packages: alias-free-torch==0.0.6, ema-pytorch==0.2.3, pytorch_lightning==2.4.0, torch>=2.0.1, torchaudio>=2.0.2, torchmetrics==0.11.4, v-diffusion-pytorch==0.0.2, vector-quantize-pytorch==1.9.14. Ensure your ComfyUI environment has these installed. Ensure your PyTorch installation matches your CUDA version (use torch.version.cuda to check).

What Python packages are required for Yuan-ManX/ComfyUI-AudioX?

To run Yuan-ManX/ComfyUI-AudioX, you need to install: aeiou, alias-free-torch==0.0.6, auraloss==0.4.0, decord==0.6.0, descript-audio-codec==1.0.0, einops, einops_exts, ema-pytorch==0.2.3, encodec==0.1.1, huggingface_hub, importlib-resources==5.12.0, k-diffusion==0.1.1, laion-clap==1.1.6, local-attention==1.8.6, pandas==2.0.2, pedalboard==0.9.14, prefigure==0.0.9, pytorch_lightning==2.4.0, pywavelets==1.4.1, safetensors, sentencepiece==0.1.99, torch>=2.0.1, torchaudio>=2.0.2, torchmetrics==0.11.4, tqdm, transformers, v-diffusion-pytorch==0.0.2, vector-quantize-pytorch==1.9.14, wandb, webdataset==0.2.48, x-transformers<1.27.0. You can install these using pip or add them to your requirements.txt file.

How can I reduce VRAM usage when running Yuan-ManX/ComfyUI-AudioX?

Yuan-ManX/ComfyUI-AudioX supports low VRAM mode. To reduce memory usage: (1) Enable --lowvram or --medvram flags in ComfyUI, (2) Reduce batch size to 1, (3) Use fp16 or fp8 precision if supported, (4) Close other GPU applications.

How do I install Yuan-ManX/ComfyUI-AudioX in ComfyUI?

To install Yuan-ManX/ComfyUI-AudioX: (1) Navigate to your ComfyUI/custom_nodes directory, (2) Clone the repository: git clone https://github.com/Yuan-ManX/ComfyUI-AudioX, (3) Install dependencies: pip install -r requirements.txt (if present), (4) Restart ComfyUI. Alternatively, use ComfyUI Manager for one-click installation.