comfyui-tts-pack

By Dlight160View on GitHub →

ComfyUI custom nodes integrating CosyVoice 2 and FishSpeech TTS engines with flexible model configuration and audio routing. (Description by CC)

How much VRAM does comfyui-tts-pack require?

Direct Answer: The ComfyUI node comfyui-tts-pack requires a minimum base VRAM of 1536MB and is optimized for GPUs with at least 6GB of VRAM. Low VRAM mode is fully supported for resource-constrained setups.

High (2-4GB)
Base VRAM:
1536MB (1.5GB)
Recommended GPU:
6GB+ VRAM
Low VRAM Mode:
✓ Supported
Estimation Confidence:
MEDIUM

Interactive VRAM Compatibility Estimator

Estimated Total VRAM: 3.00 GBTarget: 8 GB
✅ Comfortable Fit

Your GPU has plenty of headroom. You can run this node safely with your active configurations!

Deploy on High-Performance GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance GPUs on Vast.ai instantly.

🚀 Deploy on Vast.ai

Deploy on Cloud GPUs

Need more VRAM to run ComfyUI with this node? Rent low-cost, high-performance cloud GPUs on RunPod instantly.

🚀 Deploy on RunPod

What Python packages are required for comfyui-tts-pack?

Direct Answer: Running comfyui-tts-pack requires installing the following Python package dependencies: cbor2==5.9.0, conformer, deepspeed, descript-audio-codec, descript-audiotools, diffusers, einops>=0.7.0, einx, gdown, hydra-core>=1.3.2, hyperpyyaml, inflect, librosa, lightning>=2.1.0, loguru>=0.6.0, loralib>=0.1.2, modelscope, natsort>=8.4.0, numpy==1.26.4, omegaconf, onnxruntime-gpu, openai-whisper, opencc-python-reimplemented==0.1.7, pyarrow, pydantic==2.11.7, pyrootutils>=1.0.4, pytorch-lightning>=2.1.0, pyworld, safetensors>=0.4.2, soundfile, tensorboard, tensorrt-cu12==10.13.3.9; sys_platform == 'linux', tiktoken>=0.8.0, torch==2.8.0, torchaudio==2.8.0, transformers==4.57.1, vllm==0.11.0, wetext, wget, x-transformers==2.11.24. Ensure your ComfyUI environment has these packages active before launching.

requirements.txt
cbor2==5.9.0
conformer
deepspeed
descript-audio-codec
descript-audiotools
diffusers
einops>=0.7.0
einx
gdown
hydra-core>=1.3.2
hyperpyyaml
inflect
librosa
lightning>=2.1.0
loguru>=0.6.0
loralib>=0.1.2
modelscope
natsort>=8.4.0
numpy==1.26.4
omegaconf
onnxruntime-gpu
openai-whisper
opencc-python-reimplemented==0.1.7
pyarrow
pydantic==2.11.7
pyrootutils>=1.0.4
pytorch-lightning>=2.1.0
pyworld
safetensors>=0.4.2
soundfile
tensorboard
tensorrt-cu12==10.13.3.9; sys_platform == 'linux'
tiktoken>=0.8.0
torch==2.8.0
torchaudio==2.8.0
transformers==4.57.1
vllm==0.11.0
wetext
wget
x-transformers==2.11.24

Interactive Setup & Dependency Resolver

Operating System:
Environment Type:
Run this terminal command in your ComfyUI root folder:
# Loading command...

Frequently Asked Questions

How much VRAM does comfyui-tts-pack require?

comfyui-tts-pack requires a minimum of 1536MB (1.5GB) of VRAM for base operation. For optimal performance, a GPU with at least 6GB of VRAM is recommended. This node supports low VRAM mode for resource-constrained setups.

Can I run comfyui-tts-pack on an RTX 3060, RTX 4070, or RTX 4090?

✅ RTX 3060 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 (12GB): Yes, fully compatible with 9.3GB headroom. ✅ RTX 4070 Ti (16GB): Yes, fully compatible with 12.9GB headroom. ✅ RTX 4090 (24GB): Yes, fully compatible with 20.1GB headroom

What PyTorch version does comfyui-tts-pack need?

comfyui-tts-pack requires the following PyTorch-related packages: pytorch-lightning>=2.1.0, torch==2.8.0, torchaudio==2.8.0. Ensure your ComfyUI environment has these installed. Ensure your PyTorch installation matches your CUDA version (use torch.version.cuda to check).

What Python packages are required for comfyui-tts-pack?

To run comfyui-tts-pack, you need to install: cbor2==5.9.0, conformer, deepspeed, descript-audio-codec, descript-audiotools, diffusers, einops>=0.7.0, einx, gdown, hydra-core>=1.3.2, hyperpyyaml, inflect, librosa, lightning>=2.1.0, loguru>=0.6.0, loralib>=0.1.2, modelscope, natsort>=8.4.0, numpy==1.26.4, omegaconf, onnxruntime-gpu, openai-whisper, opencc-python-reimplemented==0.1.7, pyarrow, pydantic==2.11.7, pyrootutils>=1.0.4, pytorch-lightning>=2.1.0, pyworld, safetensors>=0.4.2, soundfile, tensorboard, tensorrt-cu12==10.13.3.9; sys_platform == 'linux', tiktoken>=0.8.0, torch==2.8.0, torchaudio==2.8.0, transformers==4.57.1, vllm==0.11.0, wetext, wget, x-transformers==2.11.24. You can install these using pip or add them to your requirements.txt file.

How can I reduce VRAM usage when running comfyui-tts-pack?

comfyui-tts-pack supports low VRAM mode. To reduce memory usage: (1) Enable --lowvram or --medvram flags in ComfyUI, (2) Reduce batch size to 1, (3) Use fp16 or fp8 precision if supported, (4) Close other GPU applications.

How do I install comfyui-tts-pack in ComfyUI?

To install comfyui-tts-pack: (1) Navigate to your ComfyUI/custom_nodes directory, (2) Clone the repository: git clone https://github.com/Dlight160/comfyui-tts-pack, (3) Install dependencies: pip install -r requirements.txt (if present), (4) Restart ComfyUI. Alternatively, use ComfyUI Manager for one-click installation.