35a6d0981d43a215cf6a1e355fc0bb82a005c0f3
ComfyUI Audio Nodes
ComfyUI-EdgeTTS is a powerful text-to-speech node for ComfyUI, leveraging Microsoft's Edge TTS capabilities. It enables seamless conversion of text into natural-sounding speech, supporting multiple languages and voices. Ideal for enhancing user interactions, this node is easy to integrate and customize, making it perfect for various applications.
Features
Edge TTS Node
- Edge TTS: Convert text to speech using Microsoft Edge TTS
- Multiple languages and voices support
- Adjustable speech rate and pitch
- High-quality voice synthesis
- Configurable via config.json
Speech to Text Node
- Whisper STT: High-accuracy speech recognition
- Multiple language support with auto-detection
- Multiple model sizes (tiny to large)
- Supports ComfyUI audio format
- Language detection confidence reporting
Audio File Node
- Save Audio: Export audio files
- Supports WAV, MP3, FLAC formats
- Quality presets (high/medium/low)
- Custom file naming and paths
- Automatic file numbering
Installation
Method 1. install on ComfyUI-Manager, search Comfyui-EdgeTTS and install
install requirment.txt in the ComfyUI-EdgeTTS folder
./ComfyUI/python_embeded/python -m pip install -r requirements.txt
Method 2. Clone this repository to your ComfyUI custom_nodes folder:
cd ComfyUI/custom_nodes
git clone https://github.com/1038lab/ComfyUI-EdgeTTS.git
install requirment.txt in the ComfyUI-EdgeTTS folder
./ComfyUI/python_embeded/python -m pip install -r requirements.txt
Requirements
- Python packages (see requirements.txt)
- CUDA compatible GPU (optional, for faster Whisper processing)
Usage Examples
Text to Speech
- Add Edge TTS node to workflow
- Input text and select voice
- Adjust speed and pitch if needed
- Connect to Save Audio node for export
Speech to Text
- Add Whisper STT node
- Connect audio input
- Select model size and language (or auto-detect)
- Run to get transcription
Supported Voices
| Language | Female Voices | Male Voices |
|---|---|---|
| Main Languages | ||
| Chinese | XiaoXiao (Cheerful), XiaoYi (Warm) | Yunjian (Formal), Yunxi (Casual), Yunxia (Warm), Yunyang (Pro) |
| English | Jenny (Casual), Aria (Pro), Sonia (GB), Natasha (AU) | Guy (Casual), Davis (Pro), Ryan (GB), William (AU) |
| Japanese | Nanami (Natural), Aoi (Cheerful) | Keita (Formal) |
| Korean | SunHi (Warm) | InJoon (Formal) |
| European Languages | ||
| French | Denise (Pro) | Henri (Formal) |
| German | Katja (Clear) | Conrad (Pro) |
| Spanish | Elvira (Warm) | Alvaro (Friendly) |
| Russian | Svetlana (Pro) | Dmitry (Formal) |
| Italian | Elsa (Warm) | Diego (Formal) |
| Portuguese | Francisca (BR), Raquel (PT) | Antonio (BR) |
| Dutch | Colette (Warm) | Maarten (Formal) |
| Polish | Zofia (Natural) | Marek (Formal) |
| Turkish | Emel (Warm) | Ahmet (Formal) |
| Asian Languages | ||
| Arabic | Zariyah (Warm) | Hamed (Formal) |
| Hindi | Swara (Warm) | Madhur (Formal) |
| Indonesian | Gadis (Warm) | Ardi (Formal) |
Each language provides at least one male and female voice option, allowing you to choose different voice styles based on your needs.
Credits
- Edge TTS: Microsoft Edge TTS
- Whisper: OpenAI Whisper
Languages
Python
100%