Skip to content

Dubbing Channels

Text-to-Speech (TTS) is the third step of video translation — converting translated subtitle text into spoken audio. pyVideoTrans supports 30+ dubbing channels.


Ready-to-Use (Free)

No complex configuration needed — ideal for beginners.

ProviderDescriptionRating
Edge-TTS (Free)Microsoft's free interface; natural voice; supports all languages⭐⭐⭐ Default recommendation
gTTS (Free)Google TTS; basic quality⭐⭐

⚠️ Edge-TTS may trigger rate limiting with heavy short-term usage. It is recommended to set concurrency to 1 and pause to 5–10 seconds in More Settings.


Local Built-in (Free)

Models are downloaded automatically on first use and run completely offline.

ProviderDescriptionGPU AcceleratedRating
Qwen3-TTS (Local Built-in)Alibaba open-source; supports Chinese, English, Japanese, Korean, etc.⭐⭐⭐ Recommended
MOSS-TTS-Nano (Local Built-in)Supports 20 languages⭐⭐
Piper (Local Built-in)Lightweight; supports 20 languages⭐⭐
VITS (Local Built-in)Chinese-English dubbing⭐⭐
Supertonic3 (Local Built-in)English, Korean, Spanish, French dubbing⭐⭐
ChatterBox (Local Built-in)22 languages; high quality⭐⭐⭐ Recommended

Professional Cloud Services (API Key Required)

ProviderDescriptionRating
Azure TTSMicrosoft's professional speech service⭐⭐⭐
OpenAI TTSLeading voice technology⭐⭐⭐
ByteDance TTS 2.0Natural Chinese pronunciation⭐⭐⭐
Alibaba Qwen-TTSAlibaba Cloud speech synthesis⭐⭐⭐
Gemini TTSGoogle TTS⭐⭐
Elevenlabs.ioAI audio technology company⭐⭐⭐
302.AIAggregation platform⭐⭐
MinimaxiRequires top-up to use⭐⭐
Xiaomi TTSXiaomi AI open platform⭐⭐
X.AI TTSx.ai platform⭐⭐

Local Deployment (Advanced)

ProviderDescriptionVoice CloningRating
OmniVoice-TTSSupports nearly all languages⭐⭐⭐ Recommended
GPT-SoVITSClone with minimal audio samples⭐⭐⭐ Recommended
F5-TTSChinese-English cloning⭐⭐⭐ Recommended
Index-TTSChinese-English cloning⭐⭐⭐ Recommended
Confucius-TTS14 languages⭐⭐⭐
VoxCPM-TTS10+ languages⭐⭐⭐
CosyVoiceChinese, English, Japanese, Korean, and 10+ others⭐⭐
ChatTTSChinese and English⭐⭐
Fish-TTSAll built-in languages
Kokoro-TTSChinese, English, Korean, Italian, Portuguese, German, French, Hindi
Spark-TTSEnglish⭐⭐
Dia-TTSEnglish⭐⭐
clone-voiceNo longer maintained

Using Reference Audio

Voice cloning providers require a reference audio file. Place WAV files in the f5-tts/ directory with the format filename.wav#spoken text in the audio.

For details, see Voice Cloning & Multi-Role Dubbing