Open TTS Leaderboard: Scalable Evaluation for Multilingual Text-to-Speech and Voice Cloning
TLDR π new TTS leaderboard focused on open-source and multilingual
ProofVendor claim only
Key points
- The pace of open-source text-to-speech (TTS) model releases has been incredible.
- To this end, several arena-based leaderboards have established themselves as useful reference points for the community:
- This may partly explain why open-source models are underrepresented on arena-style leaderboards: as of Sep 30, 2026, only 16 of the 92 models on Artificial Analysis are open-weights, with a similar skew on Voice Arena.
- Intelligibility: word/character error rate (WER and CER) between the prompt and the generated audio's transcript, using Qwen3 ASR (top ranking open-source model on the Open ASR Leaderboard).
Sources (1)
- [1]Open TTS Leaderboard: Scalable Evaluation for Multilingual Text-to-Speech and Voice CloningHugging Face Blog Β· Sep 30, 12:00 AM
TLDR π new TTS leaderboard focused on open-source and multilingual
The pace of open-source text-to-speech (TTS) model releases has been incredible.
Extractive summary: sentences quoted from the sources.
Before this
- Sep 29, 2026microsoft/AesCode-32B
- Sep 28, 2026Holo4: powering generalist computer-use agents
- Sep 22, 2026vllm-project/vllm v0.30.0
- Sep 9, 2026vllm-project/vllm v0.29.0
- Aug 26, 2026vllm-project/vllm v0.28.0
- Jul 25, 2026sgl-project/sglang v0.5.16