FUNDING · MODELS · #1277
Open TTS Leaderboard launches objective, scalable evaluation for multilingual TTS and voice cloning
The Open TTS Leaderboard is an open evaluation platform that ranks TTS models using objective metrics—intelligibility (WER/CER via Qwen3 ASR), speaker similarity (cosine SIM via WavLM), and speed (RTFx and TTFA measured on H200/CPU)—across Seed TTS Eval and CV3 Eval datasets. It supports multilingual and voice-cloning comparisons, provides Pareto visualizations and a Listen tab to audition outputs, and aims to speed evaluation from weeks to hours.
KEY POINTS
- The Open TTS Leaderboard is an open evaluation platform that ranks TTS models using objective metrics—intelligibility (WER/CER via Qwen3 ASR), speaker similarity (cosine SIM via WavLM), and speed (RTFx and TTFA measured on H200/CPU)—across Seed TTS Eval and CV3 Eval datasets.
- It supports multilingual and voice-cloning comparisons, provides Pareto visualizations and a Listen tab to audition outputs, and aims to speed evaluation from weeks to hours.
- It provides a fast, reproducible objective alternative to slow, voter-based TTS arenas, helping compare many open-source models and prioritize candidates for costly human preference tests.
WHY IT MATTERS
It provides a fast, reproducible objective alternative to slow, voter-based TTS arenas, helping compare many open-source models and prioritize candidates for costly human preference tests.