Text-to-speech · measured leaderboard
Only rows measured by CodeSOTA rank here.
The May hard-text rows are withheld: their manifests contain placeholder hashes and do not account for the claimed 30 audio samples. No auditable measured hard-text ranking is available until complete artifacts are restored. Blind Elo is a separate preference study and does not supply hard-text accuracy or latency results.
Separate study · not benchmark-ranked
Active blind Elo sample pool
These seven male-voice systems have the shared prompt audio ready for preference voting. They will not enter the hard-text measured ranking until ASR transcripts, diffs, entity scoring, latency logs, and artifacts are published for that benchmark.
| Model | Vendor | Voice condition | Elo clips | Measured benchmark status |
|---|---|---|---|---|
Gradium TTS | Gradium | Kent | 30 | audio ready · scoring pending |
Gradium TTS | Gradium | Damon | 30 | audio ready · scoring pending |
Gradium TTS | Gradium | Russell | 30 | audio ready · scoring pending |
Kokoro v1.0 | Hexgrad | am_michael | 30 | audio ready · scoring pending |
Speech-02 Turbo | MiniMax | English_Deep-VoicedGentleman | 30 | audio ready · scoring pending |
Speech-02 HD | MiniMax | English_Deep-VoicedGentleman | 30 | audio ready · scoring pending |
Qwen3 TTS | Qwen | Aiden | 30 | audio ready · scoring pending |
Chatterbox Turbo | Resemble AI | Andy | 30 | audio ready · scoring pending |
Chatterbox Turbo | Resemble AI | default study voice | 30 | audio ready · scoring pending |
ElevenLabs v3 | ElevenLabs | James | 30 | audio ready · scoring pending |
XTTS v2 | Coqui | Damien Black | 30 | audio ready · scoring pending |
| Rank | Model | Benchmark | Verification | Entity acc. | WER | CER | p95 TTFB | CI | Artifacts |
|---|---|---|---|---|---|---|---|---|---|
| No verified hard-text benchmark rows are currently published. Prior placeholder-backed rows are withheld pending complete audio manifests, real hashes and sample-level outputs. | |||||||||
Harness commands
codesota-tts synth --model <id> --eval <track> --out runs/<run_id> codesota-tts score --run runs/<run_id> --metrics wer,cer,entity,utmos,latency codesota-tts report --run runs/<run_id> --publish
TTS Eval v2 tracks
clean-read-en
hardtext-en
hardtext-pl
longform
cloning
controllability