click any column header to sort

TTS Bench — Samples — 2026-06-02_2126

Rig: windows-5090 — AMD Ryzen 9 9950X3D 16-Core Processor (16C) · NVIDIA GeForce RTX 5090 32GB · 126 GB RAM · Windows 11
Label: cloning · chris_hemsworth_15s · outetts — ref chris_hemsworth_15s.wav
5 prompt(s) · one section per prompt · all models ranked by warm TTFA (fastest first) within each
Each prompt section shows every model's audio output, ordered by warm TTFA (fastest first). Click any audio player to hear that model's rendering.

Reference voice

Each model below was given this clip + transcript as the voice to imitate. Source: chris_hemsworth_15s.wav

Prompt 1

[en]"Open the browser and read my email."
Rank Model Device TTFA warm Audio
1 OuteTTS 1.0 1B cuda 6.26s

Prompt 2

[en]"I'll start a new git branch, push the changes, and open a pull request when the tests pass."
Rank Model Device TTFA warm Audio
1 OuteTTS 1.0 1B cuda 13.48s

Prompt 3

[en]"The Parakeet TDT zero point six billion parameter model achieves one point six nine percent word error rate on LibriSpeech test-clean, beating Whisper Large V3 at two point seven percent while running at over two thousand times realtime on a single GPU."
Rank Model Device TTFA warm Audio
1 OuteTTS 1.0 1B cuda 39.66s

Prompt 4

[en]"Run pytest tests slash test underscore voice dot py with verbose flag and capture flag set to no."
Rank Model Device TTFA warm Audio
1 OuteTTS 1.0 1B cuda 17.15s

Prompt 5

[fr]"Bonjour, je m'appelle Cicero et je vais vous aider avec votre code aujourd'hui."
Rank Model Device TTFA warm Audio
1 OuteTTS 1.0 1B cuda 10.67s