uncloseai-speech/cloned-voices
blackops 4ac192585d F5-TTS: replace LibriSpeech ground-truth labels with real Whisper transcripts
Previous ref_texts were LibriSpeech dataset labels (ALL CAPS, no
punctuation). F5-TTS conditions on ref_text to align reference audio
prosody — commas, periods, casing matter. Now using whisper-large-v3
transcripts of the actual cloned-voices/*.wav files.

Generated via: make whisper-refs (on a GPU host).

Affects all 40 voices in both tts-1-f5 and tts-1-qwen engine blocks.
Sidecar cloned-voices/whisper_refs.json kept for traceability.
2026-05-24 14:47:25 -04:00
..
amber.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
archer.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
aria.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
atlas.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
blake.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
brooke.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
caleb.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
clara.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
cole.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
cora.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
dane.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
diana.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
eden.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
elena.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
ezra.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
faye.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
felix.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
finn.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
gemma.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
grace.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
grant.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
hazel.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
heath.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
hope.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
hugo.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
iris.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
ivan.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
ivy.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
jasper.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
jude.wav Expand to all 40 test-clean voices with idempotent registry 2026-01-27 15:30:42 -05:00
kai.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
leo.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
luna.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
marcus.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
maya.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
owen.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
ruby.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
sage.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
sofia.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
theo.wav Regenerate voices with upstream-verified genders from SPEAKERS.TXT 2026-01-27 13:22:26 -05:00
voices_metadata.json F5-TTS: replace LibriSpeech ground-truth labels with real Whisper transcripts 2026-05-24 14:47:25 -04:00
whisper_refs.json F5-TTS: replace LibriSpeech ground-truth labels with real Whisper transcripts 2026-05-24 14:47:25 -04:00