What is TTS?
Text-to-Speech — the task of generating spoken audio from written text.
TTS wants the opposite of what ASR wants: clean, stable recordings with a consistent voice. The same sentence is best recorded many times by the same speaker in the same state.
That is why audiobook-style narration is worth so much in TTS training — long sessions, one consistent voice, clear pronunciation.
Related terms
-
ASR
Automatic Speech Recognition — the task of turning recorded speech into text.
-
Read Speech
Recordings of speakers reading specified text, with clear pronunciation and known text, the base material for TTS and ASR.
-
Speaker Identification
Identifying who is speaking from the voice itself. Also called voice biometrics.
Keep reading
-
Full glossary
Every term we explain, in one list.
-
How to buy training data
Where these terms actually show up, and which ones change a quote.
-
Help center
How a project runs, from specification to delivery.
Submit a sourcing request
Tell us the language, the hours, and what the data needs to look like. You will get a real number and a real timeline — not a range. If we cannot source it well, we will tell you that instead.
- Pilot batch before the full run, so problems surface early.
- Consent documentation delivered with the data.
- No medical or clinical data. No recorded telephone calls.