Glossary → Voice & AI audio
Voice & AI audio

Text-to-speech (TTS)

Text-to-speech synthesizes spoken audio from written text — the inverse of speech recognition — with modern neural voices approaching natural prosody.

Contemporary TTS learned expressiveness from data rather than hand-built rules, enabling audiobook narration, screen readers and voice agents. Quality is judged on naturalness and correct emphasis, not just intelligibility.

Related terms