Definition·
Capabilities

Text-to-Speech

Technology that synthesizes a realistic voice from written text.

Detailed explanation

Text-to-Speech (TTS) generates natural speech, sometimes cloned from a real voice. Recent models handle prosody, emotions, multiple languages and accents, enabling voice assistants, audiobooks and automatic dubbing.

Examples

Siri or Google Assistant voice
ElevenLabs to clone a voice
Audio reading of an article
Automatic video dubbing

Frequently asked questions

Fraud risks?

Voice cloning enables CEO scams; add extra identity checks for sensitive flows.

Related terms

Last updated: 7/15/2026

Talent AI

Turn theory into practice

Post a mission or join the community of top AI, Data and Machine Learning experts.