A production-grade platform for speech, voice, dubbing, audio, and conversational agents.
Full evaluationGlossary · Updated Sep 2, 2026
Text-to-speech (TTS)
Synthesising natural spoken audio from written text, often with selectable voices, emotions, and languages.
Definition
Modern text-to-speech uses neural models that produce expressive, human-like voices with control over pace, emphasis, and style, and can stream audio with low latency for conversational agents. Products offer voice libraries, cloning, multilingual output, and APIs. Licensing determines whether generated audio can be used commercially.
Why it matters when choosing a tool
TTS quality and rights matter for narration, accessibility, and voice agents; test voices on your script and confirm commercial terms before publishing.
Where you will meet it
AI Audio & Voice, AI Video Generation, AI Learning & Tutoring
Related terms
Tools where this matters
Reviewed products in the related categories
Edit video and podcasts by editing the transcript, with AI voices, filler removal, and studio-quality fixes.
Full evaluationProfessional video editing, color grading, visual effects, and audio post-production suite
Full evaluationAI creation inside a broad collaborative design and content production suite.
Full evaluationA generative video and creative production platform for controllable AI filmmaking.
Full evaluationGoogle's source-grounded research notebook that answers only from the documents you add and makes audio overviews.
Full evaluation