Classification
2- Complexity
- High
- Impact area
- Technical
Text-to-Speech (TTS) describes automatic generation of spoken audio from text, often using neural models. It covers quality, prosody, latency, privacy and integration aspects.
360° overview
Six perspectives place the building block in context. The numbers show where each perspective continues in the reading path.
The building block at a glance
Text-to-Speech
360°
Integrations
Web Speech API / browser integration
+2
Clear separation between text preprocessing, model selection and output layer
Value stream stage: Build