The Speech SDK class used for text to speech, offering SpeakTextAsync and SpeakSsmlAsync. Created from a SpeechConfig plus an AudioOutputConfig, it sends audio to a stream, file or speaker.
Read more: Microsoft Learn
In the Ultra Transcenders books
Each book explains SpeechSynthesizer in context, with comparison tables and the common traps.
Terms in this definition
- Speech SDK
The Azure Speech client library, installed in Python as azure-cognitiveservices-speech and imported as speechsdk, providing SpeechSynthesizer, SpeechRecognizer, SpeechConfig and classes for audio configuration.
- Speech synthesis
Turning text into spoken audio that sounds natural, for example to read messages out loud.
- AudioOutputConfig
Speech SDK class choosing a single destination for synthesised speech: a file, a named device, a custom output stream or the default speaker.
Related terms
- speak_text_async()
Method of SpeechSynthesizer that synthesises plain text without blocking, returning a future that resolves to a SpeechSynthesisResult. For SSML input, speak_ssml_async() is used.