Speech is an essential part of communication. If you need to incorporate speech into your applications, the Azure AI Foundry Speech Service is a great service for the job. In this talk, we will gain an overview of some of the capabilities around Azure AI Speech, seeing how we can use the service to perform text-to-speech and speech-to-text operations. We will investigate multi-lingual speech translation, analyze the components of speech, and even dive into custom neural voices, giving your applications a unique voice. We'll also compare Azure AI Speech to OpenAI's Whisper model, learning which model performs better for specific scenarios. Along the way, we will work with the .NET and Python libraries that make this service available to a wide audience of developers. Finally, because this is a critical piece of any cloud technology conversation, we'll gain an idea of how much it all costs.
Click here to access the slides for this presentation.
The slides are licensed under Creative Commons Attribution-ShareAlike.
Click here to access demo code for this presentation.
The source code is licensed under the terms offered by the GPL.