Skip navigation

Text to Speech

A Speech service feature that converts text to lifelike speech

Bring your apps to life with natural-sounding voices

Build apps and services that speak naturally, choosing from more than 250 voices and over 70 languages and variants. Differentiate your brand with a customised voice, and access voices with different speaking styles and emotional tones to fit your use case – from text readers to customer support chatbots.

Lifelike synthesised speech

Enable fluid, natural-sounding text to speech that matches the patterns and intonation of human voices.

Customisable text-talker voices

Create a unique voice that reflects your brand’s identity.

Fine-grained audio controls

Tune voice output for your scenarios by easily adjusting rate, pitch, pronunciation, pauses and more.

Flexible deployment

Run Text to Speech anywhere – in the cloud, on-premises or at the edge in containers.

Access a wide variety of voices for every scenario

Engage global audiences by using more than 250 voices and 70 languages and variants. Bring your scenarios, such as text readers and voice-enabled assistants, to life with highly expressive and human-like voices. Neural Text to Speech supports several speaking styles, including chat, newscast and customer service, and emotions such as cheerfulness and empathy.

Try Text to Speech with this demo app, built on our JavaScript SDK

Tailor your speech output

Fine-tune synthesised speech audio to fit your scenario. Define lexicons and control speech parameters such as pronunciation, pitch, rate, pauses and intonation with Speech Synthesis Markup Language (SSML) or with the audio content creation tool.

Deploy anywhere, from the cloud to the edge

Run Text to Speech wherever your data resides. Build lifelike speech synthesis into applications optimised for both robust cloud capabilities and edge locality using containers. Speech containers support both standard and custom text-to-speech capabilities.

Build a custom voice for your brand

Differentiate your brand with a unique custom voice. Develop a highly realistic voice for more natural conversational interfaces using the Custom Neural Voice capability, starting with 30 minutes of audio. Here are a few examples of organisations that are doing this today:

Swisscom improves customer experiences with multi-lingual voice assistant

Swisscom used Speech service to create a natural-sounding custom voice assistant with voice personas that are unique to Swisscom across English, French, German and Italian.

AT&T delights customers with immersive experiences

AT&T is showcasing the power of its 5G network with an immersive experience that allows its customers to talk directly to Bugs Bunny*.

*LOONEY TUNES and all related characters and elements © and ™ Warner Bros. Entertainment Inc. (s21)

Progressive brings Flo directly to its customers

Progressive used custom neural voice to build a natural-sounding, virtual version of Flo to help customers with everything from getting a free car insurance quote to general insurance questions.

Comprehensive privacy and security

  • The Speech service, part of Azure Cognitive Services, is certified by SOC, FedRAMP, PCI DSS, HIPAA, HITECH and ISO.
  • Your data remains yours. Your text data isn’t stored during data processing or audio generation.
  • View and delete your custom voice data and synthesised speech models at any time. Your data is encrypted while it’s in storage.
  • Backed by Azure infrastructure, the Speech service offers enterprise-grade security, availability, compliance and manageability.

Flexible pricing gives you the power and control you need

Only pay for what you use, with no upfront costs. With Text to Speech, you pay as you go based on the number of characters you convert to audio.

Guidelines for building responsible synthetic voices

Documentation and resources

Explore code samples

Take a look at our sample code

See customisation resources

Customise your speech solution with Speech studio. No code required.

Built with Text to Speech

BBC innovates how it delivers trusted content

Using Azure Cognitive Services and Azure Bot Services, the BBC created an end-to-end, customised digital voice assistant that captures its brand identity and helps it establish a new conversational relationship with its broad audiences.


Swisscom improves customer experiences with multi-lingual voice assistant

Swisscom used Speech service to create a natural sounding custom voice assistant with voice personas that are unique to Swisscom across English, French, German and Italian.


Motorola helps first responders access vital data

Motorola Solutions is helping police officers and other emergency first responders gain access to important information more quickly with a voice-powered virtual assistant.

Motorola Solutions

Universal Electronics powers connected smart homes

Universal Electronics is helping manufacturers deliver voice-enabled navigation and control capabilities that work across smart home devices.

Universal Electronics

Cheetah Mobile expands international translation

Cheetah Mobile, a mobile Internet company with app users in more than 200 countries and regions, is using Text to Speech to expand accessibility of its translation device and app to international markets.

Cheetah Mobile

Get started with Speech