Description
Text & Speech – Text to Speech and Speech to Text Converter
Text & Speech – Text to Speech and Speech to Text Converter, allowing you to create various media content such as audio books, podcasts, voice contents and also applications that talk, and build entirely new categories of speech-enabled products and also allows you to transcribe audio into text in various formats, allowing you to create transcripts of any audio and voice contents, recordings, customer service calls etc simply and efficiently..
Features of Text & Speech
- Support for over +144 Languages and Dialects for Text to Speech
- Support for over +900 Different Voices and Accents for Text to Speech
- Support for over +170 Languages & Dialects for Speeh to Text
- Support for 12 Languages for Live Transcribe for Speech to Text
- Powered By:
- Amazon Web Services (TTS/STT)
- Microsoft Azure (TTS)
- Google Cloud Platform (TTS/STT)
- IBM Cloud (TTS)
- Natural sounding voices (Neural TTS)
- Google WaveNet Voices
- Various Combination of Voice Effects for Standard Voices
- Various Combination of Voice Effects for Neural Voices
- Powerful Sound Studio
- Use any of +900 voices in a single Text Synthesize Task
- Mix up to 20 voices in a single Text Synthesize Task
- Process up to 60000 characters in a single Text Synthesize Task
- Multiple Audio Output Formats (Text to Speech):
- MP3 (AWS/Azure/GCP/IBM)
- OGG (AWS/GCP/IBM/Azure)
- WAV (GCP/IBM)
- WEBM (Azure)
- Store & redistribute speech easily via social media
- Near Real-time text synthesize
- Customize & control speech output
- Optimize Your Streaming Audio
- Adjust Speaking Styles (For Neural Voices)
- Adjust Speech Rate, Pitch, and Loudness
- Adjust Speaking Emphasis
- Pronounce digits/dates/words/abbreviations properly
- Add work/phrase replacement effect
- Mute/Beep Out any part of text/sentence
- Synthesize Large Text directly to your Amazon S3 Bucket
- Store Text to Speech results in:
- Local Server
- Amazon S3
- Wasabi Storage
- Conveniently Share synthesize results or Download
- Speaker Identification up to 5 people
- GCP instant transcribe for short audio files
- Multiple Audio Input Formats (Speech to Text):
- MP3 (AWS)
- OGG (AWS)
- WAV (AWS/GCP)
- WEBM (AWS)
- MP4 (AWS)
- FLAC (AWS/GCP)


