Voice & Speech AI Companies
Voice and speech AI companies build systems for automatic speech recognition (ASR), text-to-speech (TTS), voice cloning, and real-time translation. Advances in neural codec language models have made synthetic voices nearly indistinguishable from human speech, enabling new applications in accessibility, entertainment, and enterprise communications.
132 Companies
42 Countries
Showing the top 120 of 132 Voice & Speech AI companies by capital raised. Browse all 132 →
Where the Voice & Speech AI companies are
Browse other categories
Frequently Asked Questions
- What is speech AI used for?
- Speech AI powers voice assistants (Siri, Alexa), live meeting transcription, voice cloning for content creators, multilingual customer service, accessibility tools for the hearing-impaired, and dubbing for video localisation.
- How realistic is AI-generated voice?
- Modern TTS systems achieve MOS (Mean Opinion Score) scores comparable to human speech. Leading systems from ElevenLabs, Play.ht, and LMNT can clone a voice from seconds of audio.
- Who are the top voice AI companies?
- Top companies include Spotify AI, Uniphore, Verbit, alongside ElevenLabs, Deepgram, AssemblyAI, and Whisper (OpenAI).