Home
Jobs
Events
Blogs
Library
contact us
[
Stack Layer
]
ASR models that turn caller audio into text in real time.
All layers
Submit a company
UNIO EDITORS’ PICKS
#1
AssemblyAI
Speech-to-text and audio intelligence APIs with streaming transcription.
#2
Deepgram
Fast, accurate speech-to-text APIs built for real-time voice agents.
#3
Gladia
Real-time multilingual speech-to-text API for voice products.
#4
Sarvam AI
Sovereign Indic LLMs plus STT/TTS models.
Speech-to-Text
IIT Bombay-led govt consortium; speech model Shrutam.
+LLM +TTS
India-ready
+TTS
Text-to-Speech
Expressive multilingual text-to-speech and voice cloning, plus a conversational agents platform.
+STT +Orchestration
Orchestration & Agent Platforms
Own speech models and enterprise agent platform.
+Realtime +Orchestration
+TTS +LLM
Custom Indic speech models, voice agents, secure deployments.
Voice AI cloud: one API to run any speech-to-text, text-to-speech or voice-cloning model across 20+ regions, with data residency built in.
Speech-to-text API with real-time and batch transcription across 50+ languages and accents.
Full-stack voice AI with ASR/TTS APIs and a voice-native LLM, 25+ languages.