Enterprise Voice AI platform designed for developers building voice-first products using speech-to-text, text-to-speech, or speech-to-speech APIs. Over 200,000 developers build with Deepgram's voice-native foundational models, accessed via APIs or self-managed software. Start building with $200 in free credits!
Voice Assistants — AI tools
240 tools in this categorySpeech in and speech out: text-to-speech in dozens of languages, cloned and synthetic voices, transcription of calls and interviews, and APIs for building agents that talk back. Used for narration, website accessibility and phone-line automation.
APIs, datasets, and models to add emotional intelligence to voice AI applications
Generate realistic AI voiceovers, clone voices, and create AI avatar videos in over 60 languages.
AI-powered speech recognition and healthcare solutions — now part of Microsoft
AI voice generator and text-to-speech platform — create natural-sounding voiceovers in multiple languages.
Adds realistic text-to-speech to websites, apps, and documents with 280+ AI voices in 80+ languages.
AI transcription software that turns audio, video, and live conversations into searchable, editable text in 30+ languages.
AI voice generator and text-to-speech platform — create realistic voiceovers, dubbing, and voice agents.
A lightweight Mac app that uses OpenAI's AI to transcribe your voice into formatted text in any application.
AI text-to-speech tool that reads documents, PDFs, and web pages aloud with natural-sounding voices.
AI voice companion that understands conversational context, tone, and emotion for natural dialogue
API platform for voice and SMS — build AI Voice Agents, capture conversation data, and integrate with your CRM.
Online speech recognition tool — dictate text and use voice commands for punctuation in your browser.
Speech-to-text API with real-time transcription, speaker detection, and audio understanding for developers
AI speech-to-text dictation app that auto-formats text for emails, messages, and code — open source and private.
AI celebrity voice generator — type text, get audio of famous characters saying it in seconds
AI voice intelligence platform that analyzes conversations to detect risks, emotions, and intent in real-time
Sonic is a blazing fast, lifelike generative voice API (🚀 135ms model latency). Build high quality, real time voice experiences with a diverse voice library, instant voice cloning, voice mixing, and voice design with speed and emotion control.
AI-powered speech recognition API that converts audio to text with high accuracy, even in noisy environments.
Real-time speech-to-text dictation software for Windows that lets you type with your voice.
Real-time voice AI that converts whispers into clear speech for accessible calls, private conversations, and noisy environments
Resemble AI offers cutting-edge tools for speech-to-speech, text-to-speech, and voice cloning. With seamless audio editing capabilities and advanced deepfake detection, Resemble empowers users to create, modify, and safeguard voice content effortlessly. Ideal for creators looking to enhance their audio projects with AI-driven precision.
Open source library for building hyperrealistic voice AI agents that can have natural phone conversations
AI prank call generator — enter a phone number, choose a voice, and let AI handle the hilarious conversation