Gandr
Unlimited text-to-speech and voice cloning API — flat per-stream pricing, 23 languages, low latency
| What is it | Unlimited text-to-speech and voice cloning API — flat per-stream pricing, 23 languages, low latency |
|---|---|
| Pricing | Freemium — from $10/mo |
| Free tier | Yes |
| Platform | Web Application |
| API | Yes |
| Best for | building voice-enabled AI agents, narrating audiobooks |
Data updated Aug. 12, 2026
What does Gandr do?
Gandr is a text-to-speech platform that turns written text into spoken audio. It's built for developers who need to add voice to their apps, from AI agents to audiobooks. Unlike services that charge per character, Gandr uses a flat per-stream pricing model — you pay for the stream, not the words. It supports 23 languages and can clone a voice from just ten seconds of audio, no training job required.
The platform works over WebSocket for real-time agent applications and over HTTP for narration and dubbing. It integrates with popular frameworks like LiveKit, Pipecat, Vapi, Retell, and Daily. The engine is fast: first audio byte arrives in about 146 milliseconds over the open internet, and the 95th percentile latency is still under 200 ms, making it suitable for conversational use. Every clip is watermarked with an inaudible provenance seal, and Gandr does not train on your data. Voice cloning is cached after the first use, so there's no per-voice fee.
Gandr is best for developers building voice AI agents, content creators producing audiobooks or multilingual dubbing, and enterprises that need reliable, low-latency TTS at scale. The flat pricing makes costs predictable, and the simple API lets you get started with a key in seconds.
Key features
What makes it stand outWho is Gandr for?
Who benefits most from this toolPricing
Free tier available — start without a credit cardFree
- 50,000 tokens
- 50,000 tokens to start
Take a Gandr
- 1,000,000 tokens per month
- 1 million tokens per month
- Resets monthly
Pro
- 5,000,000 tokens per month
- 5 million tokens per month
- Resets monthly
Scale
- 1 streams
- One unlimited stream
- Unmetered
Enterprise
- 500 lines and up
- Volume, burst and failover
Gallery
Click any image to enlargeAlternatives in Text-to-Speech
AI text-to-speech API with ultra-realistic voices, voice cloning, and low-latency streaming.
API platform for real-time, expressive text-to-speech, speech-to-text, and voice cloning to power AI agents.
AI voice generator and text-to-speech platform — create realistic voiceovers, dubbing, and voice agents.
High-quality, ultra-low-cost text-to-speech API for developers — 11x cheaper than ElevenLabs.
Text to Speech API by Listnr. Adding text to speech capabilities to your apps has never been easier! Listnrs' TTS API offers over 600 voices in 80 different languages, users can access all our Standard and Premium voices programmatically.
Free online AI voice generator — convert text to speech in multiple languages with natural voices, no signup required
AI text-to-speech tool — paste text or upload a PDF and get natural-sounding audio in seconds
Free online text-to-speech tool — convert text to natural-sounding AI voices in 140+ languages and download as MP3.