Miso One
Open-weights AI text-to-speech model for expressive English conversational speech and low-latency voice generation
| What is it | Open-weights AI text-to-speech model for expressive English conversational speech and low-latency voice generation |
|---|---|
| Pricing | Freemium — from $9.9/mo |
| Free tier | Yes |
| Platform | Web Application |
| API | Yes |
| Best for | Building interactive voice agents, Testing low-latency speech synthesis |
| Domain registered | 2026 |
Data updated Aug. 1, 2026
What does Miso One do?
Miso One is a text-to-speech AI tool that converts written English into natural-sounding spoken audio. It specializes in conversational speech with emotional variation, making it sound more human than robotic voice synthesis. The tool can also continue from existing audio prompts, allowing for voice cloning and consistent tone maintenance across longer audio pieces.
The system is built on Miso TTS 8B, an 8-billion parameter open-weights model that developers can download and run locally. It's designed for low-latency generation, with claims of 110ms response times suitable for voice agent applications. Unlike many cloud-based TTS services, Miso One gives users full control over deployment while requiring significant local GPU resources.
This tool is ideal for AI researchers, developers building voice assistants, and content creators needing expressive narration. It's particularly valuable for those testing voice agent latency, experimenting with voice cloning, or requiring local deployment without cloud dependencies. The open weights make it suitable for teams wanting to inspect and modify the underlying model for specific use cases.
Key features
What makes it stand outWho is Miso One for?
Who benefits most from this toolPricing
Free tier available — start without a credit cardBasic
- 800 voice credits
- 80000 TTS characters included
- 80,000 TTS characters per month
- 800 voice credits
- Maximum 1,000 characters per conversion
- Up to 40 instant voice clones
- Voice Design previews included via credits
- Private voice model creation included via credits
- Basic support by email
Pro
- 3500 voice credits
- 350000 TTS characters included
- 350,000 TTS characters per month
- 3,500 voice credits
- Maximum 1,000 characters per conversion
- Up to 175 instant voice clones
- Voice Design previews included via credits
- Private voice model creation included via credits
- Priority support for voice workflows
Enterprise
- 8000 voice credits
- 800000 TTS characters included
- 800,000 TTS characters per month
- 8,000 voice credits
- Maximum 1,000 characters per conversion
- Up to 400 instant voice clones
- Voice Design previews included via credits
- Private voice model creation included via credits
- Priority support from our team
Trust & presence
Gallery
Click any image to enlargeAlternatives in Text-to-Speech
AI voice generator — turn text into expressive speech, clone voices, and download private audio.
Small, efficient AI models for text-to-speech, speech-to-text, and conversational AI with fast response times
AI text-to-speech API with ultra-realistic voices, voice cloning, and low-latency streaming.
Convert text to realistic AI voiceovers with multiple languages, voices, and customization options
Generate realistic, human-like speech from text using OpenAI's advanced voice models.
Free frontend for OpenAI's TTS API — type text, pick a voice, get high-quality speech in seconds
Free online text-to-speech tool — convert text to natural-sounding AI voices in 140+ languages and download as MP3.
AI voice tool — generate realistic speech from text in multiple languages and voices.