Miso One is an open-weights, 8B-parameter text-to-speech (TTS) system developed by Miso Labs. It is designed specifically for producing highly realistic, expressive, and emotionally varied English conversational speech, making it ideal for voice-agent research and developer workflows. Built on a Sesame-style conversational speech model (CSM) architecture with Mimi audio codes, it features a highly optimized inference capability boasting a published low latency of 110 ms. In addition to text-to-speech generation, the model supports voice continuation and one-shot voice cloning from audio context with clear
Open-weights 8B AI text-to-speech model for expressive English speech.
- Running low-latency conversational speech generation locally for voice-agent research
- Generating high-quality, expressive voiceovers, narrations, and captions for creator content
- Creating instant voice clones and personal voice models using brief, consented audio context prompts
- Users can evaluate Miso One by reading its official model card on the repository or Hugging Face page
- trying out the hosted web demo to check voice quality
- or downloading the public 8B weights and inference code to run local benchmarks within their own CUDA environment. For hosted creator workflows
- users can sign up and choose a subscription plan based on their required annual or monthly character capacity.
Converts live stream chat messages into voice.


AI voice generator with 300 voices in 70+ languages for lifelike speech synthesis.


AI reading assistant that converts websites and local documents into natural speech.


AI-powered platform to transform text into engaging AI podcasts quickly and easily.


Free online text-to-speech tool with AI voices and multiple languages.


Guide and production API platform for advanced AI voice, speech, and music.


DesiVocal is a free AI voice generator for HD voice overs in multiple languages.


Converts articles and blog posts to natural-sounding audio with AI enhancements.


Free online AI text to speech converter with natural voices and download options.





