Fish Speech is a text-to-speech (TTS) tool developed by the creators of So-VITS-SVC and Bert-VITS2. It can synthesize natural and fluent speech from just 15 seconds of any voice, maintaining the given timbre, style, and accent. Fish Audio is a platform for audio generation, offering various voice models for users to discover and use.
Text-to-speech tool that synthesizes natural speech from short voice samples.
- Generating speech in a specific voice for audiobooks
- Creating voiceovers for videos
- Developing virtual assistants with personalized voices
- Generating speech for accessibility purposes
- Users can discover and use pre-built voice models or build their own. The platform offers a text-to-speech toolkit where users can input text and select a voice model to generate speech.

AI text-to-speech platform with 20,000+ character and celebrity voices for professional audio.


Real-time AI voice changer with vast effects for gaming, streaming, and calls.


Voice cloning and sound design app for cloning, mimicking, and designing voices.


An app for audio conversations with AI celebrity avatars.


Real-time AI voice changer with soundboard, AI cover, music generator, and audio enhancer.


Audyo creates human-quality audio from text with easy editing and voice options.


Voice cloning and AI speech studio for creating content with custom or celebrity voices.


AI-powered voice changer software for real-time and file-based voice modification.


AI video & voice generator cloning celebrity voices for captivating videos.







