Fish Speech is a text-to-speech (TTS) tool developed by the creators of So-VITS-SVC and Bert-VITS2. It can synthesize natural and fluent speech from just 15 seconds of any voice, maintaining the given timbre, style, and accent. Fish Audio is a platform for audio generation, offering various voice models for users to discover and use.
Text-to-speech tool that synthesizes natural speech from short voice samples.
- Generating speech in a specific voice for audiobooks
- Creating voiceovers for videos
- Developing virtual assistants with personalized voices
- Generating speech for accessibility purposes
- Users can discover and use pre-built voice models or build their own. The platform offers a text-to-speech toolkit where users can input text and select a voice model to generate speech.
Instantly change your voice to any celebrity voice by talking into a mic.


AI platform for voice cloning, custom synthesis, and multimedia face swapping.


An app for audio conversations with AI celebrity avatars.


Audyo creates human-quality audio from text with easy editing and voice options.


TTSLabs customizes Text to Speech for Twitch streamers with AI voices and sound clips.


AI voice generator for creating audio and videos with celebrity and character voices.


AI video & voice generator cloning celebrity voices for captivating videos.


AI-powered e-greeting card service with personalized poems, audio, and images.


Voice cloning and sound design app for cloning, mimicking, and designing voices.








