Fish Speech is a text-to-speech (TTS) tool developed by the creators of So-VITS-SVC and Bert-VITS2. It can synthesize natural and fluent speech from just 15 seconds of any voice, maintaining the given timbre, style, and accent. Fish Audio is a platform for audio generation, offering various voice models for users to discover and use.
Text-to-speech tool that synthesizes natural speech from short voice samples.
- Generating speech in a specific voice for audiobooks
- Creating voiceovers for videos
- Developing virtual assistants with personalized voices
- Generating speech for accessibility purposes
- Users can discover and use pre-built voice models or build their own. The platform offers a text-to-speech toolkit where users can input text and select a voice model to generate speech.

Create, edit, and transform audio with AI — podcasts, voiceovers, transcripts, and more — instantly in your browser.

Instantly change your voice to any celebrity voice by talking into a mic.

AI-powered celebrity voice cloning for instant voice experiences in your browser.


AI-powered platform for creating prank calls and messages with celebrity voices.


AI voice generator for creating audio and videos with celebrity and character voices.


AI text-to-speech platform with 20,000+ character and celebrity voices for professional audio.


Chrome extension to transform writing into character voices.


Real-time AI voice changer with soundboard, AI cover, music generator, and audio enhancer.


Voicemy.ai: Clone voices, train AI models, compose melodies, and share your creations.







