Fish Speech
Fish Speech is a text-to-speech (TTS) tool developed by the creators of So-VITS-SVC and Bert-VITS2. It can synthesize natural and fluent speech from just 15 seconds of any voice, maintaining the given timbre, style, and accent. Fish Audio is a platform for audio generation, offering various voice models for users to discover and use.
Text-to-speech tool that synthesizes natural speech from short voice samples.
- Generating speech in a specific voice for audiobooks
- Creating voiceovers for videos
- Developing virtual assistants with personalized voices
- Generating speech for accessibility purposes
- Users can discover and use pre-built voice models or build their own. The platform offers a text-to-speech toolkit where users can input text and select a voice model to generate speech.

AI voice generator for creating audio and videos with celebrity and character voices.


Create, edit, and transform audio with AI — podcasts, voiceovers, transcripts, and more — instantly in your browser.


Audyo creates human-quality audio from text with easy editing and voice options.


AI reading and listening assistant that converts text to audio and creates summaries.


AI voice generator that turns text into speech using celebrity voices for entertainment.


AI platform for voice cloning, custom synthesis, and multimedia face swapping.


AI voice generator for transforming text into speech with celebrity and professional voices.


AI celebrity voice generator for realistic voiceovers and lip-synced videos.


Real-time AI voice changer with soundboard, AI cover, music generator, and audio enhancer.







