Perso Ai
Perso AI is a 3-in-1 AI audio and video platform combining AI Dubbing, Speech-to-Text, and Audio Separation in a single workflow. It translates and dubs videos into 33+ languages with natural voice cloning and lip dubbing, generates speaker-separated transcripts with automatic speaker diarization in four output formats (XLSX, SRT, VTT, JSON), and isolates individual speaker voices from background audio with dual modes (vocals-only or with reactions preserved). Trusted by 460,000+ users across 80+ countries, powered by the ElevenLabs voice engine (2025 partnership), and ISO/IEC 27001 and KISA ISMS certified. Developed by ESTsoft (est. 1993, KOSDAQ: 047560). Starts at $6.99/month with up to 98% cost savings vs. traditional dubbing studios.
AI Dubbing in 33+ languages, Speech-to-Text with speaker diarization, and Audio Separation — all in one workflow.
- Corporate L&D Teams — Translate training and onboarding videos. Speech-to-Text auto-generates meeting transcripts with speaker diarization.
- YouTube Creators — Dub videos into 33+ languages without hiring voice actors. Reach global audiences with localized content.
- Marketing Agencies — Localize campaign and product-demo videos for international markets while maintaining brand voice.
- Podcast Producers — Use Audio Separation to extract individual speaker voices, remove background music, or merge selected tracks for post-production.
- E-Learning Platforms — Translate lecture videos into regional languages. Auto-generate SRT subtitles for accessibility.
- Media Production — Create dubbed versions of documentary content for international distribution at a fraction of traditional dubbing costs.
- Sign up at perso.ai and upload any video or audio file. Choose your workflow — AI Dubbing to translate videos into 33+ languages
- Lip Dubbing for natural lip-synced output
- Speech-to-Text for speaker-separated transcripts and subtitles (XLSX
- SRT
- VTT
- JSON)
- or Audio Separation to isolate individual speakers and background audio. Edit the auto-generated script in the real-time editor for instant regeneration
- then download or export. Enterprise users can access all capabilities via the API for batch processing.

AI video lipsync tool for real-time lipsync and seamless translation.


Audio-driven AI tool for talking avatars with precise lip sync.


AI-powered lip-syncing for videos with text-to-speech in 90+ languages.


Google's AI tool for generating videos with synchronized audio.


Lip syncs your mouth to make it appear you're speaking another language.



Flawless uses AI to revolutionize filmmaking with visual dubbing and AI reshoots.


AI-powered technology that animates still photos into lifelike videos.


AI-powered audio-driven full-body video dubbing and generation.


All-in-one AI video and image creation platform.


AI lip sync and video translation tool for realistic video content creation.


AI lip sync technology transforms photos into lifelike talking videos.


Free online AI tool for creating lifelike lip-synced videos easily.


Free AI video generator transforming images into professional videos with lip-sync and multilingual support.


VFX studio delivering feature-film quality VFX for TV series with innovative technology.


AI Talking Video Generator with Avatar Generator, Voice Cloning & AI Lip-Sync.


AI tool to animate photos with speech and lifelike expressions.


AI lip-sync platform for video translation, correction, and content creation.





