Hume AI is an empathic AI research lab building multimodal AI with emotional intelligence. They offer advanced AI models like Octave Text-to-Speech (TTS), which is the first LLM for text-to-speech capable of understanding context and predicting emotions, and Empathic Voice Interface (EVI), a real-time, customizable voice intelligence model for fluent, emotionally intelligent conversations. They also provide an Expression Measurement API to analyze expressions in face, voice, and language. Their goal is to create expressive AI voices and interactive personalities, with a strong focus on human well-being and ethical AI development.
Empathic AI for voice and expression with emotional intelligence.
- Generating expressive AI voices for podcasts, voiceovers, and audiobooks.
- Deploying emotionally intelligent voice agents in any application.
- Creating interactive personalities for customer service, virtual assistants, or entertainment.
- Real-time conversational AI for enhanced user experiences.
- Measuring and analyzing emotional expression in various media.
- Users can generate AI voices by providing text prompts and describing desired voice identities
- qualities
- and emotions using Octave TTS. They can also create and interact with real-time synthetic voices and personalities using EVI
- which allows for flexible prompting and voice modulation. Developers can access APIs and a full developer platform to integrate these emotionally intelligent voice agents into their own applications.

AI-powered text-to-speech converter for realistic voiceovers.


Text-to-speech solution with AI voices for personal, commercial, and educational purposes.


AI voice solution for content creation with text-to-speech, dubbing, and voice cloning.


AI video generation platform for creating engaging business videos quickly and easily.


Easy online platform for video, image, and GIF editing.


FineVoice is a versatile AI voice generator. Instantly create high-quality, royalty-free voices, SFX, and music.

AI audio platform offering text-to-speech, voice cloning, and dubbing services.


All-in-one AI app for text, image, audio, and video tasks.


AI-powered online media tools for video, audio, and photo editing.


CapCut is an AI-driven all-in-one video editor and graphic design tool.


Versatile AI voice generator for text to speech, voiceovers, and translations.


MiniMax Audio creates lifelike speech in multiple languages with diverse voices.


App for reading text aloud with high-quality voice AI.


Free online text-to-speech tool with 200+ voices and 70+ languages.


Multi-modal AI content generation for images, videos, and speech.


A platform for deploying and running machine learning models with a simple API and pay-per-use pricing.


Text-to-speech tool that synthesizes natural speech from short voice samples.


Deepgram is a Voice AI platform offering STT, TTS, and voice agent APIs for developers.





