Hume AI is an empathic AI research lab building multimodal AI with emotional intelligence. They offer advanced AI models like Octave Text-to-Speech (TTS), which is the first LLM for text-to-speech capable of understanding context and predicting emotions, and Empathic Voice Interface (EVI), a real-time, customizable voice intelligence model for fluent, emotionally intelligent conversations. They also provide an Expression Measurement API to analyze expressions in face, voice, and language. Their goal is to create expressive AI voices and interactive personalities, with a strong focus on human well-being and ethical AI development.
Empathic AI for voice and expression with emotional intelligence.
- Generating expressive AI voices for podcasts, voiceovers, and audiobooks.
- Deploying emotionally intelligent voice agents in any application.
- Creating interactive personalities for customer service, virtual assistants, or entertainment.
- Real-time conversational AI for enhanced user experiences.
- Measuring and analyzing emotional expression in various media.
- Users can generate AI voices by providing text prompts and describing desired voice identities
- qualities
- and emotions using Octave TTS. They can also create and interact with real-time synthetic voices and personalities using EVI
- which allows for flexible prompting and voice modulation. Developers can access APIs and a full developer platform to integrate these emotionally intelligent voice agents into their own applications.

All-in-one AI platform for content, images, audio, and transcription.


Free online AI text to speech generator with realistic voices and customization.


Peech is a text-to-speech reader converting text to audio in 50+ languages.


App for reading text aloud with high-quality voice AI.


All-in-one AI app for text, image, audio, and video tasks.


AI-powered platform for text-to-speech and speech-to-text services in 75+ languages.


AI-powered text-to-speech converter with human-like voiceovers and advanced customization options.


A platform for deploying and running machine learning models with a simple API and pay-per-use pricing.


Multi-modal AI content generation for images, videos, and speech.








