Summary
Voice AI that reads and expresses emotion — speech models trained on human vocal expression, not just phonemes.
Screenshots
Description
Hume AI works on the emotional layer of speech. Its models measure expression in a voice — hesitation, enthusiasm, frustration — and generate speech that carries intended emotion, rather than treating audio as text with a waveform attached.
The products
- Octave TTS, a speech model that takes acting direction. Tell it to sound sceptical, warm, or urgent and the delivery changes accordingly, including for cloned voices.
- Empathic Voice Interface, a conversational API where the assistant hears emotional cues in the user's voice and adapts its response and tone.
- Expression measurement, which analyses voice, face, and language for emotional signal, used in research and product evaluation.
Where it is used
Companion and coaching applications, customer support where tone matters, accessibility tools, research into human behaviour, and any voice product where flat delivery undermines the experience.
Access
Free credits for prototyping, then usage-based pricing, with an SDK for web and mobile. Hume publishes guidelines on responsible use of emotional inference, which is worth reading before deploying this class of model in a consequential setting.
Reviews
Similar App Suggestions
Wispr Flow
Wispr
Voice dictation that works in every app and cleans up as you speak — no filler words, correct formatting, right tone.
Superwhisper
Privacy-first voice to text for macOS, Windows, and iOS — runs offline on-device with customisable AI modes.
LALAL.AI
High-quality stem separation — split any track into vocals, drums, bass, and instruments with minimal artefacts.
Mureka
Kunlun Tech
AI music generation with lyric writing, stem separation, and reference-based style matching.
Speechify
Listen to anything — documents, articles, PDFs, and email — in natural voices, at up to several times normal speed.