Voice-first personal AI agents with continuous memory across voice and text — four distinct personalities, ultra-low latency, free on iOS.
Summary
Voice AI that reads and expresses emotion — speech models trained on human vocal expression, not just phonemes.
Screenshots
Description
Hume AI works on the emotional layer of speech. Its models measure expression in a voice — hesitation, enthusiasm, frustration — and generate speech that carries intended emotion, rather than treating audio as text with a waveform attached.
The products
- Octave TTS, a speech model that takes acting direction. Tell it to sound sceptical, warm, or urgent and the delivery changes accordingly, including for cloned voices.
- Empathic Voice Interface, a conversational API where the assistant hears emotional cues in the user's voice and adapts its response and tone.
- Expression measurement, which analyses voice, face, and language for emotional signal, used in research and product evaluation.
Where it is used
Companion and coaching applications, customer support where tone matters, accessibility tools, research into human behaviour, and any voice product where flat delivery undermines the experience.
Access
Free credits for prototyping, then usage-based pricing, with an SDK for web and mobile. Hume publishes guidelines on responsible use of emotional inference, which is worth reading before deploying this class of model in a consequential setting.
Reviews
Similar App Suggestions
Local-first meeting and interview recorder that transcribes, labels speakers and writes up your notes entirely on your own computer, for a one-time price.
Local-first Mac dictation that types into any app, plus file and link transcription across 110 spoken languages on Apple Silicon.
Voice dictation that works in every app and cleans up as you speak — no filler words, correct formatting, right tone.
Privacy-first voice to text for macOS, Windows, and iOS — runs offline on-device with customisable AI modes.