Voice-first personal AI agents with continuous memory across voice and text — four distinct personalities, ultra-low latency, free on iOS.
Audio & Voice AI Apps
AI for voice, audio processing, and music Compare features, pricing models, and use cases across the tools in this category.
Local-first meeting and interview recorder that transcribes, labels speakers and writes up your notes entirely on your own computer, for a one-time price.
Local-first Mac dictation that types into any app, plus file and link transcription across 110 spoken languages on Apple Silicon.
Voice dictation that works in every app and cleans up as you speak — no filler words, correct formatting, right tone.
Privacy-first voice to text for macOS, Windows, and iOS — runs offline on-device with customisable AI modes.
High-quality stem separation — split any track into vocals, drums, bass, and instruments with minimal artefacts.
AI music generation with lyric writing, stem separation, and reference-based style matching.
Listen to anything — documents, articles, PDFs, and email — in natural voices, at up to several times normal speed.
Speech-to-text APIs with built-in understanding — transcription plus summaries, topics, and redaction in one call.
Enterprise speech infrastructure — fast, accurate transcription, text to speech, and a full voice agent API.
Voice AI that reads and expresses emotion — speech models trained on human vocal expression, not just phonemes.
Ultra-low-latency speech models built for real-time voice agents, where every millisecond is audible.
Studio-grade text to speech and voice cloning, with an open community library of thousands of voices.
Edit video and podcasts by editing text — Descript transcribes your recording so you can cut, rearrange, and polish media as easily as a document.
AIVA is an AI music composer that generates original, royalty-free soundtracks for films, games, and videos, with editable scores and export to MP3, WAV, and MIDI.
Udio turns a text prompt into full, studio-quality songs — complete with vocals, lyrics, and instrumentation — across virtually any genre, mood, and style.
Create stunning original music in seconds using AI. Make your own masterpieces, share with friends, and discover music from artists worldwide.
Generate lifelike spoken audio with AI tools
Build audio experiences at scale with Murf's AI voice generator
Cleanvoice cleans podcasts and audio automatically.
Krisp removes background noise and echo using AI.
Wondercraft makes studio-quality podcasts with AI.
Soundraw generates royalty-free music for creators.
Resemble AI provides high-quality voice cloning.
PlayHT offers ultra-realistic AI voices and cloning.
On-device Mac dictation, rewriting, translation, and local agent — no cloud, no accounts.