Voice-first personal AI agents with continuous memory across voice and text — four distinct personalities, ultra-low latency, free on iOS.
Summary
Studio-grade text to speech and voice cloning, with an open community library of thousands of voices.
Screenshots
Description
Fish Audio is a text-to-speech and voice cloning platform built around an open model line and a community voice library. You can clone a voice from a short sample, browse thousands of voices other users have published, or use the models directly.
What it offers
- Voice cloning from a brief recording, capturing timbre and delivery rather than just pitch
- A public voice library with thousands of community-contributed voices across many languages and styles
- Multilingual synthesis with natural prosody, including cross-lingual use of a cloned voice
- Open models. The underlying speech models are published openly, so they can be self-hosted or fine-tuned
- An API with streaming support for real-time applications
Where it fits
Narration for video and audiobooks, character voices for games, accessibility tooling, and voice interfaces where per-character costs matter. The open-model approach keeps pricing well below the premium end of the market while staying competitive on quality.
Free credits are available for evaluation; paid usage is metered per character with subscription tiers for volume. Cloning someone else's voice requires their consent, and the platform's terms reflect that.
Reviews
Similar App Suggestions
Local-first meeting and interview recorder that transcribes, labels speakers and writes up your notes entirely on your own computer, for a one-time price.
Local-first Mac dictation that types into any app, plus file and link transcription across 110 spoken languages on Apple Silicon.
Voice dictation that works in every app and cleans up as you speak — no filler words, correct formatting, right tone.
Privacy-first voice to text for macOS, Windows, and iOS — runs offline on-device with customisable AI modes.