AI voice generation
10 apps ranked by verified data. See how ranking works
Cartesia builds the Sonic family of text-to-speech models using state-space models instead of transformers, delivering sub-100ms time-to-first-byte for real-time voice agents, and can clone a voice fr.
Camb.ai (branded CAMB.AI) is a localization platform for video and audio dubbing that translates content into 140+ languages while retaining the original speaker's voice, tone, and emotion via voice c.
Resemble AI provides enterprise voice cloning, real-time TTS, and streaming voice synthesis via API, alongside a deepfake-detection and watermarking product (Detect/Verify) it added as AI voice fraud.
Fish Audio is a text-to-speech and voice cloning platform built on the open-weight Fish Speech / OpenAudio models, cloning a voice from a short reference sample across 80+ languages.
Hume AI builds voice and conversational AI centered on emotional expressiveness, offering its Empathic Voice Interface (EVI) that generates and detects emotional nuance in speech.
Lovo's Genny platform combines text-to-speech, instant voice cloning, and a video/caption editor for ads, e-learning, and social content.