01
Lifelike AI speech and voice cloning in 70+ languages
Turns text into speech across 70+ languages and clones a voice from your own recordings. Used for narration and dubbing, with an API onto the same models for building voice agents.
02
Voiceover generation with 200+ voices across 30+ languages
Turns scripts into voiceovers using 200+ voices across 30+ languages, with custom pronunciation and pitch control. There is an API as well as the Studio editor.
03
P
Voice generator and text-to-speech API — service now shut down
Ran an AI voice studio and a low-latency text-to-speech API with voice cloning. The vendor's own final homepage reads "We have shut down the service", and play.ht no longer resolves.
04
Listen to your PDFs and web pages instead of reading them
Reads PDFs and web pages aloud on iOS, Android, Mac, Windows and browser extensions, with more than a thousand voices across 60+ languages and playback up to 5x.
05
Voice cloning from 10 seconds of audio, plus deepfake detection
Rapid clones train on 10 seconds of audio; professional clones need 10-25+ minutes and verifiable consent from the voice talent. The same platform covers detection and watermarking for synthetic media.
06
C
CartesiaBest for low latency
Real-time speech synthesis and transcription for voice agents
Sonic generates speech in 42 languages with sub-90ms latency, and Ink handles streaming transcription. Line is the platform for building voice agents.
07
Text-to-speech voices modelled on licensed voice-actor recordings
Turns written scripts into narration from a library of AI voices, with controls for pacing and tone. The voices also run through an API for apps and learning platforms.