Real-Time-Voice-Cloning (SV2TTS): Classic Open-Source Clone
Real-Time-Voice-Cloning is a free, open-source voice cloning tool built on the SV2TTS technique in three stages. See how it works and where it stands today.
Guides to open-source and popular AI voice models, and how to try comparable capabilities on Echora.
Real-Time-Voice-Cloning is a free, open-source voice cloning tool built on the SV2TTS technique in three stages. See how it works and where it stands today.
Speechify began as an accessible text-to-speech app for dyslexia and now offers Studio voice cloning too. See how it works, whether it is free, and pricing.
LOVO AI (Genny) combines text-to-speech, 1-minute voice cloning, and a timeline video editor into one browser workspace for voiceover-driven content.
Voice.ai clones a voice in seconds and changes it live inside games, Discord, and calls, running on-device for low latency across nearly any application.
Respeecher converts a real performance into another voice for film, TV, and games, preserving emotion and timing. See how it works, pricing, and use cases.
WellSaid Labs builds AI voice avatars from real, consenting voice actors under paid licenses, not scraped data or self-serve cloning, for enterprise narration.
Descript Overdub clones your own voice so you can fix flubbed lines by typing, built on Lyrebird AI tech inside a text-based video and podcast editor.
Murf AI pairs a business voice generator with enterprise-gated voice cloning, native PowerPoint and Canva integration, and a low-latency conversational API.
Play.ht (PlayAI) clones a voice from a short sample and streams it with ultra-low latency for live and conversational use. See pricing, sign-up, and API steps.
Resemble AI clones voices from a short sample and watermarks every output for provenance and deepfake detection. See cloning, security, pricing, and API setup.
ElevenLabs is the industry-standard AI voice cloning platform: Instant and Pro cloning, TTS, sound effects, and an API. See features, pricing, and setup steps.
Google Cloud TTS WaveNet delivers general-purpose neural speech with SSML controls, mature APIs, and $4-per-million-character pricing after its free tier.