EchoraEchora

AI Models

Guides to open-source and popular AI voice models, and how to try comparable capabilities on Echora.

7 min read

SeamlessM4T: Meta All-in-One Speech Translation Model

SeamlessM4T is a Meta open-source model unifying speech recognition, translation, and speech synthesis into a single system across up to 100 languages.

speech translationmultilingual speechspeech-to-speech translationautomatic speech recognitionSeamlessM4T
Read more
7 min read

NeuTTS Air: Lightweight On-Device Voice Cloning Model

NeuTTS Air is a Neuphonic on-device TTS model under 1B parameters that clones a voice from just 3 seconds of audio, running on phones and Raspberry Pi.

NeuTTS Airon-device TTSvoice cloninglightweight TTSGGUFRaspberry PiNeuphonic
Read more
7 min read

Kyutai Pocket TTS: 100M-Parameter CPU Real-Time Model

Kyutai Pocket TTS is a 100M-parameter, MIT-licensed model that runs real-time speech synthesis and zero-shot voice cloning on CPU alone, no GPU required.

Kyutai Pocket TTSCPU text to speechreal-time TTSzero-shot voice cloningCALMMIT license
Read more
7 min read

VoxCPM: Tokenizer-Free Bilingual Low-Latency TTS Model

VoxCPM is an OpenBMB tokenizer-free TTS model for Chinese and English, hitting a 0.17 real-time factor on a consumer GPU, with automatic context-aware prosody.

VoxCPMTTStokenizer-free speech synthesisbilingual TTSlow-latency streamingOpenBMB
Read more