#
Text-to-Speech
AI tools tagged Text-to-Speech.
CosyVoice
Alibaba's open-weight zero-shot voice-cloning and text-to-speech model
Eleven Flash v2.5
ElevenLabs' low-latency TTS model built for real-time conversational agents
Eleven Multilingual v2
ElevenLabs' flagship text-to-speech model across 29 languages
Eleven Turbo v2.5
ElevenLabs' balanced-latency TTS model between quality and speed
Eleven v3
ElevenLabs' most expressive text-to-speech model with fine-grained emotional control
NaturalSpeech 3
Microsoft's factorized diffusion TTS model for zero-shot voice and style transfer
Sonic
Cartesia's low-latency real-time text-to-speech model
Tacotron 2
Google's neural text-to-speech model pairing a spectrogram predictor with a WaveNet vocoder
VALL-E
Microsoft's neural codec voice-cloning model that mimics a speaker from a 3-second sample
VALL-E 2
Microsoft's improved zero-shot voice-cloning model reaching human parity on benchmarks
WaveNet
DeepMind's raw-audio generative model that set the foundation for modern neural TTS