Text-to-speech models and tools
Find text-to-speech models and tools for Chinese, low latency, local deployment, and open-source use cases.
Selected public KG entries
These are discovery candidates, not a paid ranking or an official product review.
Text-to-speech
KG-recorded application related to Text-to-speech; detailed English notes are not yet available.
MechanismTTS
A sequential stage in traditional voice-plus-camera architectures.
ModelQwen-Audio-3.0-TTS
KG-recorded model related to Qwen-Audio-3.0-TTS; detailed English notes are not yet available.
BenchmarkArtificial Analysis TTS leaderboard
KG-recorded benchmark related to Artificial Analysis TTS leaderboard; detailed English notes are not yet available.
ModelGemini 3.1 Flash TTS
KG-recorded model related to Gemini 3.1 Flash TTS; detailed English notes are not yet available.
ModelNVIDIA Magpie TTS
KG-recorded model related to NVIDIA Magpie TTS; detailed English notes are not yet available.
ModelCosyVoice-3.0
KG-recorded model related to CosyVoice-3.0; detailed English notes are not yet available.
MechanismEnd-to-End Architecture
Replaces cascaded ASR/VLM/TTS pipelines by processing multimodal streams in a si