Viral Wire

Tencent Cloud & Inworld AI partner for real-time voice AI with sub-130ms TTS

Inworld's #1 ranked TTS now integrated into Tencent RTC's global network.

Deep Dive

Tencent Cloud and Inworld AI announced a strategic partnership to deliver a one-stop, lifelike, real-time voice AI solution. Inworld's text-to-speech (TTS) models, ranked #1 on the Artificial Analysis Speech Arena, are deeply integrated into Tencent Real-Time Communication (Tencent RTC). This gives developers access to ultra-realistic, context-aware speech with sub-130ms first-chunk latency, over 100 languages, and instant cross-lingual conversion while preserving a consistent speaker voice. Developers can clone voices, design custom voices, control vocal style, and stream natural responses directly via the Tencent RTC console and SDK.

On the infrastructure side, Tencent RTC provides an enterprise-grade backbone with more than 3,200 global nodes and sub-300ms worldwide latency. Advanced features like AI noise suppression and weak-network resilience ensure seamless performance even in challenging connectivity environments. The partnership also includes a joint Conversational AI Demo with curated voices for different languages and scenarios. Tencent Cloud's Wison Xie emphasized that production-grade voice AI requires both a capable model and a capable network, and this partnership delivers a complete path from prototype to production for developers building emotionally intelligent voice applications.

Key Points
  • Inworld TTS is ranked #1 on the Artificial Analysis Speech Arena and delivers sub-130ms first-chunk latency.
  • Supports over 100 languages with instant cross-lingual conversion while preserving a consistent speaker voice.
  • Tencent RTC provides 3,200+ global nodes, sub-300ms latency, AI noise suppression, and weak-network resilience.

Why It Matters

This partnership gives developers a turnkey way to deploy expressive voice AI globally with enterprise-grade reliability.

📬 Get the top 10 AI stories daily