NVIDIA Magpie TTS Expands to 12 Languages With Open Weights and On-Premises Control
NVIDIA's 364M-parameter multilingual text-to-speech model gains Arabic, Korean, and Brazilian Portuguese support, enabling low-latency voice agents on private infrastructure.
PP-OCRv6 Scales Multilingual Text Detection to 50 Languages Across 1.5M–34.5M Parameters
PaddleOCR's latest model family achieves 86.2% detection accuracy in medium tier, with three deployment tiers optimized for edge to server-side workloads.
DeepL acquires Mixhalo to bring real-time translation to live events
DeepL expands into live-event audio with Mixhalo acquisition, positioning voice translation as a conferencing and sports broadcasting differentiator.
Voice Agents Struggle With Code-Switched Speech Across Four Language Pairs
ServiceNow and Hugging Face benchmark ASR models on bilingual customer interactions, revealing significant performance gaps when speakers mix languages mid-sentence.
IBM Granite Embedding Multilingual R2: 97M and 311M Parameter Models Top MTEB Multilingual Retrieval Charts
IBM releases two Apache 2.0 multilingual embedding models built on ModernBERT, with 32K-token context and coverage for 200+ languages.