NVIDIA Magpie TTS Expands to 12 Languages With Open Weights and On-Premises Control
NVIDIA's 364M-parameter multilingual text-to-speech model gains Arabic, Korean, and Brazilian Portuguese support, enabling low-latency voice agents on private infrastructure.
Hugging Face and Cerebras Demonstrate Real-Time Speech-to-Speech with Gemma 4
A modular voice AI pipeline achieves sub-second latency by pairing Google DeepMind's Gemma 4 model with Cerebras inference acceleration, powering conversational robots and assistants.
Google Launches Nano Banana 2 Lite, a Faster Image Generator at $0.034 per 1,000 Images
Google releases a stripped-down, high-speed variant of its Nano Banana image model optimized for rapid iteration and bulk workflows.
NVIDIA's Nemotron-Labs Diffusion Models Generate Multiple Tokens in Parallel, Bypassing Autoregressive Bottleneck
NVIDIA releases diffusion language models at 3B, 8B, and 14B scales that generate and refine tokens in parallel, offering latency improvements for GPU-constrained inference workloads.