NVIDIA Magpie TTS Expands to 12 Languages With Open Weights and On-Premises Control
NVIDIA's 364M-parameter multilingual text-to-speech model gains Arabic, Korean, and Brazilian Portuguese support, enabling low-latency voice agents on private infrastructure.
Late-Night AI Job Interviews Reshape Candidate Scheduling
Automated video and voice AI interviews are pushing job candidates to interview during off-hours—24% of AI-conducted interviews now happen between 10 pm and 2 am.
Ford Launches AI Assistant for Vehicle Management Across 8 Million Vehicles
Ford's new chatbot helps owners check fuel levels, towing capacity, and service needs via mobile app ahead of 2027 in-vehicle rollout.
Wispr Flow Enters the Meeting-Recording Crowdscape with New Notetaker
Voice-dictation startup Wispr Flow launches AI-powered meeting transcription and summarization, joining competitors like Granola in a rapidly expanding market.
OpenAI's GPT-Live Eliminates Turn-Based Voice Bottleneck With Full-Duplex Architecture
OpenAI's new GPT-Live voice system uses full-duplex speech models to remove latency-inducing turn detectors, enabling real-time conversational responsiveness without sacrificing reasoning depth.
Smallest.ai Lands $13M to Build Conversational Voice AI Without Perceptible Delays
The startup is developing specialized voice models that process speech in real-time, mimicking natural human dialogue instead of relying on large language models for voice interactions.
Encore AI Raises $30M Series A to Deploy Voice Agents Trained on Customer Conversations
Encore AI, a startup that mines customer interactions to train AI agents for sales and support roles, closed a $30M Series A led by Team8, with 40+ enterprise customers and 5x ARR growth.
Fish Audio Secures $50M to Expand AI Voice Generation for Creators and Enterprises
Fish Audio, a voice synthesis startup with 8M users and $21M ARR, closes seed round led by Coreline Ventures to scale customizable voice models across creative and business applications.
Smart rings emerge as the AI interface for distraction-free computing
Sandbar's Stream ring demonstrates how wearable AI devices could shift interaction patterns from screen-based to voice-first, with cross-app dictation as the key differentiator.
OpenAI brings ChatGPT Voice to desktop, enabling multi-step task automation
OpenAI's ChatGPT Voice feature, powered by GPT-Live models, now works on desktop apps to control agents and automate complex workflows.
Hugging Face Launches Real World VoiceEQ to Expose Gaps Between Voice AI Lab Performance and User Experience
A new 1M-human-rating benchmark reveals that conversational speech systems underperform on emotional nuance, speaker consistency, and accent handling despite saturation on latency metrics.
Hugging Face and Cerebras Demonstrate Real-Time Speech-to-Speech with Gemma 4
A modular voice AI pipeline achieves sub-second latency by pairing Google DeepMind's Gemma 4 model with Cerebras inference acceleration, powering conversational robots and assistants.
Amazon Expands Alexa+ Beta to India With Hindi-Language Testing
Amazon invites Indian users to test Alexa+, its generative AI assistant, in Hindi as the company pursues voice-AI adoption in a market with 600+ million native speakers.
Travelers deploys AI-powered voice claims assistant nationwide with OpenAI Realtime API
Insurance giant Travelers expands autonomous claims handling to all US states, with 85-90% of customers completing filings through AI.
Google Re-enters Smart Glasses Market with Audio-First Partnership
Google unveiled AI-powered audio glasses co-developed with Warby Parker, Gentle Monster, and Samsung, launching later in 2026.
Google Brings Voice AI to Gmail, Docs, and Keep
Gmail Live, Docs Live, and voice-driven Keep features roll out this summer for Google's AI Pro and Ultra subscribers.
Google Workspace Adds Voice Commands, Image Generation, and Gemini Spark Agent
Google introduces conversational voice features across Gmail, Docs, and Keep, plus Gemini Spark—a 24/7 AI agent for Workspace automation.
Ethos Raises $22.75M to Replace Job-Title Matching With AI Voice Interviews
London-based startup Ethos secured $22.75M Series A led by a16z to build an expert network using voice onboarding to capture professional knowledge beyond job titles.
ElevenLabs Discloses Full Series D Roster as ARR Clears $500 Million
ElevenLabs names all $500M Series D investors — BlackRock, NVIDIA, and celebrity backers among them — as the voice AI startup confirms crossing $500M in annual recurring revenue.
OpenAI's WebRTC Overhaul: Building Voice AI Infrastructure for 900 Million Users
OpenAI rebuilt its real-time audio stack with a relay-and-transceiver design to eliminate latency issues that emerge only at global scale.