Speech Technology News

Researchers Propose Alternatives to Word-Error Rate

WER improperly weighs all words in a sentence the same, researchers assert.

Insta360 Launches AI Voice Assistant for GO Ultra

Ask by Voice turns the Go Ultra into a voice assistant for travelers with real-time translation, voice and photo Q&A.

ElevenLabs Launches Dubbing v2 API

ElevenLabs' Dubbing v2 now provides direct speech-to-speech AI dubbing across 92 languages.

Speech Processing Solutions Launches Philips SpeechLive Legal AI Assistant

New SpeechLive Legal AI Assistant enables customizable document drafting powered by speech recognition tuned for legal terminology.

Modulate Partners with Scam.ai for Voice Fraud Detection

Integration brings Modulate's synthetic voice detection to the Scam.ai platform, helping organizations identify multimodal deepfake threats through a single customer experience.

Rosetta Stone Launches Sapphire Language Learning Tool

Rosetta Stone Sapphire combines its Dynamic Immersion method and TruAccent speech recognition technology.

Vidy Launches Real-Time Video AI Characters for Personalized Conversations and Storytelling

Vidy is introducing a new generation of AI interaction powered by real-time video, voice, and conversational intelligence.

Concern Grows Over AI-Powered Voice Attacks, Mutare Survey Finds

New research finds growing enterprise risk and increasing recognition that voice security must become part of every organization's cybersecurity strategy.

PolyAI Launches Dialog-RSN-1 Model

PolyAI's Dialog-RSN-1 is a dialog model that directly perceives the user's audio and fuses turn-taking, speech recognition, function calling, and response.

Smallest.ai Gets $21 Million to Build Voice 4.0

Smallest.ai's Voice 4.0 and Hydra form an asynchronous AI architecture to make AI conversations as natural, responsive, and scalable as human dialogue.

Fish Audio Raises $52 Million in Seed Funding

Funding will help Fish Audio expand its model lineup beyond text-to-speech into the full audio-native stack.

Voicelyt Launches Voice Score

As remote work and video calls drive rising vocal strain, Voicelyt offers instant, data-backed insight into vocal health.

OrcaRouter Launches OrcaDub

OrcaDub is a speech-to-speech video dubbing model for natural multilingual localization.

Deliverect Partners with SoundHound AI

Deliverect and SoundHound AI partner to turn voice into a fully automated ordering channel for restaurants.

Nabla Launches Dictation for Mac

Nabla Dictation for Mac is built specifically for medical practices operating on Apple devices.

DXC Partners with ElevenLabs

DXC and ElevenLabs collaborate to scale enterprise AI and voice innovation

Overcoming the 911 Language Barrier with Speech Automation

Language Assist gives 911 dispatchers instant, natural-sounding communication with non-English speakers—no interpreter required.

Voice-Only Outreach 'Structurally Misses' Gen Z and Millennial Debt Holders, Says Vodex AI CEO

AI-powered engagement should combine voice, SMS, email, and account intelligence to improve consumer engagement and resolution rates.

DeepL Acquires Mixhalo

DeepL adds Mixhalo's team and technology to accelerate voice AI at scale.

Deepgram Brings Nova-3 Speech Engine to Snapdragon Devices

Deepgram's Nova-3 speech-to-text model on PCs powered by Snapdragon delivers real-time voice experiences.