Speech Technology News

Fish Audio Raises $52 Million in Seed Funding

Funding will help Fish Audio expand its model lineup beyond text-to-speech into the full audio-native stack.

OrcaRouter Launches OrcaDub

OrcaDub is a speech-to-speech video dubbing model for natural multilingual localization.

Deliverect Partners with SoundHound AI

Deliverect and SoundHound AI partner to turn voice into a fully automated ordering channel for restaurants.

Nabla Launches Dictation for Mac

Nabla Dictation for Mac is built specifically for medical practices operating on Apple devices.

DXC Partners with ElevenLabs

DXC and ElevenLabs collaborate to scale enterprise AI and voice innovation

Overcoming the 911 Language Barrier with Speech Automation

Language Assist gives 911 dispatchers instant, natural-sounding communication with non-English speakers—no interpreter required.

Voice-Only Outreach 'Structurally Misses' Gen Z and Millennial Debt Holders, Says Vodex AI CEO

AI-powered engagement should combine voice, SMS, email, and account intelligence to improve consumer engagement and resolution rates.

DeepL Acquires Mixhalo

DeepL adds Mixhalo's team and technology to accelerate voice AI at scale.

Deepgram Brings Nova-3 Speech Engine to Snapdragon Devices

Deepgram's Nova-3 speech-to-text model on PCs powered by Snapdragon delivers real-time voice experiences.

Canary Speech Partners with NeuroLexIQ

New collaboration embeds Canary Speech's vocal biomarker technology into NeuroLexIQ's Concussion Probability Report, flagging potential brain injuries at a client's very first contact.

Voiskey Officially Launches

Voiskey launches with an AI voice typing app focused on expression intelligence.

LALAL.AI Launches Lynx Voice Cleanup Mode

LALAL.AI Lynx is a speech denoising model that isolates voice and removes background music, street noise, and acoustic interference from voice recordings.

VoicePing Releases VoicePing 3.0

VoicePing 3.0 combines real-time translation, AI meeting minutes, voice output, terminology dictionaries, MCP/API access, QR-based event translation, file transcription, subtitles, AI dubbing, virtual office collaboration, enterprise administration, and new VoicePing ASR and MT models.

Modulate Tops Hugging Face's Transcription Benchmark

Modulate ranked #1 out of 88 speech-to-text models evaluated by Hugging Face for accuracy, speed, and cost.

Omilia Launches Lexis TTS Model for Contact Centers

Omilia Lexis is a voice synthesis model delivered inside Omilia's Conversational Platform.

Symend Launches SymendConverse

SymendConverse's agentic voice AI personalizes every inbound and outbound call to each customer's situation.

Study Proves Assistive Technologies Improve Users' Lives

Quality of life improvements and clear economic benefits follow assistive technology deployments, a study finds.

Callie Care Collects $500K for Voice AI Development

Callie Care startup aims to bring phone-first voice AI to millions of older adults aging alone at home.

AI Voice Agents Increase Specialty Care Program Enrollment

Study finds AI voice agents increased specialty care program enrollment 340 percent in real-world clinical settings.

Emotion Detection and Recognition Market to Be Worth $43.29 Billion by 2031

The emotion detection and recognition market is expected to grow by 8.2 percent through the next five years, MarketsandMarkets predicts.