Speech Technology News

Voice AI Models Mispronounce New Drug Names

Voice AI models mispronounce a third of new drug names in Synthio benchmarking.

FineVoice Launches Aunio for Audio Creation

FineVoice's Aunio is an AI audio production agent for end-to-end audio creation.

Applied Brain Research Releases the ABR SDK for On-Device Voice Interfaces to Edge Applications

Niagara ASR and Nith TTS streaming models reach production release in a single SDK

Vosko AI Launches Video Localization Platform

Vosko AI supports video translation across 99 languages, combining AI dubbing, voice cloning, subtitle localization and emotion-preserving technology.

Lisova Launches Voice Companion for Aging Adults

Lisova checks in with loved ones every day, listens, remembers what matters to them, and helps capture their stories.

Venizum Launches Verbis Voice

Venizum's Verbis Voice brings real-time voice-to-voice translation to Salesforce Service with DeepL.

OpenAI Launches GPT-Live-1

OpenAI's GPT-Live-1 gives developers a natural voice model for building voice-enabled apps and business workflows.

Yelp and Hatch Advance Voice AI for Restaurants and Service Pros with OpenAI's GPT-Live-1

Yelp Host, now enhanced with GPT-Live-1, transforms restaurant call handling, while Hatch's business intelligence, now with GPT-Live-1, brings advanced voice AI to help service businesses book more jobs. (Featured on DestinationCRM.com.)

Everise Partners with Sanas to Deploy the One Sanas App Across 15,000 Healthcare Support Agents

Everise will run the full Sanas stack, combining accent translation, speech enhancement, on device transcription, live compliance support and translation across 34 languages.

Apple Updates Siri with Voice Customization

A new SiriAI update from Apple lets users customize the cadence, nationality, and expressivity of the voice assistant and improves dictation quality.

Zendesk Expands Contact Center Voice Support Across Languages

Real-Time Voice Translation built directly into Zendesk Contact Center helps customers and agents communicate naturally across languages. (Featured on SmartCustomerService.com.)

Voiskey Integrates Keyboard Typing into Voice

Voiskey has added a Keyboard Input feature on Voiskey Mobile.

Deepdub Launches Phantom Z 3.4 Conversational

Deepdub's Phantom Z 3.4 Conversational is multilingual text-to-speech for real customers, not just demos.

Phonely Launches Alma, a Voice LLM

Phonely's Alma LLM is trained on more than 10 million phone conversations and helps voice AI agents self-improve over time.

Mavenir Partners with Sanas

Partnership brings real-time speech enhancement, accent transformation, deepfake detection, and language understanding into operator networks.

Soniox Partners with Authentic Interactions

Partnership brings Soniox's speech recognition tech to Lookalike and StoryFile.

Wispr Raises $280 Million to Advance Its Flow AI Voice Technology

New funding will help Wispr advance its Flow speech recognition accuracy.

Google Releases Gemini 3.5 Transcribe

Google's latest speech-to-text model is designed for precise and intelligent real-time transcription.

Suki Launches Suki Dictation

Suki Dictation allows organizations to deploy AI dictation, ambient documentation, or both, natively inside Epic and MEDITECH.

HitPaw Releases Edimakor V5.2.0

HitPaw's Edimakor V5.2.0 features several new features for AI video and avatar creation.