Smallest.ai Gets $21 Million to Build Voice 4.0
Smallest.ai, a San Francisco-based research lab building real-time voice artificial intelligence infrastructure, has topped $21 million in total funding following the close of a $13 million Series A which it will use to expand its core voice AI platform across financial services, healthcare, contact centers, and business process outsourcing.
Smallest.ai characterizes its innovation as Voice 4.0, which it says is a paradigm shift toward AI architectures that process listening, reasoning, action, and response in parallel. Rather than executing these functions sequentially, Voice 4.0 enables them to happen simultaneously, allowing AI systems to respond while conversations are still unfolding.
"Voice AI has gone through three generations of innovation, but each generation has ultimately hit the same wall," said Sudarshan Kamath, founder and CEO of Smallest.ai, in a statement. "The industry has focused on making models larger when the real challenge is architectural. Humans don't wait for someone to finish speaking before they begin thinking. We listen, think, and respond simultaneously. Voice AI needs to work the same way. That's why we built Smallest.ai around a real-time architecture that processes speech as it arrives, enabling faster, more natural conversations without sacrificing intelligence. By rethinking the stack instead of simply scaling models, we're reducing latency to the point where voice interactions feel genuinely human."
Smallest.ai broke down the voice technology industry into the following three subsets:
- Voice 1.0: Interactive voice response (IVR) systems built around rigid phone trees and menu navigation.
- Voice 2.0: Machine learning-powered voice bots capable of basic intent recognition but unable to handle complexity.
- Voice 3.0: Generative AI voice agents powered by large language models that sound more natural but still rely on multiple disconnected systems working sequentially.
At the center of Voice 4.0 is Hydra, Smallest.ai's speech-to-speech model designed around asynchronous intelligence. Rather than waiting for one process to finish before starting another, Hydra performs multiple tasks in parallel, enabling real-time conversational flow, mid-conversation tool use, natural interruptions, and significantly lower latency.
Together, Hydra and Pulse STT Pro support real-time conversational interactions, with transcription latency measured in milliseconds rather than seconds. Smallest.ai's broader platform includes Pulse STT Pro and Lightning V3.1. Pulse STT Pro supports 38 languages and combines low-latency transcription with capabilities such as speaker diarization, emotion detection, code-switching, noise reduction, and built-in PII and PCI redaction.
"Voice AI is creating a real impact on life and work," said Ashish Kakran, managing partner at Seligman Ventures, the lead investor in Smallest.ai's latest funding round, in a statement. "Developers now increasingly talk to their machines instead of typing code. Smallest.ai is taking a fundamentally different approach to the category by rethinking architecture itself. Customers get an efficient vertically integrated stack and don't need to waste time stitching models together. We believe the next generation of enterprise voice will be powered by Smallest AI."