Smallest.ai Lands $21M to Build the Future of Enterprise Voice AI
Smallest.ai Secures $21 Million in Funding to Develop Voice 4.0, the Future of Enterprise Voice AI
- Secures $13 million in Series A funding led by Seligman Ventures, pushing total capital raised past $21 million
- Unveils Voice 4.0 and Hydra, an asynchronous AI framework built to deliver human-like, responsive, and scalable voice interactions
- Pulse STT and Lightning TTS rank among the top choices for enterprises on Artificial Analysis, leading on speed and cost efficiency across the global leaderboard
SAN FRANCISCO – August 05, 2026 – Smallest.ai, a San Francisco-based foundational AI research lab focused on creating next-level real-time voice infrastructure for enterprise use, has announced it has crossed the $21 million funding milestone. This follows a $13 million Series A round spearheaded by Seligman Ventures, with additional backing from Sierra Ventures and 3one4 Capital.
Industry forecasts project the global Voice AI sector to expand from $2.4 billion in 2024 to $47.5 billion by 2034, yet less than 1% of current global voice interactions utilize AI. Businesses continue to grapple with systems that sound artificial, struggle with complex real-world scenarios, suffer from latency, and present hurdles regarding reliability, governance, and compliance. Smallest.ai is scaling its voice AI platform across sectors like financial services, healthcare, contact centers, and business process outsourcing, where there is a growing demand for deploying AI-driven voice agents at scale.
"Voice AI has gone through three generations of innovation, but each generation has ultimately hit the same wall," said Sudarshan Kamath, founder and CEO of Smallest.ai. "The industry has focused on making models larger when the real challenge is architectural. Humans don't wait for someone to finish speaking before they begin thinking. We listen, think, and respond simultaneously. Voice AI needs to work the same way. That's why we built Smallest.ai around a real-time architecture that processes speech as it arrives, enabling faster, more natural conversations without sacrificing intelligence. By rethinking the stack instead of simply scaling models, we're reducing latency to the point where voice interactions feel genuinely human."" Voice technology has progressed through three distinct phases: Current voice agents typically rely on a chain of disparate technologies, including speech recognition, language models, text-to-speech systems, orchestration layers, memory systems, and guardrails. This results in high latency, fragile performance, and interactions that remain noticeably artificial. Smallest.ai defines its breakthrough as Voice 4.0, a shift toward AI architectures that handle listening, reasoning, action, and response in parallel. Instead of executing these steps one after another, Voice 4.0 allows them to occur simultaneously, enabling AI systems to formulate responses while the conversation is still in progress. At the core of Voice 4.0 is Hydra, Smallest.ai’s speech-to-speech model built on the principle of asynchronous intelligence. Rather than pausing for one process to conclude before starting the next, Hydra executes multiple tasks concurrently, facilitating real-time conversational flow, mid-conversation tool usage, natural interruptions, and significantly reduced latency. Together, Hydra and Pulse STT Pro are engineered to support real-time conversational interactions, with transcription latency measured in milliseconds rather than seconds. The broader Smallest.ai platform features Pulse STT Pro and Lightning V3.1, which are ranked among the top voice AI models on Artificial Analysis for speed, quality, and cost efficiency. Designed for enterprise-scale deployments, Pulse STT Pro supports 38 languages and integrates low-latency transcription with features such as speaker diarization, emotion detection, code-switching, noise reduction, and built-in PII and PCI redaction. Clients utilize Smallest.ai to automate enterprise voice workflows, cutting support costs by up to 80% while boosting agent productivity by as much as 10x. "Voice AI is creating a real impact on life and work," said Ashish Kakran, Managing Partner at Seligman Ventures. Developers now increasingly talk to their machines instead of typing code. Smallest.ai is taking a fundamentally different approach to the category by rethinking architecture itself. Customers get an efficient vertically integrated stack and don't need to waste time stitching models together. We believe the next generation of enterprise voice will be powered by Smallest AI." Smallest.ai currently partners with organizations managing large-scale voice operations, including RingCentral, Truecaller, Readymode, Piramal, Kogta, Pocket, and others. The company employs nearly 60 people and anticipates significant growth over the coming year as the demand for enterprise voice AI continues to rise. Smallest.ai is a foundational AI research lab building the next generation of real-time voice AI infrastructure for enterprises. The company develops speech recognition, speech generation, and speech-to-speech systems designed to enable natural, scalable AI conversations across customer service, healthcare, financial services, and other high-volume communication environments. Headquartered in San Francisco, Smallest.ai serves enterprises globally through its Voice 4.0 platform and proprietary AI models. The company is backed by Seligman Ventures, Sierra Ventures, 3one4 Capital, Better Capital, Upsparks Capital, Schema Ventures, Tiny VC, DeVC, Mission Street Capital, and other angel investors. BAM for Seligman Ventures on behalf of Smallest.ai: [email protected]
The End of Voice 3.0
Beyond Voice 3.0: Introducing Voice 4.0
Hydra: The Architecture Behind Voice 4.0
The Smallest.ai Models
About Smallest.ai
Media contact