Skip to main content

Deepgram

Category: Voice / Speech AI

A foundational AI platform providing high-performance, real-time speech-to-text and text-to-speech APIs for developers and enterprises building voice-first applications. Deepgram was founded in 2015. The company is led by Scott Stephenson. Based in San Francisco, USA. Team size: 101-500. Total funding raised: $215M total. Latest round: Series C ($130M, Jan 2026) at $1.3B valuation. Key investors include AVP, Madrona, Wing VC, NVIDIA, Tiger Global, Y Combinator, Alkeon, In-Q-Tel, BlackRock, Twilio, ServiceNow Ventures, SAP.

AMW Analysis

Deepgram provides real-time speech-to-text and text-to-speech APIs aimed at developers and enterprises building voice applications, positioning itself on accuracy, latency, and total cost of ownership. Founded in 2015 and based in San Francisco, the company has raised $215M to date. Its most recent funding event was a $130M Series C in January 2026 at a $1.3B valuation, with strategic backing from SAP and Twilio alongside existing investors including AVP, Madrona, Wing VC, NVIDIA, Tiger Global, and Y Combinator.

Alongside the raise, Deepgram acquired OfOne, a company whose technology handles a large share of drive-thru orders autonomously, indicating a move from horizontal API infrastructure toward vertically integrated solutions for specific high-volume use cases. Recent coverage describes the company as serving over 1,300 organizations, including NASA, and frames the acquisition and funding together as part of a broader push from transcription toward more autonomous voice-agent capabilities. The clustering of funding and acquisition news in a short window suggests this period represents a deliberate scaling and expansion phase for the company's platform and market reach.

AMW analysis, generated from 2 tracked news signals.

Founded
2015
Headquarters
San Francisco, USA
Team size
101-500
Total funding
$215M total

Value proposition

Delivers the industry's highest accuracy and lowest latency for voice AI at the lowest total cost of ownership, enabling human-like real-time conversational interactions.

Products and solutions

Nova-3 (Speech-to-Text): Flagship high-accuracy STT, 50+ languages, real-time multilingual code-switching, Aura-2 (Text-to-Speech): Enterprise-grade low-latency voice synthesis, Voice Agent API: End-to-end real-time conversational AI / speech-to-speech, Flux: Conversational speech recognition with turn detection and multilingual support, Saga: Voice OS, Language AI: Unified API for summarization, sentiment analysis, and NLU

Unique value

Utilizes a radical end-to-end deep learning architecture that bypasses legacy phoneme-based models, allowing for direct audio-to-text processing with higher contextual awareness.

Target customer

Software developers, enterprise engineering teams, contact center operators, and AI agent builders.

Industries served

Contact Centers & Customer Experience, Media & Entertainment, Healthcare & Telemedicine, Finance & Banking, Technology & Software Development

Technology advantage

GPU-native infrastructure and proprietary foundational models provide sub-300ms latency and superior accuracy in noisy or domain-specific environments, outperforming legacy 'Big Tech' providers in both speed and cost-efficiency.

How they differentiate

Deepgram utilizes a proprietary end-to-end deep learning architecture that bypasses legacy phoneme-based models, delivering sub-300ms latency and higher accuracy for real-time conversational AI compared to 'Big Tech' providers.

Main competitors

AssemblyAI, OpenAI (Whisper), Google Cloud Speech-to-Text

Key partnerships

NVIDIA (Strategic technology and investment partner for GPU optimization), Amazon Web Services (AWS Global Startup Programme & SageMaker integration), Twilio (Strategic investor and voice AI infrastructure partner), Qualcomm (Nova-3 on-device STT on Snapdragon X Series / Hexagon NPU, Jul 2026), ServiceNow Ventures / SAP (Series C strategic investors), Y Combinator (Alumnus and ecosystem partner)

Notable customers

Twilio, Spotify, NASA, Citibank, Granola, Vapi, Cloudflare, Sierra, Decagon, Cresta, Daily, LiveKit

Major milestones

Achieved unicorn status with $1.3B valuation via $130M Series C led by AVP (Jan 2026), Acquired YC-backed OfOne for restaurant/drive-thru voice AI (Jan 2026), Launched Powered by Deepgram program (1,300+ orgs), Acquired Poised communication AI (Jun 2024), Launched Nova-3 STT (Feb 2025) and Aura-2 TTS; Flux conversational CSR; Saga Voice OS, Expanded Nova-3 on-device to Snapdragon PCs (Jul 2026)

Growth metrics

Reached $21.8M ARR (2024 est.); cashflow positive in 2025; 200,000+ developers and 1,300+ organizations building on Deepgram APIs; processed 50,000+ years of audio / 1T+ words transcribed.

Market positioning

High-performance Voice AI platform leader targeting developers and enterprises building real-time AI agents.

Geographic focus

Global, with primary market concentration in North America, Europe, and Asia-Pacific.

Patents and IP

Patent portfolio filed since 2016 with multiple U.S. patents granted in 2025, including US 12,380,880 (End-to-End ASR with Transformer), US 12,334,075 (Hardware-Efficient ASR), and US 12,499,875 (Deep Learning Internal State Index-Based Search and Classification).

About Scott Stephenson

Scott Stephenson is a physicist-turned-entrepreneur who co-founded Deepgram after researching dark matter at the University of Michigan, where he built a lab two miles underground. He is a Y Combinator alumnus (W16) and has led Deepgram to become a $1.3 billion unicorn in the voice AI space, pioneering foundational models for speech-to-text and real-time voice agents.

Latest news about Deepgram

More Voice / Speech AI companies

Official website: