-
UK PM says not embarrassed by hoax messages with fake Trump aide
-
Syria, US say Israel struck military air base in Idlib province
-
West Bank siege exposes Israel army inaction on settler violence
-
France probes Russian disinfo campaign against centrist candidates
-
Four-time champion Roglic to ride Vuelta after training crash
-
Stocks mostly drop as Mideast hopes dim, interest rates rise
-
Ukrainian convicted in Germany over Russian sabotage plot
-
Bayern's Musiala returns to training after on-field collapse
-
Beleaguered FIFA chief Infantino rids himself of critic Lamour
-
China's Baidu, betting on AI, posts fifth straight quarterly revenue drop
-
Kane confirms Bayern extension talks set to start
-
German farmers warn of drought toll on grain harvest
-
Asian allies stress 'critical' US cooperation as Trump slashes Korea drills
-
Pant fastest to 100 sixes as India set Sri Lanka 372 to win first Test
-
Pakistan Supreme Court orders ex-PM Khan moved from jail to hospital
-
Russia rebukes Japan ambassador over disputed islands row
-
Firefighters push to 'encircle' Belgian wildfire
-
Crude extends gains, most stocks drop as Mideast hopes dim
-
ISS to host first spacewalk by Frenchwoman
-
Rift deepens as Asia football chief hits back at Qatar in Infantino saga
-
Philippine student killed in livestreamed school shooting
-
India extend lead to 294 in first Sri Lanka Test
-
More Trump-Kim 'love'? What we know about North Korea
-
Asian allies stress 'critical cooperation' as Trump slashes Korea drills
-
Major Meta trial begins as lawyers spar over witnesses, damages
-
Williams sisters fall at Cincinnati while Zverev advances
-
Trump threatens to bomb Oman if it gets in way of Iran deal
-
Transylvania's fortified churches face fragile future
-
Australian cricketer Warner ordered to use breathalyser vehicle lock
-
China mourns former premier Zhu Rongji with flags lowered, funeral
-
A Viking longship in 'The Odyssey'? It's a case of history meets Hollywood
-
Philippine student killed in rare school shooting
-
Racist abuse spurred Suzuki as Japan goalkeeper set for Premier League
-
Arsenal primed for Premier League title defence as rivals face questions
-
Japan lights torch as countdown begins to Asian Games
-
Zverev advances at Cincinnati while Williams sisters fall
-
Make it rain: Indonesia chases clouds to stem El Nino fires
-
Copper powers profit rise at mining giant BHP
-
All aboard as Asian Games preparations enter 'final stages'
-
Whales spark joy in New York but face threats from boat traffic
-
Lights out, crowds gone: Havana waits for what comes next
-
Premier League's new bosses brace for battle
-
How families in Milei's Argentina sank deep into debt
-
Zambia's Hichilema: 'cattle boy' and two-time president
-
Zambia's Hichilema wins re-election with 60% of vote
-
Elderly far more at risk from rising heat than thought: study
-
Jeanie Buss to fight siblings over sale of 17.8% stake in Lakers: reports
-
Lobe Sciences Reports Third Quarter 2026 Financial Results
-
XCF Global (NASDAQ:SAFX) Achieves Commercial Velocity: Converts Production into Sales, Shipping 55,000 Gallons of Renewable Diesel Daily
-
374Water Reports Second Quarter 2026 Financial Results; Revenue Increases More Than 280% Year-Over-Year
Modulate Launches Velma Transcribe: High-Performance Transcription For Real World Conversations at 90% Lower Cost
Modulate's ELM model architecture unlocks transcription for the masses, cutting costs by 10x while achieving industry-leading accuracy.
BOSTON, MA / ACCESS Newswire / March 18, 2026 / Modulate, the frontier conversational voice intelligence company, today announced Velma Transcribe, a speech-to-text API delivering high-accuracy, low-latency transcription at 90% lower cost per hour than other leading transcription providers. This significantly lower price point represents a fundamental shift in the economics of transcription. For a fraction of the cost, Modulate unlocks affordable speech-to-text transcription for every audio conversation in the world, empowering real-time voice agents, call center platforms, social apps, and more with industry-leading transcription tools at a global scale.

Built using Modulate's industry-leading Ensemble Listening Model (ELM) research, Velma Transcribe orchestrates an ensemble of specialized transcription models to improve accuracy, latency, and cost efficiency compared to any single model. In addition to the outstanding unit economics, Velma Transcribe achieves industry-leading results on widely used datasets, including Earnings-22 and the AMI Meeting Corpus. The result is a new standard for conversational audio transcription, combining strong accuracy on complex multi-speaker audio with dramatically improved unit economics for processing voice data at scale.
"Modulate is the world leader in using voice understanding AI, and our goal is to make the tools to understand audio available to anyone, at any scale," said Carter Huffman, CTO and Cofounder of Modulate. "Our full ensemble for conversation understanding, Velma, already outperforms LLMs in recognizing key behaviors, and now Velma Transcribe makes one of our core underlying capabilities available directly to developers who simply need accurate transcripts, not behavioral insights."
In addition, Velma Transcribe offers features built for Enterprise use cases:
Emotion detection (20+ emotions)
Accent detection (20+ accents)
Multilingual (70+ languages)
PII redaction, diarization, streaming support, and more
Lower Transcription Costs By up to 10X
Velma Transcribe reduces transcription costs to approximately $0.03 per hour of audio, more than 90% lower than leading providers. These economics make it far more cost-effective for enterprise organizations to analyze and monetize their voice data.
$0.03 - Modulate Velma Transcribe
$0.40 - ElevenLabs Scribe v2
$0.31 - Deepgram Nova-3
$0.26 - Deepgram Nova-2
$0.21 - AssemblyAI Universal-3 Pro
*Based on publicly listed pricing as of March 18, 2026
Compare the leading speech-to-text transcription companies on cost and accuracy at Speechtxt.com.

Top Marks for Conversational Audio Accuracy at Scale
Velma Transcribe is engineered for real-world conversations that challenge traditional systems, including overlapping speakers, interruptions, accents, and background noise. On the AMI Meeting Corpus dataset, a widely used benchmark for complex multi-speaker conversational audio, Velma avoids over 40% of the errors made by Eleven Labs and over 70% of the errors made by OpenAI GPT-4o-transcribe.
Huffman explains the top marks, "We've tuned Velma for conversational audio, including emotion and accent detection, leading to materially lower error rates on meeting and call data while delivering dramatic cost savings versus incumbent providers. That combination makes high-quality transcription practical at scale."
Built for Secure Enterprise Voice Production
Velma Transcribe includes all the capabilities developers expect and enterprise operations need, including:
Batch and streaming transcription endpoints with structured output and segment timestamps
Zero data stored, ensuring privacy-safe workflows
Sub-second streaming latency with partial transcripts for live applications and agent pipelines
Robust formatting optimized for conversational speech and long recordings
Broad language coverage in 70 of the world's most commonly spoken languages
Personally Identifiable Information (PII) detection and redaction
Advanced transcription enrichments, including speaker diarization, emotion detection, and accent identification
Backed by Modulate's security practices and ISO 27001 certification, these capabilities allow developers to build secure, voice-enabled applications and help organizations extract insights from large volumes of conversational data.
Models that Listen and Understand
Velma Transcribe is part of Modulate's growing family of Velma 2.0 voice analytics models built to deliver a new, context-rich listening layer for AI systems. It represents the first step in Modulate's expanding developer API strategy, with additional capabilities planned across synthetic voice detection, emotion analysis, and deeper conversational intelligence. Together, these capabilities allow developers and enterprises to move beyond transcription to understand how conversations unfold, enabling applications such as fraud detection, customer sentiment analysis, compliance monitoring, and real-time decision support.
"The industry has spent years teaching AI how to generate and respond. The next frontier is teaching it how to listen," said Mike Pappas, CEO and Cofounder of Modulate. "Most systems today rely on transcription, reducing rich conversations to flat text and losing the signals humans naturally understand. Velma is the listening layer for AI, giving developers and enterprises the 'ears' needed to build voice-native applications that can capture the nuance and intent within spoken dialogue."
Availability and Pricing
Velma Transcribe is available today with batch and sub-second streaming transcription. Modulate pricing is usage-based and optimized for high-volume workloads: https://www.modulate.ai/pricing
About Modulate
Modulate is a voice intelligence company building AI models and APIs designed to understand real-world conversational audio at scale. Its technology combines speech recognition, acoustic analysis, and conversational context to deliver reliable, explainable, and cost-effective voice intelligence for developers and enterprises.
For more information or to get started, visit modulate.ai.
Media Contact
Megan Fasy
Grithaus Agency
(e) [email protected]
(m) +1 (617) 480-3674
###
SOURCE: Modulate
View the original press release on ACCESS Newswire
P.A.Mendoza--AT