Speechmatics

Speechmatics

Cloud-based ASR platform with multilingual transcription

Overview

Speechmatics provides automatic speech recognition (ASR) technology that converts spoken language into text. It offers a platform-as-a-service (PaaS) with an API, enabling developers and businesses to embed real-time or batch transcription into their apps and workflows. Deployment options include cloud, on-premises, and on-device to support different security needs. The models are trained on millions of hours of unlabeled audio to support many languages, dialects, and accents, including a Global English model that handles major English accents with a single backbone. Distinguishing features include speaker identification, automatic punctuation, translation, and products like Ursa and Flow. The company targets media, contact centers, enterprise communications, and healthcare, and uses a usage-based pricing model with tiers and custom enterprise licensing. Its goal is to make speech-to-text accessible and accurate for diverse voices across industries.

About Speechmatics

Simplify's Rating
Why Speechmatics is rated
C+
Rated C on Competitive Edge
Rated B on Growth Potential
Rated C on Differentiation

Industries

Data & Analytics

Enterprise Software

AI & Machine Learning

Company Size

51-200

Company Stage

Series B

Total Funding

$81.6M

Headquarters

Cambridge, United Kingdom

Founded

2009

Get referred to Speechmatics

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Tresic picked Speechmatics over Google and Deepgram on September 9, 2026.
  • Boost.ai extended Speechmatics into U.S. regulated markets on February 12, 2026.
  • Stenograph and Adobe deepened deployments in July and April 2026, expanding enterprise distribution.

What critics are saying

  • Google, Deepgram, and AssemblyAI keep compressing ASR pricing and switching costs.
  • Speechmatics still depends on partners like Adobe, LiveKit, and Stenograph for distribution.
  • If Apple or OpenAI bundles superior on-device ASR, Speechmatics loses pricing power.

What makes Speechmatics unique

  • Speechmatics spans cloud, on-prem, and on-device; Adobe adopted it in Premiere in 2021.
  • Linden handles 55+ languages and strong accents through LiveKit Inference, launched August 27, 2026.
  • Melia unifies code-switching across 56+ languages and beats rivals on most FLEURS languages.

Help us improve and share your feedback! Did you find this helpful?

Funding

Total Funding

$81.6M

Above

Industry Average

Funded Over

5 Rounds

Series B funding is typically for startups that have proven their business model and need more funding to expand rapidly—often by entering new markets or adding more products. Investors are usually venture capital firms that specialize in later-stage investments.
Series B Funding Comparison
Above Average

Industry standards

$35M
$45M
Linktree
$62M
Speechmatics
$65M
Substack
$100M
ClickUp

Benefits

Health Insurance

Dental Insurance

Flexible Work Hours

Hybrid Work Options

Paid Vacation

401(k) Company Match

401(k) Retirement Plan

Family Planning Benefits

Fertility Treatment Support

Home Office Stipend

Growth & Insights and Company News

Headcount

6 month growth

0%

1 year growth

0%

2 year growth

-3%
Associated Press
Sep 9th, 2026
Tresic selects Speechmatics for speech-to-text after evaluation against Google and Deepgram

Tresic has selected Speechmatics to power speech-to-text for its Intelligence Cloud platform after evaluating it against Google and Deepgram. The platform adds conversation intelligence to business calls using the vCon standard, with Speechmatics running on Tresic's own infrastructure so audio never leaves it. Both companies operate behind other brands. Tresic sells through communications and managed service providers, whilst Speechmatics supplies speech technology for voice products built by others. This reflects a broader trend in business AI, where familiar brands use third-party models behind the scenes. Tresic's platform attaches to existing voice seats without requiring installation or affecting call quality. It produces same-day summaries, converts commitments into tasks, and alerts managers when conversations meet set conditions. Speechmatics was chosen for transcription accuracy, particularly for PCI redaction compliance and diarisation across varied audio conditions including accents and background noise.

Yahoo Finance
Aug 27th, 2026
Speechmatics launches Linden voice AI model on LiveKit for 55+ languages

Speechmatics has launched Linden, a speech-to-text model designed for voice agents, now available through LiveKit Inference across over 55 languages. Developers can integrate the model in under a minute via UI or code. The model addresses accuracy issues in real-world conditions where voice agents struggle, including strong accents, noisy environments, and alphanumeric sequences over poor-quality phone lines. Linden specialises in transcribing non-native speakers, dialects, and account numbers. LiveKit Inference manages the pipeline, routing, billing, and turn detection, eliminating the need for separate API keys or accounts. Norwegian AI company boost.ai already deploys both technologies in production for European enterprise clients. The partnership aims to reduce deployment time from weeks to minutes whilst improving transcription reliability in challenging production scenarios.

GlobeNewswire
Aug 27th, 2026
Speechmatics and LiveKit target the accuracy gap breaking voice agents in production.

Speechmatics and LiveKit target the accuracy gap breaking voice agents in production. Linden, Speechmatics' new speech-to-text model, is available now through LiveKit Inference, enabling developers to build reliable voice agents at scale across 55+ languages. CAMBRIDGE, United Kingdom, Aug. 27, 2026 (GLOBE NEWSWIRE) - From today, developers building voice agents on LiveKit Inference can switch to Speechmatics in less than a minute through the UI or code. Linden, the first Speechmatics model purpose-built for voice agents, is available through LiveKit Inference, LiveKit's platform for using leading voice AI models from one place. That means no separate API key, account or invoice. "LiveKit Inference exists so developers can ship and iterate fast, without the headache of managing API keys or rewriting their stack every time a model..." "What separates an agent people trust from one they find exhausting is whether it can understand what you said. Our new model solves for this. Linden not only..." "Voice agents live or die on the STT layer. Speechmatics gives us what actually matters in production: low-latency real-time models, fast iteration on the..." "LiveKit Inference exists so developers can ship and iterate fast, without the headache of managing API keys or rewriting their stack every time a model..." In a demo scenario, a voice agent typically hears one speaker, clean audio and a short exchange where all STT providers perform well. Real-world production is rarely that tidy: a caller with a strong accent, a noisy cafe, a 16-digit card number read out over a low quality phone line. When the transcript breaks on the words that matter, everything downstream breaks with it. Together, LiveKit and Linden solve that problem for the builder. Linden gets the words right where it is hardest: 55+ languages, non-native speakers and strong accents, and alphanumerics such as account and card numbers read out over a bad line. LiveKit Inference handles the pipeline, the routing, the billing and the turn detection that decides when the agent speaks. Adopting a speech provider used to take weeks, now it's something you can simply select. "What separates an agent people trust from one they find exhausting is whether it can understand what you said. Our new model solves for this. Linden not only excels at 55+ languages, but also transcribing non-native speakers, dialects, and accents. LiveKit built the infrastructure that made those deployments possible, and its community has worked out what production voice AI actually takes, in the real world, faster than anyone else." Ricardo Herreros-Symons, CRO, Speechmatics Norwegian conversational AI company boost.ai already runs Speechmatics and LiveKit in production across the voice agents it builds for European enterprises, on infrastructure that also powers xAI Grok and Salesforce Agentforce. "Voice agents live or die on the STT layer. Speechmatics gives us what actually matters in production: low-latency real-time models, fast iteration on the hard edges like alphanumerics and end-of-utterance, and serious language breadth across Europe. Just as importantly, the people behind it are a pleasure to work with; close, hands-on, and genuinely invested in getting things right." Rasmus Hauch, CTO, boost.ai LiveKit Inference, launched in October 2025, lets developers build a voice agent in minutes. A single API key calls the most popular STT, LLM and TTS models, and models swap with a one-line code change. Consolidated billing, rate-limit management and usage reporting all run through LiveKit Cloud. Speechmatics is available natively through LiveKit Inference, via the Speechmatics plugin for LiveKit Agents Open Source framework or through LiveKit's Agent Builder. "LiveKit Inference exists so developers can ship and iterate fast, without the headache of managing API keys or rewriting their stack every time a model changes. Adding Speechmatics equips voice agents with exceptional speaker diarization that clearly tracks who said what with any number of speakers from day one. This capability is especially powerful for physical AI and real-world multi-speaker environments." David Zhao, CTO & Co-founder, LiveKit About Speechmatics: Speechmatics builds the most accurate and inclusive speech recognition in the world, trusted by enterprises across financial services, media, contact centers and government for unmatched language coverage, accuracy, and flexibility. About LiveKit: LiveKit is the open source framework and cloud platform for voice, video, and physical AI agents. Used by developers and enterprises worldwide, LiveKit powers realtime, multimodal AI applications at scale - from customer support and telephony to robotics and consumer products. The company is headquartered in San Francisco. Learn more at livekit.com. Media Contact Mieke Smith // [email protected]

Associated Press
Jul 22nd, 2026
Stenograph integrates Speechmatics speech recognition into CATalyst VP voice reporting software

Stenograph has partnered with Speechmatics to integrate speech recognition directly into CATalyst VP, its voice reporting software for court reporters. The collaboration addresses challenges voice reporters face managing multiple technology platforms. The integration eliminates the need for separate third-party speech recognition software, reducing issues like latency and system instability. Speechmatics runs entirely on-device, functioning without internet connectivity whilst keeping sensitive proceedings on the reporter's computer. "This exclusive partnership gives voice reporters the consistency, reliability, and ease of use they've been seeking within a single solution," said Michelle McLaughlin, Stenograph's Vice President of Sales. The integration removes requirements for additional devices and per-minute charges. Speechmatics will also power CheckIt Plus, an add-on service for stenographers and voice reporters to improve realtime transcription and reduce editing time.

Speechmatics
Jul 22nd, 2026
Stenograph and Speechmatics announce industry-first on-device integration for CATalyst VP.

Stenograph and Speechmatics announce industry-first on-device integration for CATalyst VP. The move brings speech recognition directly into CATalyst VP eliminating the challenges of running multiple applications, making it the first seamless solution for the voice reporting industry. Stenograph, the global leader in court reporting technology, today announced an exclusive integration with Speechmatics, bringing industry-leading on-device & realtime speech recognition directly into CATalyst VP, its specialized voice reporting software. As voice reporting continues to grow across the legal industry, professionals face the challenges of managing multiple technology platforms to produce an accurate record. Running separate speech recognition software alongside CAT software often requires additional logins, independent update schedules, compatibility checks, and ongoing troubleshooting - creating unnecessary complexity in a profession where focus and accuracy matter most. To address these challenges, Stenograph and Speechmatics have partnered to deliver the first fully integrated speech recognition solution built directly into Voice Reporting CAT software. This exclusive integration eliminates the need for voice reporters to run a separate third-party speech engine alongside CATalyst VP, creating a more streamlined and accurate user experience while helping reduce common issues such as latency, system instability, and software version conflicts. Because Speechmatics runs entirely on-device, that reliability holds even without an internet connection, so reporters get the same performance in a fully connected courtroom or an offline deposition, with sensitive proceedings never leaving a voice reporter's computer. For decades, court reporters have trusted Stenograph to provide the tools and services they need to take down and produce the verbatim record. This exclusive partnership with Speechmatics gives voice reporters the consistency, reliability, and ease of use they've been seeking - all within a single solution. By eliminating the need for separate speech recognition software, high-spec computers, additional devices, or an internet connection, CATalyst VP with Speechmatics allows voice reporters to focus more confidently on capturing the record. It also removes the uncertainty of per-minute charges. - Michelle McLaughlin, Vice President of Sales at Stenograph. California official court reporter and CATalyst VP user, Vance Malone, CVR-M, RVR, stated: What excites me most isn't just another speech engine - it's having Speechmatics fully integrated into CATalyst VP. Native integration removes one of the biggest historical challenges for voice reporters: getting multiple applications to work together reliably in realtime. Giving reporters another high-quality recognition option, while keeping Dragon available, provides flexibility without adding complexity. In addition to powering CATalyst VP for voice reporters, Speechmatics will also power CheckIt Plus, an add-on service used by both stenographers and voice reporters to deliver cleaner realtime and rough drafts while reducing overall editing time. Court reporters need technology that truly simplifies their work. Building its speech recognition directly into Stenograph's software means legal professionals get that accuracy, connectivity and privacy they need without adding the complexity of another tool to manage. - Ricardo Herreros-Symons, CRO of Speechmatics. " You can learn more at StenographxSpeechmatics PR Media contacts: Mieke Smith [email protected] Madi Dixon [email protected]

Recently Posted Jobs

Sign up to get curated job recommendations

Speechmatics is Hiring for 4 Jobs on Simplify!

Find jobs on Simplify and start your career today

Don't see your dream role? Check out thousands of other roles on Simplify. Browse all jobs →