Speechmatics
Cambridge speech-recognition company selling transcription and real-time speech understanding as an API, deployable in its cloud, on a customer's own servers or on-device — its pitch to buyers with security and compliance constraints.
Products
Selected offerings — the AI products recorded here, not the vendor's full catalogue.
-
Speechmatics' automatic speech recognition API transcribes audio into text in 55+ languages in either real-time streaming or batch mode, with speaker diarization, custom dictionary, translation and summarization options.
-
Speechmatics' text-to-speech API generates streaming synthetic English speech from text with sub-150ms latency using four named voices (Sarah, Theo, Megan, Jack), aimed at real-time voice agent use.
-
Voice Agent API API service
Speechmatics' voice agent offering provides a real-time conversational speech API - including the Flow WebSocket endpoint that chains speech-to-text, an LLM, text-to-speech and function calling - plus a Python Voice SDK for turn detection and speaker management, and integrations with Vapi, LiveKit and Pipecat.
-
On-Device Speech to Text Developer tool
A locally-executing speech-to-text engine for Mac and Windows laptops that runs on about one CPU core plus the device's neural engine or GPU and roughly 800MB of memory, sending no audio over a network and claiming accuracy within 5% of the cloud API.
Inside Speech to Text
-
Translates a transcript into other languages in the same API call that produces it, for files or live audio. It is a feature switched on inside a Speech to Text request rather than a separate product; the docs file it under Speech to Text and return the translations alongside the transcript.
-
Generates a short summary of an audio file in the same API call that transcribes it, as paragraphs or bullets. It is a Speech Intelligence feature enabled by adding a config block to a batch Speech to Text job, not a product bought on its own.
Closest alternatives 2
Similar products 68
Closest alternatives are pairs a vendor page itself compares. Similar products are matched from what each product does; nothing in that group is asserted as a rivalry.
Sources
- Total funding
-
Two rounds are public and the company states no total: a GBP 6.35M Series A (2019) and a USD 62M Series B led by Susquehanna Growth Equity (own newsroom, 2022-06-28: "has raised $62m in Series B funding"). Summing a sterling and a dollar round would need an fx_rate/fx_date pair; left null.
- Valuation
-
Checked — none published
No valuation published.
- Employees
-
The infobox gives a RANGE, "100-250", not a number. A range cannot be stored in this field and picking a point inside it would invent precision the source does not have.
- Founded
-
Infobox: founded 2006; industry speech recognition.
- Revenue
-
Checked — none published
Private; no revenue published.
- Description
-
Official documentation · 23 Sep 2026
"A speech-to-text API with three ways to deploy it: cloud, on-prem, or on-device." The earlier clause about defence work was UNPROVEN -- no Speechmatics page read mentions defence -- and was removed.
Open Speechmatics in the directory
Something wrong here? Send a correction — quote this company id: speechmatics.