Top Builders

Explore the top contributors showcasing the highest number of app submissions within our community.

Speechmatics

Founded in 2006 by Dr. Tony Robinson, a Cambridge University speech recognition pioneer, Speechmatics builds infrastructure to understand every voice globally. The company's core mission is inclusive, multilingual speech AI, covering transcription, real-time voice agents, and on-device deployment. Speechmatics serves enterprise clients across media, healthcare, financial services, and contact centers.

General
CompanySpeechmatics
Founded2006 by Dr. Tony Robinson
CEOKaty Wigdahl
HeadquartersCambridge, United Kingdom
Websitespeechmatics.com
Documentationdocs.speechmatics.com
GitHubgithub.com/speechmatics
TypeSpeech AI / B2B SaaS

Core Products

Speechmatics API (Speech-to-Text)

The Speechmatics API provides batch and real-time transcription across 55+ languages, powered by the Ursa 2 model released in October 2024. It supports speaker diarization, custom dictionaries, automatic translation, and Voice Intelligence add-ons (summarization, sentiment analysis, entity recognition) with no retraining required.

Speechmatics Flow

Flow is a voice agent API that combines Speechmatics' speech-to-text with an LLM and text-to-speech into a single real-time pipeline. Developers connect through a single API call to build conversational AI agents with smart turn detection, interruption handling, and function calling support.


Developer Resources

Speechmatics provides SDKs for Python, JavaScript/TypeScript, and .NET, along with a developer portal and free tier to get started without a credit card.

  • Documentation — official API reference, quickstarts, and integration guides
  • GitHub — open-source SDKs and client libraries
  • Developer Portal — API key management and usage dashboard
  • Pricing — free tier and pay-as-you-go rates

Key Features

Multilingual transcription across 55+ languages Speechmatics' Ursa 2 model leads accuracy benchmarks in 62% of supported languages on the FLEURS dataset, with 7.88% WER on Kincaid46 for English, surpassing human-level accuracy on that benchmark.

Flexible deployment Speechmatics runs on private SaaS cloud, on-premises, on-device, and via Docker or Kubernetes, making it suitable for data-sensitive industries like healthcare and finance.

Voice Intelligence add-ons Summarization, sentiment analysis, topic detection, chapter generation, and entity recognition layer on top of transcription without requiring additional integration work.


Use Cases

Contact center automation Real-time transcription and sentiment analysis during calls, combined with Flow for automated voice agent handling of common queries.

Clinical transcription Speechmatics' Medical Model (launched 2024) targets 93% real-time accuracy and 96% medical keyword recall for English, German, Danish, and Norwegian.

Media and broadcast Batch transcription of audio and video files for subtitling, archiving, and content search across multiple languages.

speechmatics AI Technologies Hackathon projects

Discover innovative solutions crafted with speechmatics AI Technologies, developed by our community members during our engaging hackathons.

Speech Transcription and Recording Assistant

Speech Transcription and Recording Assistant

ASTRA, the Adaptive Speech Transcription and Recording Assistant, is a hybrid Windows desktop application designed to turn live meetings, interviews, hearings, trainings, consultations, and uploaded recordings into organized, reviewable, and exportable documentation. Before transcription begins, ASTRA prepares audio locally using FFmpeg and Silero Voice Activity Detection. Silent and non-speech portions are skipped, while useful speech is isolated and compressed before online transmission. This reduces upload size, unnecessary AI processing, and provider usage. Long recordings are divided into manageable sections, allowing users to monitor progress, replay audio, retry failed parts, resume interrupted work, and avoid restarting an entire transcription because of one failed section. Users can choose between online and offline processing. Offline mode runs Whisper locally for privacy, poor connectivity, or reduced cloud dependence. Online mode connects to the ASTRA Server through a license-protected API. The server validates access, accepts individual or batched audio clips, creates asynchronous transcription jobs, and returns job status while processing continues. It can route requests across multiple configured speech-to-text providers and automatically try another provider when the preferred service becomes unavailable. This server layer keeps provider credentials away from the desktop app and allows models or providers to be changed without rebuilding the client. After transcription, local Sherpa-ONNX speaker diarization adds anonymous speaker labels and keeps conversations easier to follow across sections. ASTRA also supports transcript polishing, summaries, timestamps, playback, speech-detection logs, processing status, and exportable output. The result is a practical transcription workflow that combines local privacy, cloud performance, provider resilience, and efficient AI resource usage for real documentation work.

Apohara Synthex

Apohara Synthex

AI agents now run on the live web, but prompt injection is the number-one risk on the OWASP LLM Top 10, and most teams cannot prove what their agents ingested, or that it was safe. Apohara Synthex fixes that. Synthex is the provenance and security layer for the web data an AI agent consumes. It fetches across the full Bright Data spectrum: Web Unlocker, the Web Scraper API, SERP API, Scraping Browser, and the MCP Server. We didn't just use Bright Data; we improved it, contributing PR #140 upstream. Every fetch runs a layered defense before anything reaches a model. A deterministic regex pass and Qwen3Guard on Featherless form a high-recall net; NVIDIA's NemoGuard, selected by a measured benchmark, is the low-false-positive block gate; and a reasoning model on the AI/ML API knows the difference between describing an attack and executing one. Clean content is classified across four lenses, then sealed into an enterprise Evidence Report. The seal is real and shipped: an Ed25519 signature, an RFC 3161 DigiCert timestamp, an offline-verifiable Sigstore Rekor transparency log, and C2PA Content Credentials. Anyone can verify it in seconds with openssl, the industry's own c2patool, and a public ledger. No trust required. Cognee adds memory across re-scrapes, TriggerWare turns it into an automated monitor, and Kiro runs our continuous test and QA hooks. Synthex spans all three tracks, Security & Compliance, Finance & Market Intelligence, and GTM Intelligence, built for the CISO, CFO, compliance lead, and underwriter who need evidence they can defend to a board or a regulator. The average data breach costs 4.44 million dollars; Synthex seals an evidence artifact for a fraction of a cent. Everything signed, nothing trusted, and every number ships with a command to reproduce it.

Apohara Synthex

Apohara Synthex

AI agents now run on the live web, but prompt injection is the number-one risk on the OWASP LLM Top 10, and most teams cannot prove what their agents ingested, or that it was safe. Apohara Synthex fixes that. Synthex is the provenance and security layer for the web data an AI agent consumes. It fetches across the full Bright Data spectrum: Web Unlocker, the Web Scraper API, SERP API, Scraping Browser, and the MCP Server. We didn't just use Bright Data; we improved it, contributing PR #140 upstream. Every fetch runs a layered defense before anything reaches a model. A deterministic regex pass and Qwen3Guard on Featherless form a high-recall net; NVIDIA's NemoGuard, selected by a measured benchmark, is the low-false-positive block gate; and a reasoning model on the AI/ML API knows the difference between describing an attack and executing one. Clean content is classified across four lenses, then sealed into an enterprise Evidence Report. The seal is real and shipped: an Ed25519 signature, an RFC 3161 DigiCert timestamp, an offline-verifiable Sigstore Rekor transparency log, and C2PA Content Credentials. Anyone can verify it in seconds with openssl, the industry's own c2patool, and a public ledger. No trust required. Cognee adds memory across re-scrapes, TriggerWare turns it into an automated monitor, and Kiro runs our continuous test and QA hooks. Synthex spans all three tracks, Security & Compliance, Finance & Market Intelligence, and GTM Intelligence, built for the CISO, CFO, compliance lead, and underwriter who need evidence they can defend to a board or a regulator. The average data breach costs 4.44 million dollars; Synthex seals an evidence artifact for a fraction of a cent. Everything signed, nothing trusted, and every number ships with a command to reproduce it.