Trint
AI transcription and an editorial workflow for journalists and media teams
A developer-first speech-to-text API with real-time and async transcription in many languages
Gladia is a developer-first speech-to-text API offering real-time and async transcription, diarization, translation, and summarization with transparent per-hour pricing.
Gladia (gladia.io) provides speech-to-text as an API for builders. It supports both asynchronous batch transcription and low-latency real-time streaming, along with features like speaker diarization, word-level timestamps, translation, and audio intelligence such as summarization and sentiment. Its multilingual coverage and accuracy make it a common choice for embedding transcription into meeting tools, contact centers, and media products. Pricing is transparent and usage-based, billed per hour of audio with no seat fees, and it includes a free monthly allowance plus starter credits for new accounts. Async transcription starts at a low per-hour rate, real-time is slightly higher, and volume discounts apply on scaling and enterprise tiers, including offline licensing for teams with data-residency needs. Gladia competes directly with speech-to-text APIs like AssemblyAI, Deepgram, and Speechmatics, positioning on price, speed, and European data options. It fits developers and product teams building transcription-powered features who want predictable per-hour costs and multilingual support, rather than individuals looking for a finished transcription interface.
Gladia is a developer-first speech-to-text API with real-time and async transcription, multilingual support, and transparent per-hour pricing, built for product teams.
Gladia (gladia.io) is a speech-to-text company providing transcription as an API. It targets developers and product teams embedding audio intelligence into their software, with a European base and data-residency options.
The company competes directly with AssemblyAI, Deepgram, and Speechmatics, differentiating on transparent per-hour pricing, speed, multilingual coverage, and offline licensing for compliance-sensitive customers.
Gladia offers async batch transcription and low-latency real-time streaming, with speaker diarization, word-level timestamps, translation, and audio intelligence such as summarization. It exposes REST, WebSocket, and SDK access with webhook delivery.
Pricing is usage-based per hour with no seat fees, including a free monthly allowance and starter credits. Scaling and enterprise tiers add volume discounts and offline licensing.
Gladia targets developers and SaaS product teams building transcription features into meeting tools, contact centers, and media products who want predictable per-hour costs and multilingual support.
Developers integrating transcription into applications.
Engineering and product leaders selecting an STT provider.
Developer communities and technical evaluators.
A product team building audio or meeting features that needs an accurate, multilingual, per-hour-priced transcription API.
Gladia is a venture-backed startup; verify specific funding details with the vendor as figures may have changed.
Gladia is a speech-to-text API for developers offering real-time and asynchronous transcription with diarization, translation, and summarization across many languages.
Gladia uses transparent per-hour pricing with no seat fees. Async transcription starts around $0.61 per hour and real-time around $0.75 per hour on Starter, with volume discounts on higher tiers.
Yes. Gladia includes a free monthly allowance of transcription hours plus starter credits for new accounts, so developers can test before scaling.
Yes. Gladia supports low-latency real-time streaming over WebSocket in addition to asynchronous batch transcription.
Gladia competes on price, speed, multilingual coverage, and European data options, positioning as a transparent, developer-first alternative to APIs like Deepgram, AssemblyAI, and Speechmatics.
Side-by-side pages for pricing, features, and best-fit use cases.
AI transcription and an editorial workflow for journalists and media teams
Query-aware prompt compression that cuts LLM input tokens by roughly 60% before inference.
Run open LLMs locally with a single command.
Developer platform for real-time conversational video AI.