AssemblyAI launches a low-latency Sync Speech-to-Text API at $0.45/hr
AssemblyAI added a seventh priced product: a Sync Speech-to-Text API at $0.45/hr that returns finished transcripts in a single synchronous call with ~134 ms p50 latency and Universal-3.5 Pro accuracy.
Six product tabs (Pre-recorded STT, Realtime STT, Voice Agent, Speech Understanding, Guardrails, LLM Gateway); no single-call synchronous transcription endpoint.
Seven product tabs — a new Sync Speech-to-Text API ($0.45/hr) joins the lineup, returning finished transcripts in one API call (no polling, WebSocket, or job to manage), up to 2 min/request, ~134 ms p50, 18 languages. All other rates unchanged.
AssemblyAI added a new Sync Speech-to-Text API priced at $0.45/hr — its seventh distinct priced product surface. The Sync API returns a finished transcript in a single synchronous API call: you POST a short clip and read the transcript off the response, with no polling, WebSocket, or job to manage.
It delivers Universal-3.5 Pro accuracy, processes up to 2 minutes per request with roughly 134 ms p50 latency, and includes per-word timing and confidence across 18 languages. The $0.45/hr rate matches Universal-3.5 Pro Realtime, reflecting the low-latency infrastructure. All other rates are unchanged: async transcription ($0.15/hr Universal-2, $0.21/hr Universal-3.5 Pro), realtime streaming ($0.15-$0.45/hr), and the Voice Agent API ($4.50/hr).
AssemblyAI added a new low-latency Sync Speech-to-Text API at $0.45/hr — a seventh priced product surface. It returns finished transcripts in a single synchronous call (no polling, WebSocket, or job to manage), delivers Universal-3.5 Pro accuracy, processes up to 2 minutes per request with ~134 ms p50 latency, and includes per-word timing and confidence across 18 languages. All other rates unchanged (async $0.15/$0.21, streaming $0.15/$0.45, Voice Agent $4.50/hr).