AssemblyAI 在 2026年9月 发布了 14 项受追踪的更新。 较 2026年8月 增长 27%。 主要主题是 feature updates(9)。 活动主要集中在 2026年9月28日 那一周。 亮点:“How to build clinical dictation on AssemblyAI”(2026年9月30日)。
浏览 AssemblyAI 的其他月份
AssemblyAI published a technical guide distinguishing clinical dictation (short-clip) from ambient scribing, clarifying when to use the Sync API (verbatim transcript) versus the Dictation API (cleaned…
AssemblyAI published a detailed runbook for operating its Voice Agent API post-launch, covering staged traffic ramps, rollback thresholds, severity definitions, and on-call ownership. The guidance add…
AssemblyAI’s blog details reproducible cases where keyterm prompting worsens transcript accuracy by overriding correct words or over-prompting, particularly with similar-sounding terms. It introduces …
AssemblyAI detailed how its ambient AI scribe technology, originally built for healthcare, adapts to veterinary and legal documentation needs. The post highlights key architectural differences, includ…
AssemblyAI released a guide on evaluating speech-to-text APIs for education platforms, emphasizing student-speech accuracy, long-form audio economics, FERPA compliance, and accessibility requirements.…
AssemblyAI released Universal-3.6 Pro Realtime, a speech-to-text model trained on tens of thousands of hours of real voice-agent and call-center conversations. It reduces errors on short replies by 45…
AssemblyAI released Universal-3.6 Pro Realtime, a major update to their streaming speech-to-text model, reducing word error rates by 16-52% across scenarios like noisy environments, background speech,…

AssemblyAI announced Universal-3.5 Pro, its top-ranked speech-to-text model, is now available on OpenRouter’s transcription endpoint. The integration allows developers to access the model through a si…

AssemblyAI introduced Speaker Diarization for pre-recorded audio, enabling automatic labeling of individual speakers in transcripts. The feature splits transcripts into utterances per speaker with con…
AssemblyAI’s LLM Gateway now lists multiple new models—including Qwen3.5 4B Fast, Claude Sonnet 4.6, and Google’s Gemini 2.5 Flash Lite—with detailed context lengths, supported parameters, and pricing…
AssemblyAI introduced the Dictation API, a new product designed to convert spoken utterances into clean, ready-to-use text by resolving self-corrections, removing filler words, and preserving user-int…

AssemblyAI introduced LLM Gateway, a unified API for accessing 25+ LLMs across providers like Anthropic, Google, OpenAI, and Qwen with automatic fallbacks and regional endpoints (US/EU). The service s…
AssemblyAI now offers official, open-source SDKs for Python and JavaScript covering pre-recorded transcription, real-time streaming, and speech understanding. The SDKs handle authentication, polling, …
AssemblyAI now supports 18 languages in its Universal-3.5 Pro model, including regional dialects like Quebecois French, Mexican Spanish, and Brazilian Portuguese, with automatic dialect recognition. T…