2026年10月
AssemblyAI released a comprehensive guide to selecting the best speech-to-text API in 2026, emphasizing benchmarking on real-world audio rather than just WER. T
AssemblyAI introduced its Dictation API, designed to return finished text (cleaned of fillers, self-corrections, and punctuated) instead of raw transcripts, add
AssemblyAI unveiled a new cloud-based voice AI infrastructure platform emphasizing unmatched price-performance, predictable usage-based pricing, and zero infras
2026年9月
AssemblyAI detailed how its ambient AI scribe technology, originally built for healthcare, adapts to veterinary and legal documentation needs. The post highligh
AssemblyAI’s blog details reproducible cases where keyterm prompting worsens transcript accuracy by overriding correct words or over-prompting, particularly wit
AssemblyAI published a technical guide distinguishing clinical dictation (short-clip) from ambient scribing, clarifying when to use the Sync API (verbatim trans
AssemblyAI published a detailed runbook for operating its Voice Agent API post-launch, covering staged traffic ramps, rollback thresholds, severity definitions,
AssemblyAI released a guide on evaluating speech-to-text APIs for education platforms, emphasizing student-speech accuracy, long-form audio economics, FERPA com
AssemblyAI released Universal-3.6 Pro Realtime, a major update to their streaming speech-to-text model, reducing word error rates by 16-52% across scenarios lik

AssemblyAI released Universal-3.6 Pro Realtime, a speech-to-text model trained on tens of thousands of hours of real voice-agent and call-center conversations.
AssemblyAI announced Universal-3.5 Pro, its top-ranked speech-to-text model, is now available on OpenRouter’s transcription endpoint. The integration allows dev

AssemblyAI introduced Speaker Diarization for pre-recorded audio, enabling automatic labeling of individual speakers in transcripts. The feature splits transcri
AssemblyAI’s LLM Gateway now lists multiple new models—including Qwen3.5 4B Fast, Claude Sonnet 4.6, and Google’s Gemini 2.5 Flash Lite—with detailed context le
AssemblyAI introduced the Dictation API, a new product designed to convert spoken utterances into clean, ready-to-use text by resolving self-corrections, removi

AssemblyAI introduced LLM Gateway, a unified API for accessing 25+ LLMs across providers like Anthropic, Google, OpenAI, and Qwen with automatic fallbacks and r
AssemblyAI now offers official, open-source SDKs for Python and JavaScript covering pre-recorded transcription, real-time streaming, and speech understanding. T
归档中还有 +116 条更新:可按上方的月份或年份浏览。
自动跟踪 AssemblyAI