AssemblyAI 在 2026年7月 发布了 16 项受追踪的更新。 较 2026年6月 下降 16%。 主要主题是 feature updates(10)。 活动主要集中在 2026年7月6日 那一周。 亮点:“AssemblyAI Universal-3.5 Pro benchmarks”(2026年7月17日)。
浏览 AssemblyAI 的其他月份
AssemblyAI’s research demonstrates that the standard diarization metric DER (Diarization Error Rate) ranks speech-to-text outputs backwards compared to human judgment. Using real audio clips and produ…
AssemblyAI released comparative benchmarks for its Universal-3.5 Pro and Universal-3.5 Pro Realtime models, demonstrating lower word error rates (WER), missed entity rates, and diarization errors than…

AssemblyAI launched Universal 3.5 Pro Realtime, a next-generation streaming speech-to-text model, and integrated it into LiveKit’s voice agent framework via the `livekit-agents` SDK (v1.6+). The integ…
AssemblyAI introduced two new prompting methods for its Universal-3.5 Pro model to improve transcription accuracy in challenging audio scenarios. Contextual prompting allows users to provide natural-l…
AssemblyAI introduced a summarization feature for audio transcripts that splits output into timestamped chapters with headlines, available in open beta. Users can choose between bullet or paragraph fo…
AssemblyAI introduced the Sync API, enabling one HTTP request to transcribe short audio clips with Universal-3.5 Pro in ~134 ms, eliminating polling, WebSockets, or chunking. The API targets use cases…

AssemblyAI introduced a webhook system for its voice agent API, enabling real-time notifications for session and call lifecycle events. Developers can subscribe to events like `session.started`, `call…
AssemblyAI’s Voice Agent API now uses semantic end-of-turn detection and adaptive pacing by default, eliminating the need for manual tuning. The system waits for complete tool values (e.g., phone numb…
AssemblyAI introduced a new `input.keyterms` feature in its Voice Agent API to boost transcription accuracy for rare or domain-specific terms like brand names, proper nouns, and jargon. Developers can…
AssemblyAI updated its Voice Agent API tools documentation to introduce parameter hints (enum, examples, pattern, format) that improve tool-calling accuracy and turn detection. Execution modes and pro…
AssemblyAI now allows voice agents to use custom OpenAI-compatible LLM endpoints instead of their managed model. Users can configure `base_url`, `model`, and `api_key` in the agent settings to route c…
AssemblyAI introduced the ability to update streaming session parameters mid-session using an `UpdateConfiguration` message without reconnecting. Users can dynamically adjust accuracy/latency modes, t…
AssemblyAI updated its Streaming WebSocket API to introduce Universal-3-5-Pro Streaming as the default model, replacing the prior default. The update adds language steering via the `language_codes` pa…
AssemblyAI introduced a new action items feature in beta that generates timestamped action items, quotes, and effort levels from audio transcripts. The feature supports US and EU regions, offers low (…
AssemblyAI released Universal-3.5 Pro, a flagship async speech-to-text model featuring native code-switching across 18 languages, the most accurate speaker diarization to date, and contextual promptin…

AssemblyAI introduced contextual awareness in Universal-3.5 Pro Realtime, enabling the model to use conversation history, speaker context, and dynamic prompts to improve transcription accuracy in nois…