AssemblyAI · 2026
116 mises à jour · 2026
En 2026, AssemblyAI a publié 116 mises à jour suivies. Soit environ 17 par mois sur 7 mois actifs. Le mois le plus actif a été avril 2026 (33). Le thème dominant : feature updates (57). À retenir : « The voice agent launch runbook: ramp, rollback, and on-call » (30 sept. 2026).
octobre 2026
AssemblyAI introduced its Dictation API, designed to return finished text (cleaned of fillers, self-corrections, and punctuated) instead of raw transcripts, addressing a gap in dictation workflows. Th…
AssemblyAI released a comprehensive guide to selecting the best speech-to-text API in 2026, emphasizing benchmarking on real-world audio rather than just WER. The post highlights AssemblyAI’s Universa…
AssemblyAI unveiled a new cloud-based voice AI infrastructure platform emphasizing unmatched price-performance, predictable usage-based pricing, and zero infrastructure management. The offering includ…
septembre 2026
AssemblyAI published a detailed runbook for operating its Voice Agent API post-launch, covering staged traffic ramps, rollback thresholds, severity definitions, and on-call ownership. The guidance add…
AssemblyAI published a technical guide distinguishing clinical dictation (short-clip) from ambient scribing, clarifying when to use the Sync API (verbatim transcript) versus the Dictation API (cleaned…
AssemblyAI detailed how its ambient AI scribe technology, originally built for healthcare, adapts to veterinary and legal documentation needs. The post highlights key architectural differences, includ…
AssemblyAI’s blog details reproducible cases where keyterm prompting worsens transcript accuracy by overriding correct words or over-prompting, particularly with similar-sounding terms. It introduces …
AssemblyAI released a guide on evaluating speech-to-text APIs for education platforms, emphasizing student-speech accuracy, long-form audio economics, FERPA compliance, and accessibility requirements.…
AssemblyAI released Universal-3.6 Pro Realtime, a speech-to-text model trained on tens of thousands of hours of real voice-agent and call-center conversations. It reduces errors on short replies by 45…
AssemblyAI released Universal-3.6 Pro Realtime, a major update to their streaming speech-to-text model, reducing word error rates by 16-52% across scenarios like noisy environments, background speech,…

AssemblyAI announced Universal-3.5 Pro, its top-ranked speech-to-text model, is now available on OpenRouter’s transcription endpoint. The integration allows developers to access the model through a si…

AssemblyAI introduced Speaker Diarization for pre-recorded audio, enabling automatic labeling of individual speakers in transcripts. The feature splits transcripts into utterances per speaker with con…
AssemblyAI’s LLM Gateway now lists multiple new models—including Qwen3.5 4B Fast, Claude Sonnet 4.6, and Google’s Gemini 2.5 Flash Lite—with detailed context lengths, supported parameters, and pricing…
AssemblyAI introduced the Dictation API, a new product designed to convert spoken utterances into clean, ready-to-use text by resolving self-corrections, removing filler words, and preserving user-int…

AssemblyAI introduced LLM Gateway, a unified API for accessing 25+ LLMs across providers like Anthropic, Google, OpenAI, and Qwen with automatic fallbacks and regional endpoints (US/EU). The service s…
AssemblyAI now offers official, open-source SDKs for Python and JavaScript covering pre-recorded transcription, real-time streaming, and speech understanding. The SDKs handle authentication, polling, …
AssemblyAI now supports 18 languages in its Universal-3.5 Pro model, including regional dialects like Quebecois French, Mexican Spanish, and Brazilian Portuguese, with automatic dialect recognition. T…
août 2026
AssemblyAI introduced a new feature enabling real-time analysis of streaming audio transcripts via its LLM Gateway API. The update allows developers to integrate 25+ models (Claude, GPT, Gemini) into …
AssemblyAI introduced a hosted Qwen3.5 4B model optimized for voice rewrite tasks in its LLM Gateway, delivering 612ms average response times and 94% lower costs than GPT-4.1. The model is purpose-bui…

AssemblyAI now allows users to preserve filler words like 'uh' and 'um' in transcripts by setting `disfluencies` to `true` in the transcription config. By default, these are removed. The change is ref…
20 premières affichées. Cliquez un mois ci-dessus pour l'archive complète de AssemblyAI.
Suivez AssemblyAI en pilote automatique