Spyingbee 的存档中收录了 Soniox 的 10 个受追踪页面,均于 2026年8月 首次记录。这些页面没有发布日期,显示的是我们首次抓取的时间。 主要主题是 feature updates(5)。 亮点:“Compare text-to-speech APIson your own text”(2026年8月30日)。
浏览 Soniox 的其他月份
Soniox introduced a live demo and pricing calculator to compare its real-time speech translation API against competitors like OpenAI. The tool highlights Soniox's unified pricing model ($180 per 1,000…
Soniox introduced a live comparison tool for text-to-speech APIs, allowing users to test Soniox, OpenAI, ElevenLabs, Google, Cartesia, and Azure on the same text in real time. The tool emphasizes Soni…

Soniox introduced a live demo and pricing calculator to compare speech-to-text APIs side-by-side using real audio, emphasizing accuracy, cost, and production-ready features like diarization and multil…

Soniox launched a data residency feature allowing customers to select regions where their audio, transcripts, and metadata are processed and stored. System data (account metadata, usage stats, billing…
Soniox published a side-by-side comparison of its real-time translation capabilities against Google's Gemini 3.5 Live Translate, emphasizing cost savings (92% cheaper at 1,000 hours/month), any-to-any…
Soniox introduced a real-time speech translation comparison tool pitting its service against OpenAI’s GPT Realtime Translate, highlighting cost savings (91% cheaper at 1,000 hours/month) and broader l…
Soniox’s Speech-to-Text AI now automatically identifies spoken languages in audio streams, including mixed-language conversations, without requiring users to pre-specify languages. The feature tags ea…
Soniox introduced speaker diarization in its Speech-to-Text AI, enabling automatic detection and separation of speakers in multi-speaker audio streams. The feature generates speaker-labeled transcript…
Soniox introduced semantic endpoint detection for real-time speech-to-text, using pauses, intonation, and context to finalize utterances earlier than silence-based methods. The feature reduces latency…
Soniox released TTS v2, introducing programmable audio tags for expressive control, support for over 60 languages with natural mixing, high-quality voice cloning from seconds of audio, and low-latency…