支持无障碍沟通与社区参与
截至 2026-10-09,AI 把这项工作做到 L3 有条件自动化。14 条更新涉及它,最强的证据:仅来自发布方。
- 级别
- L3有条件自动化
- 更新
- 14
- 公司
- 5Google · Microsoft · Alibaba
- 最强证据
- T3仅来自发布方
是什么把它推到这里
涉及这项工作的全部更新,最新的在前。
- 2026-10-06GoogleGemini Live adds Guided Vision real-time visual assistance
Gemini Live introduces Guided Vision, allowing users to share their camera for real-time audio descriptions and verbal reframing cues, built with the blind and low-vision community.
- 2026-10-01MicrosoftMicrosoft releases MAI-Transcribe-2-Streaming
Microsoft introduces MAI-Transcribe-2-Streaming for accurate streaming transcription.
- 2026-09-23GoogleGoogle launches Gemini 3.8 Flash TTS and Flash-Lite TTS
Google DeepMind introduces Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, text-to-speech models supporting custom voices across 100+ languages and 2,000+ ready-to-use voices.
- 2026-09-23AlibabaQwen-Audio-3.1 released with upgraded ASR, TTS and Realtime plus TTS-Next and ASR-Next
Alibaba released Qwen-Audio-3.1, upgrading ASR, TTS and Realtime, and adding TTS-Next and ASR-Next, with price cuts across the audio stack.
- 2026-09-19AlibabaAlibaba releases Qwen3.8-LiveTranslate real-time interpretation model
Qwen releases Qwen3.8-LiveTranslate, a next-generation real-time simultaneous interpretation model built on an Interleave architecture, reducing average lagging (LAAL) from 2.8s to 2.3s across 60 languages.
- 2026-09-18xAIxAI introduces Grok Voice Transcribe 2.0
xAI announced Grok Voice Transcribe 2.0, describing it as the world's most accurate speech transcription model.
- 2026-09-04GoogleGoogle Translate live translate gets two upgrades
Google introduced two new upgrades to live translate in the Google Translate app, covering 70+ languages.
- 2026-09-03GoogleGemini voice capabilities roll out to Gmail, Docs and Keep
Google Workspace adds conversational voice control across Gmail, Docs and Keep powered by latest Gemini voice capabilities, rolling out now.
- 2026-09-03MicrosoftMicrosoft releases MAI-Transcribe-2 on Microsoft Foundry
Microsoft announces MAI-Transcribe-2, a transcription model claimed to be 10x faster than GPT-Transcribe, now available on Microsoft Foundry.
- 2026-09-01MetaMeta launches Muse Voice Transcribe in public preview
Meta Superintelligence Labs introduces Muse Voice Transcribe, a real-time streaming ASR model with diarization for 20+ speakers, endpointing and multilingual code-switching, now in public preview on Meta Model API.
- 2026-08-27GoogleGemini 3.5 Transcribe supports 85+ languages
Google highlighted that Gemini 3.5 Transcribe can turn speech into text across more than 85 languages and switch between languages within a single transcription.
- 2026-08-26GoogleGoogle introduces Gemini 3.5 Transcribe speech-to-text model
Google released Gemini 3.5 Transcribe, a speech-to-text model supporting 85+ languages that removes filler words and handles self-corrections.
- 2026-08-25GoogleGemini app for macOS adds voice dictation and file summarization
Google adds voice dictation, file summarization and copy rewriting into any window in the Gemini app for macOS.
- 2026-08-24GoogleGemini 3.5 Live Translate enables real-time multilingual broadcast app
Google demonstrated building a real-time multilingual broadcast app using Gemini 3.5 Live Translate, LiveKit, and Google Cloud Run.