跳到正文
@GoogleAI· @GoogleAI · X·· 2026-08-27AI 评分59
AI 导读

Google 发布 Gemini 3.5 Transcribe 转录模型,支持 85+ 语言的上下文感知语音转文本。该模型会自动过滤填充词、格式化非结构化语音,并能结合屏幕上下文执行语音指令。官方演示中,它借助多模态能力把杂乱的语音输入和本地文件整理成邮件草稿。

正文

Today we’re introducing Gemini 3.5 Transcribe, our latest transcription model built for incredibly precise, smart dictation across your favorite apps and devices.

Remember when traditional speech-to-text meant shouting over background noise, constantly hitting backspace to fix misspelled words, and manually deleting every "um" and "uh"? Those days are over.

Gemini 3.5 Transcribe isn't just dictation — it’s active intelligence with precise, context-aware speech-to-text support in 85+ languages. The model automatically filters out filler words, formats unstructured speech, and even pairs with your screen context to execute voice commands.

Watch as Gemini 3.5 Transcribe removes filler words and uses multimodal capabilities to seamlessly turn messy voice input and local files into a polished email draft.

来源:@GoogleAI · x.com