AI 导读
Google 发布音频模型 Gemini 3.5 Live Translate,可在 70+ 语言间提供低延迟实时语音翻译,支持单会话多语言输入、自动识别语种、保留说话人语调和降噪。Fishjam 用 MoQ 传输直播视频并调用该模型做翻译,且按需连接,只在观众实际请求某种语言时才连接 Gemini。演示地址为 translate.fishjam.io。
正文
Fishjam 🤝 Gemini
You can now translate livestreams on demand, thanks to the newest Gemini 3.5 Live Translate!🌎
We used Fishjam to stream live video over MoQ (Media over QUIC), and Gemini to provide live translation.
Translations are fully on demand – Fishjam only connects to Gemini when a viewer actually requests a language, so you never waste tokens on translations nobody is listening to. 💪
Try it out yourself: translate.fishjam.io/
Congrats to the Gemini team on this release! 🎉
Video
Our latest audio model, Gemini 3.5 Live Translate, takes real-time speech translation to the next level for developers by delivering low-latency translation across 70+ languages. By processing speech as it streams in near real time, the model enables devs to build low-latency audio experiences with: — Multilingual input: Understands multiple languages in a single session without needing to adjust settings. — Auto-detection: Identifies the spoken language and begins translation instantly. — Native audio processing: Generates more natural-sounding speech that preserves speakers' intonation, pacing, and pitch. — Noise robustness: Filters out ambient noise for clearer conversation in loud environments. Video在 X 查看被引用的帖子
来源:@swmansion · x.com