跳到正文
@GoogleDeepMind· @GoogleDeepMind · X·· 14 天前AI 评分61
AI 导读

Google DeepMind 宣布 Gemini API 支持逐行调节语音,可控制语速、情绪以及笑声、停顿等提示,所有生成音频都带 SynthID 水印以便识别为 AI 生成。开发者可通过 Google AI Studio 使用 Gemini API 开始构建。

正文

Fine-tune the delivery line by line, shaping pacing, emotion, and cues like laughs or pauses.

All generated audio is watermarked with SynthID so it can be reliably identified as AI-generated.

Start building with the Gemini API via @GoogleAIStudio. Find out more → https://t.co/F93tMO9Fxt

来源:@GoogleDeepMind · x.com