AI 导读
Google DeepMind 宣布 Gemini API 支持逐行调节语音,可控制语速、情绪以及笑声、停顿等提示,所有生成音频都带 SynthID 水印以便识别为 AI 生成。开发者可通过 Google AI Studio 使用 Gemini API 开始构建。
正文
Fine-tune the delivery line by line, shaping pacing, emotion, and cues like laughs or pauses.
All generated audio is watermarked with SynthID so it can be reliably identified as AI-generated.
Start building with the Gemini API via @GoogleAIStudio. Find out more → https://t.co/F93tMO9Fxt
来源:@GoogleDeepMind · x.com