跳到正文
@ArtificialAnlys· @ArtificialAnlys · X·· 23 天前AI 评分28
AI 导读

在我们的语音推理基准 Big Bench Audio 上,Gemini 3.8 Live Extended Thinking (High) 得分 97.7%,领先于 Grok Voice Think Fast 2.0 High(97.2%),落后于 Qwen Audio 3.0 Realtime Plus 的 99.2%。标准版 Gemini 3.8 Live 得分为 91.7%。https://t.co/aatRk9efWw

正文

On Big Bench Audio, our speech reasoning benchmark, Gemini 3.8 Live Extended Thinking (High) scores 97.7%, ahead of Grok Voice Think Fast 2.0 High (97.2%) and behind Qwen Audio 3.0 Realtime Plus at 99.2%. The standard Gemini 3.8 Live scores 91.7%. https://t.co/aatRk9efWw

来源:@ArtificialAnlys · x.com