跳到正文
@elonmusk· @elonmusk · X·· 16 天前AI 评分55
AI 导读

Elon Musk 公布 Grok 4.7,引用 Artificial Analysis 数据显示其在 AA-Briefcase 上仅次于 Anthropic 模型、落后 Opus 5,每任务成本约为后者的 50%。在 AA-Briefcase-Lite 上,Grok 4.7 的分析质量 Elo 从 Grok 4.6 的 1698 升至 1994,呈现质量 Elo 从 1531 小幅降至 1499。生成示例 deck 的 API 成本方面,Grok 4.7(xhigh)约 8 美元,Grok 4.6(xhigh)约 4.40 美元。

正文

Grok 4.7 https://t.co/hQVyaUJTPj

引用@ArtificialAnlys@ArtificialAnlys
Grok 4.7 is behind only Anthropic models on AA-Briefcase, ranking just behind Opus 5 at ~50% of its Cost per Task Grok 4.7’s improvements over Grok 4.6 are clear in AA-Briefcase-Lite, our public due diligence scenario where models are tasked with building market models and target assessment decks. Grok 4.7 gains significantly in Analytical Quality Elo (1698 → 1994) with a slight regression in Presentation Elo (1531 → 1499). API cost to produce example decks: Grok 4.7 (xhigh) ~$8 vs. Grok 4.6 (xhigh) ~$4.40
在 X 查看被引用的帖子

来源:@elonmusk · x.com