跳到正文
@ArtificialAnlys· @ArtificialAnlys · X·· 20 天前AI 评分43
AI 导读

Artificial Analysis 的 Text to Image 评测体系覆盖 9 项能力与 10 类用例,Grok Imagine Image 2.0 在 Knowledge、Text Rendering 和 Reasoning 三项上最接近类别前沿,其次是 Material 与 Complex Compositions。

正文

Where Grok Imagine Image 2.0 is strongest in Text to Image: closest to the category frontier in Knowledge, Text Rendering and Reasoning

Our Text to Image taxonomy measures 9 capabilities and 10 use cases, each with its own leaderboard. Capabilities are the individual skills that go into making an image.

Grok Imagine Image 2.0 is closest to the category frontier in Knowledge and Text Rendering, followed by Reasoning, Material and Complex Compositions. Against grok-imagine-image-quality it closes the gap to the frontier on every capability, with the largest gains in Knowledge, Text Rendering and Lighting.

➤ Knowledge covers real landmarks, species, and domain facts across science and common sense.

➤ Text Rendering covers long text, small text, symbols, and artistic lettering.

➤ Reasoning covers entity, mathematical, spatial and logical reasoning, concept mixing, and idiom interpretation.

来源:@ArtificialAnlys · x.com