跳到正文
@rohanpaul_ai· @rohanpaul_ai · X·· 25 天前AI 评分61
AI 导读

IFM 公布 0.9B、3.7B、7B、32B 和 375B-A23B 的评测数据,称 7B 已能处理 AIME 竞赛数学,而这类任务约 1 年前需要 100B+ 模型。

正文

The small-model numbers are pretty serious too: IFM says the 7B can now handle AIME competition math, a class of work that required 100B+ models roughly 1 year ago, while the 0.9B, 3.7B and 7B results are also reflected in independent Artificial Analysis evaluation.

The 32B ranks among the top dense models below 40B, while the 375B-A23B ranks among the top models below 400B across general, reasoning, coding and agentic evaluations.

IFM also says the 375B-A23B costs roughly as much to run as a 25B dense model, which is where the sparse design starts getting very interesting.

来源:@rohanpaul_ai · x.com