Artificial Analysis 欢迎 Dario 呼吁让独立评估者获得更多前沿 AI 模型访问权限。该机构近三年来持续构建基准与基础设施,独立衡量 AI 能力,并已支持几乎所有主要 AI 实验室的前沿模型发布前基准测试。其表示在所测量的每个维度上仍持续看到能力跃升,并将继续推动独立 AI 评估生态建设。
Independent evaluation of AI models, across both capability and safety, is essential. We welcome Dario’s call to give independent evaluators greater access to frontier AI models.
For nearly three years, Artificial Analysis has been building benchmarks and infrastructure to independently measure AI capabilities. We have supported pre-launch benchmarking of frontier models with almost every major AI lab. We are continuing to see jumps in capability across every dimension we measure.
The world needs a vibrant ecosystem of independent AI evaluators. We are going to keep working to build it!
来源:@ArtificialAnlys · x.com