AI 导读
我们正在展示前沿模型在真实心理健康对话中如何持续改进,借助的是 MentalHealthBench。 这个新的开放基准在构建过程中采纳了 80 多位心理健康临床医生的意见。 我们将其公开释出,以便其他研究者可以审查方法、运行自己的评估,并在此基础上继续推进。 https://t.co/VTm5ZgxJbl
正文
We’re demonstrating how frontier models have continued to improve in realistic mental health conversations with MentalHealthBench.
This new open benchmark was built with input from more than 80 mental health clinicians.
We’re releasing it openly so other researchers can examine the methods, run their own evaluations, and build on the work.
https://t.co/VTm5ZgxJbl
来源:@OpenAI · x.com