《科学》 第391卷 第6790期 · 2026年3月12日 · 中文解读
如何判断AI是否足够聪明以从事科学研究?
How will we know if AI is smart enough to do science? · C. Zhao
约 5 分钟News在小程序里点播,20 到 60 分钟做好
这篇讲什么
本文探讨了如何通过新的基准测试来评估大型语言模型是否具备进行科学发现的能力。
原文开头
New tests gauge whether large language models can use their troves of knowledge to actually make discoveries For years, artificial intelligence (AI) researchers have dreamed of developing tools that could supercharge science by posing novel questions, designing experiments, and perhaps even carrying them out. In recent months, large language models (LLMs) have made discoveries that some AI developers claim have inched us closer to that future. But how do you test whether an AI model can truly do science? For answers, researchers turn to benchmarks: standardized sets of questions or tasks that help assess an AI’s capacities and compare it against other models. But the complexity of science makes judging their aptitude for it especially challenging. …
摘自《科学》(Science)第391卷 第6790期 · 2026年3月12日,C. Zhao。仅引用开头一小段供了解文章,版权归原刊所有,全文请阅读原刊。

微信扫码,在小程序里听完整版
不用登录先听一篇 · 或在微信搜索小程序 晨间信号