OpenAI 实测:编程智能体能把老代码提速 60 倍,但科学对错还得靠人盯。
OpenAI 与学术合作伙伴发布实地报告,用 AI 编程智能体改造被忽视的研究软件,部分任务提速最高 60 倍。参与者警告,这些系统“口齿伶俐、有说服力,但错误时自信十足且难以察觉”。人类工作重心从写代码转向耗时耗力的科学正确性验证。报告显示,智能体能高效处理代码现代化,但无法判断研究本身的科学对错。
AI coding agents can modernize research software but can't judge if the science is right
A field report from OpenAI and academic partners shows coding agents can modernize neglected research software, with speedups of up to 60x. But the systems are "eloquent, convincing, and confidently wrong in ways that are easy to miss," participants say. The effort shifts from writing code to the time-consuming work of verifying scientific correctness. The article AI coding agents can modernize research software but can't judge if the science is right appeared first on The Decoder .