AI coding agents can modernize research software but can't judge if the science is right

AI Summary
AI coding agents have demonstrated the ability to modernize outdated research software, achieving speed improvements of up to 60 times. However, they lack the capability to assess scientific accuracy, leading to a focus on verifying correctness rather than just coding.
From the source
A field report from OpenAI and academic partners shows coding agents can modernize neglected research software, with speedups of up to 60x. But the systems are "eloquent, convincing, and confidently wrong in ways that are easy to miss," participants say. The effort shifts from writing code to the time-consuming work of verifying scientific correctness. The article AI coding agents can modernize research software but can't judge if the science is right appeared first on The Decoder.
The full text couldn't be loaded here (the source may require a subscription).
View original at The Decoder