A Study of the Reliability of Agentic AI-Generated Programs
A study found that AI-generated code is often as reliable, and sometimes more reliable, than human-written code for certain utility programs. The study used fuzz testing and found that AI-generated code had fewer memory errors but more hangs. The results suggest that agentic AI can be a cost-effective way to generate sustainable software, but requires careful practice and human supervision.
Save an API key to vote.