
Survey Finds AI-Generated Code Increases Debugging and Failure Rates and Creates a Comprehension Gap
Quick Answer
A survey by Coleman Parkes for Undo reveals that while AI coding agents expedite code generation, they increase debugging time and create comprehension gaps, with 35% of generated code reaching production without full understanding.
Quick Take
Engineers now spend 42% of their week debugging, leading to significant production issues and a slower overall release cycle.
Key Points
- Teams spend 9.8 hours coding but 16.9 hours debugging weekly.
- 80% of engineers report AI struggles with complex codebases.
- 93% experienced incorrect root-cause diagnosis due to AI hallucinations.
- 79% of leaders say AI speeds up code generation but not release cycles.
- One-third of teams use AI only for simple codebases.
DeepSignal Analysis
What happened
A survey by Coleman Parkes for Undo indicates that while AI coding agents speed up code generation, they also increase debugging time and create comprehension gaps. Engineers reportedly spend 42% of their week debugging, with 35% of AI-generated code reaching production without full understanding. The survey highlights significant issues, including production incidents and incorrect root-cause diagnoses due to AI limitations.
Key evidence
- The survey found that teams spend an average of 16.9 hours per week debugging, compared to 9.8 hours spent on code production.
- Approximately 80% of respondents indicated that AI coding agents struggle with complex codebases, limiting their effectiveness.
- 81% of surveyed teams experienced production incidents or service outages affecting users in the past six months.
Why it matters
The findings raise concerns about the reliability of AI-generated code in mission-critical environments. As engineers increasingly rely on AI for code generation, the lack of understanding of the generated code could lead to significant production issues. The shift in focus from coding to debugging may hinder overall productivity and extend release cycles, challenging the perceived benefits of AI in software development.
📖 Reader Mode
~3 min readA survey conducted by independent research firm Coleman Parkes on behalf of Undo, a company focused on scaling AI-powered root-cause analysis, found that while AI coding agents have accelerated code generation, they have shifted the primary bottleneck to debugging, code comprehension, and maintenance.
The report surveyed 300 senior engineer leaders responsible for delivering mission-critical software, most of whom working with C/C++. The survey focuses specifically on mission-critical codebases where code "must be understood", with respondent identifying the most demanding task as "understanding what that code does, how it affects existing codebases, and debugging it when an application doesn’t behave the way it’s expected to".
In those environments, teams spend an average of 9.8 hours per week producing code, but 16.9 hours per week debugging issues identified during development or encountered by customers in production, accounting for 42% of the average working week.
Along with debugging getting more relevant, another critical dimension is emerging as a challenge, code comprehension:
Now that AI is generating most of the code being produced, engineers no longer have the inherent understanding they used to. That makes it easier for defects to escape, and when something inevitably goes wrong, nobody has the knowledge to trace the failure back to its root cause.
Due to the acceleration in code generation brought by AI agents, the survey found 35% of generated code reaches production before the team has fully understood it. Moreover, 80% of respondents said that coding agents struggle to solve difficult problems in complex codebases. As a result, approximately one-third of teams "use AI agents for comprehension and debugging only in straightforward codebases", while relying on additional techniques to build sufficient confidence when working with more complex systems.
Other significant problems reported by surveyed teams include production incident or service outage affecting internal users or customers (81% experienced this at least once in the previous six months, with 14% experiencing them multiple times per month), incorrect root-cause or issue diagnosing due to hallucination (93% at least once, with 18% multiple times per month), and test escapes, serious defects or poorly optimized code entering production (91% at least once, with 8% experiencing these issues multiple times per month).
Overall, 79% of engineering leaders say that AI agents can generate code significantly faster, but that the resulting shift in effort toward debugging and "unpicking" AI-generated code means the overall release cycle is "no faster than before".
Greg Law, founder and CEO of Undo, summarized the survey findings by noting that engineers "lose days trying to unravel what went wrong and why" with "code that's almost, but not quite right" and that "while agents are great at writing reams of code quickly, they're less capable at debugging it".
Since the launch and widespread adoption of AI coding agents, the software engineering community has extensively debated their benefits, limitations, and how best to use them. One recurring concern is how to manage agent speed when it outpaces humans' ability to review the generated code. This has led to several popular approaches, including the test-first red-green loop, and automated fallbacks, and others. The broader consensus, however, is that AI agents do not fundamentally change the nature of software engineering, which has never been solely about coding, but about understanding constraints, making trade-offs, and ensuring that the resulting system behaves as intended.
About the Author
Sergio De Simone
Show moreShow less
— Originally published at infoq.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from InfoQ AI, ML & Data Engineering
See more →Google Cloud Workbench Notebooks Extension Connects VS Code to Google Cloud's Jupyter Notebooks
The Google Cloud Workbench Notebooks extension for VS Code allows developers to seamlessly connect their local IDE to managed Jupyter notebook environments on Google Cloud, enhancing ML workflow efficiency. This integration eliminates context switching, enabling smooth transitions from local experimentation to high-performance cloud computing.

