🚀NEW LABGetting Started with Claude AgentsStart lab
Code · Agents

LLM-based Agents for Automated Bug Fixing

First page
LLM-based Agents for Automated Bug Fixing
Paper summary

analyzes seven leading LLM-based bug fixing systems on the SWE-bench Lite benchmark, finding MarsCode Agent (developed by ByteDance) achieved the highest success rate at 39.33%; reveals that for error localization line-level fault localization accuracy is more critical than file-level accuracy, and bug reproduction capabilities significantly impact fixing success; shows that 24/168 resolved issues could only be solved using reproduction techniques, though reproduction sometimes misled LLMs when issue descriptions were already clear; concludes that improvements are needed in both LLM reasoning capabilities and Agent workflow design to enhance automated bug fixing effectiveness.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack