[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"doc-detail-86449-en":3,"doc-seo-86449-105":30,"detail-sidebar-cat-0-en-105":95},{"code":4,"msg":5,"data":6},0,"success",{"doc_id":7,"user_id":8,"nickname":9,"user_avatar":10,"doc_module":4,"category_id":11,"category_name":12,"doc_title":13,"doc_description":14,"doc_content":15,"file_id":16,"file_url":17,"file_type":18,"file_size":19,"view_count":20,"is_deleted":4,"is_public":21,"is_downloadable":21,"audit_status":21,"page_count":22,"language":23,"language_code":24,"site_id":25,"html_lang":24,"table_of_contents":26,"faqs":27,"seo_title":13,"seo_description":14,"update_tm":28,"read_time":29},86449,2336464648322,"Aria","https://ap-avatar.wpscdn.com/avatar/2200025388227c56fec?_k=1778556882303663488",8,"Research & Report","Do These Violent Delights Have Violent Ends? Measuring the Post-Merge Fate of Agentic Code","Agentic coding tools increasingly create autonomous, repository-level changes, yet their post-merge lifecycle remains largely unknown. A longitudinal empirical study analyzes agentic versus human contributions across 182 repositories, tracking post-merge outcomes over time, characterizing subsequent modification intent, and measuring defects and vulnerabilities. Overall maintenance rates are similar, but agentic contributions demand higher corrective maintenance and introduce more security weaknesses and dependency vulnerabilities. Maintenance burden also correlates with repository characteristics, including a higher no-review rate leading to increased agentic burden.","Do These Violent Delights Have Violent Ends? Measuring the Post-Merge Fate of Agentic Code  \nChunqiu Steven Xia University of Illinois Urbana-Champaign  \nUrbana, IL, USA  \n[chunqiu2@illinois.edu](chunqiu2@illinois.edu)  \nCourtney Miller  \nGeorge Washington University Washington, DC, USA [courtney.miller@gwu.edu](courtney.miller@gwu.edu)  \narXiv :2607 .09902v 1 [ cs . SE] 10 Jul 2026  \nAbstract—Agentic coding tools are increasingly used to make autonomous repository-level changes to real-world projects. Prior work has largely evaluated these contributions at the premerge stage, through outcomes such as pull request acceptance and review effort. Far less is known about what happens to agentic code post-merge. Yet merge success alone does not reveal whether a contribution will remain stable or require bug fixes and other corrective maintenance downstream. We conduct a longitudinal empirical analysis of agentic and human contributions across 182 repositories, tracking their post-merge fate overtime, characterizing the intent of subsequent modifications, and analyzing the defects and vulnerabilities they introduce. While the overall maintenance rates are similar, agentic contributions require significantly higher rates of corrective maintenance and introduce more security weaknesses and dependency vulnerabilities. We also find statistically significant evidence that agentic maintenance burden is associated with repository characteristics. In particular, each 10 percentage-point increase in a project’sno-review rate is associated with roughly a 6% increase in agentic maintenance burden on average. As coding agents become pervasive in software development, our findings highlight the need to evaluate and design agentic tools not only to produce mergeable changes, but to produce contributions that remain secure and maintainable.  \nI. INTRODUCTION  \nThe case for adopting Generative AI (GenAI) agentic coding tools is made almost entirely based on tool-centric metrics: how much the tools produce and what their output looks like at merge. Microsoft reports GenAI produces close to 30% of its codebase [1], Meta aims for agentic coding tools to handle half of its development by late 2026 [2], and Google reports that 75% of their new code is produced by GenAI [3] . In a recent survey of more than 200 technology decision-makers, 67% claim that agentic coding tools write over half of their organization’s weekly code [4] . These are the numbers used to justify adoption. They say little about that code’s sustainability, reliability, security, or maintainability in the long term postmerge. What happens to agentic code once it becomes part of the software system remains largely unstudied.  \nThis tool-centric framing has revived evaluation metrics that are weak proxies for the quality of what actually ships: lines of code written, number of merged pull requests, percentage of code written by GenAI, and even “tokenmaxxing\" (amount of vendor tokens spent) [5] . Such metrics make for impressive growth graphs, but they reveal little about how that change behaves as it evolves with the project.  \n\n| Cumulative Incidence | 5%\u003Cbr>4%\u003Cbr>3%\u003Cbr>2%\u003Cbr>1%\u003Cbr>0% | 4.6% ~~ ~~\u003Cbr>4.0% \u003Cbr>3.5%  \u003Cbr> 3.3% \u003Cbr>2.6% 2.9%  \u003Cbr>\u003Cbr>2. 1%\u003Cbr>1.6%  Agentic\u003Cbr>Human\u003Cbr>\u003Cbr> |\n| --- | --- | --- |\n|  |  | \u003Cbr>0 100 200 300\u003Cbr>Days Since Merge |\n\nFig. 1: CIF for corrective terminations. Dots represent time points of 10/30/90/180 days since merge.  \nIn fields like medicine and aviation, performance claims about tools that carry safety risks are expected to be substantiated before the tool is relied on. Medical devices like surgical robots must clear thorough regulatory review processes for safety and effectiveness before they can get FDA approval for use. Agentic coding tools face no comparable bar before bold sweeping promises are made and they are deployed to write code for banks, hospitals, and other critical systems. And whatever scrutiny the output receives largely occurs","cbCairn2fEhwHfJQ","https://ap.wps.com/l/cbCairn2fEhwHfJQ","pdf",2477665,3,1,13,"English","en",105,"# Introduction\n## Research gap and motivation\n## Limitations of tool-centric evaluation\n## Study overview and research questions\n## Data and methodology","[{\"question\":\"What gap does the study address about agentic coding tools?\",\"answer\":\"The study targets the lack of evidence on what happens to agentic code after merge, beyond pre-merge outcomes such as pull request acceptance and review effort.\"},{\"question\":\"How do post-merge maintenance rates compare between agentic and human contributions?\",\"answer\":\"Overall maintenance rates are similar, but agentic contributions require significantly higher corrective maintenance than human code.\"},{\"question\":\"What security-related findings are reported for agentic contributions?\",\"answer\":\"Agentic contributions introduce more security weaknesses and dependency vulnerabilities, indicating increased downstream risks even when merge succeeds.\"},{\"question\":\"Which repository characteristic is associated with higher agentic maintenance burden?\",\"answer\":\"A project’s no-review rate is linked to greater agentic maintenance burden; each 10 percentage-point increase in no-review rate corresponds to roughly a 6% increase on average.\"}]",1784211807,33,{"code":4,"msg":31,"data":32},"ok",{"site_id":25,"language":24,"slug":33,"title":13,"keywords":34,"description":14,"schema_data":35,"social_meta":90,"head_meta":92,"extra_data":94,"updated_unix":28},"do-these-violent-delights-have-violent-ends-measuring-the-post-merge-fate-of-agentic-code","",{"@graph":36,"@context":89},[37,53,68],{"@type":38,"itemListElement":39},"BreadcrumbList",[40,44,48,50],{"item":41,"name":42,"@type":43,"position":21},"https://docshare.wps.com","Home","ListItem",{"item":45,"name":46,"@type":43,"position":47},"https://docshare.wps.com/document/","Document",2,{"item":49,"name":12,"@type":43,"position":20},"https://docshare.wps.com/document/research-report/",{"item":51,"name":13,"@type":43,"position":52},"https://docshare.wps.com/document/do-these-violent-delights-have-violent-ends-measuring-the-post-merge-fate-of-agentic-code/86449/",4,{"url":51,"name":13,"@type":54,"author":55,"headline":13,"publisher":57,"fileFormat":60,"inLanguage":24,"description":14,"dateModified":61,"datePublished":62,"encodingFormat":60,"isAccessibleForFree":63,"interactionStatistic":64},"DigitalDocument",{"name":9,"@type":56},"Person",{"url":41,"name":58,"@type":59},"DocShare","Organization","application/pdf","2026-07-28","2026-07-16",true,{"@type":65,"interactionType":66,"userInteractionCount":20},"InteractionCounter",{"@type":67},"ViewAction",{"@type":69,"mainEntity":70},"FAQPage",[71,77,81,85],{"name":72,"@type":73,"acceptedAnswer":74},"What gap does the study address about agentic coding tools?","Question",{"text":75,"@type":76},"The study targets the lack of evidence on what happens to agentic code after merge, beyond pre-merge outcomes such as pull request acceptance and review effort.","Answer",{"name":78,"@type":73,"acceptedAnswer":79},"How do post-merge maintenance rates compare between agentic and human contributions?",{"text":80,"@type":76},"Overall maintenance rates are similar, but agentic contributions require significantly higher corrective maintenance than human code.",{"name":82,"@type":73,"acceptedAnswer":83},"What security-related findings are reported for agentic contributions?",{"text":84,"@type":76},"Agentic contributions introduce more security weaknesses and dependency vulnerabilities, indicating increased downstream risks even when merge succeeds.",{"name":86,"@type":73,"acceptedAnswer":87},"Which repository characteristic is associated with higher agentic maintenance burden?",{"text":88,"@type":76},"A project’s no-review rate is linked to greater agentic maintenance burden; each 10 percentage-point increase in no-review rate corresponds to roughly a 6% increase on average.","https://schema.org",{"og:url":51,"og:type":91,"og:title":13,"og:site_name":58,"og:description":14},"article",{"robots":93,"canonical":51},"index,follow",{"doc_id":7,"site_id":25},{"code":4,"msg":5,"data":96},[97,101,105,109,114,119,124,127,132,135,139],{"id":21,"doc_module":4,"doc_module_name":46,"category_name":98,"show_sort_weight":99,"slug":100},"Story & Novel",90,"story-novel",{"id":47,"doc_module":4,"doc_module_name":46,"category_name":102,"show_sort_weight":103,"slug":104},"Literature",80,"literature",{"id":52,"doc_module":4,"doc_module_name":46,"category_name":106,"show_sort_weight":107,"slug":108},"Exam",70,"exam",{"id":110,"doc_module":4,"doc_module_name":46,"category_name":111,"show_sort_weight":112,"slug":113},5,"Comic",60,"comic",{"id":115,"doc_module":4,"doc_module_name":46,"category_name":116,"show_sort_weight":117,"slug":118},6,"Technology",50,"technology",{"id":120,"doc_module":4,"doc_module_name":46,"category_name":121,"show_sort_weight":122,"slug":123},7,"Healthcare",40,"healthcare",{"id":11,"doc_module":4,"doc_module_name":46,"category_name":12,"show_sort_weight":125,"slug":126},30,"research-report",{"id":128,"doc_module":4,"doc_module_name":46,"category_name":129,"show_sort_weight":130,"slug":131},9,"Religion & Spirituality",20,"religion-spirituality",{"id":130,"doc_module":4,"doc_module_name":46,"category_name":133,"show_sort_weight":130,"slug":134},"World Cup","world-cup",{"id":136,"doc_module":4,"doc_module_name":46,"category_name":137,"show_sort_weight":136,"slug":138},10,"Lifestyle","lifestyle",{"id":140,"doc_module":4,"doc_module_name":46,"category_name":141,"show_sort_weight":110,"slug":142},19,"General","general"]