Automated Program Repair: Emerging Trends Pose and Expose Problems for Benchmarks

Renzullo, Joseph; Reiter, Pemma; Weimer, Westley; Forrest, Stephanie

doi:10.1145/3704997

Citation Details

This content will become publicly available on March 22, 2026

Automated Program Repair: Emerging Trends Pose and Expose Problems for Benchmarks

Machine learning (ML) pervades the field of Automated Program Repair (APR). Algorithms deploy neural machine translation and large language models (LLMs) to generate software patches, among other tasks. But, there are important differences between these applications of ML and earlier work, which complicates the task of ensuring that results are valid and likely to generalize. A challenge is that the most popular APR evaluation benchmarks were not designed with ML techniques in mind. This is especially true for LLMs, whose large and often poorly-disclosed training datasets may include problems on which they are evaluated. This article reviews work in APR published in the field’s top five venues since 2018, emphasizing emerging trends in the field, including the dramatic rise of ML models, including LLMs. ML-based articles are categorized along structural and functional dimensions, and a variety of issues are identified that these new methods raise. Importantly, data leakage and contamination concerns arise from the challenge of validating ML-based APR using existing benchmarks, which were designed before these techniques were popular. We discuss inconsistencies in evaluation design and performance reporting and offer pointers to solutions where they are available. Finally, we highlight promising new directions that the field is already taking. more »

Award ID(s):: 2211750 2211749

PAR ID:: 10634726

Author(s) / Creator(s):: Renzullo, Joseph; Reiter, Pemma; Weimer, Westley; Forrest, Stephanie

Publisher / Repository:: ACM

Date Published:: 2025-03-22

Journal Name:: ACM Computing Surveys

Volume:: 57

Issue:: 8

ISSN:: 0360-0300

Page Range / eLocation ID:: 1 to 18

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
This content will become publicly available on March 22, 2026
Journal Article:
https://doi.org/10.1145/3704997

More Like this