Abstract
Automated Program Repair is being reorganized around the joint demands of performance with evidence quality, resource limits, and transfer across settings. This technical review examines assigning repair credit at useful granularity while controlling benchmark leakage and patch overfitting. The evidence base combines 1 focal paper with 12 independently retrieved publications whose DOI or publisher records were checked before inclusion. The analysis is organized around execution signals, credit assignment, patch validity, cross-language transfer, and benchmark design. To avoid reading metrics from unlike protocols as commensurate, the review compares problem boundaries, design logic, and conditions of validation. Across the literature, a robust inference is that advances in automated program repair become credible when representation, objective, and evaluation protocol are evaluated together and when uncertainty about distribution shift is reported explicitly. The proposed reading joins method selection to implementation risk, making transfer failures visible, and proposes a research agenda centered on traceable baselines, scenario testing, and reusable evidence.
References
Li, Y., Wang, H., Shang, X., Tang, X., Cao, Y., & Chen, X. (2026). BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models. arXiv preprint arXiv:2605.09134.
Hanna, C., Blot, A., & Petke, J. (2025). Reinforcement learning for mutation operator selection in automated program repair. Automated Software Engineering, 32(2). https://doi.org/10.1007/s10515-025-00501-z
Kumar Karne, V., Noone Srinivas,, Nagaraj Mandaloju,, & Parameshwar Reddy Kothamali, (2020). Reinforcement Learning for Optimizing Test Case Execution in Automated Testing. Innovative Research Thoughts, 6(3), 13-27. https://doi.org/10.36676/irt.v6.i3.1494
Hao, S., Shi, X., Liu, H., Yin, Y., & Chen, X. (2026). Template-guided interpretable reasoning with execution feedback for LLM-based program repair. Information and Software Technology, 193, 108058. https://doi.org/10.1016/j.infsof.2026.108058
Wan, H., Luo, H., Li, M., & Luo, X. (2024). Automated Program Repair for Introductory Programming Assignments. IEEE Transactions on Learning Technologies, 17, 1705-1720. https://doi.org/10.1109/tlt.2024.3403710
Yin, Z., Lin, W., & Kong, X. (2026). Heterogeneous multi-expert collaborative reinforcement learning for automated CAD program synthesis from engineering drawings. Discover Artificial Intelligence. https://doi.org/10.1007/s44163-026-01731-0
Jha, A. C. (2025). Automated Firewall Policy Generation with Reinforcement Learning. International journal of IoT, 5(1), 190-211. https://doi.org/10.55640/ijiot-05-01-10
Gill, S., Goolsby, B. J., & Pawluk, D. T. V. (2023). Kinesthetic Feedback for Understanding Program Execution. Sensors, 23(11), 5159. https://doi.org/10.3390/s23115159
Sukur, N. a., Milošević, N., Pracner, D., & Budimac, Z. (2023). Automated program improvement with reinforcement learning and graph neural networks. Soft Computing, 28(3), 2593-2604. https://doi.org/10.1007/s00500-023-08559-1
Xu, W., Kassim, M. S. S., & Mahmud, R. (2026). Enhancing IELTS writing automated scoring with M-LoRA fine-tuned LLAMA-3 and human feedback-driven PPO reinforcement learning. Scientific Reports, 16(1). https://doi.org/10.1038/s41598-026-43318-w
Ali, R. H., & Chyad, A. M. (2026). Automated Bug Detection and Program Repair Using Deep Learning: A Comprehensive Review. Cybernetics and Information Technologies, 26(1), 93-121. https://doi.org/10.2478/cait-2026-0006
Farrugia, F. (2026). Deep Reinforcement Learning for Automated Liquidity Management in Concentrated Automated Market Makers: A Multi-Architecture Empirical Study. Journal of Advances in Civil and Mechanical Engineering, 03(01), 01-06. https://doi.org/10.64030/3067-2457.03.01.02
Gangopadhyay, B., Soora, H., & Dasgupta, P. (2022). Hierarchical Program-Triggered Reinforcement Learning Agents for Automated Driving. IEEE Transactions on Intelligent Transportation Systems, 23(8), 10902-10911. https://doi.org/10.1109/tits.2021.3096998
