SOTAVerified

Program Repair

Task of teaching ML models to modify an existing program to fix a bug in a given code.

Papers

Showing 101–132 of 132 papers

TitleStatusHype
An Exploratory Literature Study on Sharing and Energy Use of Language Models for Source Code—0
To Err is Machine: Vulnerability Detection Challenges LLM Reasoning—0
A Multi-Dataset Evaluation of Models for Automated Vulnerability Repair—0
Repairing Bugs in Python Assignments Using Large Language Models—0
Repair Is Nearly Generation: Multilingual Program Repair with LLMs—0
Agentic Bug Reproduction for Effective Automated Program Repair at Google—0
Revisiting the Plastic Surgery Hypothesis via Large Language Models—0
Using ML filters to help automated vulnerability repairs: when it helps and when it doesn't—0
RunBugRun -- An Executable Dataset for Automated Program Repair—0
SampleFix: Learning to Generate Functionally Diverse Fixes—0
SCELMo: Source Code Embeddings from Language Models—0
Where's the Bug? Attention Probing for Scalable Fault Localization—0
SemAgent: A Semantics Aware Program Repair Agent—0
CORE: Benchmarking LLMs Code Reasoning Capabilities through Static Analysis Tasks—0
Counterexample Guided Program Repair Using Zero-Shot Learning and MaxSAT-based Fault Localization—0
Semantic-guided Search for Efficient Program Repair with Large Language Models—0
Conversational Automated Program Repair—0
DeepCode AI Fix: Fixing Security Vulnerabilities with Large Language Models—0
DeepDebug: Fixing Python Bugs Using Stack Traces, Backtranslation, and Code Skeletons—0
AdaptivePaste: Code Adaptation through Learning Semantics-aware Variable Usage Representations—0
RAP-Gen: Retrieval-Augmented Patch Generation with CodeT5 for Automatic Program Repair—0
Detect-Localize-Repair: A Unified Framework for Learning to Debug with CodeT5—0
Dissecting the SWE-Bench Leaderboards: Profiling Submitters and Architectures of LLM- and Agent-Based Repair Systems—0
SmartPaste: Learning to Adapt Source Code—0
Dynamic Neural Program Embeddings for Program Repair—0
Enabling Automatic Repair of Source Code Vulnerabilities Using Data-Driven Methods—0
ENCORE: Ensemble Learning using Convolution Neural Machine Translation for Automatic Program Repair—0
ConDefects: A New Dataset to Address the Data Leakage Concern for LLM-based Fault Localization and Program Repair—0
Enhancing Automated Program Repair with Solution Design—0
Evaluating Agent-based Program Repair at Google—0
SWE-Synth: Synthesizing Verifiable Bug-Fix Data to Enable Large Language Models in Resolving Real-World Bugs—0
Evaluating the Generalizability of LLMs in Automated Program Repair—0
Show:102550
← PrevPage 3 of 3Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1DrRepair + BIFIAverage Success Rate71.7—Unverified
2DrRepairAverage Success Rate68.2—Unverified
3SampleFixAverage Success Rate45.3—Unverified
4RLAssistAverage Success Rate26.6—Unverified
#ModelMetricClaimedVerifiedStatus
1Transformer + BIFIAccuracy (%)90.5—Unverified
2TransformerAccuracy (%)62—Unverified
#ModelMetricClaimedVerifiedStatus
1MGDebugger (DeepSeek-Coder-V2-Lite)Pass@197.6—Unverified
#ModelMetricClaimedVerifiedStatus
1TFixError Removal678—Unverified