SOTAVerified

Code Completion

Papers

Showing 51100 of 212 papers

TitleStatusHype
Scope is all you need: Transforming LLMs for HPC CodeCode1
RepoBench: Benchmarking Repository-Level Code Auto-Completion SystemsCode1
How Effective Are Neural Networks for Fixing Security VulnerabilitiesCode1
MPI-rical: Data-Driven MPI Distributed Parallelism Assistance with TransformersCode1
LLMSecEval: A Dataset of Natural Language Prompts for Security EvaluationsCode1
Learning Deep Semantics for Test CompletionCode1
Execution-based Code Generation using Deep Reinforcement LearningCode1
CoCoMIC: Code Completion By Jointly Modeling In-file and Cross-file ContextCode1
Multi-lingual Evaluation of Code Generation ModelsCode1
Reading Between the Lines: Modeling User Behavior and Costs in AI-Assisted ProgrammingCode1
MetaTPTrans: A Meta Learning Approach for Multilingual Code Representation LearningCode1
Productivity Assessment of Neural Code CompletionCode1
ReACC: A Retrieval-Augmented Code Completion FrameworkCode1
UniXcoder: Unified Cross-Modal Pre-training for Code RepresentationCode1
CodeFill: Multi-token Code Completion by Jointly Learning from Structure and Naming SequencesCode1
A Syntax-Guided Edit Decoder for Neural Program RepairCode1
Energy-Based Models for Code Generation under Compilability ConstraintsCode1
CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and GenerationCode1
A Simple Approach for Handling Out-of-Vocabulary Identifiers in Deep Learning for Source CodeCode1
Empirical Study of Transformers for Source CodeCode1
LambdaNet: Probabilistic Type Inference using Graph Neural NetworksCode1
Adversarial Robustness for CodeCode1
Beyond Autocomplete: Designing CopilotLens Towards Transparent and Explainable AI Coding Agents0
Plan for Speed -- Dilated Scheduling for Masked Diffusion Language Models0
HiLDe: Intentional Code Generation via Human-in-the-Loop Decoding0
Alignment-Augmented Speculative Decoding with Alignment Sampling and Conditional Verification0
Structure-Aware Corpus Construction and User-Perception-Aligned Metrics for Large-Language-Model Code Completion0
Can You Really Trust Code Copilots? Evaluating Large Language Models from a Code Security PerspectiveCode0
CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts0
Procedural Memory Is Not All You Need: Bridging Cognitive Gaps in LLM-Based Agents0
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks0
NoEsis: Differentially Private Knowledge Transfer in Modular LLM Adaptation0
EduBot -- Can LLMs Solve Personalized Learning and Programming Assignments?0
RTLRepoCoder: Repository-Level RTL Code Completion through the Combination of Fine-Tuning and Retrieval Augmentation0
CCCI: Code Completion with Contextual Information for Complex Data Transfer Tasks Using Large Language Models0
ObscuraCoder: Powering Efficient Code LM Pre-Training Via Obfuscation GroundingCode0
On Explaining (Large) Language Models For Code Using Global Code-Based Explanations0
Enhancing High-Quality Code Generation in Large Language Models with Comparative Prefix-TuningCode0
GenAI for Simulation Model in Model-Based Systems Engineering0
FEA-Bench: A Benchmark for Evaluating Repository-Level Code Generation for Feature Implementation0
SolBench: A Dataset and Benchmark for Evaluating Functional Correctness in Solidity Code Completion and Repair0
Alchemist: Towards the Design of Efficient Online Continual Learning System0
Automated Code Generation and Validation for Software Components of Microcontrollers0
Comparative Analysis of Large Language Models for Context-Aware Code Completion using SAFIM Framework0
Mechanistic Understanding of Language Models in Syntactic Code Completion0
MathFimer: Enhancing Mathematical Reasoning by Expanding Reasoning Steps through Fill-in-the-Middle Task0
GREEN-CODE: Learning to Optimize Energy Efficiency in LLM-based Code GenerationCode0
Improving FIM Code Completions via Context & Curriculum Based Learning0
ExecRepoBench: Multi-level Executable Code Completion Evaluation0
ContextModule: Improving Code Completion via Repository-level Contextual Information0
Show:102550
← PrevPage 2 of 5Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1deepseek-coder-33b-baseAverage69.01Unverified
2deepseek-coder-6.7b-baseAverage63.4Unverified
3starcoderbaseAverage55.54Unverified
4gpt-4-1106-previewAverage53.28Unverified
5CodeLlama-13b-hfAverage52.78Unverified
6deepseek-coder-1.3b-baseAverage52.63Unverified
7CodeLlama-34b-hfAverage49.66Unverified
8CodeLlama-7b-hfAverage45Unverified
9gpt-3.5-turbo-0301Average40.86Unverified
10incoder-6BAverage33.79Unverified
#ModelMetricClaimedVerifiedStatus
1CodeGPT-adaptedAccuracy (token-level)77.13Unverified
2CodeT5+ 770MEM (line-level)37.9Unverified
3CodeT5+ 220MEM (line-level)35.17Unverified
#ModelMetricClaimedVerifiedStatus
1CodeGPT-adaptedAccuracy (token-level)75.11Unverified
2CodeT5+ 770MEM (line-level)44.86Unverified
3CodeT5+ 220MEM (line-level)43.42Unverified
#ModelMetricClaimedVerifiedStatus
1SantaCoder-MGDCompilation Rate73.03Unverified
2SantaCoderCompilation Rate59.97Unverified
3SantaCoderCompilation Rate59.79Unverified
#ModelMetricClaimedVerifiedStatus
1RamboCompilation Rate76.47Unverified
2RepoCoderCompilation Rate74.02Unverified
#ModelMetricClaimedVerifiedStatus
1RamboCompilation Rate61.7Unverified
2RepoCoderCompilation Rate58.09Unverified