SOTAVerified

Code Completion

Papers

Showing 151175 of 212 papers

TitleStatusHype
ExecRepoBench: Multi-level Executable Code Completion Evaluation0
Exploring ChatGPT's Ability to Rank Content: A Preliminary Study on Consistency with Human Preferences0
Fast and Memory-Efficient Neural Code Completion0
FastDraft: How to Train Your Draft0
FEA-Bench: A Benchmark for Evaluating Repository-Level Code Generation for Feature Implementation0
From Copilot to Pilot: Towards AI Supported Software Development0
Full Line Code Completion: Bringing AI to Desktop0
GenAI for Simulation Model in Model-Based Systems Engineering0
Generation Probabilities Are Not Enough: Uncertainty Highlighting in AI Code Completions0
GraphCodeBERT: Pre-training Code Representations with Data Flow0
HiLDe: Intentional Code Generation via Human-in-the-Loop Decoding0
Horizon-Length Prediction: Advancing Fill-in-the-Middle Capabilities for Code Generation with Lookahead Planning0
Identifying and Mitigating the Security Risks of Generative AI0
Identifying and Mitigating Vulnerabilities in LLM-Integrated Applications0
Improving Code Autocompletion with Transfer Learning0
Improving FIM Code Completions via Context & Curriculum Based Learning0
IntelliCode Compose: Code Generation Using Transformer0
Interpretability Illusions in the Generalization of Simplified Models0
Is Next Token Prediction Sufficient for GPT? Exploration on Code Logic Comprehension0
Unveiling Code Pre-Trained Models: Investigating Syntax and Semantics Capacities0
Jailbreak Attacks and Defenses Against Large Language Models: A Survey0
JudgeRank: Leveraging Large Language Models for Reasoning-Intensive Reranking0
KV Prediction for Improved Time to First Token0
Laminar: A New Serverless Stream-based Framework with Semantic Code Search and Code Completion0
Learning to Extend Program Graphs to Work-in-Progress Code0
Show:102550
← PrevPage 7 of 9Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1deepseek-coder-33b-baseAverage69.01Unverified
2deepseek-coder-6.7b-baseAverage63.4Unverified
3starcoderbaseAverage55.54Unverified
4gpt-4-1106-previewAverage53.28Unverified
5CodeLlama-13b-hfAverage52.78Unverified
6deepseek-coder-1.3b-baseAverage52.63Unverified
7CodeLlama-34b-hfAverage49.66Unverified
8CodeLlama-7b-hfAverage45Unverified
9gpt-3.5-turbo-0301Average40.86Unverified
10incoder-6BAverage33.79Unverified
#ModelMetricClaimedVerifiedStatus
1CodeGPT-adaptedAccuracy (token-level)77.13Unverified
2CodeT5+ 770MEM (line-level)37.9Unverified
3CodeT5+ 220MEM (line-level)35.17Unverified
#ModelMetricClaimedVerifiedStatus
1CodeGPT-adaptedAccuracy (token-level)75.11Unverified
2CodeT5+ 770MEM (line-level)44.86Unverified
3CodeT5+ 220MEM (line-level)43.42Unverified
#ModelMetricClaimedVerifiedStatus
1SantaCoder-MGDCompilation Rate73.03Unverified
2SantaCoderCompilation Rate59.97Unverified
3SantaCoderCompilation Rate59.79Unverified
#ModelMetricClaimedVerifiedStatus
1RamboCompilation Rate76.47Unverified
2RepoCoderCompilation Rate74.02Unverified
#ModelMetricClaimedVerifiedStatus
1RamboCompilation Rate61.7Unverified
2RepoCoderCompilation Rate58.09Unverified