SOTAVerified

Code Completion

Papers

Showing 151–200 of 212 papers

TitleStatusHype
ExecRepoBench: Multi-level Executable Code Completion Evaluation—0
Exploring ChatGPT's Ability to Rank Content: A Preliminary Study on Consistency with Human Preferences—0
Fast and Memory-Efficient Neural Code Completion—0
FastDraft: How to Train Your Draft—0
FEA-Bench: A Benchmark for Evaluating Repository-Level Code Generation for Feature Implementation—0
From Copilot to Pilot: Towards AI Supported Software Development—0
Full Line Code Completion: Bringing AI to Desktop—0
GenAI for Simulation Model in Model-Based Systems Engineering—0
Generation Probabilities Are Not Enough: Uncertainty Highlighting in AI Code Completions—0
GraphCodeBERT: Pre-training Code Representations with Data Flow—0
HiLDe: Intentional Code Generation via Human-in-the-Loop Decoding—0
Horizon-Length Prediction: Advancing Fill-in-the-Middle Capabilities for Code Generation with Lookahead Planning—0
Identifying and Mitigating the Security Risks of Generative AI—0
Identifying and Mitigating Vulnerabilities in LLM-Integrated Applications—0
Improving Code Autocompletion with Transfer Learning—0
Improving FIM Code Completions via Context & Curriculum Based Learning—0
IntelliCode Compose: Code Generation Using Transformer—0
Interpretability Illusions in the Generalization of Simplified Models—0
Is Next Token Prediction Sufficient for GPT? Exploration on Code Logic Comprehension—0
Unveiling Code Pre-Trained Models: Investigating Syntax and Semantics Capacities—0
Jailbreak Attacks and Defenses Against Large Language Models: A Survey—0
JudgeRank: Leveraging Large Language Models for Reasoning-Intensive Reranking—0
KV Prediction for Improved Time to First Token—0
Laminar: A New Serverless Stream-based Framework with Semantic Code Search and Code Completion—0
Learning to Extend Program Graphs to Work-in-Progress Code—0
Learning to Complete Code with Sketches—0
HierarchyNet: Learning to Summarize Source Code with Heterogeneous Representations—0
LibEvolutionEval: A Benchmark and Study for Version-Specific Code Generation—0
LongCoder: A Long-Range Pre-trained Language Model for Code Completion—0
Long-Range Modeling of Source Code Files with eWASH: Extended Window Access by Syntax Hierarchy—0
M2rc-Eval: Massively Multilingual Repository-level Code Completion Evaluation—0
MarsCode Agent: AI-native Automated Bug Fixing—0
MathFimer: Enhancing Mathematical Reasoning by Expanding Reasoning Steps through Fill-in-the-Middle Task—0
Measuring memorization in RLHF for code completion—0
Mechanistic Understanding of Language Models in Syntactic Code Completion—0
Model Cascading for Code: A Cascaded Black-Box Multi-Model Framework for Cost-Efficient Code Completion with Self-Testing—0
MultiCoder: Multi-Programming-Lingual Pre-Training for Low-Resource Code Completion—0
Natural Language Generation and Understanding of Big Code for AI-Assisted Programming: A Review—0
Neural Machine Translation for Code Generation—0
Neural Models for Source Code Synthesis and Completion—0
NoEsis: Differentially Private Knowledge Transfer in Modular LLM Adaptation—0
OMPGPT: A Generative Pre-trained Transformer Model for OpenMP—0
On Explaining (Large) Language Models For Code Using Global Code-Based Explanations—0
Past as a Guide: Leveraging Retrospective Learning for Python Code Completion—0
Plan for Speed -- Dilated Scheduling for Masked Diffusion Language Models—0
Procedural Memory Is Not All You Need: Bridging Cognitive Gaps in LLM-Based Agents—0
Protect Your Secrets: Understanding and Measuring Data Exposure in VSCode Extensions—0
R2C2-Coder: Enhancing and Benchmarking Real-world Repository-level Code Completion Abilities of Code Large Language Models—0
Repoformer: Selective Retrieval for Repository-Level Code Completion—0
REPOFUSE: Repository-Level Code Completion with Fused Dual Context—0
Show:102550
← PrevPage 4 of 5Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1deepseek-coder-33b-baseAverage69.01—Unverified
2deepseek-coder-6.7b-baseAverage63.4—Unverified
3starcoderbaseAverage55.54—Unverified
4gpt-4-1106-previewAverage53.28—Unverified
5CodeLlama-13b-hfAverage52.78—Unverified
6deepseek-coder-1.3b-baseAverage52.63—Unverified
7CodeLlama-34b-hfAverage49.66—Unverified
8CodeLlama-7b-hfAverage45—Unverified
9gpt-3.5-turbo-0301Average40.86—Unverified
10incoder-6BAverage33.79—Unverified
#ModelMetricClaimedVerifiedStatus
1CodeGPT-adaptedAccuracy (token-level)77.13—Unverified
2CodeT5+ 770MEM (line-level)37.9—Unverified
3CodeT5+ 220MEM (line-level)35.17—Unverified
#ModelMetricClaimedVerifiedStatus
1CodeGPT-adaptedAccuracy (token-level)75.11—Unverified
2CodeT5+ 770MEM (line-level)44.86—Unverified
3CodeT5+ 220MEM (line-level)43.42—Unverified
#ModelMetricClaimedVerifiedStatus
1SantaCoder-MGDCompilation Rate73.03—Unverified
2SantaCoderCompilation Rate59.97—Unverified
3SantaCoderCompilation Rate59.79—Unverified
#ModelMetricClaimedVerifiedStatus
1RamboCompilation Rate76.47—Unverified
2RepoCoderCompilation Rate74.02—Unverified
#ModelMetricClaimedVerifiedStatus
1RamboCompilation Rate61.7—Unverified
2RepoCoderCompilation Rate58.09—Unverified