SOTAVerified

Text-To-SQL

Text-to-SQL is a task in natural language processing (NLP) where the goal is to automatically generate SQL queries from natural language text. The task involves converting the text input into a structured representation and then using this representation to generate a semantically correct SQL query that can be executed on a database.

( Image credit: SyntaxSQLNet )

Papers

Showing 251–300 of 424 papers

TitleStatusHype
TrustSQL: Benchmarking Text-to-SQL Reliability with Penalty-Based ScoringCode0
Schema-Aware Multi-Task Learning for Complex Text-to-SQL—0
Benchmarking the Text-to-SQL Capability of Large Language Models: A Comprehensive Evaluation—0
DFIN-SQL: Integrating Focused Schema with DIN-SQL for Superior Accuracy in Large-Scale Databases—0
Ar-Spider: Text-to-SQL in Arabic—0
R^3: "This is My SQL, Are You With Me?" A Consensus-Based Multi-Agent System for Text-to-SQL Tasks—0
Structure Guided Large Language Model for SQL Generation—0
Archer: A Human-Labeled Text-to-SQL Dataset with Arithmetic, Commonsense and Hypothetical Reasoning—0
Knowledge-to-SQL: Enhancing SQL Generation with Data Expert LLMCode0
MURRE: Multi-Hop Table Retrieval with Removal for Open-Domain Text-to-SQLCode0
Improving Demonstration Diversity by Human-Free Fusing for Text-to-SQLCode0
Improving Generalization in Semantic Parsing by Increasing Natural Language Variation—0
Evaluating the Data Model Robustness of Text-to-SQL Systems Based on Real User QueriesCode0
AraSpider: Democratizing Arabic-to-SQLCode0
Investigating the Impact of Data Contamination of Large Language Models in Text-to-SQL Translation—0
DTS-SQL: Decomposed Text-to-SQL with Small Large Language Models—0
FinSQL: Model-Agnostic LLMs-based Text-to-SQL Framework for Financial Analysis—0
Using LLM to select the right SQL Query from candidates—0
Semantic Parsing for Complex Data Retrieval: Targeting Query Plans vs. SQL for No-Code Access to Relational Databases—0
Data Transformation to Construct a Dataset for Generating Entity-Relationship Model from Natural Language—0
dIR -- Discrete Information Retrieval: Conversational Search over Unstructured (and Structured) Data with Large Language Models—0
LLM-SQL-Solver: Can LLMs Determine SQL Equivalence?Code0
Decoupling SQL Query Hardness Parsing for Text-to-SQL—0
Domain Adaptation of a State of the Art Text-to-SQL Model: Lessons Learned and Challenges Found—0
A Benchmark to Understand the Role of Knowledge Graphs on Large Language Model's Accuracy for Question Answering on Enterprise SQL Databases—0
SQLPrompt: In-Context Text-to-SQL with Minimal Labeled Data—0
Reboost Large Language Model-based Text-to-SQL, Text-to-Python, and Text-to-Function -- with Real Applications in Traffic Domain—0
ASTormer: An AST Structure-aware Transformer Decoder for Text-to-SQL—0
Evaluating Cross-Domain Text-to-SQL Models and Benchmarks—0
SQLformer: Deep Auto-Regressive Query Graph Generation for Text-to-SQL TranslationCode0
Natural Language Interfaces for Tabular Data Querying and Visualization: A Survey—0
TUR2SQL: A Cross-Domain Turkish Dataset For Text-to-SQLCode0
Benchmarking and Improving Text-to-SQL Generation under AmbiguityCode0
Semantic Decomposition of Question and SQL for Text-to-SQL ParsingCode0
MAGNIFICo: Evaluating the In-Context Learning Ability of Large Language Models to Generalize to Novel InterpretationsCode0
Battle of the Large Language Models: Dolly vs LLaMA vs Vicuna vs Guanaco vs Bard vs ChatGPT -- A Text-to-SQL Parsing Comparison—0
Selective Demonstrations for Cross-domain Text-to-SQLCode0
Enhancing Open-Domain Table Question Answering via Syntax- and Structure-aware Dense RetrievalCode0
Adapt and Decompose: Efficient Generalization of Text-to-SQL via Domain Adapted Least-To-Most Prompting—0
Retrieval-augmented GPT-3.5-based Text-to-SQL Framework with Sample-aware Prompting and Dynamic Revision Chain—0
T5-SR: A Unified Seq-to-Seq Decoding Strategy for Semantic ParsingCode0
Correcting Semantic Parses with Natural Language through Dynamic Schema EncodingCode0
Exploring the Compositional Generalization in Context Dependent Text-to-SQL Parsing—0
Federated Learning for Semantic Parsing: Task Formulation, Evaluation Setup, New AlgorithmsCode0
SQL-PaLM: Improved Large Language Model Adaptation for Text-to-SQL (extended)—0
Uncovering and Categorizing Social Biases in Text-to-SQLCode0
CSS: A Large-scale Cross-schema Chinese Text-to-SQL Medical DatasetCode0
Error Detection for Text-to-SQL Semantic ParsingCode0
Exploring Chain-of-Thought Style Prompting for Text-to-SQL—0
Enhancing Few-shot Text-to-SQL Capabilities of Large Language Models: A Study on Prompt Design Strategies—0
Show:102550
← PrevPage 6 of 9Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Human PerformanceExecution Accurarcy (Human)92.96—Unverified
2XiYan-SQLExecution Accuracy % (Test)75.63—Unverified
3DSAIR + GPT-4oExecution Accuracy % (Test)74.12—Unverified
4CHASE-SQL + GeminiExecution Accuracy % (Test)74.06—Unverified
5ExSL + granite-34b-codeExecution Accuracy % (Test)73.17—Unverified
6OpenSearch-SQL+ v2 + GPT-4oExecution Accuracy % (Test)72.28—Unverified
7Distillery + GPT-4oExecution Accuracy % (Test)71.83—Unverified
8Insights AIExecution Accuracy % (Test)70.26—Unverified
9PURPLE + RED + GPT-4oExecution Accuracy % (Test)70.21—Unverified
10MCTS-SQLExecution Accuracy % (Test)69.4—Unverified
#ModelMetricClaimedVerifiedStatus
1XiYan-SQLExecution Accuracy (Test)89.65—Unverified
2PET-SQLExecution Accuracy (Test)87.6—Unverified
3datagpt-sql-7B + InvalidSQL-FeedbackExecution Accuracy (Dev)87.2—Unverified
4DAIL-SQL + GPT-4 + Self-ConsistencyExecution Accuracy (Test)86.6—Unverified
5DIN-SQL + GPT-4Execution Accuracy (Test)85.3—Unverified
6datagpt-sql-7BExecution Accuracy (Dev)84.8—Unverified
7MSc-SQLExecution Accuracy (Test)84.7—Unverified
8MARLO + Claude 2.1Execution Accuracy (Test)84—Unverified
9C3 + ChatGPT + Zero-ShotExecution Accuracy (Test)82.3—Unverified
10code-davinci-002 175B (LEVER)Execution Accuracy (Dev)81.9—Unverified
#ModelMetricClaimedVerifiedStatus
1Spider-Agent + o1-previewSuccess Rate17.03—Unverified
2Spider-Agent + GPT-4oSuccess Rate10.13—Unverified
3Spider-Agent + Claude-3.5-SonnectSuccess Rate9.02—Unverified
4Spider-Agent + GPT-4Success Rate8.86—Unverified
5Spider-Agent + Qwen2.5-72BSuccess Rate6.17—Unverified
6Spider-Agent + DeepSeek-V2.5Success Rate5.22—Unverified
7Spider-Agent + Gemini-Pro-1.5Success Rate2.53—Unverified
8Spider-Agent + Llama-3.1-405BSuccess Rate2.21—Unverified
#ModelMetricClaimedVerifiedStatus
1RASAT+PICARDinteraction match accuracy45.2—Unverified
2RAT-SQL-TC + GAPinteraction match accuracy43.2—Unverified
3HIE-SQL + GraPPainteraction match accuracy42.9—Unverified
4RAT-SQL + SCoReinteraction match accuracy38.1—Unverified
5EditSQL + BERTinteraction match accuracy25.3—Unverified
6GAZP + BERTinteraction match accuracy23.5—Unverified
7SyntaxSQL-coninteraction match accuracy5.2—Unverified
#ModelMetricClaimedVerifiedStatus
1RAT-SQLExact Match (EM)26.77—Unverified
2Edit-SQLExact Match (EM)11.73—Unverified
#ModelMetricClaimedVerifiedStatus
1T5-LargePCM-F1 (dev)48.2—Unverified
#ModelMetricClaimedVerifiedStatus
1XiYan-SQLExecution Accuracy69.86—Unverified
#ModelMetricClaimedVerifiedStatus
1Orange-mini0-shot MRR74.17—Unverified