SOTAVerified

Natural Language Queries

Papers

Showing 1–25 of 337 papers

TitleStatusHype
A Survey of Text-to-SQL in the Era of LLMs: Where are we, and where are we going?Code5
OpenAGI: When LLM Meets Domain ExpertsCode4
Separate Anything You DescribeCode3
Feature 3DGS: Supercharging 3D Gaussian Splatting to Enable Distilled Feature FieldsCode2
E-SQL: Direct Schema Linking via Question Enrichment in Text-to-SQLCode2
Egocentric Video-Language PretrainingCode2
TR-DETR: Task-Reciprocal Transformer for Joint Moment Retrieval and Highlight DetectionCode2
Query2CAD: Generating CAD models using natural language queriesCode2
SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQLCode2
TableQuery: Querying tabular data with natural languageCode2
Fact Finder -- Enhancing Domain Expertise of Large Language Models by Incorporating Knowledge GraphsCode2
UniMD: Towards Unifying Moment Retrieval and Temporal Action DetectionCode2
LLM-Grounder: Open-Vocabulary 3D Visual Grounding with Large Language Model as an AgentCode2
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing ScenesCode2
Query-Dependent Video Representation for Moment Retrieval and Highlight DetectionCode2
FortisAVQA and MAVEN: a Benchmark Dataset and Debiasing Framework for Robust Multimodal ReasoningCode2
Datrics Text2SQL. A Framework for Natural Language to SQL Query GenerationCode2
UMT: Unified Multi-modal Transformers for Joint Video Moment Retrieval and Highlight DetectionCode2
EgoVideo: Exploring Egocentric Foundation Model and Downstream AdaptationCode2
ESTER: A Machine Reading Comprehension Dataset for Event Semantic Relation ReasoningCode1
CoSQA: 20,000+ Web Queries for Code Search and Question AnsweringCode1
Explore-And-Match: Bridging Proposal-Based and Proposal-Free With Transformer for Sentence Grounding in VideosCode1
Entity-aware Transformers for Entity SearchCode1
CoMAT: Chain of Mathematically Annotated Thought Improves Mathematical ReasoningCode1
Enhancing Network Management Using Code Generated by Large Language ModelsCode1
Show:102550
← PrevPage 1 of 14Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1EgoVideoR@1 Mean(0.3 and 0.5)23.68—Unverified
2DeCafNet-100%R@1 Mean(0.3 and 0.5)18.86—Unverified
3DeCafNet-50%R@1 Mean(0.3 and 0.5)17.93—Unverified
4RGNetR@1 Mean(0.3 and 0.5)16.55—Unverified
5DeCafNet-50% (no NaQ)R@1 Mean(0.3 and 0.5)15.32—Unverified
6InternVideoR@1 Mean(0.3 and 0.5)13.26—Unverified
7EgoVLPv2R@1 IoU=0.312.95—Unverified
8UniMD+Sync.R@1 Mean(0.3 and 0.5)12.11—Unverified
9ReLER@ZJU-AlibabaR@1 Mean(0.3 and 0.5)10.52—Unverified
10EgoVLPR@1 Mean(0.3 and 0.5)8.35—Unverified