SOTAVerified

Natural Language Queries

Papers

Showing 1–10 of 337 papers

TitleStatusHype
SPAZER: Spatial-Semantic Progressive Reasoning Agent for Zero-shot 3D Visual Grounding—0
Towards Probabilistic Question Answering Over Tabular Data—0
A Modular Multitask Reasoning Framework Integrating Spatio-temporal Models and LLMs—0
Invocable APIs derived from NL2SQL datasets for LLM Tool-Calling Evaluation—0
Improving Personalized Search with Regularized Low-Rank Parameter UpdatesCode0
Technical Report for Argoverse2 Scenario Mining Challenges on Iterative Error Correction and Spatially-Aware Prompting—0
MLVTG: Mamba-Based Feature Alignment and LLM-Driven Purification for Multi-Modal Video Temporal Grounding—0
SEED: Enhancing Text-to-SQL Performance and Practical Usability Through Automatic Evidence GenerationCode1
OSGNet @ Ego4D Episodic Memory Challenge 2025Code1
DGMO: Training-Free Audio Source Separation through Diffusion-Guided Mask Optimization—0
Show:102550
← PrevPage 1 of 34Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1EgoVideoR@1 Mean(0.3 and 0.5)23.68—Unverified
2DeCafNet-100%R@1 Mean(0.3 and 0.5)18.86—Unverified
3DeCafNet-50%R@1 Mean(0.3 and 0.5)17.93—Unverified
4RGNetR@1 Mean(0.3 and 0.5)16.55—Unverified
5DeCafNet-50% (no NaQ)R@1 Mean(0.3 and 0.5)15.32—Unverified
6InternVideoR@1 Mean(0.3 and 0.5)13.26—Unverified
7EgoVLPv2R@1 IoU=0.312.95—Unverified
8UniMD+Sync.R@1 Mean(0.3 and 0.5)12.11—Unverified
9ReLER@ZJU-AlibabaR@1 Mean(0.3 and 0.5)10.52—Unverified
10EgoVLPR@1 Mean(0.3 and 0.5)8.35—Unverified