SOTAVerified

Vision and Language Navigation

Papers

Showing 1–10 of 223 papers

TitleStatusHype
Rethinking the Embodied Gap in Vision-and-Language Navigation: A Holistic Study of Physical and Visual Disparities—0
NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous EnvironmentsCode2
Grounded Vision-Language Navigation for UAVs with Open-Vocabulary Goal Understanding—0
A Navigation Framework Utilizing Vision-Language ModelsCode0
Disrupting Vision-Language Model-Driven Navigation Services via Adversarial Object Fusion—0
Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language NavigationCode1
FlightGPT: Towards Generalizable and Interpretable UAV Vision-and-Language Navigation with Vision-Language ModelsCode2
Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language NavigationCode2
CityNavAgent: Aerial Vision-and-Language Navigation with Hierarchical Semantic Planning and Global MemoryCode1
MetaScenes: Towards Automated Replica Creation for Real-world 3D Scans—0
Show:102550
← PrevPage 1 of 23Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1FLAMETask Completion (TC)40.2—Unverified
2ORAR + junction type + heading deltaTask Completion (TC)29.1—Unverified
3ORARTask Completion (TC)24.2—Unverified
4ARC + L2STOPTask Completion (TC)16.68—Unverified
5VLN Transformer +M-50 +styleTask Completion (TC)16.2—Unverified
6VLN TransformerTask Completion (TC)14.9—Unverified
7ARCTask Completion (TC)14.13—Unverified
8Retouch-RConcatTask Completion (TC)12.8—Unverified
9Gated Attention (GA)Task Completion (TC)11.9—Unverified
10RConcatTask Completion (TC)11.8—Unverified