| CorNav: Autonomous Agent with Self-Corrected Planning for Zero-Shot Vision-and-Language Navigation | Jun 17, 2023 | Decision MakingInstruction Following | —Unverified | 0 |
| PanoGen: Text-Conditioned Panoramic Environment Generation for Vision-and-Language Navigation | May 30, 2023 | Image OutpaintingLanguage Modelling | —Unverified | 0 |
| GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot Attention for Vision-and-Language Navigation | May 26, 2023 | Vision and Language Navigation | CodeCode Available | 0 |
| Masked Path Modeling for Vision-and-Language Navigation | May 23, 2023 | Action GenerationNavigate | —Unverified | 0 |
| PASTS: Progress-Aware Spatio-Temporal Transformer Speaker For Vision-and-Language Navigation | May 19, 2023 | Data AugmentationVision and Language Navigation | —Unverified | 0 |
| Improving Vision-and-Language Navigation by Generating Future-View Image Semantics | Apr 11, 2023 | Image GenerationNavigate | —Unverified | 0 |
| HOP+: History-enhanced and Order-aware Pre-training for Vision-and-Language Navigation | Mar 20, 2023 | Decision MakingLanguage Modeling | —Unverified | 0 |
| Meta-Explore: Exploratory Hierarchical Vision-and-Language Navigation Using Scene Object Spectrum Grounding | Mar 7, 2023 | Vision and Language NavigationVisual Navigation | —Unverified | 0 |
| MLANet: Multi-Level Attention Network with Sub-instruction for Continuous Vision-and-Language Navigation | Mar 2, 2023 | NavigateVision and Language Navigation | CodeCode Available | 0 |
| Graph based Environment Representation for Vision-and-Language Navigation in Continuous Environments | Jan 11, 2023 | Objectobject-detection | —Unverified | 0 |