R-2R-CE
Emerging10papers using it
2023first seen
The 'R-2R-CE' dataset is a large-scale benchmark containing hybrid samples used to evaluate vision-and-language navigation tasks for embodied agents.
Papers using R-2R-CE (10)
- PlatonicNav: Unveiling Semantic Correspondence in Navigation with Platonic Topological MapsSpaceVLN: A Zero-Shot Vision-and-Language Navigation Agent with Online Spatial Cognitive Memory and ReasoningP2DNav: Panorama-to-Downview Reasoning for Zero-shot Vision-and-Language NavigationLatentPilot: Scene-Aware Vision-and-Language Navigation by Dreaming Ahead with Latent Visual ReasoningMapDream: Task-Driven Map Learning for Vision-Language NavigationEnhancing Vision-Language Navigation with Multimodal Event Knowledge from Real-World Indoor Tour VideosEfficient-VLN: A Training-Efficient Vision-Language Navigation ModelD3D-VLP: Dynamic 3D Vision-Language-Planning Model for Embodied Grounding and NavigationSmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language NavigationETPNav: Evolving Topological Planning for Vision-Language Navigation in
Continuous Environments