Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Taioli, Francesco, Rosa, Stefano, Castellini, Alberto, Natale, Lorenzo, Del Bue, Alessio, Farinelli, Alessandro, Cristani, Marco, Wang, Yiming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
I2EDL: Interactive Instruction Error Detection and Localization
von: Taioli, Francesco, et al.
Veröffentlicht: (2024)
von: Taioli, Francesco, et al.
Veröffentlicht: (2024)
Unsupervised Active Visual Search with Monte Carlo planning under Uncertain Detections
von: Taioli, Francesco, et al.
Veröffentlicht: (2023)
von: Taioli, Francesco, et al.
Veröffentlicht: (2023)
Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dialogues
von: Taioli, Francesco, et al.
Veröffentlicht: (2024)
von: Taioli, Francesco, et al.
Veröffentlicht: (2024)
Benchmarking Interaction, Beyond Policy: a Reproducible Benchmark for Collaborative Instance Object Navigation
von: Zorzi, Edoardo, et al.
Veröffentlicht: (2026)
von: Zorzi, Edoardo, et al.
Veröffentlicht: (2026)
Embodied Image Captioning: Self-supervised Learning Agents for Spatially Coherent Image Descriptions
von: Galliena, Tommaso, et al.
Veröffentlicht: (2025)
von: Galliena, Tommaso, et al.
Veröffentlicht: (2025)
Monte Carlo Tree Search with Velocity Obstacles for safe and efficient motion planning in dynamic environments
von: Bonanni, Lorenzo, et al.
Veröffentlicht: (2025)
von: Bonanni, Lorenzo, et al.
Veröffentlicht: (2025)
Memory-Augmented Vision-Language Agents for Persistent and Semantically Consistent Object Captioning
von: Galliena, Tommaso, et al.
Veröffentlicht: (2026)
von: Galliena, Tommaso, et al.
Veröffentlicht: (2026)
Look Around and Learn: Self-Training Object Detection by Exploration
von: Scarpellini, Gianluca, et al.
Veröffentlicht: (2023)
von: Scarpellini, Gianluca, et al.
Veröffentlicht: (2023)
Designing Control Barrier Function via Probabilistic Enumeration for Safe Reinforcement Learning Navigation
von: Marzari, Luca, et al.
Veröffentlicht: (2025)
von: Marzari, Luca, et al.
Veröffentlicht: (2025)
Depth-Constrained ASV Navigation with Deep RL and Limited Sensing
von: Zhalehmehrabi, Amirhossein, et al.
Veröffentlicht: (2025)
von: Zhalehmehrabi, Amirhossein, et al.
Veröffentlicht: (2025)
A Grasp Pose is All You Need: Learning Multi-fingered Grasping with Deep Reinforcement Learning from Vision and Touch
von: Ceola, Federico, et al.
Veröffentlicht: (2023)
von: Ceola, Federico, et al.
Veröffentlicht: (2023)
Aquatic Navigation: A Challenging Benchmark for Deep Reinforcement Learning
von: Corsi, Davide, et al.
Veröffentlicht: (2024)
von: Corsi, Davide, et al.
Veröffentlicht: (2024)
Reshaping Action Error Distributions for Reliable Vision-Language-Action Models
von: Bai, Shuanghao, et al.
Veröffentlicht: (2026)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2026)
One-Shot Open-Set Skeleton-Based Action Recognition
von: Berti, Stefano, et al.
Veröffentlicht: (2022)
von: Berti, Stefano, et al.
Veröffentlicht: (2022)
Error Decomposition for Hybrid Localization Systems
von: Flade, Benedict, et al.
Veröffentlicht: (2024)
von: Flade, Benedict, et al.
Veröffentlicht: (2024)
VISOR: VIsual Spatial Object Reasoning for Language-driven Object Navigation
von: Taioli, Francesco, et al.
Veröffentlicht: (2026)
von: Taioli, Francesco, et al.
Veröffentlicht: (2026)
Semi-Supervised Novelty Detection for Precise Ultra-Wideband Error Signal Prediction
von: Albertin, Umberto, et al.
Veröffentlicht: (2024)
von: Albertin, Umberto, et al.
Veröffentlicht: (2024)
RESPRECT: Speeding-up Multi-fingered Grasping with Residual Reinforcement Learning
von: Ceola, Federico, et al.
Veröffentlicht: (2024)
von: Ceola, Federico, et al.
Veröffentlicht: (2024)
Measuring Uncertainty in Shape Completion to Improve Grasp Quality
von: Duarte, Nuno Ferreira, et al.
Veröffentlicht: (2025)
von: Duarte, Nuno Ferreira, et al.
Veröffentlicht: (2025)
Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments
von: Hong, Haodong, et al.
Veröffentlicht: (2024)
von: Hong, Haodong, et al.
Veröffentlicht: (2024)
iCub Detecting Gazed Objects: A Pipeline Estimating Human Attention
von: Hanifi, Shiva, et al.
Veröffentlicht: (2023)
von: Hanifi, Shiva, et al.
Veröffentlicht: (2023)
Memory Unscented Particle Filter for 6-DOF Tactile Localization
von: Vezzani, Giulia, et al.
Veröffentlicht: (2016)
von: Vezzani, Giulia, et al.
Veröffentlicht: (2016)
Combining Local and Global Perception for Autonomous Navigation on Nano-UAVs
von: Lamberti, Lorenzo, et al.
Veröffentlicht: (2024)
von: Lamberti, Lorenzo, et al.
Veröffentlicht: (2024)
iCub Knows Where You Look: Exploiting Social Cues for Interactive Object Detection Learning
von: Lombardi, Maria, et al.
Veröffentlicht: (2022)
von: Lombardi, Maria, et al.
Veröffentlicht: (2022)
IFFNeRF: Initialisation Free and Fast 6DoF pose estimation from a single image and a NeRF model
von: Bortolon, Matteo, et al.
Veröffentlicht: (2024)
von: Bortolon, Matteo, et al.
Veröffentlicht: (2024)
T-ESKF: Transformed Error-State Kalman Filter for Consistent Visual-Inertial Navigation
von: Tian, Chungeng, et al.
Veröffentlicht: (2025)
von: Tian, Chungeng, et al.
Veröffentlicht: (2025)
InstruGen: Automatic Instruction Generation for Vision-and-Language Navigation Via Large Multimodal Models
von: Yan, Yu, et al.
Veröffentlicht: (2024)
von: Yan, Yu, et al.
Veröffentlicht: (2024)
IMAC-AgriVLN: Can Agricultural Vision-and-Language Navigation Agents be Aware of Instruction Mistakes?
von: Zhao, Xiaobei, et al.
Veröffentlicht: (2026)
von: Zhao, Xiaobei, et al.
Veröffentlicht: (2026)
The impact of Compositionality in Zero-shot Multi-label action recognition for Object-based tasks
von: Calabrese, Carmela, et al.
Veröffentlicht: (2024)
von: Calabrese, Carmela, et al.
Veröffentlicht: (2024)
System-Level Error Propagation and Tail-Risk Amplification in Reference-Based Robotic Navigation
von: Hu, Ning, et al.
Veröffentlicht: (2026)
von: Hu, Ning, et al.
Veröffentlicht: (2026)
Probabilistic Degeneracy Detection for Point-to-Plane Error Minimization
von: Hatleskog, Johan, et al.
Veröffentlicht: (2024)
von: Hatleskog, Johan, et al.
Veröffentlicht: (2024)
XBG: End-to-end Imitation Learning for Autonomous Behaviour in Human-Robot Interaction and Collaboration
von: Cardenas-Perez, Carlos, et al.
Veröffentlicht: (2024)
von: Cardenas-Perez, Carlos, et al.
Veröffentlicht: (2024)
NaviTrace: Evaluating Embodied Navigation of Vision-Language Models
von: Windecker, Tim, et al.
Veröffentlicht: (2025)
von: Windecker, Tim, et al.
Veröffentlicht: (2025)
LHManip: A Dataset for Long-Horizon Language-Grounded Manipulation Tasks in Cluttered Tabletop Environments
von: Ceola, Federico, et al.
Veröffentlicht: (2023)
von: Ceola, Federico, et al.
Veröffentlicht: (2023)
AED: Adaptable Error Detection for Few-shot Imitation Policy
von: Yeh, Jia-Fong, et al.
Veröffentlicht: (2024)
von: Yeh, Jia-Fong, et al.
Veröffentlicht: (2024)
Video-Based Detection and Analysis of Errors in Robotic Surgical Training
von: Lev, Hanna Kossowsky, et al.
Veröffentlicht: (2025)
von: Lev, Hanna Kossowsky, et al.
Veröffentlicht: (2025)
Pre-trained Multiple Latent Variable Generative Models are good defenders against Adversarial Attacks
von: Serez, Dario, et al.
Veröffentlicht: (2024)
von: Serez, Dario, et al.
Veröffentlicht: (2024)
A Mutual Information Perspective on Multiple Latent Variable Generative Models for Positive View Generation
von: Serez, Dario, et al.
Veröffentlicht: (2025)
von: Serez, Dario, et al.
Veröffentlicht: (2025)
GC-VLN: Instruction as Graph Constraints for Training-free Vision-and-Language Navigation
von: Yin, Hang, et al.
Veröffentlicht: (2025)
von: Yin, Hang, et al.
Veröffentlicht: (2025)
Uncertainty Aware-Predictive Control Barrier Functions: Safer Human Robot Interaction through Probabilistic Motion Forecasting
von: Busellato, Lorenzo, et al.
Veröffentlicht: (2025)
von: Busellato, Lorenzo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
I2EDL: Interactive Instruction Error Detection and Localization
von: Taioli, Francesco, et al.
Veröffentlicht: (2024) -
Unsupervised Active Visual Search with Monte Carlo planning under Uncertain Detections
von: Taioli, Francesco, et al.
Veröffentlicht: (2023) -
Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dialogues
von: Taioli, Francesco, et al.
Veröffentlicht: (2024) -
Benchmarking Interaction, Beyond Policy: a Reproducible Benchmark for Collaborative Instance Object Navigation
von: Zorzi, Edoardo, et al.
Veröffentlicht: (2026) -
Embodied Image Captioning: Self-supervised Learning Agents for Spatially Coherent Image Descriptions
von: Galliena, Tommaso, et al.
Veröffentlicht: (2025)