Benchmarking Interaction, Beyond Policy: a Reproducible Benchmark for Collaborative Instance Object Navigation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zorzi, Edoardo, Taioli, Francesco, Wang, Yiming, Cristani, Marco, Farinelli, Alessandro, Castellini, Alberto, Bazzani, Loris |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dialogues
di: Taioli, Francesco, et al.
Pubblicazione: (2024)
di: Taioli, Francesco, et al.
Pubblicazione: (2024)
VISOR: VIsual Spatial Object Reasoning for Language-driven Object Navigation
di: Taioli, Francesco, et al.
Pubblicazione: (2026)
di: Taioli, Francesco, et al.
Pubblicazione: (2026)
Multi-Level Conditioning by Pairing Localized Text and Sketch for Fashion Image Generation
di: Liu, Ziyue, et al.
Pubblicazione: (2026)
di: Liu, Ziyue, et al.
Pubblicazione: (2026)
Interactive Episodic Memory with User Feedback
di: Subedi, Nikesh, et al.
Pubblicazione: (2026)
di: Subedi, Nikesh, et al.
Pubblicazione: (2026)
Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation
di: Taioli, Francesco, et al.
Pubblicazione: (2024)
di: Taioli, Francesco, et al.
Pubblicazione: (2024)
UniCoRN: Unified Commented Retrieval Network with LMMs
di: Jaritz, Maximilian, et al.
Pubblicazione: (2025)
di: Jaritz, Maximilian, et al.
Pubblicazione: (2025)
Diffusion-based Image Generation for In-distribution Data Augmentation in Surface Defect Detection
di: Capogrosso, Luigi, et al.
Pubblicazione: (2024)
di: Capogrosso, Luigi, et al.
Pubblicazione: (2024)
I2EDL: Interactive Instruction Error Detection and Localization
di: Taioli, Francesco, et al.
Pubblicazione: (2024)
di: Taioli, Francesco, et al.
Pubblicazione: (2024)
Med-MMFL: A Multimodal Federated Learning Benchmark in Healthcare
di: Chhetri, Aavash, et al.
Pubblicazione: (2026)
di: Chhetri, Aavash, et al.
Pubblicazione: (2026)
Multi-Camera Industrial Open-Set Person Re-Identification and Tracking
di: Cunico, Federico, et al.
Pubblicazione: (2024)
di: Cunico, Federico, et al.
Pubblicazione: (2024)
CoNav: A Benchmark for Human-Centered Collaborative Navigation
di: Li, Changhao, et al.
Pubblicazione: (2024)
di: Li, Changhao, et al.
Pubblicazione: (2024)
Beyond Sequences: A Benchmark for Atomic Hand-Object Interaction Using a Static RNN Encoder
di: Movahed, Yousef Azizi, et al.
Pubblicazione: (2025)
di: Movahed, Yousef Azizi, et al.
Pubblicazione: (2025)
ToFu: Visual Tokens Reduction via Fusion for Multi-modal, Multi-patch, Multi-image Task
di: Pippi, Vittorio, et al.
Pubblicazione: (2025)
di: Pippi, Vittorio, et al.
Pubblicazione: (2025)
Learning Visual Hierarchies in Hyperbolic Space for Image Retrieval
di: Wang, Ziwei, et al.
Pubblicazione: (2024)
di: Wang, Ziwei, et al.
Pubblicazione: (2024)
Estimating the distribution of numerosity and non-numerical visual magnitudes in natural scenes using computer vision
di: Hou, Kuinan, et al.
Pubblicazione: (2024)
di: Hou, Kuinan, et al.
Pubblicazione: (2024)
Physics-Aware Video Instance Removal Benchmark
di: Li, Zirui, et al.
Pubblicazione: (2026)
di: Li, Zirui, et al.
Pubblicazione: (2026)
A New People-Object Interaction Dataset and NVS Benchmarks
di: Guo, Shuai, et al.
Pubblicazione: (2024)
di: Guo, Shuai, et al.
Pubblicazione: (2024)
Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects
di: Fan, Zicong, et al.
Pubblicazione: (2024)
di: Fan, Zicong, et al.
Pubblicazione: (2024)
Self-Supervised Iterative Refinement for Anomaly Detection in Industrial Quality Control
di: Aqeel, Muhammad, et al.
Pubblicazione: (2024)
di: Aqeel, Muhammad, et al.
Pubblicazione: (2024)
A Contrastive Learning-Guided Confident Meta-learning for Zero Shot Anomaly Detection
di: Aqeel, Muhammad, et al.
Pubblicazione: (2025)
di: Aqeel, Muhammad, et al.
Pubblicazione: (2025)
Towards Real Unsupervised Anomaly Detection Via Confident Meta-Learning
di: Aqeel, Muhammad, et al.
Pubblicazione: (2025)
di: Aqeel, Muhammad, et al.
Pubblicazione: (2025)
Robust Anomaly Detection in Industrial Environments via Meta-Learning
di: Aqeel, Muhammad, et al.
Pubblicazione: (2025)
di: Aqeel, Muhammad, et al.
Pubblicazione: (2025)
Meta Learning-Driven Iterative Refinement for Robust Anomaly Detection in Industrial Inspection
di: Aqeel, Muhammad, et al.
Pubblicazione: (2025)
di: Aqeel, Muhammad, et al.
Pubblicazione: (2025)
OoDIS: Anomaly Instance Segmentation and Detection Benchmark
di: Nekrasov, Alexey, et al.
Pubblicazione: (2024)
di: Nekrasov, Alexey, et al.
Pubblicazione: (2024)
MIRAGE: Benchmarking and Aligning Multi-Instance Image Editing
di: Liu, Ziqian, et al.
Pubblicazione: (2026)
di: Liu, Ziqian, et al.
Pubblicazione: (2026)
Towards Unconstrained Human-Object Interaction
di: Tonini, Francesco, et al.
Pubblicazione: (2026)
di: Tonini, Francesco, et al.
Pubblicazione: (2026)
Benchmarking the Reproducibility of Brain MRI Segmentation Across Scanners and Time
di: Kondrateva, Ekaterina, et al.
Pubblicazione: (2025)
di: Kondrateva, Ekaterina, et al.
Pubblicazione: (2025)
ViewFusion: Towards Multi-View Consistency via Interpolated Denoising
di: Yang, Xianghui, et al.
Pubblicazione: (2024)
di: Yang, Xianghui, et al.
Pubblicazione: (2024)
Is SAM3 ready for pathology segmentation?
di: Kong, Qiuyu, et al.
Pubblicazione: (2026)
di: Kong, Qiuyu, et al.
Pubblicazione: (2026)
Seeing the Abstract: Translating the Abstract Language for Vision Language Models
di: Talon, Davide, et al.
Pubblicazione: (2025)
di: Talon, Davide, et al.
Pubblicazione: (2025)
StructXLIP: Enhancing Vision-language Models with Multimodal Structural Cues
di: Ruan, Zanxi, et al.
Pubblicazione: (2026)
di: Ruan, Zanxi, et al.
Pubblicazione: (2026)
Egocentric Human-Object Interaction Detection: A New Benchmark and Method
di: Deng, Kunyuan, et al.
Pubblicazione: (2025)
di: Deng, Kunyuan, et al.
Pubblicazione: (2025)
UAV-ON: A Benchmark for Open-World Object Goal Navigation with Aerial Agents
di: Xiao, Jianqiang, et al.
Pubblicazione: (2025)
di: Xiao, Jianqiang, et al.
Pubblicazione: (2025)
Ins-HOI: Instance Aware Human-Object Interactions Recovery
di: Zhang, Jiajun, et al.
Pubblicazione: (2023)
di: Zhang, Jiajun, et al.
Pubblicazione: (2023)
An Instance-Centric Panoptic Occupancy Prediction Benchmark for Autonomous Driving
di: Feng, Yi, et al.
Pubblicazione: (2026)
di: Feng, Yi, et al.
Pubblicazione: (2026)
IAM: Enhancing RGB-D Instance Segmentation with New Benchmarks
di: Jung, Aecheon, et al.
Pubblicazione: (2025)
di: Jung, Aecheon, et al.
Pubblicazione: (2025)
IS-Fusion: Instance-Scene Collaborative Fusion for Multimodal 3D Object Detection
di: Yin, Junbo, et al.
Pubblicazione: (2024)
di: Yin, Junbo, et al.
Pubblicazione: (2024)
INSTINCT: Instance-Level Interaction Architecture for Query-Based Collaborative Perception
di: Xu, Yunjiang, et al.
Pubblicazione: (2025)
di: Xu, Yunjiang, et al.
Pubblicazione: (2025)
Unsupervised Active Visual Search with Monte Carlo planning under Uncertain Detections
di: Taioli, Francesco, et al.
Pubblicazione: (2023)
di: Taioli, Francesco, et al.
Pubblicazione: (2023)
Towards Small Object Editing: A Benchmark Dataset and A Training-Free Approach
di: Pan, Qihe, et al.
Pubblicazione: (2024)
di: Pan, Qihe, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dialogues
di: Taioli, Francesco, et al.
Pubblicazione: (2024) -
VISOR: VIsual Spatial Object Reasoning for Language-driven Object Navigation
di: Taioli, Francesco, et al.
Pubblicazione: (2026) -
Multi-Level Conditioning by Pairing Localized Text and Sketch for Fashion Image Generation
di: Liu, Ziyue, et al.
Pubblicazione: (2026) -
Interactive Episodic Memory with User Feedback
di: Subedi, Nikesh, et al.
Pubblicazione: (2026) -
Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation
di: Taioli, Francesco, et al.
Pubblicazione: (2024)