VBR: A Vision Benchmark in Rome
Fuente:
arXiv
Saved in:
| Main Authors: | Brizi, Leonardo, Giacomini, Emanuele, Di Giammarino, Luca, Ferrari, Simone, Salem, Omar, De Rebotti, Lorenzo, Grisetti, Giorgio |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MAD-ICP: It Is All About Matching Data -- Robust and Informed LiDAR Odometry
by: Ferrari, Simone, et al.
Published: (2024)
by: Ferrari, Simone, et al.
Published: (2024)
Splat-LOAM: Gaussian Splatting LiDAR Odometry and Mapping
by: Giacomini, Emanuele, et al.
Published: (2025)
by: Giacomini, Emanuele, et al.
Published: (2025)
Learning Where to Look: Self-supervised Viewpoint Selection for Active Localization using Geometrical Information
by: Di Giammarino, Luca, et al.
Published: (2024)
by: Di Giammarino, Luca, et al.
Published: (2024)
DRO: Doppler-Aware Direct Radar Odometry
by: Gentil, Cedric Le, et al.
Published: (2025)
by: Gentil, Cedric Le, et al.
Published: (2025)
Resolution Where It Counts: Hash-based GPU-Accelerated 3D Reconstruction via Variance-Adaptive Voxel Grids
by: De Rebotti, Lorenzo, et al.
Published: (2025)
by: De Rebotti, Lorenzo, et al.
Published: (2025)
ActLoc: Learning to Localize on the Move via Active Viewpoint Selection
by: Li, Jiajie, et al.
Published: (2025)
by: Li, Jiajie, et al.
Published: (2025)
MAD-BA: 3D LiDAR Bundle Adjustment -- from Uncertainty Modelling to Structure Optimization
by: Ćwian, Krzysztof, et al.
Published: (2025)
by: Ćwian, Krzysztof, et al.
Published: (2025)
Building Rome with Convex Optimization
by: Han, Haoyu, et al.
Published: (2025)
by: Han, Haoyu, et al.
Published: (2025)
CapNav: Benchmarking Vision Language Models on Capability-conditioned Indoor Navigation
by: Su, Xia, et al.
Published: (2026)
by: Su, Xia, et al.
Published: (2026)
Beyond Domain Randomization: Event-Inspired Perception for Visually Robust Adversarial Imitation from Videos
by: Ramazzina, Andrea, et al.
Published: (2025)
by: Ramazzina, Andrea, et al.
Published: (2025)
Exploiting Local Features and Range Images for Small Data Real-Time Point Cloud Semantic Segmentation
by: Fusaro, Daniel, et al.
Published: (2024)
by: Fusaro, Daniel, et al.
Published: (2024)
Point-Plane Projections for Accurate LiDAR Semantic Segmentation in Small Data Scenarios
by: Mosco, Simone, et al.
Published: (2025)
by: Mosco, Simone, et al.
Published: (2025)
Uncertainty Quantification for Visual Object Pose Estimation
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
Category-Level Object Shape and Pose Estimation in Less Than a Millisecond
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
Improving Robustness of Vision-Language-Action Models by Restoring Corrupted Visual Inputs
by: Orjuela, Daniel Yezid Guarnizo, et al.
Published: (2026)
by: Orjuela, Daniel Yezid Guarnizo, et al.
Published: (2026)
Towards Realistic UAV Vision-Language Navigation: Platform, Benchmark, and Methodology
by: Wang, Xiangyu, et al.
Published: (2024)
by: Wang, Xiangyu, et al.
Published: (2024)
Benchmarking Vision-Based Object Tracking for USVs in Complex Maritime Environments
by: Din, Muhayy Ud, et al.
Published: (2024)
by: Din, Muhayy Ud, et al.
Published: (2024)
Sim2Real Bilevel Adaptation for Object Surface Classification using Vision-Based Tactile Sensors
by: Caddeo, Gabriele M., et al.
Published: (2023)
by: Caddeo, Gabriele M., et al.
Published: (2023)
WasteGAN: Data Augmentation for Robotic Waste Sorting through Generative Adversarial Networks
by: Bacchin, Alberto, et al.
Published: (2024)
by: Bacchin, Alberto, et al.
Published: (2024)
Pandora: Articulated 3D Scene Graphs from Egocentric Vision
by: Yu, Alan, et al.
Published: (2026)
by: Yu, Alan, et al.
Published: (2026)
A Benchmark for Vision-Centric HD Mapping by V2I Systems
by: Fan, Miao, et al.
Published: (2025)
by: Fan, Miao, et al.
Published: (2025)
VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models
by: Zhang, Borong, et al.
Published: (2025)
by: Zhang, Borong, et al.
Published: (2025)
LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment
by: Nie, Dujun, et al.
Published: (2026)
by: Nie, Dujun, et al.
Published: (2026)
POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation
by: Gong, Ruiyan, et al.
Published: (2026)
by: Gong, Ruiyan, et al.
Published: (2026)
ArtiBench and ArtiBrain: Benchmarking Generalizable Vision-Language Articulated Object Manipulation
by: Wu, Yuhan, et al.
Published: (2025)
by: Wu, Yuhan, et al.
Published: (2025)
VLNVerse: A Benchmark for Vision-Language Navigation with Versatile, Embodied, Realistic Simulation and Evaluation
by: Lin, Sihao, et al.
Published: (2025)
by: Lin, Sihao, et al.
Published: (2025)
HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction
by: Shi, Zhonghao, et al.
Published: (2025)
by: Shi, Zhonghao, et al.
Published: (2025)
Embodied3DBench: Benchmarking Low-Level Embodied Spatial Intelligence of Vision Language Models
by: Zhang, Jiyao, et al.
Published: (2026)
by: Zhang, Jiyao, et al.
Published: (2026)
Personalized Instance-based Navigation Toward User-Specific Objects in Realistic Environments
by: Barsellotti, Luca, et al.
Published: (2024)
by: Barsellotti, Luca, et al.
Published: (2024)
AutoDrive-QA: A Multiple-Choice Benchmark for Vision-Language Evaluation in Urban Autonomous Driving
by: Khalili, Boshra, et al.
Published: (2025)
by: Khalili, Boshra, et al.
Published: (2025)
Towards Generalizable Vision-Language Robotic Manipulation: A Benchmark and LLM-guided 3D Policy
by: Garcia, Ricardo, et al.
Published: (2024)
by: Garcia, Ricardo, et al.
Published: (2024)
EagleVision: A Multi-Task Benchmark for Cross-Domain Perception in High-Speed Autonomous Racing
by: Yagudin, Zakhar, et al.
Published: (2026)
by: Yagudin, Zakhar, et al.
Published: (2026)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
by: Wang, Guokang, et al.
Published: (2024)
by: Wang, Guokang, et al.
Published: (2024)
Understanding the Impact of Geometric Foundation Models on Vision-Language-Action Models
by: Yang, Yurou, et al.
Published: (2026)
by: Yang, Yurou, et al.
Published: (2026)
Towards Generative Predictive Display for Vision-Based Teleoperation: A Zero-Shot Benchmark of Off-the-Shelf Video Models
by: Khalil, Aws, et al.
Published: (2026)
by: Khalil, Aws, et al.
Published: (2026)
FeelAnyForce: Estimating Contact Force Feedback from Tactile Sensation for Vision-Based Tactile Sensors
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2024)
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2024)
TiROD: Tiny Robotics Dataset and Benchmark for Continual Object Detection
by: Pasti, Francesco, et al.
Published: (2024)
by: Pasti, Francesco, et al.
Published: (2024)
Global Symmetry and Orthogonal Transformations from Geometrical Moment $n$-tuples
by: Tahri, Omar
Published: (2026)
by: Tahri, Omar
Published: (2026)
All Eyes, no IMU: Learning Flight Attitude from Vision Alone
by: Hagenaars, Jesse J., et al.
Published: (2025)
by: Hagenaars, Jesse J., et al.
Published: (2025)
A Comparative Analysis of Visual Odometry in Virtual and Real-World Railways Environments
by: D'Amico, Gianluca, et al.
Published: (2024)
by: D'Amico, Gianluca, et al.
Published: (2024)
Similar Items
-
MAD-ICP: It Is All About Matching Data -- Robust and Informed LiDAR Odometry
by: Ferrari, Simone, et al.
Published: (2024) -
Splat-LOAM: Gaussian Splatting LiDAR Odometry and Mapping
by: Giacomini, Emanuele, et al.
Published: (2025) -
Learning Where to Look: Self-supervised Viewpoint Selection for Active Localization using Geometrical Information
by: Di Giammarino, Luca, et al.
Published: (2024) -
DRO: Doppler-Aware Direct Radar Odometry
by: Gentil, Cedric Le, et al.
Published: (2025) -
Resolution Where It Counts: Hash-based GPU-Accelerated 3D Reconstruction via Variance-Adaptive Voxel Grids
by: De Rebotti, Lorenzo, et al.
Published: (2025)