OSMa-Bench++: Toward Open-Ended Benchmarking of Semantic Mapping for Manipulation with Prompt-Generated Synthetic Scenes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kurkova, Regina, Popov, Maxim, Kolyubin, Sergey |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OSMa-Bench: Evaluating Open Semantic Mapping Under Varying Lighting Conditions
von: Popov, Maxim, et al.
Veröffentlicht: (2025)
von: Popov, Maxim, et al.
Veröffentlicht: (2025)
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models
von: Huynh, Cuong, et al.
Veröffentlicht: (2026)
von: Huynh, Cuong, et al.
Veröffentlicht: (2026)
R5DGS: Semantic-Aware 4D Gaussian Splatting with Rigid Body Constraints for Efficient Dynamic Scene Reconstruction
von: Gridusov, Denis, et al.
Veröffentlicht: (2026)
von: Gridusov, Denis, et al.
Veröffentlicht: (2026)
RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments
von: Nasser, Zaid, et al.
Veröffentlicht: (2026)
von: Nasser, Zaid, et al.
Veröffentlicht: (2026)
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)
OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
Open-Vocabulary Online Semantic Mapping for SLAM
von: Martins, Tomas Berriel, et al.
Veröffentlicht: (2024)
von: Martins, Tomas Berriel, et al.
Veröffentlicht: (2024)
M3Bench: Benchmarking Whole-body Motion Generation for Mobile Manipulation in 3D Scenes
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
ArtiBench and ArtiBrain: Benchmarking Generalizable Vision-Language Articulated Object Manipulation
von: Wu, Yuhan, et al.
Veröffentlicht: (2025)
von: Wu, Yuhan, et al.
Veröffentlicht: (2025)
KM-ViPE: Online Tightly Coupled Vision-Language-Geometry Fusion for Open-Vocabulary Semantic SLAM
von: Nasser, Zaid, et al.
Veröffentlicht: (2025)
von: Nasser, Zaid, et al.
Veröffentlicht: (2025)
Mimicking-Bench: A Benchmark for Generalizable Humanoid-Scene Interaction Learning via Human Mimicking
von: Liu, Yun, et al.
Veröffentlicht: (2024)
von: Liu, Yun, et al.
Veröffentlicht: (2024)
Open-vocabulary Mobile Manipulation in Unseen Dynamic Environments with 3D Semantic Maps
von: Qiu, Dicong, et al.
Veröffentlicht: (2024)
von: Qiu, Dicong, et al.
Veröffentlicht: (2024)
GOTPR: General Outdoor Text-based Place Recognition Using Scene Graph Retrieval with OpenStreetMap
von: Jung, Donghwi, et al.
Veröffentlicht: (2025)
von: Jung, Donghwi, et al.
Veröffentlicht: (2025)
QueSTMaps: Queryable Semantic Topological Maps for 3D Scene Understanding
von: Mehan, Yash, et al.
Veröffentlicht: (2024)
von: Mehan, Yash, et al.
Veröffentlicht: (2024)
SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)
DISC: Dense Integrated Semantic Context for Large-Scale Open-Set Semantic Mapping
von: Igelbrink, Felix, et al.
Veröffentlicht: (2026)
von: Igelbrink, Felix, et al.
Veröffentlicht: (2026)
PhenoBench -- A Large Dataset and Benchmarks for Semantic Image Interpretation in the Agricultural Domain
von: Weyler, Jan, et al.
Veröffentlicht: (2023)
von: Weyler, Jan, et al.
Veröffentlicht: (2023)
Lift, Splat, Map: Lifting Foundation Masks for Label-Free Semantic Scene Completion
von: Zhang, Arthur, et al.
Veröffentlicht: (2024)
von: Zhang, Arthur, et al.
Veröffentlicht: (2024)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
von: Li, Huiqiong, et al.
Veröffentlicht: (2026)
von: Li, Huiqiong, et al.
Veröffentlicht: (2026)
Towards Open-World Mobile Manipulation in Homes: Lessons from the Neurips 2023 HomeRobot Open Vocabulary Mobile Manipulation Challenge
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2024)
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2024)
OpenLex3D: A Tiered Evaluation Benchmark for Open-Vocabulary 3D Scene Representations
von: Kassab, Christina, et al.
Veröffentlicht: (2025)
von: Kassab, Christina, et al.
Veröffentlicht: (2025)
LOSS-SLAM: Lightweight Open-Set Semantic Simultaneous Localization and Mapping
von: Singh, Kurran, et al.
Veröffentlicht: (2024)
von: Singh, Kurran, et al.
Veröffentlicht: (2024)
Open-Ended Instruction Realization with LLM-Enabled Multi-Planner Scheduling in Autonomous Vehicles
von: Liu, Jiawei, et al.
Veröffentlicht: (2026)
von: Liu, Jiawei, et al.
Veröffentlicht: (2026)
AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond
von: Zhang, Haiming, et al.
Veröffentlicht: (2026)
von: Zhang, Haiming, et al.
Veröffentlicht: (2026)
GSplatLoc: Grounding Keypoint Descriptors into 3D Gaussian Splatting for Improved Visual Localization
von: Sidorov, Gennady, et al.
Veröffentlicht: (2024)
von: Sidorov, Gennady, et al.
Veröffentlicht: (2024)
DeformGS: Scene Flow in Highly Deformable Scenes for Deformable Object Manipulation
von: Duisterhof, Bardienus P., et al.
Veröffentlicht: (2023)
von: Duisterhof, Bardienus P., et al.
Veröffentlicht: (2023)
Recasting Generic Pretrained Vision Transformers As Object-Centric Scene Encoders For Manipulation Policies
von: Qian, Jianing, et al.
Veröffentlicht: (2024)
von: Qian, Jianing, et al.
Veröffentlicht: (2024)
MapBERT: Bitwise Masked Modeling for Real-Time Semantic Mapping Generation
von: Deng, Yijie, et al.
Veröffentlicht: (2025)
von: Deng, Yijie, et al.
Veröffentlicht: (2025)
Articulated 3D Scene Graphs for Open-World Mobile Manipulation
von: Büchner, Martin, et al.
Veröffentlicht: (2026)
von: Büchner, Martin, et al.
Veröffentlicht: (2026)
SCENEREPLICA: Benchmarking Real-World Robot Manipulation by Creating Replicable Scenes
von: Khargonkar, Ninad, et al.
Veröffentlicht: (2023)
von: Khargonkar, Ninad, et al.
Veröffentlicht: (2023)
LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation
von: Miao, Bo, et al.
Veröffentlicht: (2026)
von: Miao, Bo, et al.
Veröffentlicht: (2026)
Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
WildScenes: A Benchmark for 2D and 3D Semantic Segmentation in Large-scale Natural Environments
von: Vidanapathirana, Kavisha, et al.
Veröffentlicht: (2023)
von: Vidanapathirana, Kavisha, et al.
Veröffentlicht: (2023)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
Zero-shot Reconstruction of In-Scene Object Manipulation from Video
von: Lin, Dixuan, et al.
Veröffentlicht: (2025)
von: Lin, Dixuan, et al.
Veröffentlicht: (2025)
SemanticVLA: Semantic-Aligned Sparsification and Enhancement for Efficient Robotic Manipulation
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
Robo-ABC: Affordance Generalization Beyond Categories via Semantic Correspondence for Robot Manipulation
von: Ju, Yuanchen, et al.
Veröffentlicht: (2024)
von: Ju, Yuanchen, et al.
Veröffentlicht: (2024)
DexGraspNet 2.0: Learning Generative Dexterous Grasping in Large-scale Synthetic Cluttered Scenes
von: Zhang, Jialiang, et al.
Veröffentlicht: (2024)
von: Zhang, Jialiang, et al.
Veröffentlicht: (2024)
Towards Generalizable Vision-Language Robotic Manipulation: A Benchmark and LLM-guided 3D Policy
von: Garcia, Ricardo, et al.
Veröffentlicht: (2024)
von: Garcia, Ricardo, et al.
Veröffentlicht: (2024)
Open-Set 3D Semantic Instance Maps for Vision Language Navigation -- O3D-SIM
von: Nanwani, Laksh, et al.
Veröffentlicht: (2024)
von: Nanwani, Laksh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
OSMa-Bench: Evaluating Open Semantic Mapping Under Varying Lighting Conditions
von: Popov, Maxim, et al.
Veröffentlicht: (2025) -
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models
von: Huynh, Cuong, et al.
Veröffentlicht: (2026) -
R5DGS: Semantic-Aware 4D Gaussian Splatting with Rigid Body Constraints for Efficient Dynamic Scene Reconstruction
von: Gridusov, Denis, et al.
Veröffentlicht: (2026) -
RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments
von: Nasser, Zaid, et al.
Veröffentlicht: (2026) -
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)