MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Pingrui, Gao, Xianqiang, Wu, Yuhan, Liu, Kehui, Wang, Dong, Wang, Zhigang, Zhao, Bin, Ding, Yan, Li, Xuelong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024)
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024)
MoMa-Pos: An Efficient Object-Kinematic-Aware Base Placement Optimization Framework for Mobile Manipulation
von: Shao, Beichen, et al.
Veröffentlicht: (2024)
von: Shao, Beichen, et al.
Veröffentlicht: (2024)
Sketch-MoMa: Teleoperation for Mobile Manipulator via Interpretation of Hand-Drawn Sketches
von: Tanada, Kosei, et al.
Veröffentlicht: (2024)
von: Tanada, Kosei, et al.
Veröffentlicht: (2024)
AIRoA MoMa Dataset: A Large-Scale Hierarchical Dataset for Mobile Manipulation
von: Takanami, Ryosuke, et al.
Veröffentlicht: (2025)
von: Takanami, Ryosuke, et al.
Veröffentlicht: (2025)
FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset
von: Liu, Kehui, et al.
Veröffentlicht: (2025)
von: Liu, Kehui, et al.
Veröffentlicht: (2025)
Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation
von: Zhang, Pingrui, et al.
Veröffentlicht: (2025)
von: Zhang, Pingrui, et al.
Veröffentlicht: (2025)
MoMa: A Modular Deep Learning Framework for Material Property Prediction
von: Wang, Botian, et al.
Veröffentlicht: (2025)
von: Wang, Botian, et al.
Veröffentlicht: (2025)
Closed-Loop Action Chunks with Dynamic Corrections for Training-Free Diffusion Policy
von: Wu, Pengyuan, et al.
Veröffentlicht: (2026)
von: Wu, Pengyuan, et al.
Veröffentlicht: (2026)
MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition
von: Yang, Yuhuan, et al.
Veröffentlicht: (2025)
von: Yang, Yuhuan, et al.
Veröffentlicht: (2025)
LiveScene: Language Embedding Interactive Radiance Fields for Physical Scene Rendering and Control
von: Qu, Delin, et al.
Veröffentlicht: (2024)
von: Qu, Delin, et al.
Veröffentlicht: (2024)
COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models
von: Liu, Kehui, et al.
Veröffentlicht: (2024)
von: Liu, Kehui, et al.
Veröffentlicht: (2024)
MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts
von: Lin, Xi Victoria, et al.
Veröffentlicht: (2024)
von: Lin, Xi Victoria, et al.
Veröffentlicht: (2024)
FastUMI: A Scalable and Hardware-Independent Universal Manipulation Interface with Dataset
von: Zhaxizhuoma, et al.
Veröffentlicht: (2024)
von: Zhaxizhuoma, et al.
Veröffentlicht: (2024)
SAGA: Open-World Mobile Manipulation via Structured Affordance Grounding
von: Fang, Kuan, et al.
Veröffentlicht: (2025)
von: Fang, Kuan, et al.
Veröffentlicht: (2025)
Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning
von: Gao, Xianqiang, et al.
Veröffentlicht: (2026)
von: Gao, Xianqiang, et al.
Veröffentlicht: (2026)
The Last Mile
von: Sinha, Amarjeet
Veröffentlicht: (2025)
von: Sinha, Amarjeet
Veröffentlicht: (2025)
AerialVG: A Challenging Benchmark for Aerial Visual Grounding by Exploring Positional Relations
von: Liu, Junli, et al.
Veröffentlicht: (2025)
von: Liu, Junli, et al.
Veröffentlicht: (2025)
Kinematic-aware Prompting for Generalizable Articulated Object Manipulation with LLMs
von: Xia, Wenke, et al.
Veröffentlicht: (2023)
von: Xia, Wenke, et al.
Veröffentlicht: (2023)
Stepping Forward on the Last Mile
von: Feng, Chen, et al.
Veröffentlicht: (2024)
von: Feng, Chen, et al.
Veröffentlicht: (2024)
On the Entropy in Last-Mile Logistics
von: Gerrits, Berry, et al.
Veröffentlicht: (2026)
von: Gerrits, Berry, et al.
Veröffentlicht: (2026)
KitchenTwin: Semantically and Geometrically Grounded 3D Kitchen Digital Twins
von: Wu, Quanyun, et al.
Veröffentlicht: (2026)
von: Wu, Quanyun, et al.
Veröffentlicht: (2026)
Manipulate-to-Navigate: Reinforcement Learning with Visual Affordances and Manipulability Priors
von: Zhang, Yuying, et al.
Veröffentlicht: (2025)
von: Zhang, Yuying, et al.
Veröffentlicht: (2025)
TeleMoMa: A Modular and Versatile Teleoperation System for Mobile Manipulation
von: Dass, Shivin, et al.
Veröffentlicht: (2024)
von: Dass, Shivin, et al.
Veröffentlicht: (2024)
MoMaStage: Skill-State Graph Guided Planning and Closed-Loop Execution for Long-Horizon Indoor Mobile Manipulation
von: Li, Chenxu, et al.
Veröffentlicht: (2026)
von: Li, Chenxu, et al.
Veröffentlicht: (2026)
No Transfers Required: Integrating Last Mile with Public Transit Using Opti-Mile
von: Altaf, Raashid, et al.
Veröffentlicht: (2023)
von: Altaf, Raashid, et al.
Veröffentlicht: (2023)
Bridging the Last Mile of Time Series Forecasting with LLM Agents
von: Liao, Yuhua, et al.
Veröffentlicht: (2026)
von: Liao, Yuhua, et al.
Veröffentlicht: (2026)
Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation
von: Yao, Yuanqi, et al.
Veröffentlicht: (2025)
von: Yao, Yuanqi, et al.
Veröffentlicht: (2025)
Affordance RAG: Hierarchical Multimodal Retrieval with Affordance-Aware Embodied Memory for Mobile Manipulation
von: Korekata, Ryosuke, et al.
Veröffentlicht: (2025)
von: Korekata, Ryosuke, et al.
Veröffentlicht: (2025)
Bridging the Last Mile of Circuit Design: PostEDA-Bench, a Hierarchical Benchmark for PPA Convergence and DRC Fixing
von: Liu, Pengju, et al.
Veröffentlicht: (2026)
von: Liu, Pengju, et al.
Veröffentlicht: (2026)
Experimental Designs for Optimizing Last-Mile Delivery
von: Rios, Nicholas, et al.
Veröffentlicht: (2024)
von: Rios, Nicholas, et al.
Veröffentlicht: (2024)
S2SServiceBench: A Multimodal Benchmark for Last-Mile S2S Climate Services
von: Li, Chenyue, et al.
Veröffentlicht: (2026)
von: Li, Chenyue, et al.
Veröffentlicht: (2026)
From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging
von: Shi, Yuling, et al.
Veröffentlicht: (2024)
von: Shi, Yuling, et al.
Veröffentlicht: (2024)
OVAL-Prompt: Open-Vocabulary Affordance Localization for Robot Manipulation through LLM Affordance-Grounding
von: Tong, Edmond, et al.
Veröffentlicht: (2024)
von: Tong, Edmond, et al.
Veröffentlicht: (2024)
MoTo: A Zero-shot Plug-in Interaction-aware Navigation for General Mobile Manipulation
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
Beyond Physical Reach: Comparing Head- and Cane-Mounted Cameras for Last-Mile Navigation by Blind Users
von: Varshney, Apurv, et al.
Veröffentlicht: (2025)
von: Varshney, Apurv, et al.
Veröffentlicht: (2025)
AffordanceSAM: Segment Anything Once More in Affordance Grounding
von: Jiang, Dengyang, et al.
Veröffentlicht: (2025)
von: Jiang, Dengyang, et al.
Veröffentlicht: (2025)
Autonomous on-Demand Shuttles for First Mile-Last Mile Connectivity: Design, Optimization, and Impact Assessment
von: Roy, Sudipta, et al.
Veröffentlicht: (2024)
von: Roy, Sudipta, et al.
Veröffentlicht: (2024)
No Last Mile: A Theory of the Human Data Market
von: Ansari, Ali, et al.
Veröffentlicht: (2026)
von: Ansari, Ali, et al.
Veröffentlicht: (2026)
Facial Features Integration in Last Mile Delivery Robots
von: Gankhuyag, Delgermaa, et al.
Veröffentlicht: (2024)
von: Gankhuyag, Delgermaa, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024) -
MoMa-Pos: An Efficient Object-Kinematic-Aware Base Placement Optimization Framework for Mobile Manipulation
von: Shao, Beichen, et al.
Veröffentlicht: (2024) -
Sketch-MoMa: Teleoperation for Mobile Manipulator via Interpretation of Hand-Drawn Sketches
von: Tanada, Kosei, et al.
Veröffentlicht: (2024) -
AIRoA MoMa Dataset: A Large-Scale Hierarchical Dataset for Mobile Manipulation
von: Takanami, Ryosuke, et al.
Veröffentlicht: (2025) -
FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset
von: Liu, Kehui, et al.
Veröffentlicht: (2025)