MOGRAS: Human Motion with Grasping in 3D Scenes

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bhosikar, Kunal, Katageri, Siddharth, Madhavaram, Vivek, Han, Kai, Sharma, Charu
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909869975535616
author Bhosikar, Kunal
Katageri, Siddharth
Madhavaram, Vivek
Han, Kai
Sharma, Charu
author_facet Bhosikar, Kunal
Katageri, Siddharth
Madhavaram, Vivek
Han, Kai
Sharma, Charu
contents Generating realistic full-body motion interacting with objects is critical for applications in robotics, virtual reality, and human-computer interaction. While existing methods can generate full-body motion within 3D scenes, they often lack the fidelity for fine-grained tasks like object grasping. Conversely, methods that generate precise grasping motions typically ignore the surrounding 3D scene. This gap, generating full-body grasping motions that are physically plausible within a 3D scene, remains a significant challenge. To address this, we introduce MOGRAS (Human MOtion with GRAsping in 3D Scenes), a large-scale dataset that bridges this gap. MOGRAS provides pre-grasping full-body walking motions and final grasping poses within richly annotated 3D indoor scenes. We leverage MOGRAS to benchmark existing full-body grasping methods and demonstrate their limitations in scene-aware generation. Furthermore, we propose a simple yet effective method to adapt existing approaches to work seamlessly within 3D scenes. Through extensive quantitative and qualitative experiments, we validate the effectiveness of our dataset and highlight the significant improvements our proposed method achieves, paving the way for more realistic human-scene interactions.
format Preprint
id arxiv_https___arxiv_org_abs_2510_22199
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle MOGRAS: Human Motion with Grasping in 3D Scenes
Bhosikar, Kunal
Katageri, Siddharth
Madhavaram, Vivek
Han, Kai
Sharma, Charu
Computer Vision and Pattern Recognition
Graphics
Robotics
Generating realistic full-body motion interacting with objects is critical for applications in robotics, virtual reality, and human-computer interaction. While existing methods can generate full-body motion within 3D scenes, they often lack the fidelity for fine-grained tasks like object grasping. Conversely, methods that generate precise grasping motions typically ignore the surrounding 3D scene. This gap, generating full-body grasping motions that are physically plausible within a 3D scene, remains a significant challenge. To address this, we introduce MOGRAS (Human MOtion with GRAsping in 3D Scenes), a large-scale dataset that bridges this gap. MOGRAS provides pre-grasping full-body walking motions and final grasping poses within richly annotated 3D indoor scenes. We leverage MOGRAS to benchmark existing full-body grasping methods and demonstrate their limitations in scene-aware generation. Furthermore, we propose a simple yet effective method to adapt existing approaches to work seamlessly within 3D scenes. Through extensive quantitative and qualitative experiments, we validate the effectiveness of our dataset and highlight the significant improvements our proposed method achieves, paving the way for more realistic human-scene interactions.
title MOGRAS: Human Motion with Grasping in 3D Scenes
topic Computer Vision and Pattern Recognition
Graphics
Robotics
url https://arxiv.org/abs/2510.22199