Object-Centric World Model for Language-Guided Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jeong, Youngjoon, Chun, Junha, Cha, Soonwoo, Kim, Taesup |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sparse Imagination for Efficient Visual World Model Planning
von: Chun, Junha, et al.
Veröffentlicht: (2025)
von: Chun, Junha, et al.
Veröffentlicht: (2025)
Learning to Act Robustly with View-Invariant Latent Actions
von: Jeong, Youngjoon, et al.
Veröffentlicht: (2026)
von: Jeong, Youngjoon, et al.
Veröffentlicht: (2026)
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation
von: Qi, Carl, et al.
Veröffentlicht: (2024)
von: Qi, Carl, et al.
Veröffentlicht: (2024)
A Survey of Embodied Learning for Object-Centric Robotic Manipulation
von: Zheng, Ying, et al.
Veröffentlicht: (2024)
von: Zheng, Ying, et al.
Veröffentlicht: (2024)
Attention-Guided Integration of CLIP and SAM for Precise Object Masking in Robotic Manipulation
von: Muttaqien, Muhammad A., et al.
Veröffentlicht: (2025)
von: Muttaqien, Muhammad A., et al.
Veröffentlicht: (2025)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
Part-Guided 3D RL for Sim2Real Articulated Object Manipulation
von: Xie, Pengwei, et al.
Veröffentlicht: (2024)
von: Xie, Pengwei, et al.
Veröffentlicht: (2024)
SynHLMA:Synthesizing Hand Language Manipulation for Articulated Object with Discrete Human Object Interaction Representation
von: zhi, Wang, et al.
Veröffentlicht: (2025)
von: zhi, Wang, et al.
Veröffentlicht: (2025)
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
von: Huang, Wenlong, et al.
Veröffentlicht: (2026)
von: Huang, Wenlong, et al.
Veröffentlicht: (2026)
Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)
SoFar: Language-Grounded Orientation Bridges Spatial Reasoning and Object Manipulation
von: Qi, Zekun, et al.
Veröffentlicht: (2025)
von: Qi, Zekun, et al.
Veröffentlicht: (2025)
Towards the Vision-Sound-Language-Action Paradigm: The HEAR Framework for Sound-Centric Manipulation
von: Nie, Chang, et al.
Veröffentlicht: (2026)
von: Nie, Chang, et al.
Veröffentlicht: (2026)
CarFormer: Self-Driving with Learned Object-Centric Representations
von: Hamdan, Shadi, et al.
Veröffentlicht: (2024)
von: Hamdan, Shadi, et al.
Veröffentlicht: (2024)
MultiPLY: A Multisensory Object-Centric Embodied Large Language Model in 3D World
von: Hong, Yining, et al.
Veröffentlicht: (2024)
von: Hong, Yining, et al.
Veröffentlicht: (2024)
HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System
von: Yang, Tianshuo, et al.
Veröffentlicht: (2026)
von: Yang, Tianshuo, et al.
Veröffentlicht: (2026)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
von: Li, Yi, et al.
Veröffentlicht: (2025)
von: Li, Yi, et al.
Veröffentlicht: (2025)
IRASim: A Fine-Grained World Model for Robot Manipulation
von: Zhu, Fangqi, et al.
Veröffentlicht: (2024)
von: Zhu, Fangqi, et al.
Veröffentlicht: (2024)
Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model
von: Xu, Wenjiang, et al.
Veröffentlicht: (2025)
von: Xu, Wenjiang, et al.
Veröffentlicht: (2025)
Driver-WM: A Driver-Centric Traffic-Conditioned Latent World Model for In-Cabin Dynamics Rollout
von: Chi, Haozhuang, et al.
Veröffentlicht: (2026)
von: Chi, Haozhuang, et al.
Veröffentlicht: (2026)
RPMArt: Towards Robust Perception and Manipulation for Articulated Objects
von: Wang, Junbo, et al.
Veröffentlicht: (2024)
von: Wang, Junbo, et al.
Veröffentlicht: (2024)
Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
von: Jiang, Guangqi, et al.
Veröffentlicht: (2024)
von: Jiang, Guangqi, et al.
Veröffentlicht: (2024)
FindAnything: Open-Vocabulary and Object-Centric Mapping for Robot Exploration in Any Environment
von: Laina, Sebastián Barbas, et al.
Veröffentlicht: (2025)
von: Laina, Sebastián Barbas, et al.
Veröffentlicht: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
PlaceIt3D: Language-Guided Object Placement in Real 3D Scenes
von: Abdelreheem, Ahmed, et al.
Veröffentlicht: (2025)
von: Abdelreheem, Ahmed, et al.
Veröffentlicht: (2025)
AnyPlace: Learning Generalized Object Placement for Robot Manipulation
von: Zhao, Yuchi, et al.
Veröffentlicht: (2025)
von: Zhao, Yuchi, et al.
Veröffentlicht: (2025)
Pri4R: Learning World Dynamics for Vision-Language-Action Models with Privileged 4D Representation
von: Kim, Jisoo, et al.
Veröffentlicht: (2026)
von: Kim, Jisoo, et al.
Veröffentlicht: (2026)
PhysWorld: From Real Videos to World Models of Deformable Objects via Physics-Aware Demonstration Synthesis
von: Yang, Yu, et al.
Veröffentlicht: (2025)
von: Yang, Yu, et al.
Veröffentlicht: (2025)
World Models for Learning Dexterous Hand-Object Interactions from Human Videos
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2025)
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2025)
SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World
von: Ehsani, Kiana, et al.
Veröffentlicht: (2023)
von: Ehsani, Kiana, et al.
Veröffentlicht: (2023)
OC-SOP: Enhancing Vision-Based 3D Semantic Occupancy Prediction by Object-Centric Awareness
von: Cao, Helin, et al.
Veröffentlicht: (2025)
von: Cao, Helin, et al.
Veröffentlicht: (2025)
TFusionOcc: T-Primitive Based Object-Centric Multi-Sensor Fusion Framework for 3D Occupancy Prediction
von: Ming, Zhenxing, et al.
Veröffentlicht: (2026)
von: Ming, Zhenxing, et al.
Veröffentlicht: (2026)
Language-Conditioned World Modeling for Visual Navigation
von: Dong, Yifei, et al.
Veröffentlicht: (2026)
von: Dong, Yifei, et al.
Veröffentlicht: (2026)
Articulated 3D Scene Graphs for Open-World Mobile Manipulation
von: Büchner, Martin, et al.
Veröffentlicht: (2026)
von: Büchner, Martin, et al.
Veröffentlicht: (2026)
WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
Learning Environment-Aware Affordance for 3D Articulated Object Manipulation under Occlusions
von: Wu, Ruihai, et al.
Veröffentlicht: (2023)
von: Wu, Ruihai, et al.
Veröffentlicht: (2023)
From Scene to Object: Text-Guided Dual-Gaze Prediction
von: Ke, Zehong, et al.
Veröffentlicht: (2026)
von: Ke, Zehong, et al.
Veröffentlicht: (2026)
ContactHandover: Contact-Guided Robot-to-Human Object Handover
von: Wang, Zixi, et al.
Veröffentlicht: (2024)
von: Wang, Zixi, et al.
Veröffentlicht: (2024)
FlowBot3D: Learning 3D Articulation Flow to Manipulate Articulated Objects
von: Eisner, Ben, et al.
Veröffentlicht: (2022)
von: Eisner, Ben, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Sparse Imagination for Efficient Visual World Model Planning
von: Chun, Junha, et al.
Veröffentlicht: (2025) -
Learning to Act Robustly with View-Invariant Latent Actions
von: Jeong, Youngjoon, et al.
Veröffentlicht: (2026) -
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation
von: Qi, Carl, et al.
Veröffentlicht: (2024) -
A Survey of Embodied Learning for Object-Centric Robotic Manipulation
von: Zheng, Ying, et al.
Veröffentlicht: (2024) -
Attention-Guided Integration of CLIP and SAM for Precise Object Masking in Robotic Manipulation
von: Muttaqien, Muhammad A., et al.
Veröffentlicht: (2025)