How Physics and Background Attributes Impact Video Transformers in Robotic Manipulation: A Case Study on Planar Pushing
Fuente:
arXiv
Salvato in:
| Autori principali: | Jin, Shutong, Wang, Ruiyu, Zahid, Muhammad, Pokorny, Florian T. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PACA: Perspective-Aware Cross-Attention Representation for Zero-Shot Scene Rearrangement
di: Jin, Shutong, et al.
Pubblicazione: (2024)
di: Jin, Shutong, et al.
Pubblicazione: (2024)
RealCraft: Attention Control as A Tool for Zero-Shot Consistent Video Editing
di: Jin, Shutong, et al.
Pubblicazione: (2023)
di: Jin, Shutong, et al.
Pubblicazione: (2023)
RoboLight: A Dataset with Linearly Composable Illumination for Robotic Manipulation
di: Jin, Shutong, et al.
Pubblicazione: (2026)
di: Jin, Shutong, et al.
Pubblicazione: (2026)
R900: Understanding the Cost-Effectiveness of Random Exploration from 900 Hours of Robotic Data Collection
di: Jin, Shutong, et al.
Pubblicazione: (2025)
di: Jin, Shutong, et al.
Pubblicazione: (2025)
ViTA-Seg: Vision Transformer for Amodal Segmentation in Robotics
di: Caramia, Donato, et al.
Pubblicazione: (2025)
di: Caramia, Donato, et al.
Pubblicazione: (2025)
Physically-based Lighting Generation for Robotic Manipulation
di: Jin, Shutong, et al.
Pubblicazione: (2025)
di: Jin, Shutong, et al.
Pubblicazione: (2025)
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation
di: Xie, Senwei, et al.
Pubblicazione: (2025)
di: Xie, Senwei, et al.
Pubblicazione: (2025)
RoboPearls: Editable Video Simulation for Robot Manipulation
di: Tang, Tao, et al.
Pubblicazione: (2025)
di: Tang, Tao, et al.
Pubblicazione: (2025)
Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations
di: Patel, Shivansh, et al.
Pubblicazione: (2025)
di: Patel, Shivansh, et al.
Pubblicazione: (2025)
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation
di: Li, Gen, et al.
Pubblicazione: (2024)
di: Li, Gen, et al.
Pubblicazione: (2024)
6D Object Pose Tracking in Internet Videos for Robotic Manipulation
di: Ponimatkin, Georgy, et al.
Pubblicazione: (2025)
di: Ponimatkin, Georgy, et al.
Pubblicazione: (2025)
GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation
di: Zhou, Kaichen, et al.
Pubblicazione: (2026)
di: Zhou, Kaichen, et al.
Pubblicazione: (2026)
MAPLE: Encoding Dexterous Robotic Manipulation Priors Learned From Egocentric Videos
di: Gavryushin, Alexey, et al.
Pubblicazione: (2025)
di: Gavryushin, Alexey, et al.
Pubblicazione: (2025)
Distracted Robot: How Visual Clutter Undermine Robotic Manipulation
di: Rasouli, Amir, et al.
Pubblicazione: (2025)
di: Rasouli, Amir, et al.
Pubblicazione: (2025)
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
di: Shen, Yichao, et al.
Pubblicazione: (2025)
di: Shen, Yichao, et al.
Pubblicazione: (2025)
Beyond Dense Futures: World Models as Structured Planners for Robotic Manipulation
di: Jin, Minghao, et al.
Pubblicazione: (2026)
di: Jin, Minghao, et al.
Pubblicazione: (2026)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
di: Jiang, Zebin, et al.
Pubblicazione: (2025)
di: Jiang, Zebin, et al.
Pubblicazione: (2025)
VidMan: Exploiting Implicit Dynamics from Video Diffusion Model for Effective Robot Manipulation
di: Wen, Youpeng, et al.
Pubblicazione: (2024)
di: Wen, Youpeng, et al.
Pubblicazione: (2024)
Garbage Segmentation and Attribute Analysis by Robotic Dogs
di: Xu, Nuo, et al.
Pubblicazione: (2024)
di: Xu, Nuo, et al.
Pubblicazione: (2024)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
di: Wu, Yiming, et al.
Pubblicazione: (2025)
di: Wu, Yiming, et al.
Pubblicazione: (2025)
Track2Act: Predicting Point Tracks from Internet Videos enables Generalizable Robot Manipulation
di: Bharadhwaj, Homanga, et al.
Pubblicazione: (2024)
di: Bharadhwaj, Homanga, et al.
Pubblicazione: (2024)
You Only Teach Once: Learn One-Shot Bimanual Robotic Manipulation from Video Demonstrations
di: Zhou, Huayi, et al.
Pubblicazione: (2025)
di: Zhou, Huayi, et al.
Pubblicazione: (2025)
TASTE-Rob: Advancing Video Generation of Task-Oriented Hand-Object Interaction for Generalizable Robotic Manipulation
di: Zhao, Hongxiang, et al.
Pubblicazione: (2025)
di: Zhao, Hongxiang, et al.
Pubblicazione: (2025)
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
di: Chen, Yuzhi, et al.
Pubblicazione: (2026)
di: Chen, Yuzhi, et al.
Pubblicazione: (2026)
Fast Visuomotor Policy for Robotic Manipulation
di: Jia, Jingkai, et al.
Pubblicazione: (2025)
di: Jia, Jingkai, et al.
Pubblicazione: (2025)
From Generated Human Videos to Physically Plausible Robot Trajectories
di: Ni, James, et al.
Pubblicazione: (2025)
di: Ni, James, et al.
Pubblicazione: (2025)
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning
di: Chen, Hao, et al.
Pubblicazione: (2026)
di: Chen, Hao, et al.
Pubblicazione: (2026)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
di: Guo, Heyu, et al.
Pubblicazione: (2025)
di: Guo, Heyu, et al.
Pubblicazione: (2025)
An Integrated Approach to Robotic Object Grasping and Manipulation
di: Ahmed, Owais, et al.
Pubblicazione: (2024)
di: Ahmed, Owais, et al.
Pubblicazione: (2024)
Improving Generalization of Language-Conditioned Robot Manipulation
di: Cui, Chenglin, et al.
Pubblicazione: (2025)
di: Cui, Chenglin, et al.
Pubblicazione: (2025)
Object-Centric Instruction Augmentation for Robotic Manipulation
di: Wen, Junjie, et al.
Pubblicazione: (2024)
di: Wen, Junjie, et al.
Pubblicazione: (2024)
Towards Generalizable Robotic Manipulation in Dynamic Environments
di: Fang, Heng, et al.
Pubblicazione: (2026)
di: Fang, Heng, et al.
Pubblicazione: (2026)
ZeroMimic: Distilling Robotic Manipulation Skills from Web Videos
di: Shi, Junyao, et al.
Pubblicazione: (2025)
di: Shi, Junyao, et al.
Pubblicazione: (2025)
PASG: A Closed-Loop Framework for Automated Geometric Primitive Extraction and Semantic Anchoring in Robotic Manipulation
di: Zhu, Zhihao, et al.
Pubblicazione: (2025)
di: Zhu, Zhihao, et al.
Pubblicazione: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
di: Gao, Jensen, et al.
Pubblicazione: (2023)
di: Gao, Jensen, et al.
Pubblicazione: (2023)
Mitigating the Human-Robot Domain Discrepancy in Visual Pre-training for Robotic Manipulation
di: Zhou, Jiaming, et al.
Pubblicazione: (2024)
di: Zhou, Jiaming, et al.
Pubblicazione: (2024)
MIND-V: Hierarchical World Model for Long-Horizon Robotic Manipulation with RL-based Physical Alignment
di: Zhang, Ruicheng, et al.
Pubblicazione: (2025)
di: Zhang, Ruicheng, et al.
Pubblicazione: (2025)
Planar Velocity Estimation for Fast-Moving Mobile Robots Using Event-Based Optical Flow
di: Boyle, Liam, et al.
Pubblicazione: (2025)
di: Boyle, Liam, et al.
Pubblicazione: (2025)
Patch-Based Spatial Authorship Attribution in Human-Robot Collaborative Paintings
di: Chen, Eric, et al.
Pubblicazione: (2026)
di: Chen, Eric, et al.
Pubblicazione: (2026)
Closed Loop Interactive Embodied Reasoning for Robot Manipulation
di: Nazarczuk, Michal, et al.
Pubblicazione: (2024)
di: Nazarczuk, Michal, et al.
Pubblicazione: (2024)
Documenti analoghi
-
PACA: Perspective-Aware Cross-Attention Representation for Zero-Shot Scene Rearrangement
di: Jin, Shutong, et al.
Pubblicazione: (2024) -
RealCraft: Attention Control as A Tool for Zero-Shot Consistent Video Editing
di: Jin, Shutong, et al.
Pubblicazione: (2023) -
RoboLight: A Dataset with Linearly Composable Illumination for Robotic Manipulation
di: Jin, Shutong, et al.
Pubblicazione: (2026) -
R900: Understanding the Cost-Effectiveness of Random Exploration from 900 Hours of Robotic Data Collection
di: Jin, Shutong, et al.
Pubblicazione: (2025) -
ViTA-Seg: Vision Transformer for Amodal Segmentation in Robotics
di: Caramia, Donato, et al.
Pubblicazione: (2025)