Redundancy-aware Action Spaces for Robot Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Mazzaglia, Pietro, Backshall, Nicholas, Ma, Xiao, James, Stephen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BiGym: A Demo-Driven Mobile Bi-Manual Manipulation Benchmark
por: Chernyadev, Nikita, et al.
Publicado: (2024)
por: Chernyadev, Nikita, et al.
Publicado: (2024)
Information-driven Affordance Discovery for Efficient Robotic Manipulation
por: Mazzaglia, Pietro, et al.
Publicado: (2024)
por: Mazzaglia, Pietro, et al.
Publicado: (2024)
Hybrid Training for Vision-Language-Action Models
por: Mazzaglia, Pietro, et al.
Publicado: (2025)
por: Mazzaglia, Pietro, et al.
Publicado: (2025)
Focusing on What Matters: Object-Agent-centric Tokenization for Vision Language Action models
por: Bendikas, Rokas, et al.
Publicado: (2025)
por: Bendikas, Rokas, et al.
Publicado: (2025)
GenRL: Multimodal-foundation world models for generalization in embodied agents
por: Mazzaglia, Pietro, et al.
Publicado: (2024)
por: Mazzaglia, Pietro, et al.
Publicado: (2024)
Green Screen Augmentation Enables Scene Generalisation in Robotic Manipulation
por: Teoh, Eugene, et al.
Publicado: (2024)
por: Teoh, Eugene, et al.
Publicado: (2024)
Hierarchical Diffusion Policy for Kinematics-Aware Multi-Task Robotic Manipulation
por: Ma, Xiao, et al.
Publicado: (2024)
por: Ma, Xiao, et al.
Publicado: (2024)
Render and Diffuse: Aligning Image and Action Spaces for Diffusion-based Behaviour Cloning
por: Vosylius, Vitalis, et al.
Publicado: (2024)
por: Vosylius, Vitalis, et al.
Publicado: (2024)
Geometry-aware 4D Video Generation for Robot Manipulation
por: Liu, Zeyi, et al.
Publicado: (2025)
por: Liu, Zeyi, et al.
Publicado: (2025)
Generative Image as Action Models
por: Shridhar, Mohit, et al.
Publicado: (2024)
por: Shridhar, Mohit, et al.
Publicado: (2024)
CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining
por: Liu, I-Chun Arthur, et al.
Publicado: (2026)
por: Liu, I-Chun Arthur, et al.
Publicado: (2026)
Point Cloud Matters: Rethinking the Impact of Different Observation Spaces on Robot Learning
por: Zhu, Haoyi, et al.
Publicado: (2024)
por: Zhu, Haoyi, et al.
Publicado: (2024)
Code-as-Monitor: Constraint-aware Visual Programming for Reactive and Proactive Robotic Failure Detection
por: Zhou, Enshen, et al.
Publicado: (2024)
por: Zhou, Enshen, et al.
Publicado: (2024)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
por: Wang, Beichen, et al.
Publicado: (2024)
por: Wang, Beichen, et al.
Publicado: (2024)
mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs
por: Pai, Jonas, et al.
Publicado: (2025)
por: Pai, Jonas, et al.
Publicado: (2025)
RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation
por: Jiang, Hanxiao, et al.
Publicado: (2024)
por: Jiang, Hanxiao, et al.
Publicado: (2024)
Vision-Language-Action Models for Robotics: A Review Towards Real-World Applications
por: Kawaharazuka, Kento, et al.
Publicado: (2025)
por: Kawaharazuka, Kento, et al.
Publicado: (2025)
Learning to Visually Connect Actions and their Effects
por: Parmar, Paritosh, et al.
Publicado: (2024)
por: Parmar, Paritosh, et al.
Publicado: (2024)
Grounding Foundational Vision Models with 3D Human Poses for Robust Action Recognition
por: Babey, Nicholas, et al.
Publicado: (2025)
por: Babey, Nicholas, et al.
Publicado: (2025)
Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos
por: Li, Qixiu, et al.
Publicado: (2025)
por: Li, Qixiu, et al.
Publicado: (2025)
ViPRA: Video Prediction for Robot Actions
por: Routray, Sandeep, et al.
Publicado: (2025)
por: Routray, Sandeep, et al.
Publicado: (2025)
Imperative Learning: A Self-supervised Neuro-Symbolic Learning Framework for Robot Autonomy
por: Wang, Chen, et al.
Publicado: (2024)
por: Wang, Chen, et al.
Publicado: (2024)
AdaWorld: Learning Adaptable World Models with Latent Actions
por: Gao, Shenyuan, et al.
Publicado: (2025)
por: Gao, Shenyuan, et al.
Publicado: (2025)
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
por: Li, Qixiu, et al.
Publicado: (2024)
por: Li, Qixiu, et al.
Publicado: (2024)
Semantically Controllable Augmentations for Generalizable Robot Learning
por: Chen, Zoey, et al.
Publicado: (2024)
por: Chen, Zoey, et al.
Publicado: (2024)
On the Evaluation of Generative Robotic Simulations
por: Chen, Feng, et al.
Publicado: (2024)
por: Chen, Feng, et al.
Publicado: (2024)
Learning Visual Feature-Based World Models via Residual Latent Action
por: Zhang, Xinyu, et al.
Publicado: (2026)
por: Zhang, Xinyu, et al.
Publicado: (2026)
Learned Visual Navigation for Under-Canopy Agricultural Robots
por: Sivakumar, Arun Narenthiran, et al.
Publicado: (2021)
por: Sivakumar, Arun Narenthiran, et al.
Publicado: (2021)
EmbodiSwap for Zero-Shot Robot Imitation Learning
por: Dessalene, Eadom, et al.
Publicado: (2025)
por: Dessalene, Eadom, et al.
Publicado: (2025)
Learning by Watching: A Review of Video-based Learning Approaches for Robot Manipulation
por: Eze, Chrisantus, et al.
Publicado: (2024)
por: Eze, Chrisantus, et al.
Publicado: (2024)
FALCON: Future-Aware Learning with Contextual Object-Centric Pretraining for UAV Action Recognition
por: Xian, Ruiqi, et al.
Publicado: (2024)
por: Xian, Ruiqi, et al.
Publicado: (2024)
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
por: Yang, Ruihan, et al.
Publicado: (2025)
por: Yang, Ruihan, et al.
Publicado: (2025)
A Survey of Embodied Learning for Object-Centric Robotic Manipulation
por: Zheng, Ying, et al.
Publicado: (2024)
por: Zheng, Ying, et al.
Publicado: (2024)
Theia: Distilling Diverse Vision Foundation Models for Robot Learning
por: Shang, Jinghuan, et al.
Publicado: (2024)
por: Shang, Jinghuan, et al.
Publicado: (2024)
From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors
por: Zhang, Zhengshen, et al.
Publicado: (2025)
por: Zhang, Zhengshen, et al.
Publicado: (2025)
TidyBot++: An Open-Source Holonomic Mobile Manipulator for Robot Learning
por: Wu, Jimmy, et al.
Publicado: (2024)
por: Wu, Jimmy, et al.
Publicado: (2024)
PerAct2: Benchmarking and Learning for Robotic Bimanual Manipulation Tasks
por: Grotz, Markus, et al.
Publicado: (2024)
por: Grotz, Markus, et al.
Publicado: (2024)
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
por: Ren, Pengzhen, et al.
Publicado: (2023)
por: Ren, Pengzhen, et al.
Publicado: (2023)
Synchronous vs Asynchronous Reinforcement Learning in a Real World Robot
por: Parsaee, Ali, et al.
Publicado: (2025)
por: Parsaee, Ali, et al.
Publicado: (2025)
ExoPredicator: Learning Abstract Models of Dynamic Worlds for Robot Planning
por: Liang, Yichao, et al.
Publicado: (2025)
por: Liang, Yichao, et al.
Publicado: (2025)
Ejemplares similares
-
BiGym: A Demo-Driven Mobile Bi-Manual Manipulation Benchmark
por: Chernyadev, Nikita, et al.
Publicado: (2024) -
Information-driven Affordance Discovery for Efficient Robotic Manipulation
por: Mazzaglia, Pietro, et al.
Publicado: (2024) -
Hybrid Training for Vision-Language-Action Models
por: Mazzaglia, Pietro, et al.
Publicado: (2025) -
Focusing on What Matters: Object-Agent-centric Tokenization for Vision Language Action models
por: Bendikas, Rokas, et al.
Publicado: (2025) -
GenRL: Multimodal-foundation world models for generalization in embodied agents
por: Mazzaglia, Pietro, et al.
Publicado: (2024)