What Matters to You? Towards Visual Representation Alignment for Robot Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Tian, Ran, Xu, Chenfeng, Tomizuka, Masayoshi, Malik, Jitendra, Bajcsy, Andrea |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
por: Tian, Ran, et al.
Publicado: (2024)
por: Tian, Ran, et al.
Publicado: (2024)
Reimagination with Test-time Observation Interventions: Distractor-Robust World Model Predictions for Visual Model Predictive Control
por: Chen, Yuxin, et al.
Publicado: (2025)
por: Chen, Yuxin, et al.
Publicado: (2025)
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
por: Yuan, Puzhen, et al.
Publicado: (2025)
por: Yuan, Puzhen, et al.
Publicado: (2025)
PhyGrasp: Generalizing Robotic Grasping with Physics-informed Large Multimodal Models
por: Guo, Dingkun, et al.
Publicado: (2024)
por: Guo, Dingkun, et al.
Publicado: (2024)
Spatially Visual Perception for End-to-End Robotic Learning
por: Davies, Travis, et al.
Publicado: (2024)
por: Davies, Travis, et al.
Publicado: (2024)
LanguageMPC: Large Language Models as Decision Makers for Autonomous Driving
por: Sha, Hao, et al.
Publicado: (2023)
por: Sha, Hao, et al.
Publicado: (2023)
StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement
por: Seo, Junwon, et al.
Publicado: (2026)
por: Seo, Junwon, et al.
Publicado: (2026)
What Matters for Scalable and Robust Learning in End-to-End Driving Planners?
por: Holtz, David, et al.
Publicado: (2026)
por: Holtz, David, et al.
Publicado: (2026)
NeRFs in Robotics: A Survey
por: Wang, Guangming, et al.
Publicado: (2024)
por: Wang, Guangming, et al.
Publicado: (2024)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
por: Jiang, Guangqi, et al.
Publicado: (2024)
por: Jiang, Guangqi, et al.
Publicado: (2024)
Your Robot Will Feel You Now: Empathy in Robots and Embodied Agents
por: Lim, Angelica, et al.
Publicado: (2026)
por: Lim, Angelica, et al.
Publicado: (2026)
OK-Robot: What Really Matters in Integrating Open-Knowledge Models for Robotics
por: Liu, Peiqi, et al.
Publicado: (2024)
por: Liu, Peiqi, et al.
Publicado: (2024)
MATRIX: Multi-Agent Trajectory Generation with Diverse Contexts
por: Xu, Zhuo, et al.
Publicado: (2024)
por: Xu, Zhuo, et al.
Publicado: (2024)
Spatio-Temporal Graph Dual-Attention Network for Multi-Agent Prediction and Tracking
por: Li, Jiachen, et al.
Publicado: (2021)
por: Li, Jiachen, et al.
Publicado: (2021)
Hand-Object Interaction Pretraining from Videos
por: Singh, Himanshu Gaurav, et al.
Publicado: (2024)
por: Singh, Himanshu Gaurav, et al.
Publicado: (2024)
Robotic Visual Instruction
por: Li, Yanbang, et al.
Publicado: (2025)
por: Li, Yanbang, et al.
Publicado: (2025)
Focus On What Matters: Separated Models For Visual-Based RL Generalization
por: Zhang, Di, et al.
Publicado: (2024)
por: Zhang, Di, et al.
Publicado: (2024)
Memorize What Matters: Emergent Scene Decomposition from Multitraverse
por: Li, Yiming, et al.
Publicado: (2024)
por: Li, Yiming, et al.
Publicado: (2024)
Learning Visuotactile Skills with Two Multifingered Hands
por: Lin, Toru, et al.
Publicado: (2024)
por: Lin, Toru, et al.
Publicado: (2024)
ViTaS: Visual Tactile Soft Fusion Contrastive Learning for Visuomotor Learning
por: Tian, Yufeng, et al.
Publicado: (2026)
por: Tian, Yufeng, et al.
Publicado: (2026)
Distracted Robot: How Visual Clutter Undermine Robotic Manipulation
por: Rasouli, Amir, et al.
Publicado: (2025)
por: Rasouli, Amir, et al.
Publicado: (2025)
Visual Homing in Outdoor Robots Using Mushroom Body Circuits and Learning Walks
por: Gattaux, Gabriel G., et al.
Publicado: (2025)
por: Gattaux, Gabriel G., et al.
Publicado: (2025)
ROSA: Harnessing Robot States for Vision-Language and Action Alignment
por: Wen, Yuqing, et al.
Publicado: (2025)
por: Wen, Yuqing, et al.
Publicado: (2025)
DSLO: Deep Sequence LiDAR Odometry Based on Inconsistent Spatio-temporal Propagation
por: Zhang, Huixin, et al.
Publicado: (2024)
por: Zhang, Huixin, et al.
Publicado: (2024)
CL3R: 3D Reconstruction and Contrastive Learning for Enhanced Robotic Manipulation Representations
por: Cui, Wenbo, et al.
Publicado: (2025)
por: Cui, Wenbo, et al.
Publicado: (2025)
Toward Aligning Human and Robot Actions via Multi-Modal Demonstration Learning
por: Zahid, Azizul, et al.
Publicado: (2025)
por: Zahid, Azizul, et al.
Publicado: (2025)
DEXOP: A Device for Robotic Transfer of Dexterous Human Manipulation
por: Fang, Hao-Shu, et al.
Publicado: (2025)
por: Fang, Hao-Shu, et al.
Publicado: (2025)
Twisting Lids Off with Two Hands
por: Lin, Toru, et al.
Publicado: (2024)
por: Lin, Toru, et al.
Publicado: (2024)
Visual IRL for Human-Like Robotic Manipulation
por: Asali, Ehsan, et al.
Publicado: (2024)
por: Asali, Ehsan, et al.
Publicado: (2024)
Towards Realistic Scene Generation with LiDAR Diffusion Models
por: Ran, Haoxi, et al.
Publicado: (2024)
por: Ran, Haoxi, et al.
Publicado: (2024)
Sim-to-Real Reinforcement Learning for Vision-Based Dexterous Manipulation on Humanoids
por: Lin, Toru, et al.
Publicado: (2025)
por: Lin, Toru, et al.
Publicado: (2025)
DexGarmentLab: Dexterous Garment Manipulation Environment with Generalizable Policy
por: Wang, Yuran, et al.
Publicado: (2025)
por: Wang, Yuran, et al.
Publicado: (2025)
ContraMap: Contrastive Uncertainty Mapping for Robot Environment Representation
por: Le, Chi Cuong, et al.
Publicado: (2026)
por: Le, Chi Cuong, et al.
Publicado: (2026)
Adaptive Visual Imitation Learning for Robotic Assisted Feeding Across Varied Bowl Configurations and Food Types
por: Liu, Rui, et al.
Publicado: (2024)
por: Liu, Rui, et al.
Publicado: (2024)
Visual SLAMMOT Considering Multiple Motion Models
por: Tian, Peilin, et al.
Publicado: (2024)
por: Tian, Peilin, et al.
Publicado: (2024)
Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models
por: Shi, Junyao, et al.
Publicado: (2024)
por: Shi, Junyao, et al.
Publicado: (2024)
Visual Foresight for Robotic Stow: A Diffusion-Based World Model from Sparse Snapshots
por: Zhang, Lijun, et al.
Publicado: (2026)
por: Zhang, Lijun, et al.
Publicado: (2026)
RoboVIP: Multi-View Video Generation with Visual Identity Prompting Augments Robot Manipulation
por: Wang, Boyang, et al.
Publicado: (2026)
por: Wang, Boyang, et al.
Publicado: (2026)
DynaRend: Learning 3D Dynamics via Masked Future Rendering for Robotic Manipulation
por: Tian, Jingyi, et al.
Publicado: (2025)
por: Tian, Jingyi, et al.
Publicado: (2025)
Learning to Manipulate Anywhere: A Visual Generalizable Framework For Reinforcement Learning
por: Yuan, Zhecheng, et al.
Publicado: (2024)
por: Yuan, Zhecheng, et al.
Publicado: (2024)
Ejemplares similares
-
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
por: Tian, Ran, et al.
Publicado: (2024) -
Reimagination with Test-time Observation Interventions: Distractor-Robust World Model Predictions for Visual Model Predictive Control
por: Chen, Yuxin, et al.
Publicado: (2025) -
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
por: Yuan, Puzhen, et al.
Publicado: (2025) -
PhyGrasp: Generalizing Robotic Grasping with Physics-informed Large Multimodal Models
por: Guo, Dingkun, et al.
Publicado: (2024) -
Spatially Visual Perception for End-to-End Robotic Learning
por: Davies, Travis, et al.
Publicado: (2024)