PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Jesse, Memmel, Marius, Kim, Kevin, Fox, Dieter, Thomason, Jesse, Ramos, Fabio, Bıyık, Erdem, Gupta, Abhishek, Li, Anqi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViSaRL: Visual Reinforcement Learning Guided by Human Saliency
by: Liang, Anthony, et al.
Published: (2024)
by: Liang, Anthony, et al.
Published: (2024)
HAND Me the Data: Fast Robot Adaptation via Hand Path Retrieval
by: Hong, Matthew, et al.
Published: (2025)
by: Hong, Matthew, et al.
Published: (2025)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
by: Li, Yi, et al.
Published: (2025)
by: Li, Yi, et al.
Published: (2025)
ASID: Active Exploration for System Identification in Robotic Manipulation
by: Memmel, Marius, et al.
Published: (2024)
by: Memmel, Marius, et al.
Published: (2024)
ReWiND: Language-Guided Rewards Teach Robot Policies without New Demonstrations
by: Zhang, Jiahui, et al.
Published: (2025)
by: Zhang, Jiahui, et al.
Published: (2025)
Contrast Sets for Evaluating Language-Guided Robot Policies
by: Anwar, Abrar, et al.
Published: (2024)
by: Anwar, Abrar, et al.
Published: (2024)
THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation
by: Pumacay, Wilbert, et al.
Published: (2024)
by: Pumacay, Wilbert, et al.
Published: (2024)
STRAP: Robot Sub-Trajectory Retrieval for Augmented Policy Learning
by: Memmel, Marius, et al.
Published: (2024)
by: Memmel, Marius, et al.
Published: (2024)
Words that make SENSE: Sensorimotor Norms in Learned Lexical Token Representations
by: Gupta, Abhinav, et al.
Published: (2026)
by: Gupta, Abhinav, et al.
Published: (2026)
EXTRACT: Efficient Policy Learning by Extracting Transferable Robot Skills from Offline Data
by: Zhang, Jesse, et al.
Published: (2024)
by: Zhang, Jesse, et al.
Published: (2024)
Training robots with natural and lightweight human feedback
by: Erdem Bıyık
Published: (2026)
by: Erdem Bıyık
Published: (2026)
ManiFlow: A General Robot Manipulation Policy via Consistency Flow Training
by: Yan, Ge, et al.
Published: (2025)
by: Yan, Ge, et al.
Published: (2025)
Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection
by: Anwar, Abrar, et al.
Published: (2025)
by: Anwar, Abrar, et al.
Published: (2025)
URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images
by: Chen, Zoey, et al.
Published: (2024)
by: Chen, Zoey, et al.
Published: (2024)
RobotFleet: An Open-Source Framework for Centralized Multi-Robot Task Planning
by: Gupta, Rohan, et al.
Published: (2025)
by: Gupta, Rohan, et al.
Published: (2025)
Phonological Representation Learning for Isolated Signs Improves Out-of-Vocabulary Generalization
by: Kezar, Lee, et al.
Published: (2025)
by: Kezar, Lee, et al.
Published: (2025)
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
by: Liang, Anthony, et al.
Published: (2026)
by: Liang, Anthony, et al.
Published: (2026)
Learning to Deliberate: Meta-policy Collaboration for Agentic LLMs with Multi-agent Reinforcement Learning
by: Yang, Wei, et al.
Published: (2025)
by: Yang, Wei, et al.
Published: (2025)
Adjust for Trust: Mitigating Trust-Induced Inappropriate Reliance on AI Assistance
by: Srinivasan, Tejas, et al.
Published: (2025)
by: Srinivasan, Tejas, et al.
Published: (2025)
MILE: Model-based Intervention Learning
by: Korkmaz, Yigit, et al.
Published: (2025)
by: Korkmaz, Yigit, et al.
Published: (2025)
Religion and spirituality in counselor education: Do we really need to talk about this?
by: Jesse Fox
Published: (2024)
by: Jesse Fox
Published: (2024)
SyncTwin: Fast Digital Twin Construction and Synchronization for Safe Robotic Manipulation
by: Huang, Ruopeng, et al.
Published: (2026)
by: Huang, Ruopeng, et al.
Published: (2026)
Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning
by: Park, Chan Young, et al.
Published: (2025)
by: Park, Chan Young, et al.
Published: (2025)
Expert Personas Improve LLM Alignment but Damage Accuracy: Bootstrapping Intent-Based Persona Routing with PRISM
by: Hu, Zizhao, et al.
Published: (2026)
by: Hu, Zizhao, et al.
Published: (2026)
M3PT: A Transformer for Multimodal, Multi-Party Social Signal Prediction with Person-aware Blockwise Attention
by: Tang, Yiming, et al.
Published: (2025)
by: Tang, Yiming, et al.
Published: (2025)
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP
by: Cai, Yuliang, et al.
Published: (2025)
by: Cai, Yuliang, et al.
Published: (2025)
Multi-modal Synthetic Data Training and Model Collapse: Insights from VLMs and Diffusion Models
by: Hu, Zizhao, et al.
Published: (2025)
by: Hu, Zizhao, et al.
Published: (2025)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
by: Wang, Yufei, et al.
Published: (2024)
by: Wang, Yufei, et al.
Published: (2024)
TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning
by: Hong, Matthew M., et al.
Published: (2026)
by: Hong, Matthew M., et al.
Published: (2026)
Distributional Successor Features Enable Zero-Shot Policy Optimization
by: Zhu, Chuning, et al.
Published: (2024)
by: Zhu, Chuning, et al.
Published: (2024)
Point Bridge: 3D Representations for Cross Domain Policy Learning
by: Haldar, Siddhant, et al.
Published: (2026)
by: Haldar, Siddhant, et al.
Published: (2026)
Do Localization Methods Actually Localize Memorized Data in LLMs? A Tale of Two Benchmarks
by: Chang, Ting-Yun, et al.
Published: (2023)
by: Chang, Ting-Yun, et al.
Published: (2023)
When Parts Are Greater Than Sums: Individual LLM Components Can Outperform Full Models
by: Chang, Ting-Yun, et al.
Published: (2024)
by: Chang, Ting-Yun, et al.
Published: (2024)
Robot Policy Evaluation for Sim-to-Real Transfer: A Benchmarking Perspective
by: Yang, Xuning, et al.
Published: (2025)
by: Yang, Xuning, et al.
Published: (2025)
Batch Active Learning of Reward Functions from Human Preferences
by: Bıyık, Erdem, et al.
Published: (2024)
by: Bıyık, Erdem, et al.
Published: (2024)
Judgelight: Trajectory-Level Post-Optimization for Multi-Agent Path Finding via Closed-Subwalk Collapsing
by: Tang, Yimin, et al.
Published: (2026)
by: Tang, Yimin, et al.
Published: (2026)
Zero-Shot Visual Generalization in Robot Manipulation
by: Batra, Sumeet, et al.
Published: (2025)
by: Batra, Sumeet, et al.
Published: (2025)
Value Explicit Pretraining for Learning Transferable Representations
by: Lekkala, Kiran, et al.
Published: (2023)
by: Lekkala, Kiran, et al.
Published: (2023)
Language Models can Infer Action Semantics for Symbolic Planners from Environment Feedback
by: Zhu, Wang, et al.
Published: (2024)
by: Zhu, Wang, et al.
Published: (2024)
WinoViz: Probing Visual Properties of Objects Under Different States
by: Jin, Woojeong, et al.
Published: (2024)
by: Jin, Woojeong, et al.
Published: (2024)
Similar Items
-
ViSaRL: Visual Reinforcement Learning Guided by Human Saliency
by: Liang, Anthony, et al.
Published: (2024) -
HAND Me the Data: Fast Robot Adaptation via Hand Path Retrieval
by: Hong, Matthew, et al.
Published: (2025) -
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
by: Li, Yi, et al.
Published: (2025) -
ASID: Active Exploration for System Identification in Robotic Manipulation
by: Memmel, Marius, et al.
Published: (2024) -
ReWiND: Language-Guided Rewards Teach Robot Policies without New Demonstrations
by: Zhang, Jiahui, et al.
Published: (2025)