Feature Extractor or Decision Maker: Rethinking the Role of Visual Encoders in Visuomotor Policies
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Ruiyu, Zhuang, Zheyu, Jin, Shutong, Ingelhag, Nils, Kragic, Danica, Pokorny, Florian T. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PALM: Enhanced Generalizability for Local Visuomotor Policies via Perception Alignment
von: Wang, Ruiyu, et al.
Veröffentlicht: (2026)
von: Wang, Ruiyu, et al.
Veröffentlicht: (2026)
Real-Time Operator Takeover for Visuomotor Diffusion Policy Training
von: Moletta, Marco, et al.
Veröffentlicht: (2025)
von: Moletta, Marco, et al.
Veröffentlicht: (2025)
Raising Body Ownership in End-to-End Visuomotor Policy Learning via Robot-Centric Pooling
von: Zhuang, Zheyu, et al.
Veröffentlicht: (2024)
von: Zhuang, Zheyu, et al.
Veröffentlicht: (2024)
Preference Aligned Visuomotor Diffusion Policies for Deformable Object Manipulation
von: Moletta, Marco, et al.
Veröffentlicht: (2026)
von: Moletta, Marco, et al.
Veröffentlicht: (2026)
A Robotic Skill Learning System Built Upon Diffusion Policies and Foundation Models
von: Ingelhag, Nils, et al.
Veröffentlicht: (2024)
von: Ingelhag, Nils, et al.
Veröffentlicht: (2024)
PACA: Perspective-Aware Cross-Attention Representation for Zero-Shot Scene Rearrangement
von: Jin, Shutong, et al.
Veröffentlicht: (2024)
von: Jin, Shutong, et al.
Veröffentlicht: (2024)
How Physics and Background Attributes Impact Video Transformers in Robotic Manipulation: A Case Study on Planar Pushing
von: Jin, Shutong, et al.
Veröffentlicht: (2023)
von: Jin, Shutong, et al.
Veröffentlicht: (2023)
R900: Understanding the Cost-Effectiveness of Random Exploration from 900 Hours of Robotic Data Collection
von: Jin, Shutong, et al.
Veröffentlicht: (2025)
von: Jin, Shutong, et al.
Veröffentlicht: (2025)
Learning Dexterous In-Hand Manipulation with Multifingered Hands via Visuomotor Diffusion
von: Koczy, Piotr, et al.
Veröffentlicht: (2025)
von: Koczy, Piotr, et al.
Veröffentlicht: (2025)
Hybrid-Diffusion Models: Combining Open-loop Routines with Visuomotor Diffusion Policies
von: Van Haastregt, Jonne, et al.
Veröffentlicht: (2025)
von: Van Haastregt, Jonne, et al.
Veröffentlicht: (2025)
Speculative Policy Orchestration: A Latency-Resilient Framework for Cloud-Robotic Manipulation
von: Nguyen, Chanh, et al.
Veröffentlicht: (2026)
von: Nguyen, Chanh, et al.
Veröffentlicht: (2026)
Physically-based Lighting Generation for Robotic Manipulation
von: Jin, Shutong, et al.
Veröffentlicht: (2025)
von: Jin, Shutong, et al.
Veröffentlicht: (2025)
RoboLight: A Dataset with Linearly Composable Illumination for Robotic Manipulation
von: Jin, Shutong, et al.
Veröffentlicht: (2026)
von: Jin, Shutong, et al.
Veröffentlicht: (2026)
Grasping a Handful: Sequential Multi-Object Dexterous Grasp Generation
von: Lu, Haofei, et al.
Veröffentlicht: (2025)
von: Lu, Haofei, et al.
Veröffentlicht: (2025)
Real-Time Iteration Scheme for Diffusion Policy
von: Duan, Yufei, et al.
Veröffentlicht: (2025)
von: Duan, Yufei, et al.
Veröffentlicht: (2025)
RealCraft: Attention Control as A Tool for Zero-Shot Consistent Video Editing
von: Jin, Shutong, et al.
Veröffentlicht: (2023)
von: Jin, Shutong, et al.
Veröffentlicht: (2023)
VBM-NET: Visual Base Pose Learning for Mobile Manipulation using Equivariant TransporterNet and GNNs
von: Naik, Lakshadeep, et al.
Veröffentlicht: (2025)
von: Naik, Lakshadeep, et al.
Veröffentlicht: (2025)
Augmented Reality for RObots (ARRO): Pointing Visuomotor Policies Towards Visual Robustness
von: Mirjalili, Reihaneh, et al.
Veröffentlicht: (2025)
von: Mirjalili, Reihaneh, et al.
Veröffentlicht: (2025)
Perception Stitching: Zero-Shot Perception Encoder Transfer for Visuomotor Robot Policies
von: Jian, Pingcheng, et al.
Veröffentlicht: (2024)
von: Jian, Pingcheng, et al.
Veröffentlicht: (2024)
Vision Beyond Boundaries: An Initial Design Space of Domain-specific Large Vision Models in Human-robot Interaction
von: Zhang, Yuchong, et al.
Veröffentlicht: (2024)
von: Zhang, Yuchong, et al.
Veröffentlicht: (2024)
Reframing Human-Robot Interaction Through Extended Reality: Unlocking Safer, Smarter, and More Empathic Interactions with Virtual Robots and Foundation Models
von: Zhang, Yuchong, et al.
Veröffentlicht: (2025)
von: Zhang, Yuchong, et al.
Veröffentlicht: (2025)
Pushing Everything Everywhere All At Once: Probabilistic Prehensile Pushing
von: Perugini, Patrizio, et al.
Veröffentlicht: (2025)
von: Perugini, Patrizio, et al.
Veröffentlicht: (2025)
CAPGrasp: An $\mathbb{R}^3\times \text{SO(2)-equivariant}$ Continuous Approach-Constrained Generative Grasp Sampler
von: Weng, Zehang, et al.
Veröffentlicht: (2023)
von: Weng, Zehang, et al.
Veröffentlicht: (2023)
Visual Action Planning with Multiple Heterogeneous Agents
von: Lippi, Martina, et al.
Veröffentlicht: (2024)
von: Lippi, Martina, et al.
Veröffentlicht: (2024)
AdaFold: Adapting Folding Trajectories of Cloths via Feedback-loop Manipulation
von: Longhini, Alberta, et al.
Veröffentlicht: (2024)
von: Longhini, Alberta, et al.
Veröffentlicht: (2024)
DexDiffuser: Generating Dexterous Grasps with Diffusion Models
von: Weng, Zehang, et al.
Veröffentlicht: (2024)
von: Weng, Zehang, et al.
Veröffentlicht: (2024)
Will You Participate? Exploring the Potential of Robotics Competitions on Human-centric Topics
von: Zhang, Yuchong, et al.
Veröffentlicht: (2024)
von: Zhang, Yuchong, et al.
Veröffentlicht: (2024)
Puppeteer Your Robot: Augmented Reality Leader-Follower Teleoperation
von: van Haastregt, Jonne, et al.
Veröffentlicht: (2024)
von: van Haastregt, Jonne, et al.
Veröffentlicht: (2024)
Spatial-Temporal Aware Visuomotor Diffusion Policy Learning
von: Liu, Zhenyang, et al.
Veröffentlicht: (2025)
von: Liu, Zhenyang, et al.
Veröffentlicht: (2025)
S$^2$-Diffusion: Generalizing from Instance-level to Category-level Skills in Robot Manipulation
von: Yang, Quantao, et al.
Veröffentlicht: (2025)
von: Yang, Quantao, et al.
Veröffentlicht: (2025)
Bridging the Sim2Real Gap: Vision Encoder Pre-Training for Visuomotor Policy Transfer
von: Yardi, Yash, et al.
Veröffentlicht: (2025)
von: Yardi, Yash, et al.
Veröffentlicht: (2025)
Robot-DIFT: Distilling Diffusion Features for Geometrically Consistent Visuomotor Control
von: Deng, Yu, et al.
Veröffentlicht: (2026)
von: Deng, Yu, et al.
Veröffentlicht: (2026)
History-Aware Visuomotor Policy Learning via Point Tracking
von: Chen, Jingjing, et al.
Veröffentlicht: (2025)
von: Chen, Jingjing, et al.
Veröffentlicht: (2025)
Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
von: Chi, Cheng, et al.
Veröffentlicht: (2023)
von: Chi, Cheng, et al.
Veröffentlicht: (2023)
FLAME: A Federated Learning Benchmark for Robotic Manipulation
von: Betran, Santiago Bou, et al.
Veröffentlicht: (2025)
von: Betran, Santiago Bou, et al.
Veröffentlicht: (2025)
On the Generalization Capabilities, Design Choices and Limitations of Keypoint Imitation Learning
von: Lips, Thomas, et al.
Veröffentlicht: (2026)
von: Lips, Thomas, et al.
Veröffentlicht: (2026)
Falcon: Fast Visuomotor Policies via Partial Denoising
von: Chen, Haojun, et al.
Veröffentlicht: (2025)
von: Chen, Haojun, et al.
Veröffentlicht: (2025)
AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies
von: Dai, Yinpei, et al.
Veröffentlicht: (2025)
von: Dai, Yinpei, et al.
Veröffentlicht: (2025)
Learning Generalizable Visuomotor Policy through Dynamics-Alignment
von: Lee, Dohyeok, et al.
Veröffentlicht: (2025)
von: Lee, Dohyeok, et al.
Veröffentlicht: (2025)
Reduced-order Control and Geometric Structure of Learned Lagrangian Latent Dynamics
von: Friedl, Katharina, et al.
Veröffentlicht: (2026)
von: Friedl, Katharina, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PALM: Enhanced Generalizability for Local Visuomotor Policies via Perception Alignment
von: Wang, Ruiyu, et al.
Veröffentlicht: (2026) -
Real-Time Operator Takeover for Visuomotor Diffusion Policy Training
von: Moletta, Marco, et al.
Veröffentlicht: (2025) -
Raising Body Ownership in End-to-End Visuomotor Policy Learning via Robot-Centric Pooling
von: Zhuang, Zheyu, et al.
Veröffentlicht: (2024) -
Preference Aligned Visuomotor Diffusion Policies for Deformable Object Manipulation
von: Moletta, Marco, et al.
Veröffentlicht: (2026) -
A Robotic Skill Learning System Built Upon Diffusion Policies and Foundation Models
von: Ingelhag, Nils, et al.
Veröffentlicht: (2024)