Viewpoint Matters: Dynamically Optimizing Viewpoints with Masked Autoencoder for Visual Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Yi, Pengfei, Han, Yifan, Li, Junyan, Liu, Litao, Lian, Wenzhao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FoAM: Foresight-Augmented Multi-Task Imitation Policy for Robotic Manipulation
by: Liu, Litao, et al.
Published: (2024)
by: Liu, Litao, et al.
Published: (2024)
AVR: Active Vision-Driven Precise Robot Manipulation with Viewpoint and Focal Length Optimization
by: Liu, Yushan, et al.
Published: (2025)
by: Liu, Yushan, et al.
Published: (2025)
Viewpoint-Agnostic Manipulation Policies with Strategic Vantage Selection
by: Vasudevan, Sreevishakh, et al.
Published: (2025)
by: Vasudevan, Sreevishakh, et al.
Published: (2025)
Optimizing Active Perception for Learning Simultaneous Viewpoint Selection and Manipulation with Diffusion Policy
by: Sun, Xiatao, et al.
Published: (2024)
by: Sun, Xiatao, et al.
Published: (2024)
FSAG: Enhancing Human-to-Dexterous-Hand Finger-Specific Affordance Grounding via Diffusion Models
by: Han, Yifan, et al.
Published: (2026)
by: Han, Yifan, et al.
Published: (2026)
EgoDemoGen: Egocentric Demonstration Generation for Viewpoint Generalization in Robotic Manipulation
by: Xu, Yuan, et al.
Published: (2025)
by: Xu, Yuan, et al.
Published: (2025)
SAGE: Scene Graph-Aware Guidance and Execution for Long-Horizon Manipulation Tasks
by: Li, Jialiang, et al.
Published: (2025)
by: Li, Jialiang, et al.
Published: (2025)
Robust Imitation Learning for Mobile Manipulator Focusing on Task-Related Viewpoints and Regions
by: Ishida, Yutaro, et al.
Published: (2024)
by: Ishida, Yutaro, et al.
Published: (2024)
Robot Guided Evacuation with Viewpoint Constraints
by: Chen, Gong, et al.
Published: (2024)
by: Chen, Gong, et al.
Published: (2024)
Beyond Viewpoint Generalization: What Multi-View Demonstrations Offer and How to Synthesize Them for Robot Manipulation?
by: Cai, Boyang, et al.
Published: (2026)
by: Cai, Boyang, et al.
Published: (2026)
A General One-Shot Multimodal Active Perception Framework for Robotic Manipulation: Learning to Predict Optimal Viewpoint
by: Qin, Deyun, et al.
Published: (2026)
by: Qin, Deyun, et al.
Published: (2026)
Open-Source Multi-Viewpoint Surgical Telerobotics
by: Caccianiga, Guido, et al.
Published: (2025)
by: Caccianiga, Guido, et al.
Published: (2025)
Where to Look Next: Learning Viewpoint Recommendations for Informative Trajectory Planning
by: Lodel, Max, et al.
Published: (2022)
by: Lodel, Max, et al.
Published: (2022)
SPOT: Point Cloud Based Stereo Visual Place Recognition for Similar and Opposing Viewpoints
by: Carmichael, Spencer, et al.
Published: (2024)
by: Carmichael, Spencer, et al.
Published: (2024)
RoVi-Aug: Robot and Viewpoint Augmentation for Cross-Embodiment Robot Learning
by: Chen, Lawrence Yunliang, et al.
Published: (2024)
by: Chen, Lawrence Yunliang, et al.
Published: (2024)
Skill Transfer and Discovery for Sim-to-Real Learning: A Representation-Based Viewpoint
by: Ma, Haitong, et al.
Published: (2024)
by: Ma, Haitong, et al.
Published: (2024)
BridgeACT: Bridging Human Demonstrations to Robot Actions via Unified Tool-Target Affordances
by: Han, Yifan, et al.
Published: (2026)
by: Han, Yifan, et al.
Published: (2026)
DyNaVLM: Zero-Shot Vision-Language Navigation System with Dynamic Viewpoints and Self-Refining Graph Memory
by: Ji, Zihe, et al.
Published: (2025)
by: Ji, Zihe, et al.
Published: (2025)
Multi-Layered Reasoning from a Single Viewpoint for Learning See-Through Grasping
by: Wan, Fang, et al.
Published: (2023)
by: Wan, Fang, et al.
Published: (2023)
Viewpoint-Agnostic Grasp Pipeline using VLM and Partial Observations
by: Almeida, Dilermando, et al.
Published: (2026)
by: Almeida, Dilermando, et al.
Published: (2026)
DexHiL: A Human-in-the-Loop Framework for Vision-Language-Action Model Post-Training in Dexterous Manipulation
by: Han, Yifan, et al.
Published: (2026)
by: Han, Yifan, et al.
Published: (2026)
Object Pose Estimation by Camera Arm Control Based on the Next Viewpoint Estimation
by: Mizuno, Tomoki, et al.
Published: (2025)
by: Mizuno, Tomoki, et al.
Published: (2025)
ActLoc: Learning to Localize on the Move via Active Viewpoint Selection
by: Li, Jiajie, et al.
Published: (2025)
by: Li, Jiajie, et al.
Published: (2025)
Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments
by: Li, Zerui, et al.
Published: (2025)
by: Li, Zerui, et al.
Published: (2025)
Look, Zoom, Understand: The Robotic Eyeball for Embodied Perception
by: Yang, Jiashu, et al.
Published: (2025)
by: Yang, Jiashu, et al.
Published: (2025)
AnyCamVLA: Zero-Shot Camera Adaptation for Viewpoint Robust Vision-Language-Action Models
by: Heo, Hyeongjun, et al.
Published: (2026)
by: Heo, Hyeongjun, et al.
Published: (2026)
Act, Sense, Act: Learning Non-Markovian Active Perception Strategies from Large-Scale Egocentric Human Data
by: Li, Jialiang, et al.
Published: (2026)
by: Li, Jialiang, et al.
Published: (2026)
GS-NBV: a Geometry-based, Semantics-aware Viewpoint Planning Algorithm for Avocado Harvesting under Occlusions
by: Song, Xiao'ao, et al.
Published: (2025)
by: Song, Xiao'ao, et al.
Published: (2025)
CLOVER: Context-aware Long-term Object Viewpoint- and Environment- Invariant Representation Learning
by: Lee, Dongmyeong, et al.
Published: (2024)
by: Lee, Dongmyeong, et al.
Published: (2024)
A Path Planning Algorithm for UAV 3D Surface Inspection Based on Normal Vector Filtering and Integrated Viewpoint Evaluation
by: Yunlong Wang, et al.
Published: (2025)
by: Yunlong Wang, et al.
Published: (2025)
Learning Where to Look: Self-supervised Viewpoint Selection for Active Localization using Geometrical Information
by: Di Giammarino, Luca, et al.
Published: (2024)
by: Di Giammarino, Luca, et al.
Published: (2024)
BORA: Bridging Offline Reinforcement Learning and Online Residual Adaptation for Real-World Dexterous VLA Models
by: Chen, Zhongxi, et al.
Published: (2026)
by: Chen, Zhongxi, et al.
Published: (2026)
Beyond the Patch: Exploring Vulnerabilities of Visuomotor Policies via Viewpoint-Consistent 3D Adversarial Object
by: Lee, Chanmi, et al.
Published: (2026)
by: Lee, Chanmi, et al.
Published: (2026)
Context-Nav: Context-Driven Exploration and Viewpoint-Aware 3D Spatial Reasoning for Instance Navigation
by: Jang, Won Shik, et al.
Published: (2026)
by: Jang, Won Shik, et al.
Published: (2026)
CLASH: Collision Learning via Augmented Sim-to-real Hybridization to Bridge the Reality Gap
by: He, Haotian, et al.
Published: (2026)
by: He, Haotian, et al.
Published: (2026)
RoboDexVLM: Visual Language Model-Enabled Task Planning and Motion Control for Dexterous Robot Manipulation
by: Liu, Haichao, et al.
Published: (2025)
by: Liu, Haichao, et al.
Published: (2025)
Sim2real Image Translation Enables Viewpoint-Robust Policies from Fixed-Camera Datasets
by: Coholich, Jeremiah, et al.
Published: (2026)
by: Coholich, Jeremiah, et al.
Published: (2026)
EEG-Driven AR-Robot System for Zero-Touch Grasping Manipulation
by: Wang, Junzhe, et al.
Published: (2025)
by: Wang, Junzhe, et al.
Published: (2025)
MaskedManipulator: Versatile Whole-Body Manipulation
by: Tessler, Chen, et al.
Published: (2025)
by: Tessler, Chen, et al.
Published: (2025)
GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation
by: Chai, Ying, et al.
Published: (2025)
by: Chai, Ying, et al.
Published: (2025)
Similar Items
-
FoAM: Foresight-Augmented Multi-Task Imitation Policy for Robotic Manipulation
by: Liu, Litao, et al.
Published: (2024) -
AVR: Active Vision-Driven Precise Robot Manipulation with Viewpoint and Focal Length Optimization
by: Liu, Yushan, et al.
Published: (2025) -
Viewpoint-Agnostic Manipulation Policies with Strategic Vantage Selection
by: Vasudevan, Sreevishakh, et al.
Published: (2025) -
Optimizing Active Perception for Learning Simultaneous Viewpoint Selection and Manipulation with Diffusion Policy
by: Sun, Xiatao, et al.
Published: (2024) -
FSAG: Enhancing Human-to-Dexterous-Hand Finger-Specific Affordance Grounding via Diffusion Models
by: Han, Yifan, et al.
Published: (2026)