Do You Know Where Your Camera Is? View-Invariant Policy Learning with Camera Conditioning
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Tianchong, Ji, Jingtian, Tan, Xiangshan, Fang, Jiading, Bhattad, Anand, Guizilini, Vitor, Walter, Matthew R. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HapCompass: A Rotational Haptic Device for Contact-Rich Robotic Teleoperation
by: Tan, Xiangshan, et al.
Published: (2026)
by: Tan, Xiangshan, et al.
Published: (2026)
FlashBack: Consistency Model-Accelerated Shared Autonomy
by: Sun, Luzhe, et al.
Published: (2025)
by: Sun, Luzhe, et al.
Published: (2025)
View-Invariant Policy Learning via Zero-Shot Novel View Synthesis
by: Tian, Stephen, et al.
Published: (2024)
by: Tian, Stephen, et al.
Published: (2024)
Statler: State-Maintaining Language Models for Embodied Reasoning
by: Yoneda, Takuma, et al.
Published: (2023)
by: Yoneda, Takuma, et al.
Published: (2023)
SplArt: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian Splatting
by: Lin, Shengjie, et al.
Published: (2025)
by: Lin, Shengjie, et al.
Published: (2025)
Efficient Camera Pose Augmentation for View Generalization in Robotic Policy Learning
by: Wang, Sen, et al.
Published: (2026)
by: Wang, Sen, et al.
Published: (2026)
Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning
by: Fang, Jiading
Published: (2025)
by: Fang, Jiading
Published: (2025)
StackGen: Generating Stable Structures from Silhouettes via Diffusion
by: Sun, Luzhe, et al.
Published: (2024)
by: Sun, Luzhe, et al.
Published: (2024)
Transcrib3D: 3D Referring Expression Resolution through Large Language Models
by: Fang, Jiading, et al.
Published: (2024)
by: Fang, Jiading, et al.
Published: (2024)
Coverage Optimization for Camera View Selection
by: Chen, Timothy, et al.
Published: (2026)
by: Chen, Timothy, et al.
Published: (2026)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
by: Bhatt, Neel P., et al.
Published: (2024)
by: Bhatt, Neel P., et al.
Published: (2024)
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement
by: Ayalew, Tewodros, et al.
Published: (2024)
by: Ayalew, Tewodros, et al.
Published: (2024)
iCub Knows Where You Look: Exploiting Social Cues for Interactive Object Detection Learning
by: Lombardi, Maria, et al.
Published: (2022)
by: Lombardi, Maria, et al.
Published: (2022)
Cross-Modal Semi-Dense 6-DoF Tracking of an Event Camera in Challenging Conditions
by: Zuo, Yi-Fan, et al.
Published: (2024)
by: Zuo, Yi-Fan, et al.
Published: (2024)
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
by: Du, Xiaodan, et al.
Published: (2023)
by: Du, Xiaodan, et al.
Published: (2023)
Multi-Camera View Scaling for Data-Efficient Robot Imitation Learning
by: Xie, Yichen, et al.
Published: (2026)
by: Xie, Yichen, et al.
Published: (2026)
Towards Realistic Scene Generation with LiDAR Diffusion Models
by: Ran, Haoxi, et al.
Published: (2024)
by: Ran, Haoxi, et al.
Published: (2024)
Targetless Extrinsic Calibration of Stereo Cameras, Thermal Cameras, and Laser Sensors in the Wild
by: Fu, Taimeng, et al.
Published: (2021)
by: Fu, Taimeng, et al.
Published: (2021)
Unify Robot Actions in Camera Frame
by: Xie, Sicheng, et al.
Published: (2025)
by: Xie, Sicheng, et al.
Published: (2025)
CtRNet-X: Camera-to-Robot Pose Estimation in Real-world Conditions Using a Single Camera
by: Lu, Jingpei, et al.
Published: (2024)
by: Lu, Jingpei, et al.
Published: (2024)
I Know You're Listening: Adaptive Voice for HRI
by: Tuttösí, Paige
Published: (2025)
by: Tuttösí, Paige
Published: (2025)
I Know Your Feelings Before You Do: Predicting Future Affective Reactions in Human-Computer Dialogue
by: Li, Yuanchao, et al.
Published: (2023)
by: Li, Yuanchao, et al.
Published: (2023)
Motion-Aware Optical Camera Communication with Event Cameras
by: Su, Hang, et al.
Published: (2024)
by: Su, Hang, et al.
Published: (2024)
CRPlace: Camera-Radar Fusion with BEV Representation for Place Recognition
by: Fu, Shaowei, et al.
Published: (2024)
by: Fu, Shaowei, et al.
Published: (2024)
Adversarial Constrained Policy Optimization: Improving Constrained Reinforcement Learning by Adapting Budgets
by: Ma, Jianmina, et al.
Published: (2024)
by: Ma, Jianmina, et al.
Published: (2024)
Do You Know the Way? Human-in-the-Loop Understanding for Fast Traversability Estimation in Mobile Robotics
by: Schreiber, Andre, et al.
Published: (2025)
by: Schreiber, Andre, et al.
Published: (2025)
Self-Supervised Geometry-Guided Initialization for Robust Monocular Visual Odometry
by: Kanai, Takayuki, et al.
Published: (2024)
by: Kanai, Takayuki, et al.
Published: (2024)
Improved Extrinsic Calibration of Acoustic Cameras via Batch Optimization
by: Li, Zhi, et al.
Published: (2025)
by: Li, Zhi, et al.
Published: (2025)
Opt-in Camera: Person Identification in Video via UWB Localization and Its Application to Opt-in Systems
by: Ishige, Matthew, et al.
Published: (2024)
by: Ishige, Matthew, et al.
Published: (2024)
Autonomous Field-of-View Adjustment Using Adaptive Kinematic Constrained Control with Robot-Held Microscopic Camera Feedback
by: Lin, Hung-Ching, et al.
Published: (2023)
by: Lin, Hung-Ching, et al.
Published: (2023)
Benchmarking Multi-View BEV Object Detection with Mixed Pinhole and Fisheye Cameras
by: Liu, Xiangzhong, et al.
Published: (2026)
by: Liu, Xiangzhong, et al.
Published: (2026)
What You Don't Know Can Hurt You: How Well do Latent Safety Filters Understand Partially Observable Safety Constraints?
by: Kim, Matthew, et al.
Published: (2025)
by: Kim, Matthew, et al.
Published: (2025)
CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection
by: Kuang, Zhaonian, et al.
Published: (2026)
by: Kuang, Zhaonian, et al.
Published: (2026)
Fiducial Exoskeletons: Image-Centric Robot State Estimation
by: Smith, Cameron, et al.
Published: (2026)
by: Smith, Cameron, et al.
Published: (2026)
CLAIM: Camera-LiDAR Alignment with Intensity and Monodepth
by: Zhang, Zhuo, et al.
Published: (2025)
by: Zhang, Zhuo, et al.
Published: (2025)
StereoNavNet: Learning to Navigate using Stereo Cameras with Auxiliary Occupancy Voxels
by: Li, Hongyu, et al.
Published: (2024)
by: Li, Hongyu, et al.
Published: (2024)
Deep Visual Odometry for Stereo Event Cameras
by: Zhong, Sheng, et al.
Published: (2025)
by: Zhong, Sheng, et al.
Published: (2025)
Two-Stage Camera Calibration Method for Multi-Camera Systems Using Scene Geometry
by: Abramov, Aleksandr
Published: (2025)
by: Abramov, Aleksandr
Published: (2025)
Moving On, Even When You're Broken: Fail-Active Trajectory Generation via Diffusion Policies Conditioned on Embodiment and Task
by: Briscoe-Martinez, Gilberto G., et al.
Published: (2026)
by: Briscoe-Martinez, Gilberto G., et al.
Published: (2026)
End-to-end Autonomous Vehicle Following System using Monocular Fisheye Camera
by: Zhang, Jiale, et al.
Published: (2025)
by: Zhang, Jiale, et al.
Published: (2025)
Similar Items
-
HapCompass: A Rotational Haptic Device for Contact-Rich Robotic Teleoperation
by: Tan, Xiangshan, et al.
Published: (2026) -
FlashBack: Consistency Model-Accelerated Shared Autonomy
by: Sun, Luzhe, et al.
Published: (2025) -
View-Invariant Policy Learning via Zero-Shot Novel View Synthesis
by: Tian, Stephen, et al.
Published: (2024) -
Statler: State-Maintaining Language Models for Embodied Reasoning
by: Yoneda, Takuma, et al.
Published: (2023) -
SplArt: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian Splatting
by: Lin, Shengjie, et al.
Published: (2025)