AnyCamVLA: Zero-Shot Camera Adaptation for Viewpoint Robust Vision-Language-Action Models
Fuente:
arXiv
Saved in:
| Main Authors: | Heo, Hyeongjun, Woo, Seungyeon, Kim, Sang Min, Kim, Junho, Lee, Junho, Lee, Yonghyeon, Kim, Young Min |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Point2Act: Efficient 3D Distillation of Multimodal LLMs for Zero-Shot Context-Aware Grasping
by: Kim, Sang Min, et al.
Published: (2025)
by: Kim, Sang Min, et al.
Published: (2025)
DreamGrasp: Zero-Shot 3D Multi-Object Reconstruction from Partial-View Images for Robotic Manipulation
by: Kim, Young Hun, et al.
Published: (2025)
by: Kim, Young Hun, et al.
Published: (2025)
Motion Manifold Flow Primitives for Task-Conditioned Trajectory Generation under Complex Task-Motion Dependencies
by: Lee, Yonghyeon, et al.
Published: (2024)
by: Lee, Yonghyeon, et al.
Published: (2024)
ScrewSplat: An End-to-End Method for Articulated Object Recognition
by: Kim, Seungyeon, et al.
Published: (2025)
by: Kim, Seungyeon, et al.
Published: (2025)
RoEL: Robust Event-based 3D Line Reconstruction
by: Bae, Gwangtak, et al.
Published: (2026)
by: Bae, Gwangtak, et al.
Published: (2026)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
by: Kim, Ju-Young, et al.
Published: (2025)
by: Kim, Ju-Young, et al.
Published: (2025)
Scanning Bot: Efficient Scan Planning using Panoramic Cameras
by: Lee, Euijeong, et al.
Published: (2025)
by: Lee, Euijeong, et al.
Published: (2025)
Why Look at It at All?: Vision-Free Multifingered Blind Grasping Using Uniaxial Fingertip Force Sensing
by: Lee, Edgar, et al.
Published: (2026)
by: Lee, Edgar, et al.
Published: (2026)
SaWa-ML: Structure-Aware Pose Correction and Weight Adaptation-Based Robust Multi-Robot Localization
by: Choi, Junho, et al.
Published: (2025)
by: Choi, Junho, et al.
Published: (2025)
MoS-VLA: A Vision-Language-Action Model with One-Shot Skill Adaptation
by: Zhao, Ruihan, et al.
Published: (2025)
by: Zhao, Ruihan, et al.
Published: (2025)
Finding 3D Scene Analogies with Multimodal Foundation Models
by: Kim, Junho, et al.
Published: (2025)
by: Kim, Junho, et al.
Published: (2025)
I$^2$-SLAM: Inverting Imaging Process for Robust Photorealistic Dense SLAM
by: Bae, Gwangtak, et al.
Published: (2024)
by: Bae, Gwangtak, et al.
Published: (2024)
Calibrating Panoramic Depth Estimation for Practical Localization and Mapping
by: Kim, Junho, et al.
Published: (2023)
by: Kim, Junho, et al.
Published: (2023)
CleaR: Towards Robust and Generalized Parameter-Efficient Fine-Tuning for Noisy Label Learning
by: Kim, Yeachan, et al.
Published: (2024)
by: Kim, Yeachan, et al.
Published: (2024)
LTGS: Long-Term Gaussian Scene Chronology From Sparse View Updates
by: Kim, Minkwan, et al.
Published: (2025)
by: Kim, Minkwan, et al.
Published: (2025)
AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement
by: Hu, Zhaofeng, et al.
Published: (2026)
by: Hu, Zhaofeng, et al.
Published: (2026)
Differentiable Motion Manifold Primitives for Reactive Motion Generation under Kinodynamic Constraints
by: Lee, Yonghyeon
Published: (2024)
by: Lee, Yonghyeon
Published: (2024)
HBRB-BoW: A Retrained Bag-of-Words Vocabulary for ORB-SLAM via Hierarchical BRB-KMeans
by: Lee, Minjae, et al.
Published: (2026)
by: Lee, Minjae, et al.
Published: (2026)
PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models
by: Guo, Xinyu, et al.
Published: (2026)
by: Guo, Xinyu, et al.
Published: (2026)
Depth Any Camera: Zero-Shot Metric Depth Estimation from Any Camera
by: Guo, Yuliang, et al.
Published: (2025)
by: Guo, Yuliang, et al.
Published: (2025)
MMP++: Motion Manifold Primitives with Parametric Curve Models
by: Lee, Yonghyeon
Published: (2023)
by: Lee, Yonghyeon
Published: (2023)
High-Bandwidth Tactile-Reactive Control for Grasp Adjustment
by: Lee, Yonghyeon, et al.
Published: (2025)
by: Lee, Yonghyeon, et al.
Published: (2025)
Hierarchical Reactive Grasping via Task-Space Velocity Fields and Joint-Space Quadratic Programming
by: Lee, Yonghyeon, et al.
Published: (2025)
by: Lee, Yonghyeon, et al.
Published: (2025)
Mentor-KD: Making Small Language Models Better Multi-step Reasoners
by: Lee, Hojae, et al.
Published: (2024)
by: Lee, Hojae, et al.
Published: (2024)
RetoVLA: Reusing Register Tokens for Spatial Reasoning in Vision-Language-Action Models
by: Koo, Jiyeon, et al.
Published: (2025)
by: Koo, Jiyeon, et al.
Published: (2025)
Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm
by: Kim, Hyeonjun, et al.
Published: (2025)
by: Kim, Hyeonjun, et al.
Published: (2025)
CPO: Change Robust Panorama to Point Cloud Localization
by: Kim, Junho, et al.
Published: (2022)
by: Kim, Junho, et al.
Published: (2022)
StableVLA: Towards Robust Vision-Language-Action Models without Extra Data
by: Fu, Yiyang, et al.
Published: (2026)
by: Fu, Yiyang, et al.
Published: (2026)
Fully Geometric Panoramic Localization
by: Kim, Junho, et al.
Published: (2024)
by: Kim, Junho, et al.
Published: (2024)
Adaptive Capacity Allocation for Vision Language Action Fine-tuning
by: Kim, Donghoon, et al.
Published: (2026)
by: Kim, Donghoon, et al.
Published: (2026)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
by: Ding, Pengxiang, et al.
Published: (2023)
by: Ding, Pengxiang, et al.
Published: (2023)
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
by: Luo, Jingzhou, et al.
Published: (2026)
by: Luo, Jingzhou, et al.
Published: (2026)
Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models
by: Jin, Ruofan, et al.
Published: (2026)
by: Jin, Ruofan, et al.
Published: (2026)
Enhancing Rating-Based Reinforcement Learning to Effectively Leverage Feedback from Large Vision-Language Models
by: Luu, Tung Minh, et al.
Published: (2025)
by: Luu, Tung Minh, et al.
Published: (2025)
OpenVLA: An Open-Source Vision-Language-Action Model
by: Kim, Moo Jin, et al.
Published: (2024)
by: Kim, Moo Jin, et al.
Published: (2024)
Frequency Response Data-Driven Disturbance Observer Design for Flexible Joint Robots
by: Lee, Deokjin, et al.
Published: (2025)
by: Lee, Deokjin, et al.
Published: (2025)
Learning 3D Scene Analogies with Neural Contextual Scene Maps
by: Kim, Junho, et al.
Published: (2025)
by: Kim, Junho, et al.
Published: (2025)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
by: Zhang, Hongyin, et al.
Published: (2025)
by: Zhang, Hongyin, et al.
Published: (2025)
TacVLA: Contact-Aware Tactile Fusion for Robust Vision-Language-Action Manipulation
by: Zhang, Kaidi, et al.
Published: (2026)
by: Zhang, Kaidi, et al.
Published: (2026)
DynaVINS++: Robust Visual-Inertial State Estimator in Dynamic Environments by Adaptive Truncated Least Squares and Stable State Recovery
by: Song, Seungwon, et al.
Published: (2024)
by: Song, Seungwon, et al.
Published: (2024)
Similar Items
-
Point2Act: Efficient 3D Distillation of Multimodal LLMs for Zero-Shot Context-Aware Grasping
by: Kim, Sang Min, et al.
Published: (2025) -
DreamGrasp: Zero-Shot 3D Multi-Object Reconstruction from Partial-View Images for Robotic Manipulation
by: Kim, Young Hun, et al.
Published: (2025) -
Motion Manifold Flow Primitives for Task-Conditioned Trajectory Generation under Complex Task-Motion Dependencies
by: Lee, Yonghyeon, et al.
Published: (2024) -
ScrewSplat: An End-to-End Method for Articulated Object Recognition
by: Kim, Seungyeon, et al.
Published: (2025) -
RoEL: Robust Event-based 3D Line Reconstruction
by: Bae, Gwangtak, et al.
Published: (2026)