Observing and Controlling Features in Vision-Language-Action Models
Fuente:
arXiv
Saved in:
| Main Authors: | Buurmeijer, Hugo, Alonso, Carmen Amo, Swann, Aiden, Pavone, Marco |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sparse Autoencoders Reveal Interpretable and Steerable Features in VLA Models
by: Swann, Aiden, et al.
Published: (2026)
by: Swann, Aiden, et al.
Published: (2026)
Graph Neural Model Predictive Control for High-Dimensional Systems
by: Eberhard, Patrick Benito, et al.
Published: (2026)
by: Eberhard, Patrick Benito, et al.
Published: (2026)
Taming High-Dimensional Dynamics: Learning Optimal Projections onto Spectral Submanifolds
by: Buurmeijer, Hugo, et al.
Published: (2025)
by: Buurmeijer, Hugo, et al.
Published: (2025)
Learning Actuator-Aware Spectral Submanifolds for Precise Control of Continuum Robots
by: Wolff, Paul Leonard, et al.
Published: (2026)
by: Wolff, Paul Leonard, et al.
Published: (2026)
NARRATE: Versatile Language Architecture for Optimal Control in Robotics
by: Ismail, Seif, et al.
Published: (2024)
by: Ismail, Seif, et al.
Published: (2024)
Safe, Task-Consistent Manipulation with Operational Space Control Barrier Functions
by: Morton, Daniel, et al.
Published: (2025)
by: Morton, Daniel, et al.
Published: (2025)
RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models
by: Kwok, Jacky, et al.
Published: (2025)
by: Kwok, Jacky, et al.
Published: (2025)
Counterfactual VLA: Self-Reflective Vision-Language-Action Model with Adaptive Reasoning
by: Peng, Zhenghao "Mark", et al.
Published: (2025)
by: Peng, Zhenghao "Mark", et al.
Published: (2025)
DEMONSTRATE: Zero-shot Language to Robotic Control via Multi-task Demonstration Learning
by: Rickenbach, Rahel, et al.
Published: (2025)
by: Rickenbach, Rahel, et al.
Published: (2025)
Scaling Verification Can Be More Effective than Scaling Policy Learning for Vision-Language-Action Alignment
by: Kwok, Jacky, et al.
Published: (2026)
by: Kwok, Jacky, et al.
Published: (2026)
SAFER-Splat: A Control Barrier Function for Safe Navigation with Online Gaussian Splatting Maps
by: Chen, Timothy, et al.
Published: (2024)
by: Chen, Timothy, et al.
Published: (2024)
Semantic-Metric Bayesian Risk Fields: Learning Robot Safety from Human Videos with a VLM Prior
by: Chen, Timothy, et al.
Published: (2025)
by: Chen, Timothy, et al.
Published: (2025)
Not All Features Are Created Equal: A Mechanistic Study of Vision-Language-Action Models
by: Grant, Bryce, et al.
Published: (2026)
by: Grant, Bryce, et al.
Published: (2026)
Next Best Sense: Guiding Vision and Touch with FisherRF for 3D Gaussian Splatting
by: Strong, Matthew, et al.
Published: (2024)
by: Strong, Matthew, et al.
Published: (2024)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
by: Wang, Guokang, et al.
Published: (2024)
by: Wang, Guokang, et al.
Published: (2024)
frax: Fast Robot Kinematics and Dynamics in JAX
by: Morton, Daniel, et al.
Published: (2026)
by: Morton, Daniel, et al.
Published: (2026)
Constrained Hierarchical Monte Carlo Belief-State Planning
by: Jamgochian, Arec, et al.
Published: (2023)
by: Jamgochian, Arec, et al.
Published: (2023)
TensorTouch: Calibration of Tactile Sensors for High Resolution Stress Tensor and Deformation for Dexterous Manipulation
by: Do, Won Kyung, et al.
Published: (2025)
by: Do, Won Kyung, et al.
Published: (2025)
Health-Conditioned Vision-Language-Action Models for Malfunction-Aware Robot Control
by: Arslan, Hüseyin, et al.
Published: (2026)
by: Arslan, Hüseyin, et al.
Published: (2026)
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment
by: Kong, Weijie, et al.
Published: (2026)
by: Kong, Weijie, et al.
Published: (2026)
Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference
by: Liu, Yudong, et al.
Published: (2026)
by: Liu, Yudong, et al.
Published: (2026)
Space-LLaVA: a Vision-Language Model Adapted to Extraterrestrial Applications
by: Foutter, Matthew, et al.
Published: (2024)
by: Foutter, Matthew, et al.
Published: (2024)
Reshaping Action Error Distributions for Reliable Vision-Language-Action Models
by: Bai, Shuanghao, et al.
Published: (2026)
by: Bai, Shuanghao, et al.
Published: (2026)
Adaptive Action Chunking at Inference-time for Vision-Language-Action Models
by: Liang, Yuanchang, et al.
Published: (2026)
by: Liang, Yuanchang, et al.
Published: (2026)
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
by: Zhong, Linqing, et al.
Published: (2026)
by: Zhong, Linqing, et al.
Published: (2026)
A Survey on Vision-Language-Action Models: An Action Tokenization Perspective
by: Zhong, Yifan, et al.
Published: (2025)
by: Zhong, Yifan, et al.
Published: (2025)
Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment
by: Wulff, Theodor, et al.
Published: (2026)
by: Wulff, Theodor, et al.
Published: (2026)
Action Hallucination in Generative Vision-Language-Action Models
by: Soh, Harold, et al.
Published: (2026)
by: Soh, Harold, et al.
Published: (2026)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
by: Deng, Shengliang, et al.
Published: (2025)
by: Deng, Shengliang, et al.
Published: (2025)
Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control
by: Chen, William, et al.
Published: (2026)
by: Chen, William, et al.
Published: (2026)
LatBot: Distilling Universal Latent Actions for Vision-Language-Action Models
by: Li, Zuolei, et al.
Published: (2025)
by: Li, Zuolei, et al.
Published: (2025)
Bilevel MPC for Linear Systems: A Tractable Reduction and Continuous Connection to Hierarchical MPC
by: Moriyasu, Ryuta, et al.
Published: (2026)
by: Moriyasu, Ryuta, et al.
Published: (2026)
Embodiment Transfer Learning for Vision-Language-Action Models
by: Li, Chengmeng, et al.
Published: (2025)
by: Li, Chengmeng, et al.
Published: (2025)
Phys2Real: Fusing VLM Priors with Interactive Online Adaptation for Uncertainty-Aware Sim-to-Real Manipulation
by: Wang, Maggie, et al.
Published: (2025)
by: Wang, Maggie, et al.
Published: (2025)
FAST: Efficient Action Tokenization for Vision-Language-Action Models
by: Pertsch, Karl, et al.
Published: (2025)
by: Pertsch, Karl, et al.
Published: (2025)
TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments
by: Huang, Zhiyu, et al.
Published: (2026)
by: Huang, Zhiyu, et al.
Published: (2026)
ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models
by: Li, Puhao, et al.
Published: (2025)
by: Li, Puhao, et al.
Published: (2025)
Stable Language Guidance for Vision-Language-Action Models
by: Zhan, Zhihao, et al.
Published: (2026)
by: Zhan, Zhihao, et al.
Published: (2026)
OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction
by: Huang, Huang, et al.
Published: (2025)
by: Huang, Huang, et al.
Published: (2025)
Task-Driven Manipulation with Reconfigurable Parallel Robots
by: Morton, Daniel, et al.
Published: (2024)
by: Morton, Daniel, et al.
Published: (2024)
Similar Items
-
Sparse Autoencoders Reveal Interpretable and Steerable Features in VLA Models
by: Swann, Aiden, et al.
Published: (2026) -
Graph Neural Model Predictive Control for High-Dimensional Systems
by: Eberhard, Patrick Benito, et al.
Published: (2026) -
Taming High-Dimensional Dynamics: Learning Optimal Projections onto Spectral Submanifolds
by: Buurmeijer, Hugo, et al.
Published: (2025) -
Learning Actuator-Aware Spectral Submanifolds for Precise Control of Continuum Robots
by: Wolff, Paul Leonard, et al.
Published: (2026) -
NARRATE: Versatile Language Architecture for Optimal Control in Robotics
by: Ismail, Seif, et al.
Published: (2024)