Overcoming Visual Clutter in Vision Language Action Models via Concept-Gated Visual Distillation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Sangmim, Kodagoda, Sarath, Carmichael, Marc, Thiyagarajan, Karthick |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Guide-LLM: An Embodied LLM Agent and Text-Based Topological Map for Robotic Guidance of People with Visual Impairments
von: Song, Sangmim, et al.
Veröffentlicht: (2024)
von: Song, Sangmim, et al.
Veröffentlicht: (2024)
Observability Conditions and Filter Design for Visual Pose Estimation via Dual Quaternions
von: Andrews, Nicholas B., et al.
Veröffentlicht: (2026)
von: Andrews, Nicholas B., et al.
Veröffentlicht: (2026)
Inline Photometrically Calibrated Hybrid Visual SLAM
von: Abboud, Nicolas, et al.
Veröffentlicht: (2024)
von: Abboud, Nicolas, et al.
Veröffentlicht: (2024)
TRACE: A Self-Improving Framework for Robot Behavior Forecasting with Vision-Language Models
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2025)
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2025)
Visual Servoing for Robotic On-Orbit Servicing: A Survey
von: Amaya-Mejía, Lina María, et al.
Veröffentlicht: (2024)
von: Amaya-Mejía, Lina María, et al.
Veröffentlicht: (2024)
Hybrid Visual Servoing of Tendon-driven Continuum Robots
von: Danesh, Rana, et al.
Veröffentlicht: (2025)
von: Danesh, Rana, et al.
Veröffentlicht: (2025)
Design of a Visual Pose Estimation Algorithm for Moon Landing
von: Süslü, Atakan, et al.
Veröffentlicht: (2025)
von: Süslü, Atakan, et al.
Veröffentlicht: (2025)
Lite VLA: Efficient Vision-Language-Action Control on CPU-Bound Edge Robots
von: Williams, Justin, et al.
Veröffentlicht: (2025)
von: Williams, Justin, et al.
Veröffentlicht: (2025)
Enhancing Feature Tracking Reliability for Visual Navigation using Real-Time Safety Filter
von: Kim, Dabin, et al.
Veröffentlicht: (2025)
von: Kim, Dabin, et al.
Veröffentlicht: (2025)
An Intelligent Water-Saving Irrigation System Based on Multi-Sensor Fusion and Visual Servoing Control
von: Huang, ZhengKai, et al.
Veröffentlicht: (2025)
von: Huang, ZhengKai, et al.
Veröffentlicht: (2025)
VISION-SLS: Safe Perception-Based Control from Learned Visual Representations via System Level Synthesis
von: Leeman, Antoine P., et al.
Veröffentlicht: (2026)
von: Leeman, Antoine P., et al.
Veröffentlicht: (2026)
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)
Visual Heading Prediction for Autonomous Aerial Vehicles
von: Ahmari, Reza, et al.
Veröffentlicht: (2025)
von: Ahmari, Reza, et al.
Veröffentlicht: (2025)
Vision-Based Adaptive Robotics for Autonomous Surface Crack Repair
von: Genova, Joshua, et al.
Veröffentlicht: (2024)
von: Genova, Joshua, et al.
Veröffentlicht: (2024)
Tube-Based Robust Control Strategy for Vision-Guided Autonomous Vehicles
von: Lee, Der-Hau
Veröffentlicht: (2025)
von: Lee, Der-Hau
Veröffentlicht: (2025)
Vision-based control for landing an aerial vehicle on a marine vessel
von: Dong, Haohua
Veröffentlicht: (2024)
von: Dong, Haohua
Veröffentlicht: (2024)
Continuous Wrist Control on the Hannes Prosthesis: a Vision-based Shared Autonomy Framework
von: Vasile, Federico, et al.
Veröffentlicht: (2025)
von: Vasile, Federico, et al.
Veröffentlicht: (2025)
Vision-Based System Identification of a Quadrotor
von: Iz, Selim Ahmet, et al.
Veröffentlicht: (2025)
von: Iz, Selim Ahmet, et al.
Veröffentlicht: (2025)
SOUS VIDE: Cooking Visual Drone Navigation Policies in a Gaussian Splatting Vacuum
von: Low, JunEn, et al.
Veröffentlicht: (2024)
von: Low, JunEn, et al.
Veröffentlicht: (2024)
Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient
von: You, Haoxiang, et al.
Veröffentlicht: (2026)
von: You, Haoxiang, et al.
Veröffentlicht: (2026)
PVI: Plug-in Visual Injection for Vision-Language-Action Models
von: Zhang, Zezhou, et al.
Veröffentlicht: (2026)
von: Zhang, Zezhou, et al.
Veröffentlicht: (2026)
Mitigating Error Accumulation in Continuous Navigation via Memory-Augmented Kalman Filtering
von: Tang, Yin, et al.
Veröffentlicht: (2026)
von: Tang, Yin, et al.
Veröffentlicht: (2026)
LIR-LIVO: A Lightweight,Robust LiDAR/Vision/Inertial Odometry with Illumination-Resilient Deep Features
von: Zhou, Shujie, et al.
Veröffentlicht: (2025)
von: Zhou, Shujie, et al.
Veröffentlicht: (2025)
Designing Latent Safety Filters using Pre-Trained Vision Models
von: Tabbara, Ihab, et al.
Veröffentlicht: (2025)
von: Tabbara, Ihab, et al.
Veröffentlicht: (2025)
Improving Robustness of Vision-Language-Action Models by Restoring Corrupted Visual Inputs
von: Orjuela, Daniel Yezid Guarnizo, et al.
Veröffentlicht: (2026)
von: Orjuela, Daniel Yezid Guarnizo, et al.
Veröffentlicht: (2026)
Diffusion Dynamics Models with Generative State Estimation for Cloth Manipulation
von: Tian, Tongxuan, et al.
Veröffentlicht: (2025)
von: Tian, Tongxuan, et al.
Veröffentlicht: (2025)
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models
von: Ye, Wencheng, et al.
Veröffentlicht: (2025)
von: Ye, Wencheng, et al.
Veröffentlicht: (2025)
NavigScene: Bridging Local Perception and Global Navigation for Beyond-Visual-Range Autonomous Driving
von: Peng, Qucheng, et al.
Veröffentlicht: (2025)
von: Peng, Qucheng, et al.
Veröffentlicht: (2025)
Leave No Observation Behind: Real-time Correction for VLA Action Chunks
von: Sendai, Kohei, et al.
Veröffentlicht: (2025)
von: Sendai, Kohei, et al.
Veröffentlicht: (2025)
OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction
von: Huang, Huang, et al.
Veröffentlicht: (2025)
von: Huang, Huang, et al.
Veröffentlicht: (2025)
DualCoT-VLA: Visual-Linguistic Chain of Thought via Parallel Reasoning for Vision-Language-Action Models
von: Zhong, Zhide, et al.
Veröffentlicht: (2026)
von: Zhong, Zhide, et al.
Veröffentlicht: (2026)
ObjectReact: Learning Object-Relative Control for Visual Navigation
von: Garg, Sourav, et al.
Veröffentlicht: (2025)
von: Garg, Sourav, et al.
Veröffentlicht: (2025)
Detecting and Mitigating System-Level Anomalies of Vision-Based Controllers
von: Gupta, Aryaman, et al.
Veröffentlicht: (2023)
von: Gupta, Aryaman, et al.
Veröffentlicht: (2023)
Vision-Based Safety System for Barrierless Human-Robot Collaboration
von: Amaya-Mejía, Lina María, et al.
Veröffentlicht: (2022)
von: Amaya-Mejía, Lina María, et al.
Veröffentlicht: (2022)
Test-Time Training for Visual Foresight Vision-Language-Action Models
von: Park, Sangwu, et al.
Veröffentlicht: (2026)
von: Park, Sangwu, et al.
Veröffentlicht: (2026)
Vision-Based Runtime Monitoring under Varying Specifications using Semantic Latent Representations
von: Hoxha, Bardh, et al.
Veröffentlicht: (2026)
von: Hoxha, Bardh, et al.
Veröffentlicht: (2026)
AI for Green Spaces: Leveraging Autonomous Navigation and Computer Vision for Park Litter Removal
von: Kao, Christopher, et al.
Veröffentlicht: (2026)
von: Kao, Christopher, et al.
Veröffentlicht: (2026)
Tiny-DroNeRF: Tiny Neural Radiance Fields aboard Federated Learning-enabled Nano-drones
von: Carboni, Ilenia, et al.
Veröffentlicht: (2026)
von: Carboni, Ilenia, et al.
Veröffentlicht: (2026)
ADAPT: An Autonomous Forklift for Construction Site Operation
von: Huemer, Johannes, et al.
Veröffentlicht: (2025)
von: Huemer, Johannes, et al.
Veröffentlicht: (2025)
Holistic Fusion: Task- and Setup-Agnostic Robot Localization and State Estimation with Factor Graphs
von: Nubert, Julian, et al.
Veröffentlicht: (2025)
von: Nubert, Julian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Guide-LLM: An Embodied LLM Agent and Text-Based Topological Map for Robotic Guidance of People with Visual Impairments
von: Song, Sangmim, et al.
Veröffentlicht: (2024) -
Observability Conditions and Filter Design for Visual Pose Estimation via Dual Quaternions
von: Andrews, Nicholas B., et al.
Veröffentlicht: (2026) -
Inline Photometrically Calibrated Hybrid Visual SLAM
von: Abboud, Nicolas, et al.
Veröffentlicht: (2024) -
TRACE: A Self-Improving Framework for Robot Behavior Forecasting with Vision-Language Models
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2025) -
Visual Servoing for Robotic On-Orbit Servicing: A Survey
von: Amaya-Mejía, Lina María, et al.
Veröffentlicht: (2024)