Selective Perception for Robot: Task-Aware Attention in Multimodal VLA
Fuente:
arXiv
Saved in:
| Main Authors: | Son, Young-Chae, Lee, Jung-Woo, Choi, Yoon-Ji, Ko, Dae-Kwan, Lim, Soo-Chul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making
by: Son, Young-Chae, et al.
Published: (2026)
by: Son, Young-Chae, et al.
Published: (2026)
MATT-GS: Masked Attention-based 3DGS for Robot Perception and Object Detection
by: Lee, Jee Won, et al.
Published: (2025)
by: Lee, Jee Won, et al.
Published: (2025)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
by: Guo, Heyu, et al.
Published: (2025)
by: Guo, Heyu, et al.
Published: (2025)
ProgVLA: Progress-Aware Robot Manipulation Skill Learning
by: Kim, Seungsu, et al.
Published: (2026)
by: Kim, Seungsu, et al.
Published: (2026)
Contact Tooling Manipulation Control for Robotic Repair Platform
by: Lee, Joong-Ku, et al.
Published: (2024)
by: Lee, Joong-Ku, et al.
Published: (2024)
Pixel2Catch: Multi-Agent Sim-to-Real Transfer for Agile Manipulation with a Single RGB Camera
by: Kim, Seongyong, et al.
Published: (2026)
by: Kim, Seongyong, et al.
Published: (2026)
Uncertainty-Aware Multi-Robot Task Allocation With Strongly Coupled Inter-Robot Rewards
by: Rossano, Ben, et al.
Published: (2025)
by: Rossano, Ben, et al.
Published: (2025)
VLM-driven Skill Selection for Robotic Assembly Tasks
by: Kim, Jeong-Jung, et al.
Published: (2025)
by: Kim, Jeong-Jung, et al.
Published: (2025)
Dual-Arm Telerobotic Platform for Robotic Hotbox Operations for Nuclear Waste Disposition in EM Sites
by: Lee, Joong-Ku, et al.
Published: (2024)
by: Lee, Joong-Ku, et al.
Published: (2024)
DexTouch: Learning to Seek and Manipulate Objects with Tactile Dexterity
by: Lee, Kang-Won, et al.
Published: (2024)
by: Lee, Kang-Won, et al.
Published: (2024)
Teaching Robots to Handle Nuclear Waste: A Teleoperation-Based Learning Approach<
by: Lee, Joong-Ku, et al.
Published: (2025)
by: Lee, Joong-Ku, et al.
Published: (2025)
TGM-VLA: Task-Guided Mixup for Sampling-Efficient and Robust Robotic Manipulation
by: Pu, Fanqi, et al.
Published: (2026)
by: Pu, Fanqi, et al.
Published: (2026)
AnyCamVLA: Zero-Shot Camera Adaptation for Viewpoint Robust Vision-Language-Action Models
by: Heo, Hyeongjun, et al.
Published: (2026)
by: Heo, Hyeongjun, et al.
Published: (2026)
SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models
by: Li, Meng, et al.
Published: (2025)
by: Li, Meng, et al.
Published: (2025)
Attention-Based Neural-Augmented Kalman Filter for Legged Robot State Estimation
by: Lee, Seokju, et al.
Published: (2026)
by: Lee, Seokju, et al.
Published: (2026)
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization
by: Jia, Xiaosong, et al.
Published: (2026)
by: Jia, Xiaosong, et al.
Published: (2026)
LiteVLA-Edge: Quantized On-Device Multimodal Control for Embedded Robotics
by: Williams, Justin, et al.
Published: (2026)
by: Williams, Justin, et al.
Published: (2026)
Predicting Human Perceptions of Robot Performance During Navigation Tasks
by: Zhang, Qiping, et al.
Published: (2023)
by: Zhang, Qiping, et al.
Published: (2023)
ST-VLA: Enabling 4D-Aware Spatiotemporal Understanding for General Robot Manipulation
by: Wu, You, et al.
Published: (2026)
by: Wu, You, et al.
Published: (2026)
Terrain-Aware Model Predictive Control of Heterogeneous Bipedal and Aerial Robot Coordination for Search and Rescue Tasks
by: Shamsah, Abdulaziz, et al.
Published: (2024)
by: Shamsah, Abdulaziz, et al.
Published: (2024)
Reactive Model Predictive Contouring Control for Robot Manipulators
by: Yoon, Junheon, et al.
Published: (2025)
by: Yoon, Junheon, et al.
Published: (2025)
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
by: Wei, Xiangyi, et al.
Published: (2025)
by: Wei, Xiangyi, et al.
Published: (2025)
ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning
by: Van Vo, Tuan, et al.
Published: (2026)
by: Van Vo, Tuan, et al.
Published: (2026)
ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation
by: Liu, Zhenyang, et al.
Published: (2026)
by: Liu, Zhenyang, et al.
Published: (2026)
SimVLA: A Simple VLA Baseline for Robotic Manipulation
by: Luo, Yuankai, et al.
Published: (2026)
by: Luo, Yuankai, et al.
Published: (2026)
Task-Aware Positioning for Improvisational Tasks in Mobile Construction Robots via an AI Agent with Multi-LMM Modules
by: Jang, Seongju, et al.
Published: (2026)
by: Jang, Seongju, et al.
Published: (2026)
Designing Robot Identity: The Role of Voice, Clothing, and Task on Robot Gender Perception
by: Dennler, Nathaniel S., et al.
Published: (2024)
by: Dennler, Nathaniel S., et al.
Published: (2024)
GUARD: Toward a Compromise between Traditional Control and Learning for Safe Robot Systems
by: Gaus, Johannes A., et al.
Published: (2025)
by: Gaus, Johannes A., et al.
Published: (2025)
ManualVLA: A Unified VLA Model for Chain-of-Thought Manual Generation and Robotic Manipulation
by: Gu, Chenyang, et al.
Published: (2025)
by: Gu, Chenyang, et al.
Published: (2025)
DAM-VLA: A Dynamic Action Model-Based Vision-Language-Action Framework for Robot Manipulation
by: Peng, Xiongfeng, et al.
Published: (2026)
by: Peng, Xiongfeng, et al.
Published: (2026)
Reflection-Based Task Adaptation for Self-Improving VLA
by: Li, Baicheng, et al.
Published: (2025)
by: Li, Baicheng, et al.
Published: (2025)
Phase-Aware Policy Learning for Skateboard Riding of Quadruped Robots via Feature-wise Linear Modulation
by: Yoon, Minsung, et al.
Published: (2026)
by: Yoon, Minsung, et al.
Published: (2026)
TSP-Bot: Robotic TSP Pen Art using High-DoF Manipulators
by: Song, Daeun, et al.
Published: (2022)
by: Song, Daeun, et al.
Published: (2022)
ConsisVLA-4D: Advancing Spatiotemporal Consistency in Efficient 3D-Perception and 4D-Reasoning for Robotic Manipulation
by: Li, Wei, et al.
Published: (2026)
by: Li, Wei, et al.
Published: (2026)
Multimodal Behaviour Trees for Robotic Laboratory Task Automation
by: Fakhruldeen, Hatem, et al.
Published: (2025)
by: Fakhruldeen, Hatem, et al.
Published: (2025)
GazeVLA: Learning Human Intention for Robotic Manipulation
by: Li, Chengyang, et al.
Published: (2026)
by: Li, Chengyang, et al.
Published: (2026)
Sci-VLA: Agentic VLA Inference Plugin for Long-Horizon Tasks in Scientific Experiments
by: Pang, Yiwen, et al.
Published: (2026)
by: Pang, Yiwen, et al.
Published: (2026)
Multi-Robot Motion Planning from Vision and Language using Heat-Inspired Diffusion
by: Chae, Jebeom, et al.
Published: (2025)
by: Chae, Jebeom, et al.
Published: (2025)
Gaze-based Human-Robot Interaction System for Infrastructure Inspections
by: Choi, Sunwoong, et al.
Published: (2024)
by: Choi, Sunwoong, et al.
Published: (2024)
SaWa-ML: Structure-Aware Pose Correction and Weight Adaptation-Based Robust Multi-Robot Localization
by: Choi, Junho, et al.
Published: (2025)
by: Choi, Junho, et al.
Published: (2025)
Similar Items
-
ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making
by: Son, Young-Chae, et al.
Published: (2026) -
MATT-GS: Masked Attention-based 3DGS for Robot Perception and Object Detection
by: Lee, Jee Won, et al.
Published: (2025) -
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
by: Guo, Heyu, et al.
Published: (2025) -
ProgVLA: Progress-Aware Robot Manipulation Skill Learning
by: Kim, Seungsu, et al.
Published: (2026) -
Contact Tooling Manipulation Control for Robotic Repair Platform
by: Lee, Joong-Ku, et al.
Published: (2024)