AIR-VLA: Vision-Language-Action Systems for Aerial Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Jianli, Tian, Bin, Zhang, Qiyao, Li, Chengxiang, Song, Zihan, Cui, Zhiyong, Lv, Yisheng, Tian, Yonglin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models
von: Zhang, Qiyao, et al.
Veröffentlicht: (2026)
von: Zhang, Qiyao, et al.
Veröffentlicht: (2026)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
von: Zhao, Han, et al.
Veröffentlicht: (2025)
von: Zhao, Han, et al.
Veröffentlicht: (2025)
UnderwaterVLA: Dual-brain Vision-Language-Action architecture for Autonomous Underwater Navigation
von: Wang, Zhangyuan, et al.
Veröffentlicht: (2025)
von: Wang, Zhangyuan, et al.
Veröffentlicht: (2025)
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
von: Hu, Yucheng, et al.
Veröffentlicht: (2026)
von: Hu, Yucheng, et al.
Veröffentlicht: (2026)
TacVLA: Contact-Aware Tactile Fusion for Robust Vision-Language-Action Manipulation
von: Zhang, Kaidi, et al.
Veröffentlicht: (2026)
von: Zhang, Kaidi, et al.
Veröffentlicht: (2026)
DroneVLA: VLA based Aerial Manipulation
von: Mehboob, Fawad, et al.
Veröffentlicht: (2026)
von: Mehboob, Fawad, et al.
Veröffentlicht: (2026)
DAM-VLA: A Dynamic Action Model-Based Vision-Language-Action Framework for Robot Manipulation
von: Peng, Xiongfeng, et al.
Veröffentlicht: (2026)
von: Peng, Xiongfeng, et al.
Veröffentlicht: (2026)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
VLA-AN: An Efficient and Onboard Vision-Language-Action Framework for Aerial Navigation in Complex Environments
von: Wu, Yuze, et al.
Veröffentlicht: (2025)
von: Wu, Yuze, et al.
Veröffentlicht: (2025)
RationalVLA: A Rational Vision-Language-Action Model with Dual System
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
von: Deng, Shengliang, et al.
Veröffentlicht: (2025)
von: Deng, Shengliang, et al.
Veröffentlicht: (2025)
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
von: Gbagbe, Koffivi Fidèle, et al.
Veröffentlicht: (2024)
von: Gbagbe, Koffivi Fidèle, et al.
Veröffentlicht: (2024)
SG-VLA: Learning Spatially-Grounded Vision-Language-Action Models for Mobile Manipulation
von: Tu, Ruisen, et al.
Veröffentlicht: (2026)
von: Tu, Ruisen, et al.
Veröffentlicht: (2026)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
von: Miao, Cui, et al.
Veröffentlicht: (2025)
von: Miao, Cui, et al.
Veröffentlicht: (2025)
GeoAware-VLA: Implicit Geometry Aware Vision-Language-Action Model
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
von: Khan, Muhamamd Haris, et al.
Veröffentlicht: (2025)
von: Khan, Muhamamd Haris, et al.
Veröffentlicht: (2025)
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
von: Sun, Lin, et al.
Veröffentlicht: (2025)
von: Sun, Lin, et al.
Veröffentlicht: (2025)
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
von: Wei, Xiangyi, et al.
Veröffentlicht: (2025)
von: Wei, Xiangyi, et al.
Veröffentlicht: (2025)
CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
UrbanVLA: A Vision-Language-Action Model for Urban Micromobility
von: Li, Anqi, et al.
Veröffentlicht: (2025)
von: Li, Anqi, et al.
Veröffentlicht: (2025)
SELF-VLA: A Skill Enhanced Agentic Vision-Language-Action Framework for Contact-Rich Disassembly
von: Liu, Chang, et al.
Veröffentlicht: (2026)
von: Liu, Chang, et al.
Veröffentlicht: (2026)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
TaF-VLA: Tactile-Force Alignment in Vision-Language-Action Models for Force-aware Manipulation
von: Huang, Yuzhe, et al.
Veröffentlicht: (2026)
von: Huang, Yuzhe, et al.
Veröffentlicht: (2026)
SteerVLA: Steering Vision-Language-Action Models in Long-Tail Driving Scenarios
von: Gao, Tian, et al.
Veröffentlicht: (2026)
von: Gao, Tian, et al.
Veröffentlicht: (2026)
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
von: Yu, Wenda, et al.
Veröffentlicht: (2026)
von: Yu, Wenda, et al.
Veröffentlicht: (2026)
Offline Reinforcement Learning with Discrete Diffusion Skills
von: Qiao, RuiXi, et al.
Veröffentlicht: (2025)
von: Qiao, RuiXi, et al.
Veröffentlicht: (2025)
VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
Counterfactual VLA: Self-Reflective Vision-Language-Action Model with Adaptive Reasoning
von: Peng, Zhenghao "Mark", et al.
Veröffentlicht: (2025)
von: Peng, Zhenghao "Mark", et al.
Veröffentlicht: (2025)
Long-VLA: Unleashing Long-Horizon Capability of Vision Language Action Model for Robot Manipulation
von: Fan, Yiguo, et al.
Veröffentlicht: (2025)
von: Fan, Yiguo, et al.
Veröffentlicht: (2025)
PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation
von: Su, Taiyi, et al.
Veröffentlicht: (2026)
von: Su, Taiyi, et al.
Veröffentlicht: (2026)
CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
CollabVLA: Self-Reflective Vision-Language-Action Model Dreaming Together with Human
von: Sun, Nan, et al.
Veröffentlicht: (2025)
von: Sun, Nan, et al.
Veröffentlicht: (2025)
PokeVLA: Empowering Pocket-Sized Vision-Language-Action Model with Comprehensive World Knowledge Guidance
von: Zheng, Yupeng, et al.
Veröffentlicht: (2026)
von: Zheng, Yupeng, et al.
Veröffentlicht: (2026)
AnchorVLA4D: an Anchor-Based Spatial-Temporal Vision-Language-Action Model for Robotic Manipulation
von: Zhu, Juan, et al.
Veröffentlicht: (2026)
von: Zhu, Juan, et al.
Veröffentlicht: (2026)
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
MAP-VLA: Memory-Augmented Prompting for Vision-Language-Action Model in Robotic Manipulation
von: Li, Runhao, et al.
Veröffentlicht: (2025)
von: Li, Runhao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models
von: Zhang, Qiyao, et al.
Veröffentlicht: (2026) -
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
von: Yang, Shuai, et al.
Veröffentlicht: (2025) -
VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
von: Zhao, Han, et al.
Veröffentlicht: (2025) -
UnderwaterVLA: Dual-brain Vision-Language-Action architecture for Autonomous Underwater Navigation
von: Wang, Zhangyuan, et al.
Veröffentlicht: (2025) -
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
von: Hu, Yucheng, et al.
Veröffentlicht: (2026)