LiteVLA-Edge: Quantized On-Device Multimodal Control for Embedded Robotics
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Williams, Justin, Gupta, Kishor Datta, George, Roy, Sarkar, Mrinmoy |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Lite VLA: Efficient Vision-Language-Action Control on CPU-Bound Edge Robots
par: Williams, Justin, et autres
Publié: (2025)
par: Williams, Justin, et autres
Publié: (2025)
LiteVLA-H: Dual-Rate Vision-Language-Action Inference for Onboard Aerial Guidance and Semantic Perception
par: williams, Justin, et autres
Publié: (2026)
par: williams, Justin, et autres
Publié: (2026)
EdgeNav-QE: QLoRA Quantization and Dynamic Early Exit for LAM-based Navigation on Edge Devices
par: Liu, Mengyun, et autres
Publié: (2026)
par: Liu, Mengyun, et autres
Publié: (2026)
ReconVLA: An Uncertainty-Guided and Failure-Aware Vision-Language-Action Framework for Robotic Control
par: Chen, Lingling, et autres
Publié: (2026)
par: Chen, Lingling, et autres
Publié: (2026)
EvoScene-VLA: Evolving Scene Beliefs Inside the Action Decoder for Chunked Robot Control
par: Zhang, Chushan, et autres
Publié: (2026)
par: Zhang, Chushan, et autres
Publié: (2026)
VacuumVLA: Boosting VLA Capabilities via a Unified Suction and Gripping Tool for Complex Robotic Manipulation
par: Zhou, Hui, et autres
Publié: (2025)
par: Zhou, Hui, et autres
Publié: (2025)
Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs
par: Chiang, Hao-Tien Lewis, et autres
Publié: (2024)
par: Chiang, Hao-Tien Lewis, et autres
Publié: (2024)
DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
par: Yue, Yang, et autres
Publié: (2024)
par: Yue, Yang, et autres
Publié: (2024)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
par: Li, Haoyun, et autres
Publié: (2025)
par: Li, Haoyun, et autres
Publié: (2025)
Towards Accessible Physical AI: LoRA-Based Fine-Tuning of VLA Models for Real-World Robot Control
par: Omaisan, Abdullah Yahya Abdullah, et autres
Publié: (2025)
par: Omaisan, Abdullah Yahya Abdullah, et autres
Publié: (2025)
DroneVLA: VLA based Aerial Manipulation
par: Mehboob, Fawad, et autres
Publié: (2026)
par: Mehboob, Fawad, et autres
Publié: (2026)
AsyncShield: A Plug-and-Play Edge Adapter for Asynchronous Cloud-based VLA Navigation
par: Yang, Kai, et autres
Publié: (2026)
par: Yang, Kai, et autres
Publié: (2026)
VLA-RAIL: A Real-Time Asynchronous Inference Linker for VLA Models and Robots
par: Zhao, Yongsheng, et autres
Publié: (2025)
par: Zhao, Yongsheng, et autres
Publié: (2025)
ProgressVLA: Progress-Guided Diffusion Policy for Vision-Language Robotic Manipulation
par: Yan, Hongyu, et autres
Publié: (2026)
par: Yan, Hongyu, et autres
Publié: (2026)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
par: Lu, Guanxing, et autres
Publié: (2025)
par: Lu, Guanxing, et autres
Publié: (2025)
Jailbreaking LLM-Controlled Robots
par: Robey, Alexander, et autres
Publié: (2024)
par: Robey, Alexander, et autres
Publié: (2024)
Bridging Speech, Emotion, and Motion: a VLM-based Multimodal Edge-deployable Framework for Humanoid Robots
par: Yang, Songhua, et autres
Publié: (2026)
par: Yang, Songhua, et autres
Publié: (2026)
TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies
par: Zheng, Ruijie, et autres
Publié: (2024)
par: Zheng, Ruijie, et autres
Publié: (2024)
RaceVLA: VLA-based Racing Drone Navigation with Human-like Behaviour
par: Serpiva, Valerii, et autres
Publié: (2025)
par: Serpiva, Valerii, et autres
Publié: (2025)
TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control
par: Sun, Yuteng, et autres
Publié: (2026)
par: Sun, Yuteng, et autres
Publié: (2026)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
par: Miao, Cui, et autres
Publié: (2025)
par: Miao, Cui, et autres
Publié: (2025)
Sci-VLA: Agentic VLA Inference Plugin for Long-Horizon Tasks in Scientific Experiments
par: Pang, Yiwen, et autres
Publié: (2026)
par: Pang, Yiwen, et autres
Publié: (2026)
TMR-VLA:Vision-Language-Action Model for Magnetic Motion Control of Tri-leg Silicone-based Soft Robot
par: Tang, Ruijie, et autres
Publié: (2026)
par: Tang, Ruijie, et autres
Publié: (2026)
Characterizing VLA Models: Identifying the Action Generation Bottleneck for Edge AI Architectures
par: Vishwanathan, Manoj, et autres
Publié: (2026)
par: Vishwanathan, Manoj, et autres
Publié: (2026)
Sample-Efficient Robot Skill Learning for Construction Tasks: Benchmarking Hierarchical Reinforcement Learning and Vision-Language-Action VLA Model
par: Hu, Zhaofeng, et autres
Publié: (2025)
par: Hu, Zhaofeng, et autres
Publié: (2025)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
par: Zhang, Rongyu, et autres
Publié: (2025)
par: Zhang, Rongyu, et autres
Publié: (2025)
Embedding Autonomous Agents in Resource-Constrained Robotic Platforms
par: Halakou, Negar, et autres
Publié: (2026)
par: Halakou, Negar, et autres
Publié: (2026)
Vision-Language Models on the Edge for Real-Time Robotic Perception
par: Ahmad, Sarat, et autres
Publié: (2026)
par: Ahmad, Sarat, et autres
Publié: (2026)
UrbanInsight: A Distributed Edge Computing Framework with LLM-Powered Data Filtering for Smart City Digital Twins
par: Gupta, Kishor Datta, et autres
Publié: (2025)
par: Gupta, Kishor Datta, et autres
Publié: (2025)
Industrial Internet Robot Collaboration System and Edge Computing Optimization
par: Zhao, Haopeng, et autres
Publié: (2025)
par: Zhao, Haopeng, et autres
Publié: (2025)
Towards Deploying VLA without Fine-Tuning: Plug-and-Play Inference-Time VLA Policy Steering via Embodied Evolutionary Diffusion
par: Li, Zhuo, et autres
Publié: (2025)
par: Li, Zhuo, et autres
Publié: (2025)
Ablation Study of Multimodal Perception, Language Grounding, and Control for Human-Robot Interaction in an Object Detection and Grasping Task
par: Tian, Zi, et autres
Publié: (2026)
par: Tian, Zi, et autres
Publié: (2026)
Multimodal Coherent Explanation Generation of Robot Failures
par: Pramanick, Pradip, et autres
Publié: (2024)
par: Pramanick, Pradip, et autres
Publié: (2024)
Safe Multimodal Communication in Human-Robot Collaboration
par: Ferrari, Davide, et autres
Publié: (2023)
par: Ferrari, Davide, et autres
Publié: (2023)
Advanced Tool Learning and Selection System (ATLASS): A Closed-Loop Framework Using LLM
par: Haque, Mohd Ariful, et autres
Publié: (2025)
par: Haque, Mohd Ariful, et autres
Publié: (2025)
AnywhereVLA: Language-Conditioned Exploration and Mobile Manipulation
par: Gubernatorov, Konstantin, et autres
Publié: (2025)
par: Gubernatorov, Konstantin, et autres
Publié: (2025)
WorldVLA: Towards Autoregressive Action World Model
par: Cen, Jun, et autres
Publié: (2025)
par: Cen, Jun, et autres
Publié: (2025)
GRACE: Generalizing Robot-Assisted Caregiving with User Functionality Embeddings
par: Liu, Ziang, et autres
Publié: (2025)
par: Liu, Ziang, et autres
Publié: (2025)
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
par: Jiang, Shuo, et autres
Publié: (2025)
par: Jiang, Shuo, et autres
Publié: (2025)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
par: Wang, Qiuyue, et autres
Publié: (2026)
par: Wang, Qiuyue, et autres
Publié: (2026)
Documents similaires
-
Lite VLA: Efficient Vision-Language-Action Control on CPU-Bound Edge Robots
par: Williams, Justin, et autres
Publié: (2025) -
LiteVLA-H: Dual-Rate Vision-Language-Action Inference for Onboard Aerial Guidance and Semantic Perception
par: williams, Justin, et autres
Publié: (2026) -
EdgeNav-QE: QLoRA Quantization and Dynamic Early Exit for LAM-based Navigation on Edge Devices
par: Liu, Mengyun, et autres
Publié: (2026) -
ReconVLA: An Uncertainty-Guided and Failure-Aware Vision-Language-Action Framework for Robotic Control
par: Chen, Lingling, et autres
Publié: (2026) -
EvoScene-VLA: Evolving Scene Beliefs Inside the Action Decoder for Chunked Robot Control
par: Zhang, Chushan, et autres
Publié: (2026)