Environment-Aware Adaptive Pruning with Interleaved Inference Orchestration for Vision-Language-Action Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Yuting, Ding, Leilei, Tang, Zhipeng, Zhu, Zenghuan, Deng, Jiajun, Lin, Xinrui, Liu, Shuo, Ren, Haojie, Ji, Jianmin, Zhang, Yanyong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Framework for Benchmarking and Aligning Task-Planning Safety in LLM-Based Embodied Agents
von: Huang, Yuting, et al.
Veröffentlicht: (2025)
von: Huang, Yuting, et al.
Veröffentlicht: (2025)
VLMPlanner: Integrating Visual Language Models with Motion Planning
von: Tang, Zhipeng, et al.
Veröffentlicht: (2025)
von: Tang, Zhipeng, et al.
Veröffentlicht: (2025)
OA-DET3D: Embedding Object Awareness as a General Plug-in for Multi-Camera 3D Object Detection
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2023)
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2023)
UrgenGo: Urgency-Aware Transparent GPU Kernel Launching for Autonomous Driving
von: Zhu, Hanqi, et al.
Veröffentlicht: (2025)
von: Zhu, Hanqi, et al.
Veröffentlicht: (2025)
CLMASP: Coupling Large Language Models with Answer Set Programming for Robotic Task Planning
von: Lin, Xinrui, et al.
Veröffentlicht: (2024)
von: Lin, Xinrui, et al.
Veröffentlicht: (2024)
SpatialSplat: Efficient Semantic 3D from Sparse Unposed Images
von: Sheng, Yu, et al.
Veröffentlicht: (2025)
von: Sheng, Yu, et al.
Veröffentlicht: (2025)
GraspCoT: Integrating Physical Property Reasoning for 6-DoF Grasping under Flexible Language Instructions
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2025)
PRANCE: Joint Token-Optimization and Structural Channel-Pruning for Adaptive ViT Inference
von: Li, Ye, et al.
Veröffentlicht: (2024)
von: Li, Ye, et al.
Veröffentlicht: (2024)
Ghost Points Matter: Far-Range Vehicle Detection with a Single mmWave Radar in Tunnel
von: He, Chenming, et al.
Veröffentlicht: (2025)
von: He, Chenming, et al.
Veröffentlicht: (2025)
Adaptive Action Chunking at Inference-time for Vision-Language-Action Models
von: Liang, Yuanchang, et al.
Veröffentlicht: (2026)
von: Liang, Yuanchang, et al.
Veröffentlicht: (2026)
Environment-Aware Dynamic Pruning for Pipelined Edge Inference
von: O'Quinn, Austin, et al.
Veröffentlicht: (2025)
von: O'Quinn, Austin, et al.
Veröffentlicht: (2025)
SDT‐MCS: Topology‐Aware Microservice Orchestration With Adaptive Learning in Cloud‐Edge Environments
von: Jianyong Zhu, et al.
Veröffentlicht: (2025)
von: Jianyong Zhu, et al.
Veröffentlicht: (2025)
Towards Generalized Routing: Model and Agent Orchestration for Adaptive and Efficient Inference
von: Guo, Xiyu, et al.
Veröffentlicht: (2025)
von: Guo, Xiyu, et al.
Veröffentlicht: (2025)
SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning
von: Wang, Hanzhen, et al.
Veröffentlicht: (2025)
von: Wang, Hanzhen, et al.
Veröffentlicht: (2025)
ElectricSight: 3D Hazard Monitoring for Power Lines Using Low-Cost Sensors
von: Li, Xingchen, et al.
Veröffentlicht: (2025)
von: Li, Xingchen, et al.
Veröffentlicht: (2025)
INTERLACE: Interleaved Layer Pruning and Efficient Adaptation in Large Vision-Language Models
von: Madinei, Parsa, et al.
Veröffentlicht: (2025)
von: Madinei, Parsa, et al.
Veröffentlicht: (2025)
SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation
von: Ma, Shilin, et al.
Veröffentlicht: (2026)
von: Ma, Shilin, et al.
Veröffentlicht: (2026)
Act, Think or Abstain: Complexity-Aware Adaptive Inference for Vision-Language-Action Models
von: Izzo, Riccardo Andrea, et al.
Veröffentlicht: (2026)
von: Izzo, Riccardo Andrea, et al.
Veröffentlicht: (2026)
A-IO: Adaptive Inference Orchestration for Memory-Bound NPUs
von: Zhang, Chen, et al.
Veröffentlicht: (2026)
von: Zhang, Chen, et al.
Veröffentlicht: (2026)
Learning Surgical Robotic Manipulation with 3D Spatial Priors
von: Sheng, Yu, et al.
Veröffentlicht: (2026)
von: Sheng, Yu, et al.
Veröffentlicht: (2026)
Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference
von: Liu, Ziyan, et al.
Veröffentlicht: (2025)
von: Liu, Ziyan, et al.
Veröffentlicht: (2025)
CAFE-AD: Cross-Scenario Adaptive Feature Enhancement for Trajectory Planning in Autonomous Driving
von: Zhang, Junrui, et al.
Veröffentlicht: (2025)
von: Zhang, Junrui, et al.
Veröffentlicht: (2025)
Motion-Aware Adaptive Pixel Pruning for Efficient Local Motion Deblurring
von: Shang, Wei, et al.
Veröffentlicht: (2025)
von: Shang, Wei, et al.
Veröffentlicht: (2025)
FAVLA: A Force-Adaptive Fast-Slow VLA model for Contact-Rich Robotic Manipulation
von: Li, Yao, et al.
Veröffentlicht: (2026)
von: Li, Yao, et al.
Veröffentlicht: (2026)
BFA++: Hierarchical Best-Feature-Aware Token Prune for Multi-View Vision Language Action Model
von: Li, Haosheng, et al.
Veröffentlicht: (2026)
von: Li, Haosheng, et al.
Veröffentlicht: (2026)
AdaptInfer: Adaptive Token Pruning for Vision-Language Model Inference with Dynamical Text Guidance
von: Zhang, Weichen, et al.
Veröffentlicht: (2025)
von: Zhang, Weichen, et al.
Veröffentlicht: (2025)
VisPCO: Visual Token Pruning Configuration Optimization via Budget-Aware Pareto-Frontier Learning for Vision-Language Models
von: Ji, Huawei, et al.
Veröffentlicht: (2026)
von: Ji, Huawei, et al.
Veröffentlicht: (2026)
$π_0$-EqM: Equilibrium Matching for Closed-Loop Vision-Language-Action Control
von: Liu, Huanming, et al.
Veröffentlicht: (2026)
von: Liu, Huanming, et al.
Veröffentlicht: (2026)
Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving
von: Wang, Shumin, et al.
Veröffentlicht: (2025)
von: Wang, Shumin, et al.
Veröffentlicht: (2025)
CORP: A Multi-Modal Dataset for Campus-Oriented Roadside Perception Tasks
von: Wang, Beibei, et al.
Veröffentlicht: (2024)
von: Wang, Beibei, et al.
Veröffentlicht: (2024)
Vector Contrastive Learning For Pixel-Wise Pretraining In Medical Vision
von: He, Yuting, et al.
Veröffentlicht: (2025)
von: He, Yuting, et al.
Veröffentlicht: (2025)
DGR: A General Graph Desmoothing Framework for Recommendation via Global and Local Perspectives
von: Ding, Leilei, et al.
Veröffentlicht: (2024)
von: Ding, Leilei, et al.
Veröffentlicht: (2024)
Inference Load-Aware Orchestration for Hierarchical Federated Learning
von: Lackinger, Anna, et al.
Veröffentlicht: (2024)
von: Lackinger, Anna, et al.
Veröffentlicht: (2024)
HALO: Hierarchical Autonomous Logic-Oriented Orchestration for Multi-Agent LLM Systems
von: Hou, Zhipeng, et al.
Veröffentlicht: (2025)
von: Hou, Zhipeng, et al.
Veröffentlicht: (2025)
Action-aware Dynamic Pruning for Efficient Vision-Language-Action Manipulation
von: Pei, Xiaohuan, et al.
Veröffentlicht: (2025)
von: Pei, Xiaohuan, et al.
Veröffentlicht: (2025)
Adaptive Pruning for Large Language Models with Structural Importance Awareness
von: Zheng, Haotian, et al.
Veröffentlicht: (2024)
von: Zheng, Haotian, et al.
Veröffentlicht: (2024)
RAP: Runtime Adaptive Pruning for LLM Inference
von: Liu, Huanrong, et al.
Veröffentlicht: (2025)
von: Liu, Huanrong, et al.
Veröffentlicht: (2025)
PMC-InterCPT: Rethinking Biomedical Interleaved Data for Multimodal Continued Pretraining
von: Zhu, Guanghao, et al.
Veröffentlicht: (2026)
von: Zhu, Guanghao, et al.
Veröffentlicht: (2026)
Twilight: Adaptive Attention Sparsity with Hierarchical Top-$p$ Pruning
von: Lin, Chaofan, et al.
Veröffentlicht: (2025)
von: Lin, Chaofan, et al.
Veröffentlicht: (2025)
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
von: Yu, Wenda, et al.
Veröffentlicht: (2026)
von: Yu, Wenda, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Framework for Benchmarking and Aligning Task-Planning Safety in LLM-Based Embodied Agents
von: Huang, Yuting, et al.
Veröffentlicht: (2025) -
VLMPlanner: Integrating Visual Language Models with Motion Planning
von: Tang, Zhipeng, et al.
Veröffentlicht: (2025) -
OA-DET3D: Embedding Object Awareness as a General Plug-in for Multi-Camera 3D Object Detection
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2023) -
UrgenGo: Urgency-Aware Transparent GPU Kernel Launching for Autonomous Driving
von: Zhu, Hanqi, et al.
Veröffentlicht: (2025) -
CLMASP: Coupling Large Language Models with Answer Set Programming for Robotic Task Planning
von: Lin, Xinrui, et al.
Veröffentlicht: (2024)