Avi: Action from Volumetric Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Harris, Le, Long |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Equivariant Volumetric Grasping
by: Song, Pinhao, et al.
Published: (2025)
by: Song, Pinhao, et al.
Published: (2025)
Adaptive Action Chunking at Inference-time for Vision-Language-Action Models
by: Liang, Yuanchang, et al.
Published: (2026)
by: Liang, Yuanchang, et al.
Published: (2026)
Volumetric Ergodic Control
by: Kwon, Jueun, et al.
Published: (2025)
by: Kwon, Jueun, et al.
Published: (2025)
BLURR: A Boosted Low-Resource Inference for Vision-Language-Action Models
by: Ma, Xiaoyu, et al.
Published: (2025)
by: Ma, Xiaoyu, et al.
Published: (2025)
VolumeDP: Modeling Volumetric Representation for Manipulation Policy Learning
by: Zhou, Tianxing, et al.
Published: (2026)
by: Zhou, Tianxing, et al.
Published: (2026)
Uncertainty-Aware Visual-Inertial SLAM with Volumetric Occupancy Mapping
by: Jung, Jaehyung, et al.
Published: (2024)
by: Jung, Jaehyung, et al.
Published: (2024)
Towards Versatile Opti-Acoustic Sensor Fusion and Volumetric Mapping
by: Collado-Gonzalez, Ivana, et al.
Published: (2026)
by: Collado-Gonzalez, Ivana, et al.
Published: (2026)
Volumetric Occupancy Mapping With Probabilistic Depth Completion for Robotic Navigation
by: Popovic, Marija, et al.
Published: (2020)
by: Popovic, Marija, et al.
Published: (2020)
Boosting Action-Information via a Variational Bottleneck on Unlabelled Robot Videos
by: Zhang, Haoyu, et al.
Published: (2025)
by: Zhang, Haoyu, et al.
Published: (2025)
Concept-Based Dictionary Learning for Inference-Time Safety in Vision Language Action Models
by: Wen, Siqi, et al.
Published: (2026)
by: Wen, Siqi, et al.
Published: (2026)
DYMO-Hair: Generalizable Volumetric Dynamics Modeling for Robot Hair Manipulation
by: Zhao, Chengyang, et al.
Published: (2025)
by: Zhao, Chengyang, et al.
Published: (2025)
Threading Optimization for Vision-Language-Action Model Inference in Low-Cost Smart Agricultural Manipulation
by: Truongcao, Keith, et al.
Published: (2026)
by: Truongcao, Keith, et al.
Published: (2026)
Imagination at Inference: Synthesizing In-Hand Views for Robust Visuomotor Policy Inference
by: Ding, Haoran, et al.
Published: (2025)
by: Ding, Haoran, et al.
Published: (2025)
Hydraulic Volumetric Soft Everting Vine Robot Steering Mechanism for Underwater Exploration
by: Kaleel, Danyaal, et al.
Published: (2024)
by: Kaleel, Danyaal, et al.
Published: (2024)
Sampling-Based Model Predictive Control for Volumetric Ablation in Robotic Laser Surgery
by: Wang, Vincent Y., et al.
Published: (2024)
by: Wang, Vincent Y., et al.
Published: (2024)
Vision in Action: Learning Active Perception from Human Demonstrations
by: Xiong, Haoyu, et al.
Published: (2025)
by: Xiong, Haoyu, et al.
Published: (2025)
Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference
by: Liu, Yudong, et al.
Published: (2026)
by: Liu, Yudong, et al.
Published: (2026)
Programmable Deformation Design of Porous Soft Actuator through Volumetric-Pattern-Induced Anisotropy
by: Meng, Canqi, et al.
Published: (2025)
by: Meng, Canqi, et al.
Published: (2025)
DepthCache: Depth-Guided Training-Free Visual Token Merging for Vision-Language-Action Model Inference
by: Li, Yuquan, et al.
Published: (2026)
by: Li, Yuquan, et al.
Published: (2026)
Open-Ended Goal Inference through Actions and Language for Human-Robot Collaboration
by: Ghose, Debasmita, et al.
Published: (2025)
by: Ghose, Debasmita, et al.
Published: (2025)
Volumetric Reconstruction From Partial Views for Task-Oriented Grasping
by: Yan, Fujian, et al.
Published: (2025)
by: Yan, Fujian, et al.
Published: (2025)
See, Plan, Cut: MPC-Based Autonomous Volumetric Robotic Laser Surgery with OCT Guidance
by: Prakash, Ravi, et al.
Published: (2025)
by: Prakash, Ravi, et al.
Published: (2025)
DB-TSDF: Directional Bitmask-based Truncated Signed Distance Fields for Efficient Volumetric Mapping
by: Maese, Jose E., et al.
Published: (2025)
by: Maese, Jose E., et al.
Published: (2025)
Tightly-Coupled LiDAR-Visual-Inertial SLAM and Large-Scale Volumetric Occupancy Mapping
by: Boche, Simon, et al.
Published: (2024)
by: Boche, Simon, et al.
Published: (2024)
HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing
by: Gubernatorov, Konstantin, et al.
Published: (2026)
by: Gubernatorov, Konstantin, et al.
Published: (2026)
IA-VLA: Input Augmentation for Vision-Language-Action models in settings with semantically complex tasks
by: Hannus, Eric, et al.
Published: (2025)
by: Hannus, Eric, et al.
Published: (2025)
Reality Fusion: Robust Real-time Immersive Mobile Robot Teleoperation with Volumetric Visual Data Fusion
by: Li, Ke, et al.
Published: (2024)
by: Li, Ke, et al.
Published: (2024)
Action Deviation-Aware Inference for Low-Latency Wireless Robots
by: Park, Jeyoung, et al.
Published: (2025)
by: Park, Jeyoung, et al.
Published: (2025)
Understanding Asynchronous Inference Methods for Vision-Language-Action Models
by: Agouzoul, Ayoub
Published: (2026)
by: Agouzoul, Ayoub
Published: (2026)
Legs Over Arms: On the Predictive Value of Lower-Body Pose for Human Trajectory Prediction from Egocentric Robot Perception
by: Le, Nhat, et al.
Published: (2026)
by: Le, Nhat, et al.
Published: (2026)
DySL-VLA: Efficient Vision-Language-Action Model Inference via Dynamic-Static Layer-Skipping for Robot Manipulation
by: Yang, Zebin, et al.
Published: (2026)
by: Yang, Zebin, et al.
Published: (2026)
Temporal Action Selection for Action Chunking
by: Weng, Yueyang, et al.
Published: (2025)
by: Weng, Yueyang, et al.
Published: (2025)
EnergyAction: Unimanual to Bimanual Composition with Energy-Based Models
by: Song, Mingchen, et al.
Published: (2026)
by: Song, Mingchen, et al.
Published: (2026)
CRADMap: Applied Distributed Volumetric Mapping with 5G-Connected Multi-Robots and 4D Radar Perception
by: Qureshi, Maaz, et al.
Published: (2025)
by: Qureshi, Maaz, et al.
Published: (2025)
Towards Autonomous Robotic Kidney Ultrasound: Spatial-Efficient Volumetric Imaging via Template Guided Optimal Pivoting
by: Ma, Xihan, et al.
Published: (2026)
by: Ma, Xihan, et al.
Published: (2026)
Neural Implicit Action Fields: From Discrete Waypoints to Continuous Functions for Vision-Language-Action Models
by: Liu, Haoyun, et al.
Published: (2026)
by: Liu, Haoyun, et al.
Published: (2026)
From Inference Efficiency to Embodied Efficiency: Revisiting Efficiency Metrics for Vision-Language-Action Models
by: Li, Zhuofan, et al.
Published: (2026)
by: Li, Zhuofan, et al.
Published: (2026)
GeoAware-VLA: Implicit Geometry Aware Vision-Language-Action Model
by: Abouzeid, Ali, et al.
Published: (2025)
by: Abouzeid, Ali, et al.
Published: (2025)
Language-Grounded Decoupled Action Representation for Robotic Manipulation
by: Weng, Wuding, et al.
Published: (2026)
by: Weng, Wuding, et al.
Published: (2026)
Adaptive Motion Planning via Contact-Based Intent Inference for Human-Robot Collaboration
by: Song, Jiurun, et al.
Published: (2025)
by: Song, Jiurun, et al.
Published: (2025)
Similar Items
-
Equivariant Volumetric Grasping
by: Song, Pinhao, et al.
Published: (2025) -
Adaptive Action Chunking at Inference-time for Vision-Language-Action Models
by: Liang, Yuanchang, et al.
Published: (2026) -
Volumetric Ergodic Control
by: Kwon, Jueun, et al.
Published: (2025) -
BLURR: A Boosted Low-Resource Inference for Vision-Language-Action Models
by: Ma, Xiaoyu, et al.
Published: (2025) -
VolumeDP: Modeling Volumetric Representation for Manipulation Policy Learning
by: Zhou, Tianxing, et al.
Published: (2026)