Vision-Language Model-based Physical Reasoning for Robot Liquid Perception
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lai, Wenqiang, Gao, Yuan, Lam, Tin Lun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Real-Time Polygonal Semantic Mapping for Humanoid Robot Stair Climbing
von: Bin, Teng, et al.
Veröffentlicht: (2024)
von: Bin, Teng, et al.
Veröffentlicht: (2024)
MODUR: A Modular Dual-reconfigurable Robot
von: Gu, Jie, et al.
Veröffentlicht: (2025)
von: Gu, Jie, et al.
Veröffentlicht: (2025)
RhoMorph: Rhombus-shaped Deformable Modular Robots for Stable, Medium-Independent Reconfiguration Motion
von: Gu, Jie, et al.
Veröffentlicht: (2026)
von: Gu, Jie, et al.
Veröffentlicht: (2026)
Robotic Perception with a Large Tactile-Vision-Language Model for Physical Property Inference
von: Guo, Zexiang, et al.
Veröffentlicht: (2025)
von: Guo, Zexiang, et al.
Veröffentlicht: (2025)
Motion planning for highly-dynamic unconditioned reflexes based on chained Signed Distance Functions
von: Lin, Ken, et al.
Veröffentlicht: (2025)
von: Lin, Ken, et al.
Veröffentlicht: (2025)
Vision-Language Models on the Edge for Real-Time Robotic Perception
von: Ahmad, Sarat, et al.
Veröffentlicht: (2026)
von: Ahmad, Sarat, et al.
Veröffentlicht: (2026)
Semantic Area Graph Reasoning for Multi-Robot Language-Guided Search
von: Wang, Ruiyang, et al.
Veröffentlicht: (2026)
von: Wang, Ruiyang, et al.
Veröffentlicht: (2026)
AHA: A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
RePLan: Robotic Replanning with Perception and Language Models
von: Skreta, Marta, et al.
Veröffentlicht: (2024)
von: Skreta, Marta, et al.
Veröffentlicht: (2024)
Graph-Fused Vision-Language-Action for Policy Reasoning in Multi-Arm Robotic Manipulation
von: Li, Shunlei, et al.
Veröffentlicht: (2025)
von: Li, Shunlei, et al.
Veröffentlicht: (2025)
Assessing Vision-Language Models for Perception in Autonomous Underwater Robotic Software
von: Yousaf, Muhammad, et al.
Veröffentlicht: (2026)
von: Yousaf, Muhammad, et al.
Veröffentlicht: (2026)
ELHPlan: Efficient Long-Horizon Task Planning for Multi-Agent Collaboration
von: Ling, Shaobin, et al.
Veröffentlicht: (2025)
von: Ling, Shaobin, et al.
Veröffentlicht: (2025)
Cover Image, Volume 43, Number 3, May 2026
von: Zhenliang Zheng, et al.
Veröffentlicht: (2026)
von: Zhenliang Zheng, et al.
Veröffentlicht: (2026)
A Cascaded Strategy With Embodied Artificial Intelligence: Forward Kinematics Solutions for CCRobot‐S
von: Zhenliang Zheng, et al.
Veröffentlicht: (2025)
von: Zhenliang Zheng, et al.
Veröffentlicht: (2025)
Commonsense Reasoning for Legged Robot Adaptation with Vision-Language Models
von: Chen, Annie S., et al.
Veröffentlicht: (2024)
von: Chen, Annie S., et al.
Veröffentlicht: (2024)
ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making
von: Son, Young-Chae, et al.
Veröffentlicht: (2026)
von: Son, Young-Chae, et al.
Veröffentlicht: (2026)
R4: Retrieval-Augmented Reasoning for Vision-Language Models in 4D Spatio-Temporal Space
von: Sohn, Tin Stribor, et al.
Veröffentlicht: (2025)
von: Sohn, Tin Stribor, et al.
Veröffentlicht: (2025)
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
von: Khan, Muhamamd Haris, et al.
Veröffentlicht: (2025)
von: Khan, Muhamamd Haris, et al.
Veröffentlicht: (2025)
RoboStream: Weaving Spatio-Temporal Reasoning with Memory in Vision-Language Models for Robotics
von: Huang, Yuzhi, et al.
Veröffentlicht: (2026)
von: Huang, Yuzhi, et al.
Veröffentlicht: (2026)
Self-Refining Vision Language Model for Robotic Failure Detection and Reasoning
von: Qi, Carl, et al.
Veröffentlicht: (2026)
von: Qi, Carl, et al.
Veröffentlicht: (2026)
SimLiquid: A Simulation‐Based Liquid Perception Pipeline for Robot Liquid Manipulation
von: Yan Huang, et al.
Veröffentlicht: (2025)
von: Yan Huang, et al.
Veröffentlicht: (2025)
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
von: Wei, Xiangyi, et al.
Veröffentlicht: (2025)
von: Wei, Xiangyi, et al.
Veröffentlicht: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
Information-Theoretic Graph Fusion with Vision-Language-Action Model for Policy Reasoning and Dual Robotic Control
von: Li, Shunlei, et al.
Veröffentlicht: (2025)
von: Li, Shunlei, et al.
Veröffentlicht: (2025)
ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation
von: Liu, Zhenyang, et al.
Veröffentlicht: (2026)
von: Liu, Zhenyang, et al.
Veröffentlicht: (2026)
Robot Navigation Using Physically Grounded Vision-Language Models in Outdoor Environments
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2024)
von: Elnoor, Mohamed, et al.
Veröffentlicht: (2024)
COMRES-VLM: Coordinated Multi-Robot Exploration and Search using Vision Language Models
von: Wang, Ruiyang, et al.
Veröffentlicht: (2025)
von: Wang, Ruiyang, et al.
Veröffentlicht: (2025)
InSpire: Vision-Language-Action Models with Intrinsic Spatial Reasoning
von: Zhang, Ji, et al.
Veröffentlicht: (2025)
von: Zhang, Ji, et al.
Veröffentlicht: (2025)
Vision-based Perception System for Automated Delivery Robot-Pedestrians Interactions
von: Tushe, Ergi, et al.
Veröffentlicht: (2025)
von: Tushe, Ergi, et al.
Veröffentlicht: (2025)
HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model
von: Zhu, Xiang, et al.
Veröffentlicht: (2026)
von: Zhu, Xiang, et al.
Veröffentlicht: (2026)
MERGE: Guided Vision-Language Models for Multi-Actor Event Reasoning and Grounding in Human-Robot Interaction
von: Deigmoeller, Joerg, et al.
Veröffentlicht: (2026)
von: Deigmoeller, Joerg, et al.
Veröffentlicht: (2026)
VLMgineer: Vision Language Models as Robotic Toolsmiths
von: Gao, George Jiayuan, et al.
Veröffentlicht: (2025)
von: Gao, George Jiayuan, et al.
Veröffentlicht: (2025)
From Perception to Symbolic Task Planning: Vision-Language Guided Human-Robot Collaborative Structured Assembly
von: Chen, Yanyi, et al.
Veröffentlicht: (2026)
von: Chen, Yanyi, et al.
Veröffentlicht: (2026)
OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning
von: Lin, Fanqi, et al.
Veröffentlicht: (2025)
von: Lin, Fanqi, et al.
Veröffentlicht: (2025)
Vision Language Model-based Testing of Industrial Autonomous Mobile Robots
von: Wu, Jiahui, et al.
Veröffentlicht: (2025)
von: Wu, Jiahui, et al.
Veröffentlicht: (2025)
Task-oriented Robotic Manipulation with Vision Language Models
von: Guran, Nurhan Bulus, et al.
Veröffentlicht: (2024)
von: Guran, Nurhan Bulus, et al.
Veröffentlicht: (2024)
RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics
von: Zhou, Enshen, et al.
Veröffentlicht: (2025)
von: Zhou, Enshen, et al.
Veröffentlicht: (2025)
Look-to-Touch: A Vision-Enhanced Proximity and Tactile Sensor for Distance and Geometry Perception in Robotic Manipulation
von: Dong, Yueshi, et al.
Veröffentlicht: (2025)
von: Dong, Yueshi, et al.
Veröffentlicht: (2025)
DDBot: Differentiable Physics-based Digging Robot for Unknown Granular Materials
von: Yang, Xintong, et al.
Veröffentlicht: (2025)
von: Yang, Xintong, et al.
Veröffentlicht: (2025)
RoboNurse-VLA: Robotic Scrub Nurse System based on Vision-Language-Action Model
von: Li, Shunlei, et al.
Veröffentlicht: (2024)
von: Li, Shunlei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Real-Time Polygonal Semantic Mapping for Humanoid Robot Stair Climbing
von: Bin, Teng, et al.
Veröffentlicht: (2024) -
MODUR: A Modular Dual-reconfigurable Robot
von: Gu, Jie, et al.
Veröffentlicht: (2025) -
RhoMorph: Rhombus-shaped Deformable Modular Robots for Stable, Medium-Independent Reconfiguration Motion
von: Gu, Jie, et al.
Veröffentlicht: (2026) -
Robotic Perception with a Large Tactile-Vision-Language Model for Physical Property Inference
von: Guo, Zexiang, et al.
Veröffentlicht: (2025) -
Motion planning for highly-dynamic unconditioned reflexes based on chained Signed Distance Functions
von: Lin, Ken, et al.
Veröffentlicht: (2025)