ICPRL: Acquiring Physical Intuition from Interactive Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Xinrun, Bu, Pi, Wang, Ye, Karlsson, Börje F., Wang, Ziming, Song, Tengtao, Zhu, Qi, Song, Jun, Zhang, Shuo, Ding, Zhiming, Zheng, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning
von: Xu, Xinrun, et al.
Veröffentlicht: (2025)
von: Xu, Xinrun, et al.
Veröffentlicht: (2025)
MLLM as Retriever: Interactively Learning Multimodal Retrieval for Embodied Agents
von: Yue, Junpeng, et al.
Veröffentlicht: (2024)
von: Yue, Junpeng, et al.
Veröffentlicht: (2024)
A Survey on Game Playing Agents and Large Models: Methods, Applications, and Challenges
von: Xu, Xinrun, et al.
Veröffentlicht: (2024)
von: Xu, Xinrun, et al.
Veröffentlicht: (2024)
MindRef: Mimicking Human Memory for Hierarchical Reference Retrieval with Fine-Grained Location Awareness
von: Wang, Ye, et al.
Veröffentlicht: (2024)
von: Wang, Ye, et al.
Veröffentlicht: (2024)
Market-GAN: Adding Control to Financial Market Data Generation with Semantic Context
von: Xia, Haochong, et al.
Veröffentlicht: (2023)
von: Xia, Haochong, et al.
Veröffentlicht: (2023)
Can VLMs Play Action Role-Playing Games? Take Black Myth Wukong as a Study Case
von: Chen, Peng, et al.
Veröffentlicht: (2024)
von: Chen, Peng, et al.
Veröffentlicht: (2024)
CombatVLA: An Efficient Vision-Language-Action Model for Combat Tasks in 3D Action Role-Playing Games
von: Chen, Peng, et al.
Veröffentlicht: (2025)
von: Chen, Peng, et al.
Veröffentlicht: (2025)
"See the World, Discover Knowledge": A Chinese Factuality Evaluation for Large Vision Language Models
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control
von: Zheng, Longtao, et al.
Veröffentlicht: (2023)
von: Zheng, Longtao, et al.
Veröffentlicht: (2023)
Mobile-R1: Towards Interactive Capability for VLM-Based Mobile Agent via Systematic Training
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
How Foundational Skills Influence VLM-based Embodied Agents:A Native Perspective
von: Peng, Bo, et al.
Veröffentlicht: (2026)
von: Peng, Bo, et al.
Veröffentlicht: (2026)
ExoActor: Exocentric Video Generation as Generalizable Interactive Humanoid Control
von: Zhou, Yanghao, et al.
Veröffentlicht: (2026)
von: Zhou, Yanghao, et al.
Veröffentlicht: (2026)
X-DiffVLA: X-Embodied Diffusion Action Heads for Vision-Language-Action Models
von: Li, Boyu, et al.
Veröffentlicht: (2026)
von: Li, Boyu, et al.
Veröffentlicht: (2026)
Acquiring Human-Like Mechanics Intuition from Scarce Observations via Deep Reinforcement Learning
von: Peng, Jingruo, et al.
Veröffentlicht: (2026)
von: Peng, Jingruo, et al.
Veröffentlicht: (2026)
Taking Notes Brings Focus? Towards Multi-Turn Multimodal Dialogue Learning
von: Liu, Jiazheng, et al.
Veröffentlicht: (2025)
von: Liu, Jiazheng, et al.
Veröffentlicht: (2025)
SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios
von: Lin, Jieru, et al.
Veröffentlicht: (2025)
von: Lin, Jieru, et al.
Veröffentlicht: (2025)
IMWM: Intuition Models Complement World Models for Latent Planning
von: Gao, Baoqi, et al.
Veröffentlicht: (2026)
von: Gao, Baoqi, et al.
Veröffentlicht: (2026)
SecAgent: Efficient Mobile GUI Agent with Semantic Context
von: Xie, Yiping, et al.
Veröffentlicht: (2026)
von: Xie, Yiping, et al.
Veröffentlicht: (2026)
Token Preference Optimization with Self-Calibrated Visual-Anchored Rewards for Hallucination Mitigation
von: Gu, Jihao, et al.
Veröffentlicht: (2024)
von: Gu, Jihao, et al.
Veröffentlicht: (2024)
Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
GeoSense: Evaluating Identification and Application of Geometric Principles in Multimodal Reasoning
von: Xu, Liangyu, et al.
Veröffentlicht: (2025)
von: Xu, Liangyu, et al.
Veröffentlicht: (2025)
RL from Physical Feedback: Aligning Large Motion Models with Humanoid Control
von: Yue, Junpeng, et al.
Veröffentlicht: (2025)
von: Yue, Junpeng, et al.
Veröffentlicht: (2025)
正方形内接试证明(不一定为真,先存证)
von: Song, Ziming
Veröffentlicht: (2026)
von: Song, Ziming
Veröffentlicht: (2026)
Reinforcement Learning with Maskable Stock Representation for Portfolio Management in Customizable Stock Pools
von: Zhang, Wentao, et al.
Veröffentlicht: (2023)
von: Zhang, Wentao, et al.
Veröffentlicht: (2023)
GDBA Revisited: Unleashing the Power of Guided Local Search for Distributed Constraint Optimization
von: Deng, Yanchen, et al.
Veröffentlicht: (2025)
von: Deng, Yanchen, et al.
Veröffentlicht: (2025)
Resolving Latency and Inventory Risk in Market Making with Reinforcement Learning
von: Jiang, Junzhe, et al.
Veröffentlicht: (2025)
von: Jiang, Junzhe, et al.
Veröffentlicht: (2025)
Variational Learning of Physical Intuition from a Few Observations
von: Peng, Jingruo, et al.
Veröffentlicht: (2025)
von: Peng, Jingruo, et al.
Veröffentlicht: (2025)
Catastrophic Forgetting Mitigation via Discrepancy-Weighted Experience Replay
von: Xu, Xinrun, et al.
Veröffentlicht: (2025)
von: Xu, Xinrun, et al.
Veröffentlicht: (2025)
Cradle: Empowering Foundation Agents Towards General Computer Control
von: Tan, Weihao, et al.
Veröffentlicht: (2024)
von: Tan, Weihao, et al.
Veröffentlicht: (2024)
RANGER: A Monocular Zero-Shot Semantic Navigation Framework through Visual Contextual Adaptation
von: Yu, Ming-Ming, et al.
Veröffentlicht: (2025)
von: Yu, Ming-Ming, et al.
Veröffentlicht: (2025)
VStyle: A Benchmark for Voice Style Adaptation with Spoken Instructions
von: Zhan, Jun, et al.
Veröffentlicht: (2025)
von: Zhan, Jun, et al.
Veröffentlicht: (2025)
A Bibliography of Writings on Distance Education.
von: Holmberg, Borje
Veröffentlicht: (1990)
von: Holmberg, Borje
Veröffentlicht: (1990)
Microphytobenthic productivity in mangrove areas of different replanting regimes, Gazi Bay, Kenya.
von: Borje, Annika
Veröffentlicht: (2004)
von: Borje, Annika
Veröffentlicht: (2004)
Physically Interpretable Emulation of a Moist Convecting Atmosphere with a Recurrent Neural Network
von: Song, Qiyu, et al.
Veröffentlicht: (2025)
von: Song, Qiyu, et al.
Veröffentlicht: (2025)
Control Map Distribution using Map Query Bank for Online Map Generation
von: Liu, Ziming, et al.
Veröffentlicht: (2025)
von: Liu, Ziming, et al.
Veröffentlicht: (2025)
Extract Information from Hybrid Long Documents Leveraging LLMs: A Framework and Dataset
von: Yue, Chongjian, et al.
Veröffentlicht: (2024)
von: Yue, Chongjian, et al.
Veröffentlicht: (2024)
Acquiring and Modelling Abstract Commonsense Knowledge via Conceptualization
von: He, Mutian, et al.
Veröffentlicht: (2022)
von: He, Mutian, et al.
Veröffentlicht: (2022)
INTENTION: Inferring Tendencies of Humanoid Robot Motion Through Interactive Intuition and Grounded VLM
von: Wang, Jin, et al.
Veröffentlicht: (2025)
von: Wang, Jin, et al.
Veröffentlicht: (2025)
On Truthful Item-Acquiring Mechanisms for Reward Maximization
von: Shan, Liang, et al.
Veröffentlicht: (2024)
von: Shan, Liang, et al.
Veröffentlicht: (2024)
Acquire and then Adapt: Squeezing out Text-to-Image Model for Image Restoration
von: Deng, Junyuan, et al.
Veröffentlicht: (2025)
von: Deng, Junyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning
von: Xu, Xinrun, et al.
Veröffentlicht: (2025) -
MLLM as Retriever: Interactively Learning Multimodal Retrieval for Embodied Agents
von: Yue, Junpeng, et al.
Veröffentlicht: (2024) -
A Survey on Game Playing Agents and Large Models: Methods, Applications, and Challenges
von: Xu, Xinrun, et al.
Veröffentlicht: (2024) -
MindRef: Mimicking Human Memory for Hierarchical Reference Retrieval with Fine-Grained Location Awareness
von: Wang, Ye, et al.
Veröffentlicht: (2024) -
Market-GAN: Adding Control to Financial Market Data Generation with Semantic Context
von: Xia, Haochong, et al.
Veröffentlicht: (2023)