Salvato in:
| Autori principali: | Zhou, Weijie, Xiong, Xuangtang, Tian, Ye, Yue, Lijun, Wu, Xinyu, Li, Wei, Zhao, Chaoyang, Dong, Honghui, Tang, Ming, Wang, Jinqiao, Zhang, Zhengyou |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2512.18571 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Listening with the Eyes: Benchmarking Egocentric Co-Speech Grounding across Space and Time
di: Zhou, Weijie, et al.
Pubblicazione: (2026)
di: Zhou, Weijie, et al.
Pubblicazione: (2026)
LightPlanner: Unleashing the Reasoning Capabilities of Lightweight Large Language Models in Task Planning
di: Zhou, Weijie, et al.
Pubblicazione: (2025)
di: Zhou, Weijie, et al.
Pubblicazione: (2025)
PhysVLM-AVR: Active Visual Reasoning for Multimodal Large Language Models in Physical Environments
di: Zhou, Weijie, et al.
Pubblicazione: (2025)
di: Zhou, Weijie, et al.
Pubblicazione: (2025)
PhysVLM: Enabling Visual Language Models to Understand Robotic Physical Reachability
di: Zhou, Weijie, et al.
Pubblicazione: (2025)
di: Zhou, Weijie, et al.
Pubblicazione: (2025)
MLLM as Retriever: Interactively Learning Multimodal Retrieval for Embodied Agents
di: Yue, Junpeng, et al.
Pubblicazione: (2024)
di: Yue, Junpeng, et al.
Pubblicazione: (2024)
ProAct: A Benchmark and Multimodal Framework for Structure-Aware Proactive Response
di: Zhu, Xiaomeng, et al.
Pubblicazione: (2026)
di: Zhu, Xiaomeng, et al.
Pubblicazione: (2026)
EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents
di: Zhu, Zihao, et al.
Pubblicazione: (2024)
di: Zhu, Zihao, et al.
Pubblicazione: (2024)
BREATH: A Bio-Radar Embodied Agent for Tonal and Human-Aware Diffusion Music Generation
di: Wang, Yunzhe, et al.
Pubblicazione: (2025)
di: Wang, Yunzhe, et al.
Pubblicazione: (2025)
UniBYD: A Unified Framework for Learning Robotic Manipulation Across Embodiments Beyond Imitation of Human Demonstrations
di: Yuan, Tingyu, et al.
Pubblicazione: (2025)
di: Yuan, Tingyu, et al.
Pubblicazione: (2025)
Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning
di: Zhan, Yufei, et al.
Pubblicazione: (2025)
di: Zhan, Yufei, et al.
Pubblicazione: (2025)
VRAgent-R1: Boosting Video Recommendation with MLLM-based Agents via Reinforcement Learning
di: Chen, Siran, et al.
Pubblicazione: (2025)
di: Chen, Siran, et al.
Pubblicazione: (2025)
VIKI-R: Coordinating Embodied Multi-Agent Cooperation via Reinforcement Learning
di: Kang, Li, et al.
Pubblicazione: (2025)
di: Kang, Li, et al.
Pubblicazione: (2025)
FOCUS: Fine-grained Optimization with Semantic Guided Understanding for Pedestrian Attributes Recognition
di: An, Hongyan, et al.
Pubblicazione: (2025)
di: An, Hongyan, et al.
Pubblicazione: (2025)
TraceVision: Trajectory-Aware Vision-Language Model for Human-Like Spatial Understanding
di: Yang, Fan, et al.
Pubblicazione: (2026)
di: Yang, Fan, et al.
Pubblicazione: (2026)
Deep Reinforcement Learning Empowered Activity-Aware Dynamic Health Monitoring Systems
di: Ye, Ziqiaing, et al.
Pubblicazione: (2024)
di: Ye, Ziqiaing, et al.
Pubblicazione: (2024)
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own
di: Ye, Weirui, et al.
Pubblicazione: (2023)
di: Ye, Weirui, et al.
Pubblicazione: (2023)
ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering
di: Liu, Zexi, et al.
Pubblicazione: (2025)
di: Liu, Zexi, et al.
Pubblicazione: (2025)
ERA: Transforming VLMs into Embodied Agents via Embodied Prior Learning and Online Reinforcement Learning
di: Chen, Hanyang, et al.
Pubblicazione: (2025)
di: Chen, Hanyang, et al.
Pubblicazione: (2025)
ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing
di: An, Yongqi, et al.
Pubblicazione: (2026)
di: An, Yongqi, et al.
Pubblicazione: (2026)
AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
di: Xiong, Chuyan, et al.
Pubblicazione: (2024)
di: Xiong, Chuyan, et al.
Pubblicazione: (2024)
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning
di: Rui, Shaohao, et al.
Pubblicazione: (2025)
di: Rui, Shaohao, et al.
Pubblicazione: (2025)
Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control
di: Jiang, Sicong, et al.
Pubblicazione: (2024)
di: Jiang, Sicong, et al.
Pubblicazione: (2024)
Dual-Agent Reinforcement Learning for Adaptive and Cost-Aware Visual-Inertial Odometry
di: Pan, Feiyang, et al.
Pubblicazione: (2025)
di: Pan, Feiyang, et al.
Pubblicazione: (2025)
Manifold-Aware Exploration for Reinforcement Learning in Video Generation
di: Zheng, Mingzhe, et al.
Pubblicazione: (2026)
di: Zheng, Mingzhe, et al.
Pubblicazione: (2026)
Tempo-R0: A Video-MLLM for Temporal Video Grounding through Efficient Temporal Sensing Reinforcement Learning
di: Yue, Feng, et al.
Pubblicazione: (2025)
di: Yue, Feng, et al.
Pubblicazione: (2025)
Memo: Training Memory-Efficient Embodied Agents with Reinforcement Learning
di: Gupta, Gunshi, et al.
Pubblicazione: (2025)
di: Gupta, Gunshi, et al.
Pubblicazione: (2025)
GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning
di: Duan, Chengqi, et al.
Pubblicazione: (2025)
di: Duan, Chengqi, et al.
Pubblicazione: (2025)
SPA: 3D Spatial-Awareness Enables Effective Embodied Representation
di: Zhu, Haoyi, et al.
Pubblicazione: (2024)
di: Zhu, Haoyi, et al.
Pubblicazione: (2024)
SEEA-R1: Tree-Structured Reinforcement Fine-Tuning for Self-Evolving Embodied Agents
di: Tian, Wanxin, et al.
Pubblicazione: (2025)
di: Tian, Wanxin, et al.
Pubblicazione: (2025)
Quality-Aware Language-Conditioned Local Auto-Regressive Anomaly Synthesis and Detection
di: Qian, Long, et al.
Pubblicazione: (2025)
di: Qian, Long, et al.
Pubblicazione: (2025)
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
di: Yuan, Yifu, et al.
Pubblicazione: (2025)
di: Yuan, Yifu, et al.
Pubblicazione: (2025)
Hybrid Differential Reward: Combining Temporal Difference and Action Gradients for Efficient Multi-Agent Reinforcement Learning in Cooperative Driving
di: Han, Ye, et al.
Pubblicazione: (2025)
di: Han, Ye, et al.
Pubblicazione: (2025)
Efficient Masked Autoencoders with Self-Consistency
di: Li, Zhaowen, et al.
Pubblicazione: (2023)
di: Li, Zhaowen, et al.
Pubblicazione: (2023)
Mind the Third Eye! Benchmarking Privacy Awareness in MLLM-powered Smartphone Agents
di: Lin, Zhixin, et al.
Pubblicazione: (2025)
di: Lin, Zhixin, et al.
Pubblicazione: (2025)
Drive-R1: Bridging Reasoning and Planning in VLMs for Autonomous Driving with Reinforcement Learning
di: Li, Yue, et al.
Pubblicazione: (2025)
di: Li, Yue, et al.
Pubblicazione: (2025)
OPTAGENT: Optimizing Multi-Agent LLM Interactions Through Verbal Reinforcement Learning for Enhanced Reasoning
di: Bi, Zhenyu, et al.
Pubblicazione: (2025)
di: Bi, Zhenyu, et al.
Pubblicazione: (2025)
Robot-R1: Reinforcement Learning for Enhanced Embodied Reasoning in Robotics
di: Kim, Dongyoung, et al.
Pubblicazione: (2025)
di: Kim, Dongyoung, et al.
Pubblicazione: (2025)
Embodied-R: Collaborative Framework for Activating Embodied Spatial Reasoning in Foundation Models via Reinforcement Learning
di: Zhao, Baining, et al.
Pubblicazione: (2025)
di: Zhao, Baining, et al.
Pubblicazione: (2025)
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
di: Jin, Bowen, et al.
Pubblicazione: (2025)
di: Jin, Bowen, et al.
Pubblicazione: (2025)
Improving Generalization in LLM Structured Pruning via Function-Aware Neuron Grouping
di: Yu, Tao, et al.
Pubblicazione: (2025)
di: Yu, Tao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Listening with the Eyes: Benchmarking Egocentric Co-Speech Grounding across Space and Time
di: Zhou, Weijie, et al.
Pubblicazione: (2026) -
LightPlanner: Unleashing the Reasoning Capabilities of Lightweight Large Language Models in Task Planning
di: Zhou, Weijie, et al.
Pubblicazione: (2025) -
PhysVLM-AVR: Active Visual Reasoning for Multimodal Large Language Models in Physical Environments
di: Zhou, Weijie, et al.
Pubblicazione: (2025) -
PhysVLM: Enabling Visual Language Models to Understand Robotic Physical Reachability
di: Zhou, Weijie, et al.
Pubblicazione: (2025) -
MLLM as Retriever: Interactively Learning Multimodal Retrieval for Embodied Agents
di: Yue, Junpeng, et al.
Pubblicazione: (2024)