ProAct: Agentic Lookahead in Interactive Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Yangbin, Yang, Mingyu, Li, Junyou, Gao, Yiming, Liu, Feiyu, Yang, Yijun, Lin, Zichuan, Lyu, Jiafei, Liu, Yicheng, Lu, Zhicong, Ye, Deheng, Jiang, Jie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UI-Voyager: A Self-Evolving GUI Agent Learning via Failed Experience
von: Lin, Zichuan, et al.
Veröffentlicht: (2026)
von: Lin, Zichuan, et al.
Veröffentlicht: (2026)
ProAct: A Dual-System Framework for Proactive Embodied Social Agents
von: Zhang, Zeyi, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyi, et al.
Veröffentlicht: (2026)
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
ProAct: A Benchmark and Multimodal Framework for Structure-Aware Proactive Response
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2026)
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2026)
Playable Game Generation
von: Yang, Mingyu, et al.
Veröffentlicht: (2024)
von: Yang, Mingyu, et al.
Veröffentlicht: (2024)
Affordable Generative Agents
von: Yu, Yangbin, et al.
Veröffentlicht: (2024)
von: Yu, Yangbin, et al.
Veröffentlicht: (2024)
More Agents Is All You Need
von: Li, Junyou, et al.
Veröffentlicht: (2024)
von: Li, Junyou, et al.
Veröffentlicht: (2024)
EntroPIC: Towards Stable Long-Term Training of LLMs via Entropy Stabilization with Proportional-Integral Control
von: Yang, Kai, et al.
Veröffentlicht: (2025)
von: Yang, Kai, et al.
Veröffentlicht: (2025)
ProAct: Progressive Training for Hybrid Clipped Activation Function to Enhance Resilience of DNNs
von: Mousavi, Seyedhamidreza, et al.
Veröffentlicht: (2024)
von: Mousavi, Seyedhamidreza, et al.
Veröffentlicht: (2024)
SeeNav-Agent: Enhancing Vision-Language Navigation with Visual Prompt and Step-Level Policy Optimization
von: Wang, Zhengcheng, et al.
Veröffentlicht: (2025)
von: Wang, Zhengcheng, et al.
Veröffentlicht: (2025)
HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation
von: Zhao, He, et al.
Veröffentlicht: (2026)
von: Zhao, He, et al.
Veröffentlicht: (2026)
Debiased Model-based Representations for Sample-efficient Continuous Control
von: Lyu, Jiafei, et al.
Veröffentlicht: (2026)
von: Lyu, Jiafei, et al.
Veröffentlicht: (2026)
Cross-Domain Offline Policy Adaptation via Selective Transition Correction
von: Yan, Mengbei, et al.
Veröffentlicht: (2026)
von: Yan, Mengbei, et al.
Veröffentlicht: (2026)
HISR: Hindsight Information Modulated Segmental Process Rewards For Multi-turn Agentic Reinforcement Learning
von: Lu, Zhicong, et al.
Veröffentlicht: (2026)
von: Lu, Zhicong, et al.
Veröffentlicht: (2026)
Temporal Difference Learning with Constrained Initial Representations
von: Lyu, Jiafei, et al.
Veröffentlicht: (2026)
von: Lyu, Jiafei, et al.
Veröffentlicht: (2026)
PROF: An LLM-based Reward Code Preference Optimization Framework for Offline Imitation Learning
von: Sun, Shengjie, et al.
Veröffentlicht: (2025)
von: Sun, Shengjie, et al.
Veröffentlicht: (2025)
Yan: Foundational Interactive Video Generation
von: Ye, Deheng, et al.
Veröffentlicht: (2025)
von: Ye, Deheng, et al.
Veröffentlicht: (2025)
Boosting Vulnerability Detection of LLMs via Curriculum Preference Optimization with Synthetic Reasoning Data
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
Active Hypothesis Testing for Correlated Combinatorial Anomaly Detection
von: Yang, Zichuan, et al.
Veröffentlicht: (2026)
von: Yang, Zichuan, et al.
Veröffentlicht: (2026)
GTR-Turbo: Merged Checkpoint is Secretly a Free Teacher for Agentic VLM Training
von: Wei, Tong, et al.
Veröffentlicht: (2025)
von: Wei, Tong, et al.
Veröffentlicht: (2025)
Vul-R2: A Reasoning LLM for Automated Vulnerability Repair
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
Improving Sample Efficiency of Reinforcement Learning with Background Knowledge from Large Language Models
von: Zhang, Fuxiang, et al.
Veröffentlicht: (2024)
von: Zhang, Fuxiang, et al.
Veröffentlicht: (2024)
Faithful-MR1: Faithful Multimodal Reasoning via Anchoring and Reinforcing Visual Attention
von: Tian, Changyuan, et al.
Veröffentlicht: (2026)
von: Tian, Changyuan, et al.
Veröffentlicht: (2026)
Learning Versatile Skills with Curriculum Masking
von: Tang, Yao, et al.
Veröffentlicht: (2024)
von: Tang, Yao, et al.
Veröffentlicht: (2024)
Lookahead Q-Cache: Achieving More Consistent KV Cache Eviction via Pseudo Query
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
PIPCFR: Pseudo-outcome Imputation with Post-treatment Variables for Individual Treatment Effect Estimation
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
MetaColloc: Optimization-Free PDE Solving via Meta-Learned Basis Functions
von: Yang, Zichuan
Veröffentlicht: (2026)
von: Yang, Zichuan
Veröffentlicht: (2026)
The Lifecycle Principle: Stabilizing Dynamic Neural Networks with State Memory
von: Yang, Zichuan
Veröffentlicht: (2025)
von: Yang, Zichuan
Veröffentlicht: (2025)
A Test of Lookahead Bias in LLM Forecasts
von: Gao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Gao, Zhenyu, et al.
Veröffentlicht: (2025)
CausalMACE: Causality Empowered Multi-Agents in Minecraft Cooperative Tasks
von: Chai, Qi, et al.
Veröffentlicht: (2025)
von: Chai, Qi, et al.
Veröffentlicht: (2025)
Lookahead Exploration with Neural Radiance Representation for Continuous Vision-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
WALL-E: World Alignment by Rule Learning Improves World Model-based LLM Agents
von: Zhou, Siyu, et al.
Veröffentlicht: (2024)
von: Zhou, Siyu, et al.
Veröffentlicht: (2024)
WALL-E 2.0: World Alignment by NeuroSymbolic Learning improves World Model-based LLM Agents
von: Zhou, Siyu, et al.
Veröffentlicht: (2025)
von: Zhou, Siyu, et al.
Veröffentlicht: (2025)
GTR: Guided Thought Reinforcement Prevents Thought Collapse in RL-based VLM Agent Training
von: Wei, Tong, et al.
Veröffentlicht: (2025)
von: Wei, Tong, et al.
Veröffentlicht: (2025)
Nonlinear kernel-free quadratic hyper-surface support vector machine with 0-1 loss function
von: Wu, Mingyang, et al.
Veröffentlicht: (2024)
von: Wu, Mingyang, et al.
Veröffentlicht: (2024)
MolAct: An Agentic RL Framework for Molecular Editing and Property Optimization
von: Yang, Zhuo, et al.
Veröffentlicht: (2025)
von: Yang, Zhuo, et al.
Veröffentlicht: (2025)
IC-World: In-Context Generation for Shared World Modeling
von: Wu, Fan, et al.
Veröffentlicht: (2025)
von: Wu, Fan, et al.
Veröffentlicht: (2025)
Multi-agent In-context Coordination via Decentralized Memory Retrieval
von: Jiang, Tao, et al.
Veröffentlicht: (2025)
von: Jiang, Tao, et al.
Veröffentlicht: (2025)
RAM-SD: Retrieval-Augmented Multi-agent framework for Sarcasm Detection
von: Zhou, Ziyang, et al.
Veröffentlicht: (2026)
von: Zhou, Ziyang, et al.
Veröffentlicht: (2026)
Lookahead Path Likelihood Optimization for Diffusion LLMs
von: Liu, Xuejie, et al.
Veröffentlicht: (2026)
von: Liu, Xuejie, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
UI-Voyager: A Self-Evolving GUI Agent Learning via Failed Experience
von: Lin, Zichuan, et al.
Veröffentlicht: (2026) -
ProAct: A Dual-System Framework for Proactive Embodied Social Agents
von: Zhang, Zeyi, et al.
Veröffentlicht: (2026) -
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
von: Lin, Zichuan, et al.
Veröffentlicht: (2025) -
ProAct: A Benchmark and Multimodal Framework for Structure-Aware Proactive Response
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2026) -
Playable Game Generation
von: Yang, Mingyu, et al.
Veröffentlicht: (2024)