Look Before Leap: Look-Ahead Planning with Uncertainty in Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yongshuai, Liu, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adventurer: Exploration with BiGAN for Deep Reinforcement Learning
by: Liu, Yongshuai, et al.
Published: (2025)
by: Liu, Yongshuai, et al.
Published: (2025)
Evidence of Learned Look-Ahead in a Chess-Playing Neural Network
by: Jenner, Erik, et al.
Published: (2024)
by: Jenner, Erik, et al.
Published: (2024)
Streaming Looking Ahead with Token-level Self-reward
by: Zhang, Hongming, et al.
Published: (2025)
by: Zhang, Hongming, et al.
Published: (2025)
Looking Ahead to Avoid Being Late: Solving Hard-Constrained Traveling Salesman Problem
by: Chen, Jingxiao, et al.
Published: (2024)
by: Chen, Jingxiao, et al.
Published: (2024)
Look Before You Leap: Autonomous Exploration for LLM Agents
by: Ye, Ziang, et al.
Published: (2026)
by: Ye, Ziang, et al.
Published: (2026)
Towards Fully Automated Decision-Making Systems for Greenhouse Control: Challenges and Opportunities
by: Liu, Yongshuai, et al.
Published: (2025)
by: Liu, Yongshuai, et al.
Published: (2025)
Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance
by: Benhenda, Mostapha
Published: (2026)
by: Benhenda, Mostapha
Published: (2026)
Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models
by: Huang, Yuheng, et al.
Published: (2023)
by: Huang, Yuheng, et al.
Published: (2023)
Beyond Prediction: Reinforcement Learning as the Defining Leap in Healthcare AI
by: Perera, Dilruk, et al.
Published: (2025)
by: Perera, Dilruk, et al.
Published: (2025)
Look Before You Leap: Using Serialized State Machine for Language Conditioned Robotic Manipulation
by: Mu, Tong, et al.
Published: (2025)
by: Mu, Tong, et al.
Published: (2025)
On the Hardness of Reinforcement Learning with Transition Look-Ahead
by: Pla, Corentin, et al.
Published: (2025)
by: Pla, Corentin, et al.
Published: (2025)
LookAhead Tuning: Safer Language Models via Partial Answer Previews
by: Liu, Kangwei, et al.
Published: (2025)
by: Liu, Kangwei, et al.
Published: (2025)
Local Look-Ahead Guidance via Verifier-in-the-Loop for Automated Theorem Proving
by: Rajaee, Sara, et al.
Published: (2025)
by: Rajaee, Sara, et al.
Published: (2025)
Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models
by: Li, Weixian Waylon, et al.
Published: (2026)
by: Li, Weixian Waylon, et al.
Published: (2026)
Can a Small Model Learn to Look Before It Leaps? Dynamic Learning and Proactive Correction for Hallucination Detection
by: Bao, Zepeng, et al.
Published: (2025)
by: Bao, Zepeng, et al.
Published: (2025)
LookPlanGraph: Embodied Instruction Following Method with VLM Graph Augmentation
by: Onishchenko, Anatoly O., et al.
Published: (2025)
by: Onishchenko, Anatoly O., et al.
Published: (2025)
ProSpec RL: Plan Ahead, then Execute
by: Liu, Liangliang, et al.
Published: (2024)
by: Liu, Liangliang, et al.
Published: (2024)
Optimal Look-back Horizon for Time Series Forecasting in Federated Learning
by: Tang, Dahao, et al.
Published: (2025)
by: Tang, Dahao, et al.
Published: (2025)
A Closer Look at the Application of Causal Inference in Graph Representation Learning
by: Gao, Hang, et al.
Published: (2026)
by: Gao, Hang, et al.
Published: (2026)
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings
by: Alver, Safa, et al.
Published: (2022)
by: Alver, Safa, et al.
Published: (2022)
Looking beyond the next token
by: Thankaraj, Abitha, et al.
Published: (2025)
by: Thankaraj, Abitha, et al.
Published: (2025)
LookAlike: Consistent Distractor Generation in Math MCQs
by: Parikh, Nisarg, et al.
Published: (2025)
by: Parikh, Nisarg, et al.
Published: (2025)
This Probably Looks Exactly Like That: An Invertible Prototypical Network
by: Carmichael, Zachariah, et al.
Published: (2024)
by: Carmichael, Zachariah, et al.
Published: (2024)
Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM
by: Zhang, Shaoqing, et al.
Published: (2024)
by: Zhang, Shaoqing, et al.
Published: (2024)
A Closer Look at Adversarial Suffix Learning for Jailbreaking LLMs: Augmented Adversarial Trigger Learning
by: Wang, Zhe, et al.
Published: (2025)
by: Wang, Zhe, et al.
Published: (2025)
Look Further Ahead: Testing the Limits of GPT-4 in Path Planning
by: Aghzal, Mohamed, et al.
Published: (2024)
by: Aghzal, Mohamed, et al.
Published: (2024)
Memorization: A Close Look at Books
by: Ma, Iris, et al.
Published: (2025)
by: Ma, Iris, et al.
Published: (2025)
Look Before You Leap: A GUI-Critic-R1 Model for Pre-Operative Error Diagnosis in GUI Automation
by: Wanyan, Yuyang, et al.
Published: (2025)
by: Wanyan, Yuyang, et al.
Published: (2025)
Look Before You Leap: Problem Elaboration Prompting Improves Mathematical Reasoning in Large Language Models
by: Liao, Haoran, et al.
Published: (2024)
by: Liao, Haoran, et al.
Published: (2024)
Hierarchical Reinforcement Learning for Swarm Confrontation with High Uncertainty
by: Wu, Qizhen, et al.
Published: (2024)
by: Wu, Qizhen, et al.
Published: (2024)
These Are Not All the Features You Are Looking For: A Fundamental Bottleneck in Supervised Pretraining
by: Yang, Xingyu Alice, et al.
Published: (2025)
by: Yang, Xingyu Alice, et al.
Published: (2025)
A Sobering Look at Tabular Data Generation via Probabilistic Circuits
by: Scassola, Davide, et al.
Published: (2026)
by: Scassola, Davide, et al.
Published: (2026)
Continual Reinforcement Learning by Planning with Online World Models
by: Liu, Zichen, et al.
Published: (2025)
by: Liu, Zichen, et al.
Published: (2025)
Base Models Look Human To AI Detectors
by: Xu, Yixuan Even, et al.
Published: (2026)
by: Xu, Yixuan Even, et al.
Published: (2026)
Look-Ahead Reasoning on Learning Platforms
by: Zhu, Haiqing, et al.
Published: (2025)
by: Zhu, Haiqing, et al.
Published: (2025)
To See Far, Look Close: Evolutionary Forecasting for Long-term Time Series
by: Ma, Jiaming, et al.
Published: (2026)
by: Ma, Jiaming, et al.
Published: (2026)
Interpreting What Typical Fault Signals Look Like via Prototype-matching
by: Chen, Qian, et al.
Published: (2024)
by: Chen, Qian, et al.
Published: (2024)
Learning Future Representation with Synthetic Observations for Sample-efficient Reinforcement Learning
by: Liu, Xin, et al.
Published: (2024)
by: Liu, Xin, et al.
Published: (2024)
Back to the Future: Look-ahead Augmentation and Parallel Self-Refinement for Time Series Forecasting
by: Kim, Sunho, et al.
Published: (2026)
by: Kim, Sunho, et al.
Published: (2026)
Look Globally and Reason: Two-stage Path Reasoning over Sparse Knowledge Graphs
by: Guan, Saiping, et al.
Published: (2024)
by: Guan, Saiping, et al.
Published: (2024)
Similar Items
-
Adventurer: Exploration with BiGAN for Deep Reinforcement Learning
by: Liu, Yongshuai, et al.
Published: (2025) -
Evidence of Learned Look-Ahead in a Chess-Playing Neural Network
by: Jenner, Erik, et al.
Published: (2024) -
Streaming Looking Ahead with Token-level Self-reward
by: Zhang, Hongming, et al.
Published: (2025) -
Looking Ahead to Avoid Being Late: Solving Hard-Constrained Traveling Salesman Problem
by: Chen, Jingxiao, et al.
Published: (2024) -
Look Before You Leap: Autonomous Exploration for LLM Agents
by: Ye, Ziang, et al.
Published: (2026)