When Remembering and Planning are Worth it: Navigating under Change
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Madani, Omid, Burns, J. Brian, Eghbali, Reza, Dean, Thomas L. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tracking Changing Probabilities via Dynamic Learners
von: Madani, Omid
Veröffentlicht: (2024)
von: Madani, Omid
Veröffentlicht: (2024)
Programming by Backprop: An Instruction is Worth 100 Examples When Finetuning LLMs
von: Cook, Jonathan, et al.
Veröffentlicht: (2025)
von: Cook, Jonathan, et al.
Veröffentlicht: (2025)
Choosing How to Remember: Adaptive Memory Structures for LLM Agents
von: Lu, Mingfei, et al.
Veröffentlicht: (2026)
von: Lu, Mingfei, et al.
Veröffentlicht: (2026)
Vision HgNN: An Electron-Micrograph is Worth Hypergraph of Hypernodes
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2024)
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2024)
Learning to Remember: End-to-End Training of Memory Agents for Long-Context Reasoning
von: Zhang, Kehao, et al.
Veröffentlicht: (2026)
von: Zhang, Kehao, et al.
Veröffentlicht: (2026)
The Pitfalls of Memorization: When Memorization Hurts Generalization
von: Bayat, Reza, et al.
Veröffentlicht: (2024)
von: Bayat, Reza, et al.
Veröffentlicht: (2024)
Teaching AI to Remember: Insights from Brain-Inspired Replay in Continual Learning
von: Kim, Jina
Veröffentlicht: (2025)
von: Kim, Jina
Veröffentlicht: (2025)
When Engineering Outruns Intelligence: Rethinking Instruction-Guided Navigation
von: Aghaei, Matin, et al.
Veröffentlicht: (2025)
von: Aghaei, Matin, et al.
Veröffentlicht: (2025)
A Graph is Worth 1-bit Spikes: When Graph Contrastive Learning Meets Spiking Neural Networks
von: Li, Jintang, et al.
Veröffentlicht: (2023)
von: Li, Jintang, et al.
Veröffentlicht: (2023)
When can transformers reason with abstract symbols?
von: Boix-Adsera, Enric, et al.
Veröffentlicht: (2023)
von: Boix-Adsera, Enric, et al.
Veröffentlicht: (2023)
A Wave is Worth 100 Words: Investigating Cross-Domain Transferability in Time Series
von: Ma, Xiangkai, et al.
Veröffentlicht: (2024)
von: Ma, Xiangkai, et al.
Veröffentlicht: (2024)
PROPEL: Supervised and Reinforcement Learning for Large-Scale Supply Chain Planning
von: Akhlaghi, Vahid Eghbal, et al.
Veröffentlicht: (2025)
von: Akhlaghi, Vahid Eghbal, et al.
Veröffentlicht: (2025)
A Reinforcement Learning-Based Task Mapping Method to Improve the Reliability of Clustered Manycores
von: Hossein-Khani, Fatemeh, et al.
Veröffentlicht: (2024)
von: Hossein-Khani, Fatemeh, et al.
Veröffentlicht: (2024)
MASP: Scalable GNN-based Planning for Multi-Agent Navigation
von: Yang, Xinyi, et al.
Veröffentlicht: (2023)
von: Yang, Xinyi, et al.
Veröffentlicht: (2023)
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
von: Alamdari, Parand A., et al.
Veröffentlicht: (2023)
von: Alamdari, Parand A., et al.
Veröffentlicht: (2023)
A Picture is Worth A Thousand Numbers: Enabling LLMs Reason about Time Series via Visualization
von: Liu, Haoxin, et al.
Veröffentlicht: (2024)
von: Liu, Haoxin, et al.
Veröffentlicht: (2024)
Is Escalation Worth It? A Decision-Theoretic Characterization of LLM Cascades
von: Bouchard, Dylan
Veröffentlicht: (2026)
von: Bouchard, Dylan
Veröffentlicht: (2026)
Hybrid Motion Planning with Deep Reinforcement Learning for Mobile Robot Navigation
von: Kolomeytsev, Yury, et al.
Veröffentlicht: (2025)
von: Kolomeytsev, Yury, et al.
Veröffentlicht: (2025)
Preferential subspace identification (PSID) with forward-backward smoothing
von: Sani, Omid G., et al.
Veröffentlicht: (2025)
von: Sani, Omid G., et al.
Veröffentlicht: (2025)
Bayesian Optimization for Function-Valued Responses under Min-Max Criteria
von: Ahadi, Pouya, et al.
Veröffentlicht: (2025)
von: Ahadi, Pouya, et al.
Veröffentlicht: (2025)
When Planning Fails Despite Correct Execution: On Epistemic Calibration for LLM-Based Multi-Agent Systems
von: Wang, Zehao, et al.
Veröffentlicht: (2026)
von: Wang, Zehao, et al.
Veröffentlicht: (2026)
A Noise is Worth Diffusion Guidance
von: Ahn, Donghoon, et al.
Veröffentlicht: (2024)
von: Ahn, Donghoon, et al.
Veröffentlicht: (2024)
When Sensors Fail: Temporal Sequence Models for Robust PPO under Sensor Drift
von: Vogt-Lowell, Kevin, et al.
Veröffentlicht: (2026)
von: Vogt-Lowell, Kevin, et al.
Veröffentlicht: (2026)
Turn Waste into Worth: Rectifying Top-$k$ Router of MoE
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)
Thinking Forward and Backward: Effective Backward Planning with Large Language Models
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
Provably Robust Bayesian Counterfactual Explanations under Model Changes
von: Duell, Jamie, et al.
Veröffentlicht: (2026)
von: Duell, Jamie, et al.
Veröffentlicht: (2026)
Remember This Event That Year? Assessing Temporal Information and Reasoning in Large Language Models
von: Beniwal, Himanshu, et al.
Veröffentlicht: (2024)
von: Beniwal, Himanshu, et al.
Veröffentlicht: (2024)
Augmenting The Weather: A Hybrid Counterfactual-SMOTE Algorithm for Improving Crop Growth Prediction When Climate Changes
von: Temraz, Mohammed, et al.
Veröffentlicht: (2025)
von: Temraz, Mohammed, et al.
Veröffentlicht: (2025)
When is Tree Search Useful for LLM Planning? It Depends on the Discriminator
von: Chen, Ziru, et al.
Veröffentlicht: (2024)
von: Chen, Ziru, et al.
Veröffentlicht: (2024)
How Many Human Survey Respondents is a Large Language Model Worth? An Uncertainty Quantification Perspective
von: Huang, Chengpiao, et al.
Veröffentlicht: (2025)
von: Huang, Chengpiao, et al.
Veröffentlicht: (2025)
Causal Unlearning in Collaborative Optimization: Exact and Approximate Influence Reversal under Adversarial Contributions
von: Mahdavi, Ali, et al.
Veröffentlicht: (2026)
von: Mahdavi, Ali, et al.
Veröffentlicht: (2026)
Pareto-Optimal Offline Reinforcement Learning via Smooth Tchebysheff Scalarization
von: Bhatnagar, Aadyot, et al.
Veröffentlicht: (2026)
von: Bhatnagar, Aadyot, et al.
Veröffentlicht: (2026)
Online Multi-Label Classification under Noisy and Changing Label Distribution
von: Zou, Yizhang, et al.
Veröffentlicht: (2024)
von: Zou, Yizhang, et al.
Veröffentlicht: (2024)
Partner Modelling Emerges in Recurrent Agents (But Only When It Matters)
von: Mon-Williams, Ruaridh, et al.
Veröffentlicht: (2025)
von: Mon-Williams, Ruaridh, et al.
Veröffentlicht: (2025)
When and How Human Curation Backfires: Preference Alignment under Multi-Model Self-Consuming Loop
von: Zhang, Yang, et al.
Veröffentlicht: (2026)
von: Zhang, Yang, et al.
Veröffentlicht: (2026)
Different Victims, Same Layout: Email Visual Similarity Detection for Enhanced Email Protection
von: Shukla, Sachin, et al.
Veröffentlicht: (2024)
von: Shukla, Sachin, et al.
Veröffentlicht: (2024)
The Mean is the Mirage: Entropy-Adaptive Model Merging under Heterogeneous Domain Shifts in Medical Imaging
von: Ambekar, Sameer, et al.
Veröffentlicht: (2026)
von: Ambekar, Sameer, et al.
Veröffentlicht: (2026)
Towards Understanding Subliminal Learning: When and How Hidden Biases Transfer
von: Schrodi, Simon, et al.
Veröffentlicht: (2025)
von: Schrodi, Simon, et al.
Veröffentlicht: (2025)
What is Your Data Worth to GPT? LLM-Scale Data Valuation with Influence Functions
von: Choe, Sang Keun, et al.
Veröffentlicht: (2024)
von: Choe, Sang Keun, et al.
Veröffentlicht: (2024)
Is Scaling Learned Optimizers Worth It? Evaluating The Value of VeLO's 4000 TPU Months
von: Rezk, Fady, et al.
Veröffentlicht: (2023)
von: Rezk, Fady, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Tracking Changing Probabilities via Dynamic Learners
von: Madani, Omid
Veröffentlicht: (2024) -
Programming by Backprop: An Instruction is Worth 100 Examples When Finetuning LLMs
von: Cook, Jonathan, et al.
Veröffentlicht: (2025) -
Choosing How to Remember: Adaptive Memory Structures for LLM Agents
von: Lu, Mingfei, et al.
Veröffentlicht: (2026) -
Vision HgNN: An Electron-Micrograph is Worth Hypergraph of Hypernodes
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2024) -
Learning to Remember: End-to-End Training of Memory Agents for Long-Context Reasoning
von: Zhang, Kehao, et al.
Veröffentlicht: (2026)