Bounding-Box Inference for Error-Aware Model-Based Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Talvitie, Erin J., Shao, Zilei, Li, Huiying, Hu, Jinghan, Boerma, Jacob, Zhao, Rory, Wang, Xintong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mitigating Value Hallucination in Dyna Planning via Multistep Predecessor Models
by: Aminmansour, Farzane, et al.
Published: (2020)
by: Aminmansour, Farzane, et al.
Published: (2020)
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
by: Zhao, Heng, et al.
Published: (2026)
by: Zhao, Heng, et al.
Published: (2026)
Enhancing Robustness in Deep Reinforcement Learning: A Lyapunov Exponent Approach
by: Young, Rory, et al.
Published: (2024)
by: Young, Rory, et al.
Published: (2024)
Optimizing the Unknown: Black Box Bayesian Optimization with Energy-Based Model and Reinforcement Learning
by: Miao, Ruiyao, et al.
Published: (2025)
by: Miao, Ruiyao, et al.
Published: (2025)
Connecting the Dots: Collaborative Fine-tuning for Black-Box Vision-Language Models
by: Wang, Zhengbo, et al.
Published: (2024)
by: Wang, Zhengbo, et al.
Published: (2024)
SEBA: Sample-Efficient Black-Box Attacks on Visual Reinforcement Learning
by: Huang, Tairan, et al.
Published: (2025)
by: Huang, Tairan, et al.
Published: (2025)
A Survey on Neural Architecture Search Based on Reinforcement Learning
by: Shao, Wenzhu
Published: (2024)
by: Shao, Wenzhu
Published: (2024)
Breaking the Factorization Barrier in Diffusion Language Models
by: Li, Ian, et al.
Published: (2026)
by: Li, Ian, et al.
Published: (2026)
Adaptive Layer Splitting for Wireless LLM Inference in Edge Computing: A Model-Based Reinforcement Learning Approach
by: Chen, Yuxuan, et al.
Published: (2024)
by: Chen, Yuxuan, et al.
Published: (2024)
Hierarchical Scoring for Machine Learning Classifier Error Impact Evaluation
by: Lanus, Erin, et al.
Published: (2025)
by: Lanus, Erin, et al.
Published: (2025)
Striking a Balance in Fairness for Dynamic Systems Through Reinforcement Learning
by: Hu, Yaowei, et al.
Published: (2024)
by: Hu, Yaowei, et al.
Published: (2024)
Diffusion-Modeled Reinforcement Learning for Carbon and Risk-Aware Microgrid Optimization
by: Zhao, Yunyi, et al.
Published: (2025)
by: Zhao, Yunyi, et al.
Published: (2025)
BET: Explaining Deep Reinforcement Learning through The Error-Prone Decisions
by: Liu, Xiao, et al.
Published: (2024)
by: Liu, Xiao, et al.
Published: (2024)
From Observations to Events: Event-Aware World Model for Reinforcement Learning
by: Peng, Zhao-Han, et al.
Published: (2026)
by: Peng, Zhao-Han, et al.
Published: (2026)
Adversarial Tokenization
by: Geh, Renato Lui, et al.
Published: (2025)
by: Geh, Renato Lui, et al.
Published: (2025)
Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling
by: Zhao, Chenran, et al.
Published: (2026)
by: Zhao, Chenran, et al.
Published: (2026)
A New Error Temporal Difference Algorithm for Deep Reinforcement Learning in Microgrid Optimization
by: Yao, Fulong, et al.
Published: (2025)
by: Yao, Fulong, et al.
Published: (2025)
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
by: Bethell, Daniel, et al.
Published: (2024)
by: Bethell, Daniel, et al.
Published: (2024)
Is Prior-Free Black-Box Non-Stationary Reinforcement Learning Feasible?
by: Gerogiannis, Argyrios, et al.
Published: (2024)
by: Gerogiannis, Argyrios, et al.
Published: (2024)
Sharpness-Aware Black-Box Optimization
by: Ye, Feiyang, et al.
Published: (2024)
by: Ye, Feiyang, et al.
Published: (2024)
A Note on Loss Functions and Error Compounding in Model-based Reinforcement Learning
by: Jiang, Nan
Published: (2024)
by: Jiang, Nan
Published: (2024)
On The Sample Complexity Bounds In Bilevel Reinforcement Learning
by: Gaur, Mudit, et al.
Published: (2025)
by: Gaur, Mudit, et al.
Published: (2025)
BandPO: Bridging Trust Regions and Ratio Clipping via Probability-Aware Bounds for LLM Reinforcement Learning
by: Li, Yuan, et al.
Published: (2026)
by: Li, Yuan, et al.
Published: (2026)
Exact Stiefel Optimization for Probabilistic PLS: Closed-Form Updates, Error Bounds, and Calibrated Uncertainty
by: Hu, Haoran, et al.
Published: (2026)
by: Hu, Haoran, et al.
Published: (2026)
Boosting Soft Q-Learning by Bounding
by: Adamczyk, Jacob, et al.
Published: (2024)
by: Adamczyk, Jacob, et al.
Published: (2024)
Self-Improving Safety Performance of Reinforcement Learning Based Driving with Black-Box Verification Algorithms
by: Dagdanov, Resul, et al.
Published: (2022)
by: Dagdanov, Resul, et al.
Published: (2022)
Provable Multi-Party Reinforcement Learning with Diverse Human Feedback
by: Zhong, Huiying, et al.
Published: (2024)
by: Zhong, Huiying, et al.
Published: (2024)
Evaluation-Aware Reinforcement Learning
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2025)
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2025)
Beyond Error-Based Optimization: Experience-Driven Symbolic Regression with Goal-Conditioned Reinforcement Learning
by: Sun, Jianwen, et al.
Published: (2026)
by: Sun, Jianwen, et al.
Published: (2026)
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance
by: Jin, Luozhijie, et al.
Published: (2025)
by: Jin, Luozhijie, et al.
Published: (2025)
Object-Centric World Models for Causality-Aware Reinforcement Learning
by: Nishimoto, Yosuke, et al.
Published: (2025)
by: Nishimoto, Yosuke, et al.
Published: (2025)
Kernelized Reinforcement Learning with Order Optimal Regret Bounds
by: Vakili, Sattar, et al.
Published: (2023)
by: Vakili, Sattar, et al.
Published: (2023)
Kernel-Based Distributed Q-Learning: A Scalable Reinforcement Learning Approach for Dynamic Treatment Regimes
by: Wang, Di, et al.
Published: (2023)
by: Wang, Di, et al.
Published: (2023)
eFedLLM: Efficient LLM Inference Based on Federated Learning
by: Ding, Shengwen, et al.
Published: (2024)
by: Ding, Shengwen, et al.
Published: (2024)
Value Bonuses using Ensemble Errors for Exploration in Reinforcement Learning
by: Wahab, Abdul, et al.
Published: (2026)
by: Wahab, Abdul, et al.
Published: (2026)
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
by: Rojas, Juan Sebastian, et al.
Published: (2026)
by: Rojas, Juan Sebastian, et al.
Published: (2026)
Thermodynamics of Reinforcement Learning Curricula
by: Adamczyk, Jacob, et al.
Published: (2026)
by: Adamczyk, Jacob, et al.
Published: (2026)
Uncertainty-Aware Multi-Objective Reinforcement Learning-Guided Diffusion Models for 3D De Novo Molecular Design
by: Chen, Lianghong, et al.
Published: (2025)
by: Chen, Lianghong, et al.
Published: (2025)
Entropy-Aware Model Initialization for Effective Exploration in Deep Reinforcement Learning
by: Jang, Sooyoung, et al.
Published: (2021)
by: Jang, Sooyoung, et al.
Published: (2021)
$λ$-models: Effective Decision-Aware Reinforcement Learning with Latent Models
by: Voelcker, Claas A, et al.
Published: (2023)
by: Voelcker, Claas A, et al.
Published: (2023)
Similar Items
-
Mitigating Value Hallucination in Dyna Planning via Multistep Predecessor Models
by: Aminmansour, Farzane, et al.
Published: (2020) -
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
by: Zhao, Heng, et al.
Published: (2026) -
Enhancing Robustness in Deep Reinforcement Learning: A Lyapunov Exponent Approach
by: Young, Rory, et al.
Published: (2024) -
Optimizing the Unknown: Black Box Bayesian Optimization with Energy-Based Model and Reinforcement Learning
by: Miao, Ruiyao, et al.
Published: (2025) -
Connecting the Dots: Collaborative Fine-tuning for Black-Box Vision-Language Models
by: Wang, Zhengbo, et al.
Published: (2024)