Bayesian Conservative Policy Optimization (BCPO): A Novel Uncertainty-Calibrated Offline Reinforcement Learning with Credible Lower Bounds
Fuente:
arXiv
Saved in:
| Main Author: | Chatterjee, Debashis |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bayesian Linear Programming under Learned Uncertainty: Posterior Feasibility Guarantees, Scenario Certification, and Applications
by: Chatterjee, Debashis
Published: (2026)
by: Chatterjee, Debashis
Published: (2026)
Can a Bayesian Oracle Prevent Harm from an Agent?
by: Bengio, Yoshua, et al.
Published: (2024)
by: Bengio, Yoshua, et al.
Published: (2024)
Semantic-Constrained Federated Aggregation: Convergence Theory and Privacy-Utility Bounds for Knowledge-Enhanced Distributed Learning
by: Arafat, Jahidul
Published: (2025)
by: Arafat, Jahidul
Published: (2025)
Maximizing Rollout Informativeness under a Fixed Budget: A Submodular View of Tree Search for Tool-Use Agentic Reinforcement Learning
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
Classifier Calibration at Scale: An Empirical Study of Model-Agnostic Post-Hoc Methods
by: Manokhin, Valery, et al.
Published: (2026)
by: Manokhin, Valery, et al.
Published: (2026)
Machine Learning Algorithms for Improving Black Box Optimization Solvers
by: Kimiaei, Morteza, et al.
Published: (2025)
by: Kimiaei, Morteza, et al.
Published: (2025)
Actor-Curator: Co-adaptive Curriculum Learning via Policy-Improvement Bandits for RL Post-Training
by: Gu, Zhengyao, et al.
Published: (2026)
by: Gu, Zhengyao, et al.
Published: (2026)
Accelerated Multi-objective Task Learning using Modified Q-learning Algorithm
by: Rajamohan, Varun Prakash, et al.
Published: (2024)
by: Rajamohan, Varun Prakash, et al.
Published: (2024)
Autonomous AI Agents for Real-Time Affordable Housing Site Selection: Multi-Objective Reinforcement Learning Under Regulatory Constraints
by: Imanov, Olaf Yunus Laitinen, et al.
Published: (2026)
by: Imanov, Olaf Yunus Laitinen, et al.
Published: (2026)
Emotion-Inspired Learning Signals (EILS): A Homeostatic Framework for Adaptive Autonomous Agents
by: Tiwari, Dhruv
Published: (2025)
by: Tiwari, Dhruv
Published: (2025)
Long-Term Electricity Demand Prediction Using Non-negative Tensor Factorization and Genetic Algorithm-Driven Temporal Modeling
by: Masaki, Toma, et al.
Published: (2025)
by: Masaki, Toma, et al.
Published: (2025)
Gaussian Ensemble Belief Propagation for Efficient Inference in High-Dimensional Systems
by: MacKinlay, Dan, et al.
Published: (2024)
by: MacKinlay, Dan, et al.
Published: (2024)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
by: Xu, Zhe
Published: (2026)
by: Xu, Zhe
Published: (2026)
Ice Cream Doesn't Cause Drowning: Benchmarking LLMs Against Statistical Pitfalls in Causal Inference
by: Du, Jin, et al.
Published: (2025)
by: Du, Jin, et al.
Published: (2025)
RoboMoRe: LLM-based Robot Co-design via Joint Optimization of Morphology and Reward
by: Fang, Jiawei, et al.
Published: (2025)
by: Fang, Jiawei, et al.
Published: (2025)
The Six Sigma Agent: Achieving Enterprise-Grade Reliability in LLM Systems Through Consensus-Driven Decomposed Execution
by: Patel, Khush, et al.
Published: (2026)
by: Patel, Khush, et al.
Published: (2026)
Ontology Neural Network and ORTSF: A Framework for Topological Reasoning and Delay-Robust Control
by: Oh, Jaehong
Published: (2025)
by: Oh, Jaehong
Published: (2025)
SALE-Based Offline Reinforcement Learning with Ensemble Q-Networks
by: Chun, Zheng
Published: (2025)
by: Chun, Zheng
Published: (2025)
ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation
by: Chen, Kewei, et al.
Published: (2026)
by: Chen, Kewei, et al.
Published: (2026)
Closed-Form Beta Distribution Estimation from Sparse Statistics with Random Forest Implicit Regularization
by: Landers, Jonathan R.
Published: (2025)
by: Landers, Jonathan R.
Published: (2025)
Bayesian Generalized Nonlinear Models Offer Basis Free SINDy With Model Uncertainty
by: Hubin, Aliaksandr
Published: (2025)
by: Hubin, Aliaksandr
Published: (2025)
A Bayesian Framework for Regularized Estimation in Multivariate Models Integrating Approximate Computing Concepts
by: Kalina, Jan
Published: (2025)
by: Kalina, Jan
Published: (2025)
Diffusion-MPC in Discrete Domains: Feasibility Constraints, Horizon Effects, and Critic Alignment: Case study with Tetris
by: Wang, Haochuan Kevin
Published: (2026)
by: Wang, Haochuan Kevin
Published: (2026)
STACHE: Local Black-Box Explanations for Reinforcement Learning Policies
by: Elashkin, Andrew, et al.
Published: (2025)
by: Elashkin, Andrew, et al.
Published: (2025)
Multi-level meta-reinforcement learning with skill-based curriculum
by: Yang, Sichen, et al.
Published: (2026)
by: Yang, Sichen, et al.
Published: (2026)
Murphys Laws of AI Alignment: Why the Gap Always Wins
by: Gaikwad, Madhava
Published: (2025)
by: Gaikwad, Madhava
Published: (2025)
SCOPE: Selective Conformal Optimized Pairwise LLM Judging
by: Badshah, Sher, et al.
Published: (2026)
by: Badshah, Sher, et al.
Published: (2026)
When Are Two RLHF Objectives the Same?
by: Gaikwad, Madhava
Published: (2025)
by: Gaikwad, Madhava
Published: (2025)
A Structural Threshold in Decision Capacity Governs Collapse in Self-Play Reinforcement Learning
by: Kujur, Arahan
Published: (2026)
by: Kujur, Arahan
Published: (2026)
Pre-trained Visual Representations Generalize Where it Matters in Model-Based Reinforcement Learning
by: Jones, Scott, et al.
Published: (2025)
by: Jones, Scott, et al.
Published: (2025)
The Geometry of Thought: Disclosing the Transformer as a Tropical Polynomial Circuit
by: Alpay, Faruk, et al.
Published: (2026)
by: Alpay, Faruk, et al.
Published: (2026)
Learning coordinated badminton skills for legged manipulators
by: Ma, Yuntao, et al.
Published: (2025)
by: Ma, Yuntao, et al.
Published: (2025)
Scalable Bayesian Clustering for Integrative Analysis of Multi-View Data
by: Cabral, Rafael, et al.
Published: (2024)
by: Cabral, Rafael, et al.
Published: (2024)
Batched Nonparametric Bandits via k-Nearest Neighbor UCB
by: Arya, Sakshi
Published: (2025)
by: Arya, Sakshi
Published: (2025)
Explainable Bayesian deep learning through input-skip Latent Binary Bayesian Neural Networks
by: Høyheim, Eirik, et al.
Published: (2025)
by: Høyheim, Eirik, et al.
Published: (2025)
Discovering Algorithms with Computational Language Processing
by: Bourdais, Theo, et al.
Published: (2025)
by: Bourdais, Theo, et al.
Published: (2025)
TOPSIS-like metaheuristic for LABS problem
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
Enhancing Diversity in Multi-objective Feature Selection
by: Miyandoab, Sevil Zanjani, et al.
Published: (2024)
by: Miyandoab, Sevil Zanjani, et al.
Published: (2024)
Practical Quantum CIM Empowerment via All-Domestic-Core Agentic Large Model
by: Rui, Wang, et al.
Published: (2026)
by: Rui, Wang, et al.
Published: (2026)
FBMS: An R Package for Flexible Bayesian Model Selection and Model Averaging
by: Frommlet, Florian, et al.
Published: (2025)
by: Frommlet, Florian, et al.
Published: (2025)
Similar Items
-
Bayesian Linear Programming under Learned Uncertainty: Posterior Feasibility Guarantees, Scenario Certification, and Applications
by: Chatterjee, Debashis
Published: (2026) -
Can a Bayesian Oracle Prevent Harm from an Agent?
by: Bengio, Yoshua, et al.
Published: (2024) -
Semantic-Constrained Federated Aggregation: Convergence Theory and Privacy-Utility Bounds for Knowledge-Enhanced Distributed Learning
by: Arafat, Jahidul
Published: (2025) -
Maximizing Rollout Informativeness under a Fixed Budget: A Submodular View of Tree Search for Tool-Use Agentic Reinforcement Learning
by: Hu, Yuelin, et al.
Published: (2026) -
Classifier Calibration at Scale: An Empirical Study of Model-Agnostic Post-Hoc Methods
by: Manokhin, Valery, et al.
Published: (2026)