$κ$-Explorer: A Unified Framework for Active Model Estimation in MDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Xihe, Mitra, Urbashi, Javidi, Tara |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Relative Entropy to Minimax: A Unified Framework for Coverage in MDPs
by: Gu, Xihe, et al.
Published: (2026)
by: Gu, Xihe, et al.
Published: (2026)
Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning
by: Bozkus, Talha, et al.
Published: (2024)
by: Bozkus, Talha, et al.
Published: (2024)
Trojan Cleansing with Neural Collapse
by: Gu, Xihe, et al.
Published: (2024)
by: Gu, Xihe, et al.
Published: (2024)
A Multi-Agent Multi-Environment Mixed Q-Learning for Partially Decentralized Wireless Network Optimization
by: Bozkus, Talha, et al.
Published: (2024)
by: Bozkus, Talha, et al.
Published: (2024)
ModShift: Model Privacy via Designed Shifts
by: Kherani, Nomaan A., et al.
Published: (2025)
by: Kherani, Nomaan A., et al.
Published: (2025)
Multi-Timescale Ensemble Q-learning for Markov Decision Process Policy Optimization
by: Bozkus, Talha, et al.
Published: (2024)
by: Bozkus, Talha, et al.
Published: (2024)
Partially Decentralized Multi-Agent Q-Learning via Digital Cousins for Wireless Networks
by: Bozkus, Talha, et al.
Published: (2025)
by: Bozkus, Talha, et al.
Published: (2025)
Consequences of Kernel Regularity for Bandit Optimization
by: Lee, Madison, et al.
Published: (2025)
by: Lee, Madison, et al.
Published: (2025)
Leveraging Digital Cousins for Ensemble Q-Learning in Large-Scale Wireless Networks
by: Bozkus, Talha, et al.
Published: (2024)
by: Bozkus, Talha, et al.
Published: (2024)
Coverage Analysis of Multi-Environment Q-Learning Algorithms for Wireless Network Optimization
by: Bozkus, Talha, et al.
Published: (2024)
by: Bozkus, Talha, et al.
Published: (2024)
Learning to Ask: Decision Transformers for Adaptive Quantitative Group Testing
by: Soleymani, Mahdi, et al.
Published: (2025)
by: Soleymani, Mahdi, et al.
Published: (2025)
Asymmetric Graph Error Control with Low Complexity in Causal Bandits
by: Peng, Chen, et al.
Published: (2024)
by: Peng, Chen, et al.
Published: (2024)
Convex and Non-convex Federated Learning with Stale Stochastic Gradients: Diminishing Step Size is All You Need
by: Zheng, Xinran, et al.
Published: (2026)
by: Zheng, Xinran, et al.
Published: (2026)
CI-CBM: Class-Incremental Concept Bottleneck Model for Interpretable Continual Learning
by: Javadi, Amirhosein, et al.
Published: (2026)
by: Javadi, Amirhosein, et al.
Published: (2026)
An Improved Model-Free Decision-Estimation Coefficient with Applications in Adversarial MDPs
by: Liu, Haolin, et al.
Published: (2025)
by: Liu, Haolin, et al.
Published: (2025)
Primal Dual Continual Learning: Balancing Stability and Plasticity through Adaptive Memory Allocation
by: Elenter, Juan, et al.
Published: (2023)
by: Elenter, Juan, et al.
Published: (2023)
Reward Guidance for Reinforcement Learning Tasks Based on Large Language Models: The LMGT Framework
by: Deng, Yongxin, et al.
Published: (2024)
by: Deng, Yongxin, et al.
Published: (2024)
Demystifying Linear MDPs and Novel Dynamics Aggregation Framework
by: Lee, Joongkyu, et al.
Published: (2024)
by: Lee, Joongkyu, et al.
Published: (2024)
Adaptive Estimation and Optimal Control in Offline Contextual MDPs without Stationarity
by: Bhattacharyya, Riddhiman, et al.
Published: (2026)
by: Bhattacharyya, Riddhiman, et al.
Published: (2026)
Evolutionary Multi-Task Optimization for LLM-Guided Program Discovery
by: Gozeten, Halil Alperen, et al.
Published: (2026)
by: Gozeten, Halil Alperen, et al.
Published: (2026)
Beyond Scalar Rewards: An Axiomatic Framework for Lexicographic MDPs
by: Shakerinava, Mehran, et al.
Published: (2025)
by: Shakerinava, Mehran, et al.
Published: (2025)
Optimal Horizon-Free Reward-Free Exploration for Linear Mixture MDPs
by: Zhang, Junkai, et al.
Published: (2023)
by: Zhang, Junkai, et al.
Published: (2023)
Time-Constrained Robust MDPs
by: Zouitine, Adil, et al.
Published: (2024)
by: Zouitine, Adil, et al.
Published: (2024)
SureFED: Robust Federated Learning via Uncertainty-Aware Inward and Outward Inspection
by: Heydaribeni, Nasimeh, et al.
Published: (2023)
by: Heydaribeni, Nasimeh, et al.
Published: (2023)
Hybrid Atomic Norm Sparse/Diffuse Channel Estimation
by: Lyu, Lei, et al.
Published: (2025)
by: Lyu, Lei, et al.
Published: (2025)
Reinforcement Learning for Infinite-Horizon Average-Reward Linear MDPs via Approximation by Discounted-Reward MDPs
by: Hong, Kihyuk, et al.
Published: (2024)
by: Hong, Kihyuk, et al.
Published: (2024)
Bring Your Own (Non-Robust) Algorithm to Solve Robust MDPs by Estimating The Worst Kernel
by: Wang, Kaixin, et al.
Published: (2023)
by: Wang, Kaixin, et al.
Published: (2023)
Pragmatic Curiosity: A Unified Framework for Hybrid Learning and Optimization via Active Inference
by: Li, Yingke, et al.
Published: (2026)
by: Li, Yingke, et al.
Published: (2026)
Robust Parameter Learning for Uncertain MDPs
by: Schnitzer, Yannik, et al.
Published: (2026)
by: Schnitzer, Yannik, et al.
Published: (2026)
On Value Iteration Convergence in Connected MDPs
by: Mustafin, Arsenii, et al.
Published: (2024)
by: Mustafin, Arsenii, et al.
Published: (2024)
MDPs with a State Sensing Cost
by: Kapoor, Vansh, et al.
Published: (2025)
by: Kapoor, Vansh, et al.
Published: (2025)
Truly No-Regret Learning in Constrained MDPs
by: Müller, Adrian, et al.
Published: (2024)
by: Müller, Adrian, et al.
Published: (2024)
A Unified Framework for Estimation of High-dimensional Conditional Factor Models
by: Chen, Qihui
Published: (2022)
by: Chen, Qihui
Published: (2022)
Sample Complexity Bounds for Linear Constrained MDPs with a Generative Model
by: Liu, Xingtu, et al.
Published: (2025)
by: Liu, Xingtu, et al.
Published: (2025)
Soft Robust MDPs and Risk-Sensitive MDPs: Equivalence, Policy Gradient, and Sample Complexity
by: Zhang, Runyu, et al.
Published: (2023)
by: Zhang, Runyu, et al.
Published: (2023)
Efficient Model-Free Exploration in Low-Rank MDPs
by: Mhammedi, Zakaria, et al.
Published: (2023)
by: Mhammedi, Zakaria, et al.
Published: (2023)
Causal-EPIG: A Prediction-Oriented Active Learning Framework for CATE Estimation
by: Gao, Erdun, et al.
Published: (2025)
by: Gao, Erdun, et al.
Published: (2025)
Learning Adversarial MDPs with Stochastic Hard Constraints
by: Stradi, Francesco Emanuele, et al.
Published: (2024)
by: Stradi, Francesco Emanuele, et al.
Published: (2024)
Solving Robust MDPs through No-Regret Dynamics
by: Guha, Etash Kumar
Published: (2023)
by: Guha, Etash Kumar
Published: (2023)
Eluder-based Regret for Stochastic Contextual MDPs
by: Levy, Orin, et al.
Published: (2022)
by: Levy, Orin, et al.
Published: (2022)
Similar Items
-
From Relative Entropy to Minimax: A Unified Framework for Coverage in MDPs
by: Gu, Xihe, et al.
Published: (2026) -
Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning
by: Bozkus, Talha, et al.
Published: (2024) -
Trojan Cleansing with Neural Collapse
by: Gu, Xihe, et al.
Published: (2024) -
A Multi-Agent Multi-Environment Mixed Q-Learning for Partially Decentralized Wireless Network Optimization
by: Bozkus, Talha, et al.
Published: (2024) -
ModShift: Model Privacy via Designed Shifts
by: Kherani, Nomaan A., et al.
Published: (2025)