To Switch or Not to Switch? Balanced Policy Switching in Offline Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Tao, Yang, Xuzhi, Szabo, Zoltan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluation-Time Policy Switching for Offline Reinforcement Learning
by: Neggatu, Natinael Solomon, et al.
Published: (2025)
by: Neggatu, Natinael Solomon, et al.
Published: (2025)
Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning
by: Chemingui, Yassine, et al.
Published: (2024)
by: Chemingui, Yassine, et al.
Published: (2024)
Contextual Control without Memory Growth in a Context-Switching Task
by: Kim, Song-Ju
Published: (2026)
by: Kim, Song-Ju
Published: (2026)
Switching the Loss Reduces the Cost in Batch (Offline) Reinforcement Learning
by: Ayoub, Alex, et al.
Published: (2024)
by: Ayoub, Alex, et al.
Published: (2024)
Settling the Sample Complexity of Model-Based Offline Reinforcement Learning
by: Li, Gen, et al.
Published: (2022)
by: Li, Gen, et al.
Published: (2022)
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
by: Yan, Yuling, et al.
Published: (2022)
by: Yan, Yuling, et al.
Published: (2022)
Learning When to Switch: Adaptive Policy Selection via Reinforcement Learning
by: Tava, Chris
Published: (2025)
by: Tava, Chris
Published: (2025)
Load Balancing in Federated Learning
by: Javani, Alireza, et al.
Published: (2024)
by: Javani, Alireza, et al.
Published: (2024)
Policy-Guided Causal State Representation for Offline Reinforcement Learning Recommendation
by: Wang, Siyu, et al.
Published: (2025)
by: Wang, Siyu, et al.
Published: (2025)
Switched Feedback for the Multiple-Access Channel
by: Kosut, Oliver, et al.
Published: (2025)
by: Kosut, Oliver, et al.
Published: (2025)
Optimal Single-Policy Sample Complexity and Transient Coverage for Average-Reward Offline RL
by: Zurek, Matthew, et al.
Published: (2025)
by: Zurek, Matthew, et al.
Published: (2025)
Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability
by: Zhao, Qingyue, et al.
Published: (2026)
by: Zhao, Qingyue, et al.
Published: (2026)
Nyström $M$-Hilbert-Schmidt Independence Criterion
by: Kalinke, Florian, et al.
Published: (2023)
by: Kalinke, Florian, et al.
Published: (2023)
Model-free Reinforcement Learning of Semantic Communication by Stochastic Policy Gradient
by: Beck, Edgar, et al.
Published: (2023)
by: Beck, Edgar, et al.
Published: (2023)
Balancing Client Participation in Federated Learning Using AoI
by: Javani, Alireza, et al.
Published: (2025)
by: Javani, Alireza, et al.
Published: (2025)
Improved Offline Contextual Bandits with Second-Order Bounds: Betting and Freezing
by: Ryu, J. Jon, et al.
Published: (2025)
by: Ryu, J. Jon, et al.
Published: (2025)
Best Arm Identification with Possibly Biased Offline Data
by: Yang, Le, et al.
Published: (2025)
by: Yang, Le, et al.
Published: (2025)
SwitchTab: Switched Autoencoders Are Effective Tabular Learners
by: Wu, Jing, et al.
Published: (2024)
by: Wu, Jing, et al.
Published: (2024)
Continual Deep Reinforcement Learning for Decentralized Satellite Routing
by: Lozano-Cuadra, Federico, et al.
Published: (2024)
by: Lozano-Cuadra, Federico, et al.
Published: (2024)
Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning
by: Stojanovic, Stefan, et al.
Published: (2026)
by: Stojanovic, Stefan, et al.
Published: (2026)
Multi-Agent Deep Reinforcement Learning for Distributed Satellite Routing
by: Lozano-Cuadra, Federico, et al.
Published: (2024)
by: Lozano-Cuadra, Federico, et al.
Published: (2024)
Switched Flow Matching: Eliminating Singularities via Switching ODEs
by: Zhu, Qunxi, et al.
Published: (2024)
by: Zhu, Qunxi, et al.
Published: (2024)
A Memory-Based Reinforcement Learning Approach to Integrated Sensing and Communication
by: Nikbakht, Homa, et al.
Published: (2024)
by: Nikbakht, Homa, et al.
Published: (2024)
Information-Theoretic Minimax Regret Bounds for Reinforcement Learning based on Duality
by: Bongole, Raghav, et al.
Published: (2024)
by: Bongole, Raghav, et al.
Published: (2024)
Koopman-Based Generalization of Deep Reinforcement Learning With Application to Wireless Communications
by: Termehchi, Atefeh, et al.
Published: (2025)
by: Termehchi, Atefeh, et al.
Published: (2025)
On the Exponential Convergence for Offline RLHF with Pairwise Comparisons
by: Chen, Zhirui, et al.
Published: (2024)
by: Chen, Zhirui, et al.
Published: (2024)
Maximizing the Promptness of Metaverse Systems using Edge Computing by Deep Reinforcement Learning
by: Thi-Thanh, Tam Ninh, et al.
Published: (2025)
by: Thi-Thanh, Tam Ninh, et al.
Published: (2025)
Mode Switching-based STAR-RIS with Discrete Phase Shifters
by: Alishahi, MohammadHossein, et al.
Published: (2025)
by: Alishahi, MohammadHossein, et al.
Published: (2025)
Identifying Reasons for Contraceptive Switching from Real-World Data Using Large Language Models
by: Miao, Brenda Y., et al.
Published: (2024)
by: Miao, Brenda Y., et al.
Published: (2024)
The Curious Price of Distributional Robustness in Reinforcement Learning with a Generative Model
by: Shi, Laixi, et al.
Published: (2023)
by: Shi, Laixi, et al.
Published: (2023)
Deep Learning-Based Detection for Marker Codes over Insertion and Deletion Channels
by: Ma, Guochen, et al.
Published: (2024)
by: Ma, Guochen, et al.
Published: (2024)
Finite-Horizon Quickest Change Detection Balancing Latency with False Alarm Probability
by: Huang, Yu-Han, et al.
Published: (2025)
by: Huang, Yu-Han, et al.
Published: (2025)
Dynamic Switch Layers For Unsupervised Learning
by: Li, Haiguang, et al.
Published: (2024)
by: Li, Haiguang, et al.
Published: (2024)
Reconfigurable Intelligent Surfaces for THz: Hardware Impairments and Switching Technologies
by: Matos, Sérgio, et al.
Published: (2024)
by: Matos, Sérgio, et al.
Published: (2024)
AdaSwitch: An Adaptive Switching Meta-Algorithm for Learning-Augmented Bounded-Influence Problems
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
A Novel Deep Reinforcement Learning Method for Computation Offloading in Multi-User Mobile Edge Computing with Decentralization
by: Long, Nguyen Chi, et al.
Published: (2025)
by: Long, Nguyen Chi, et al.
Published: (2025)
Adapting Language Balance in Code-Switching Speech
by: Ugan, Enes Yavuz, et al.
Published: (2025)
by: Ugan, Enes Yavuz, et al.
Published: (2025)
Federated Multi-Agent Reinforcement Learning for Privacy-Preserving and Energy-Aware Resource Management in 6G Edge Networks
by: Andong, Francisco Javier Esono Nkulu, et al.
Published: (2025)
by: Andong, Francisco Javier Esono Nkulu, et al.
Published: (2025)
Generative Actor-Critic with Soft Bridge Policies
by: He, Ke, et al.
Published: (2026)
by: He, Ke, et al.
Published: (2026)
Random pairing MLE for estimation of item parameters in Rasch model
by: Yang, Yuepeng, et al.
Published: (2024)
by: Yang, Yuepeng, et al.
Published: (2024)
Similar Items
-
Evaluation-Time Policy Switching for Offline Reinforcement Learning
by: Neggatu, Natinael Solomon, et al.
Published: (2025) -
Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning
by: Chemingui, Yassine, et al.
Published: (2024) -
Contextual Control without Memory Growth in a Context-Switching Task
by: Kim, Song-Ju
Published: (2026) -
Switching the Loss Reduces the Cost in Batch (Offline) Reinforcement Learning
by: Ayoub, Alex, et al.
Published: (2024) -
Settling the Sample Complexity of Model-Based Offline Reinforcement Learning
by: Li, Gen, et al.
Published: (2022)