Contrastive UCB: Provably Efficient Contrastive Self-Supervised Learning in Online Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Qiu, Shuang, Wang, Lingxiao, Bai, Chenjia, Yang, Zhuoran, Wang, Zhaoran |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Embed to Control Partially Observed Systems: Representation Learning with Provable Sample Efficiency
by: Wang, Lingxiao, et al.
Published: (2022)
by: Wang, Lingxiao, et al.
Published: (2022)
Reinforcement Learning from Partial Observation: Linear Function Approximation with Provable Sample Efficiency
by: Cai, Qi, et al.
Published: (2022)
by: Cai, Qi, et al.
Published: (2022)
Pessimism in the Face of Confounders: Provably Efficient Offline Reinforcement Learning in Partially Observable Markov Decision Processes
by: Lu, Miao, et al.
Published: (2022)
by: Lu, Miao, et al.
Published: (2022)
Variational Dynamic for Self-Supervised Exploration in Deep Reinforcement Learning
by: Bai, Chenjia, et al.
Published: (2020)
by: Bai, Chenjia, et al.
Published: (2020)
Provably Efficient Exploration in Policy Optimization
by: Cai, Qi, et al.
Published: (2019)
by: Cai, Qi, et al.
Published: (2019)
Pessimistic Value Iteration for Multi-Task Data Sharing in Offline Reinforcement Learning
by: Bai, Chenjia, et al.
Published: (2024)
by: Bai, Chenjia, et al.
Published: (2024)
Self-Supervised Contrastive Learning is Approximately Supervised Contrastive Learning
by: Luthra, Achleshwar, et al.
Published: (2025)
by: Luthra, Achleshwar, et al.
Published: (2025)
Optimistic Policy Optimization is Provably Efficient in Non-stationary MDPs
by: Zhong, Han, et al.
Published: (2021)
by: Zhong, Han, et al.
Published: (2021)
Provably Efficient Information-Directed Sampling Algorithms for Multi-Agent Reinforcement Learning
by: Zhang, Qiaosheng, et al.
Published: (2024)
by: Zhang, Qiaosheng, et al.
Published: (2024)
Contrastive Representation for Data Filtering in Cross-Domain Offline Reinforcement Learning
by: Wen, Xiaoyu, et al.
Published: (2024)
by: Wen, Xiaoyu, et al.
Published: (2024)
Learning Dynamic Mechanisms in Unknown Environments: A Reinforcement Learning Approach
by: Qiu, Shuang, et al.
Published: (2022)
by: Qiu, Shuang, et al.
Published: (2022)
Provably Efficient UCB-type Algorithms For Learning Predictive State Representations
by: Huang, Ruiquan, et al.
Published: (2023)
by: Huang, Ruiquan, et al.
Published: (2023)
Bayesian Self-Supervised Contrastive Learning
by: Liu, Bin, et al.
Published: (2023)
by: Liu, Bin, et al.
Published: (2023)
Diverse Randomized Value Functions: A Provably Pessimistic Approach for Offline Reinforcement Learning
by: Yu, Xudong, et al.
Published: (2024)
by: Yu, Xudong, et al.
Published: (2024)
Provable Training for Graph Contrastive Learning
by: Yu, Yue, et al.
Published: (2023)
by: Yu, Yue, et al.
Published: (2023)
Efficient Cross-Domain Offline Reinforcement Learning with Dynamics- and Value-Aligned Data Filtering
by: Qiao, Zhongjian, et al.
Published: (2025)
by: Qiao, Zhongjian, et al.
Published: (2025)
Unifying Value Alignment and Assignment in Cross-Domain Offline Reinforcement Learning with Heterogeneous Datasets
by: Qiao, Zhongjian, et al.
Published: (2026)
by: Qiao, Zhongjian, et al.
Published: (2026)
A Generalized Learning Framework for Self-Supervised Contrastive Learning
by: Si, Lingyu, et al.
Published: (2025)
by: Si, Lingyu, et al.
Published: (2025)
An Augmentation-Aware Theory for Self-Supervised Contrastive Learning
by: Cui, Jingyi, et al.
Published: (2025)
by: Cui, Jingyi, et al.
Published: (2025)
Towards Robust Offline-to-Online Reinforcement Learning via Uncertainty and Smoothness
by: Wen, Xiaoyu, et al.
Published: (2023)
by: Wen, Xiaoyu, et al.
Published: (2023)
Instrumental Variable Value Iteration for Causal Offline Reinforcement Learning
by: Liao, Luofeng, et al.
Published: (2021)
by: Liao, Luofeng, et al.
Published: (2021)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Ensemble Successor Representations for Task Generalization in Offline-to-Online Reinforcement Learning
by: Wang, Changhong, et al.
Published: (2024)
by: Wang, Changhong, et al.
Published: (2024)
On the Alignment Between Supervised and Self-Supervised Contrastive Learning
by: Luthra, Achleshwar, et al.
Published: (2025)
by: Luthra, Achleshwar, et al.
Published: (2025)
Traversing Pareto Optimal Policies: Provably Efficient Multi-Objective Reinforcement Learning
by: Qiu, Shuang, et al.
Published: (2024)
by: Qiu, Shuang, et al.
Published: (2024)
Efficient Availability Attacks against Supervised and Contrastive Learning Simultaneously
by: Wang, Yihan, et al.
Published: (2024)
by: Wang, Yihan, et al.
Published: (2024)
Provable Optimization for Adversarial Fair Self-supervised Contrastive Learning
by: Qi, Qi, et al.
Published: (2024)
by: Qi, Qi, et al.
Published: (2024)
SelfBC: Self Behavior Cloning for Offline Reinforcement Learning
by: Liu, Shirong, et al.
Published: (2024)
by: Liu, Shirong, et al.
Published: (2024)
Contrastive Self-Supervised Learning at the Edge: An Energy Perspective
by: Famá, Fernanda, et al.
Published: (2025)
by: Famá, Fernanda, et al.
Published: (2025)
Collapse-Proof Non-Contrastive Self-Supervised Learning
by: Sansone, Emanuele, et al.
Published: (2024)
by: Sansone, Emanuele, et al.
Published: (2024)
Self-Reinforced Graph Contrastive Learning
by: Hsieh, Chou-Ying, et al.
Published: (2025)
by: Hsieh, Chou-Ying, et al.
Published: (2025)
Online Preference Alignment for Language Models via Count-based Exploration
by: Bai, Chenjia, et al.
Published: (2025)
by: Bai, Chenjia, et al.
Published: (2025)
Provable Contrastive Continual Learning
by: Wen, Yichen, et al.
Published: (2024)
by: Wen, Yichen, et al.
Published: (2024)
A Reinforcement Learning Approach in Multi-Phase Second-Price Auction Design
by: Ai, Rui, et al.
Published: (2022)
by: Ai, Rui, et al.
Published: (2022)
Variational Supervised Contrastive Learning
by: Wang, Ziwen, et al.
Published: (2025)
by: Wang, Ziwen, et al.
Published: (2025)
Can Temporal-Difference and Q-Learning Learn Representation? A Mean-Field Theory
by: Zhang, Yufeng, et al.
Published: (2020)
by: Zhang, Yufeng, et al.
Published: (2020)
Understanding Self-supervised Contrastive Learning through Supervised Objectives
by: Lee, Byeongchan
Published: (2025)
by: Lee, Byeongchan
Published: (2025)
Why the Agent Made that Decision: Contrastive Explanation Learning for Reinforcement Learning
by: Zuo, Rui, et al.
Published: (2024)
by: Zuo, Rui, et al.
Published: (2024)
Self-Supervised Contrastive Learning for Long-term Forecasting
by: Park, Junwoo, et al.
Published: (2024)
by: Park, Junwoo, et al.
Published: (2024)
Dual Perspectives on Non-Contrastive Self-Supervised Learning
by: Ponce, Jean, et al.
Published: (2025)
by: Ponce, Jean, et al.
Published: (2025)
Similar Items
-
Embed to Control Partially Observed Systems: Representation Learning with Provable Sample Efficiency
by: Wang, Lingxiao, et al.
Published: (2022) -
Reinforcement Learning from Partial Observation: Linear Function Approximation with Provable Sample Efficiency
by: Cai, Qi, et al.
Published: (2022) -
Pessimism in the Face of Confounders: Provably Efficient Offline Reinforcement Learning in Partially Observable Markov Decision Processes
by: Lu, Miao, et al.
Published: (2022) -
Variational Dynamic for Self-Supervised Exploration in Deep Reinforcement Learning
by: Bai, Chenjia, et al.
Published: (2020) -
Provably Efficient Exploration in Policy Optimization
by: Cai, Qi, et al.
Published: (2019)