MICRO: Model-Based Offline Reinforcement Learning with a Conservative Bellman Operator
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xiao-Yin, Zhou, Xiao-Hu, Li, Guotao, Li, Hao, Gui, Mei-Jiang, Xiang, Tian-Yu, Huang, De-Xing, Hou, Zeng-Guang |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LEASE: Offline Preference-based Reinforcement Learning with High Sample Efficiency
by: Liu, Xiao-Yin, et al.
Published: (2024)
by: Liu, Xiao-Yin, et al.
Published: (2024)
CROP: Conservative Reward for Model-based Offline Policy Optimization
by: Li, Hao, et al.
Published: (2023)
by: Li, Hao, et al.
Published: (2023)
A Weight-aware-based Multi-source Unsupervised Domain Adaptation Method for Human Motion Intention Recognition
by: Liu, Xiao-Yin, et al.
Published: (2024)
by: Liu, Xiao-Yin, et al.
Published: (2024)
CAS-GAN for Contrast-free Angiography Synthesis
by: Huang, De-Xing, et al.
Published: (2024)
by: Huang, De-Xing, et al.
Published: (2024)
CAS-IQA: Teaching Vision-Language Models for Synthetic Angiography Quality Assessment
by: Wang, Bo, et al.
Published: (2025)
by: Wang, Bo, et al.
Published: (2025)
DOMAIN: MilDly COnservative Model-BAsed OfflINe Reinforcement Learning
by: Liu, Xiao-Yin, et al.
Published: (2023)
by: Liu, Xiao-Yin, et al.
Published: (2023)
VasoMIM: Vascular Anatomy-Aware Masked Image Modeling for Vessel Segmentation
by: Huang, De-Xing, et al.
Published: (2025)
by: Huang, De-Xing, et al.
Published: (2025)
SPIRONet: Spatial-Frequency Learning and Topological Channel Interaction Network for Vessel Segmentation
by: Huang, De-Xing, et al.
Published: (2024)
by: Huang, De-Xing, et al.
Published: (2024)
Online Adaptation via Dual-Stage Alignment and Self-Supervision for Fast-Calibration Brain-Computer Interfaces
by: Duan, Sheng-Bin, et al.
Published: (2025)
by: Duan, Sheng-Bin, et al.
Published: (2025)
Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Vascular anatomy-aware self-supervised pre-training for X-ray angiogram analysis
by: Huang, De-Xing, et al.
Published: (2026)
by: Huang, De-Xing, et al.
Published: (2026)
Gradual Transition from Bellman Optimality Operator to Bellman Operator in Online Reinforcement Learning
by: Omura, Motoki, et al.
Published: (2025)
by: Omura, Motoki, et al.
Published: (2025)
REASON: Probability map-guided dual-branch fusion framework for gastric content assessment
by: Xiao, Nu-Fnag, et al.
Published: (2025)
by: Xiao, Nu-Fnag, et al.
Published: (2025)
MOSformer: Momentum encoder-based inter-slice fusion transformer for medical image segmentation
by: Huang, De-Xing, et al.
Published: (2024)
by: Huang, De-Xing, et al.
Published: (2024)
Assessment of Angiography‐Based Renal Quantitative Flow Ratio Measurement in Patients with Atherosclerotic Renal Artery Stenosis
by: Xiang Huang, et al.
Published: (2024)
by: Xiang Huang, et al.
Published: (2024)
Target-Aligned Bellman Backup for Cross-domain Offline Reinforcement Learning
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
VLA Model Post-Training via Action-Chunked PPO and Self Behavior Cloning
by: Wang, Si-Cheng, et al.
Published: (2025)
by: Wang, Si-Cheng, et al.
Published: (2025)
Two-Step Offline Preference-Based Reinforcement Learning with Constrained Actions
by: Xu, Yinglun, et al.
Published: (2023)
by: Xu, Yinglun, et al.
Published: (2023)
Task-Oriented Learning for Automatic EEG Denoising
by: Xiang, Tian-Yu, et al.
Published: (2025)
by: Xiang, Tian-Yu, et al.
Published: (2025)
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
by: Golowich, Noah, et al.
Published: (2024)
by: Golowich, Noah, et al.
Published: (2024)
Mildly Conservative Q-Learning for Offline Reinforcement Learning
by: Lyu, Jiafei, et al.
Published: (2022)
by: Lyu, Jiafei, et al.
Published: (2022)
GraSP-STL: A Graph-Based Framework for Zero-Shot Signal Temporal Logic Planning via Offline Goal-Conditioned Reinforcement Learning
by: Hou, Ancheng, et al.
Published: (2026)
by: Hou, Ancheng, et al.
Published: (2026)
Offline-to-Online Reinforcement Learning with Classifier-Free Diffusion Generation
by: Huang, Xiao, et al.
Published: (2025)
by: Huang, Xiao, et al.
Published: (2025)
CDSA: Conservative Denoising Score-based Algorithm for Offline Reinforcement Learning
by: Liu, Zeyuan, et al.
Published: (2024)
by: Liu, Zeyuan, et al.
Published: (2024)
Structural Information-based Hierarchical Diffusion for Offline Reinforcement Learning
by: Zeng, Xianghua, et al.
Published: (2025)
by: Zeng, Xianghua, et al.
Published: (2025)
Policy-Based Trajectory Clustering in Offline Reinforcement Learning
by: Hu, Hao, et al.
Published: (2025)
by: Hu, Hao, et al.
Published: (2025)
Theoretical Barriers in Bellman-Based Reinforcement Learning
by: Pinon, Brieuc, et al.
Published: (2025)
by: Pinon, Brieuc, et al.
Published: (2025)
The Bellman Function for Level Sets of Sparse Operators
by: Fay, Irina Holmes, et al.
Published: (2025)
by: Fay, Irina Holmes, et al.
Published: (2025)
Distributional Bellman Operators over Mean Embeddings
by: Wenliang, Li Kevin, et al.
Published: (2023)
by: Wenliang, Li Kevin, et al.
Published: (2023)
Parameterized Projected Bellman Operator
by: Vincent, Théo, et al.
Published: (2023)
by: Vincent, Théo, et al.
Published: (2023)
VLA Model-Expert Collaboration for Bi-directional Manipulation Learning
by: Xiang, Tian-Yu, et al.
Published: (2025)
by: Xiang, Tian-Yu, et al.
Published: (2025)
Learning Novel Skills from Language-Generated Demonstrations
by: Jin, Ao-Qun, et al.
Published: (2024)
by: Jin, Ao-Qun, et al.
Published: (2024)
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL
by: Luo, Qin-Wen, et al.
Published: (2025)
by: Luo, Qin-Wen, et al.
Published: (2025)
ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
Mildly Conservative Regularized Evaluation for Offline Reinforcement Learning
by: Chen, Haohui, et al.
Published: (2025)
by: Chen, Haohui, et al.
Published: (2025)
Path-Coupled Bellman Flows for Distributional Reinforcement Learning
by: Xu, Boyang, et al.
Published: (2026)
by: Xu, Boyang, et al.
Published: (2026)
Search-Based Credit Assignment for Offline Preference-Based Reinforcement Learning
by: Gao, Xiancheng, et al.
Published: (2025)
by: Gao, Xiancheng, et al.
Published: (2025)
SAGAS: Semantic-Aware Graph-Assisted Stitching for Offline Temporal Logic Planning
by: Liu, Ruijia, et al.
Published: (2025)
by: Liu, Ruijia, et al.
Published: (2025)
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures
by: Yin, Ming, et al.
Published: (2025)
by: Yin, Ming, et al.
Published: (2025)
On Piecewise Affine Reachability with Bellman Operators
by: Varonka, Anton, et al.
Published: (2025)
by: Varonka, Anton, et al.
Published: (2025)
Similar Items
-
LEASE: Offline Preference-based Reinforcement Learning with High Sample Efficiency
by: Liu, Xiao-Yin, et al.
Published: (2024) -
CROP: Conservative Reward for Model-based Offline Policy Optimization
by: Li, Hao, et al.
Published: (2023) -
A Weight-aware-based Multi-source Unsupervised Domain Adaptation Method for Human Motion Intention Recognition
by: Liu, Xiao-Yin, et al.
Published: (2024) -
CAS-GAN for Contrast-free Angiography Synthesis
by: Huang, De-Xing, et al.
Published: (2024) -
CAS-IQA: Teaching Vision-Language Models for Synthetic Angiography Quality Assessment
by: Wang, Bo, et al.
Published: (2025)