Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ghanem, Abdelghani, Ghogho, Mounir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Confidence-gated training for efficient early-exit neural networks
von: Mokssit, Saad, et al.
Veröffentlicht: (2025)
von: Mokssit, Saad, et al.
Veröffentlicht: (2025)
Applications of machine learning and IoT for Outdoor Air Pollution Monitoring and Prediction: A Systematic Literature Review
von: Gryech, Ihsane, et al.
Veröffentlicht: (2024)
von: Gryech, Ihsane, et al.
Veröffentlicht: (2024)
Mutual Information Regularized Offline Reinforcement Learning
von: Ma, Xiao, et al.
Veröffentlicht: (2022)
von: Ma, Xiao, et al.
Veröffentlicht: (2022)
Mildly Conservative Regularized Evaluation for Offline Reinforcement Learning
von: Chen, Haohui, et al.
Veröffentlicht: (2025)
von: Chen, Haohui, et al.
Veröffentlicht: (2025)
Discrete Flow Matching for Offline-to-Online Reinforcement Learning
von: Khan, Fairoz Nower, et al.
Veröffentlicht: (2026)
von: Khan, Fairoz Nower, et al.
Veröffentlicht: (2026)
Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning
von: Gao, Chen-Xiao, et al.
Veröffentlicht: (2025)
von: Gao, Chen-Xiao, et al.
Veröffentlicht: (2025)
Flow Matching with Injected Noise for Offline-to-Online Reinforcement Learning
von: Shin, Yongjae, et al.
Veröffentlicht: (2026)
von: Shin, Yongjae, et al.
Veröffentlicht: (2026)
Adjoint Sampling: Highly Scalable Diffusion Samplers via Adjoint Matching
von: Havens, Aaron, et al.
Veröffentlicht: (2025)
von: Havens, Aaron, et al.
Veröffentlicht: (2025)
In-Dataset Trajectory Return Regularization for Offline Preference-based Reinforcement Learning
von: Tu, Songjun, et al.
Veröffentlicht: (2024)
von: Tu, Songjun, et al.
Veröffentlicht: (2024)
Epigraph-Guided Flow Matching for Safe and Performant Offline Reinforcement Learning
von: Tayal, Manan, et al.
Veröffentlicht: (2026)
von: Tayal, Manan, et al.
Veröffentlicht: (2026)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
von: Liu, Tenglong, et al.
Veröffentlicht: (2024)
von: Liu, Tenglong, et al.
Veröffentlicht: (2024)
Q-learning with Adjoint Matching
von: Li, Qiyang, et al.
Veröffentlicht: (2026)
von: Li, Qiyang, et al.
Veröffentlicht: (2026)
Early Exiting Predictive Coding Neural Networks for Edge AI
von: Zniber, Alaa, et al.
Veröffentlicht: (2023)
von: Zniber, Alaa, et al.
Veröffentlicht: (2023)
Improving Offline-to-Online Reinforcement Learning with Q Conditioned State Entropy Exploration
von: Zhang, Ziqi, et al.
Veröffentlicht: (2023)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2023)
Stable CDE Autoencoders with Acuity Regularization for Offline Reinforcement Learning in Sepsis Treatment
von: Gao, Yue
Veröffentlicht: (2025)
von: Gao, Yue
Veröffentlicht: (2025)
RAMAC: Multimodal Risk-Aware Offline Reinforcement Learning and the Role of Behavior Regularization
von: Fukazawa, Kai, et al.
Veröffentlicht: (2025)
von: Fukazawa, Kai, et al.
Veröffentlicht: (2025)
Robust Offline Reinforcement Learning with Linearly Structured f-Divergence Regularization
von: Tang, Cheng, et al.
Veröffentlicht: (2024)
von: Tang, Cheng, et al.
Veröffentlicht: (2024)
Offline Reinforcement Learning with Wasserstein Regularization via Optimal Transport Maps
von: Omura, Motoki, et al.
Veröffentlicht: (2025)
von: Omura, Motoki, et al.
Veröffentlicht: (2025)
Trust Region Q Adjoint Matching
von: Dong, Yonghoon, et al.
Veröffentlicht: (2026)
von: Dong, Yonghoon, et al.
Veröffentlicht: (2026)
Offline Trajectory Optimization for Offline Reinforcement Learning
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning
von: Wang, Qingjun, et al.
Veröffentlicht: (2026)
von: Wang, Qingjun, et al.
Veröffentlicht: (2026)
Matching Problems to Solutions: An Explainable Way of Solving Machine Learning Problems
von: Saleh, Lokman, et al.
Veröffentlicht: (2024)
von: Saleh, Lokman, et al.
Veröffentlicht: (2024)
Revisiting Neighborhood Aggregation in Graph Neural Networks for Node Classification using Statistical Signal Processing
von: Ghogho, Mounir
Veröffentlicht: (2024)
von: Ghogho, Mounir
Veröffentlicht: (2024)
Boosting Maximum Entropy Reinforcement Learning via One-Step Flow Matching
von: Li, Zeqiao, et al.
Veröffentlicht: (2026)
von: Li, Zeqiao, et al.
Veröffentlicht: (2026)
The Role of Deep Learning Regularizations on Actors in Offline RL
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
Preference Elicitation for Offline Reinforcement Learning
von: Pace, Alizée, et al.
Veröffentlicht: (2024)
von: Pace, Alizée, et al.
Veröffentlicht: (2024)
Offline Reinforcement Learning with Imbalanced Datasets
von: Jiang, Li, et al.
Veröffentlicht: (2023)
von: Jiang, Li, et al.
Veröffentlicht: (2023)
Simple Ingredients for Offline Reinforcement Learning
von: Cetin, Edoardo, et al.
Veröffentlicht: (2024)
von: Cetin, Edoardo, et al.
Veröffentlicht: (2024)
State-Constrained Offline Reinforcement Learning
von: Hepburn, Charles A., et al.
Veröffentlicht: (2024)
von: Hepburn, Charles A., et al.
Veröffentlicht: (2024)
The Generalization Gap in Offline Reinforcement Learning
von: Mediratta, Ishita, et al.
Veröffentlicht: (2023)
von: Mediratta, Ishita, et al.
Veröffentlicht: (2023)
Dataset Distillation for Offline Reinforcement Learning
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
Offline Reinforcement Learning with Imputed Rewards
von: Romeo, Carlo, et al.
Veröffentlicht: (2024)
von: Romeo, Carlo, et al.
Veröffentlicht: (2024)
FLAC: Maximum Entropy RL via Kinetic Energy Regularized Bridge Matching
von: Lv, Lei, et al.
Veröffentlicht: (2026)
von: Lv, Lei, et al.
Veröffentlicht: (2026)
OffSim: Offline Simulator for Model-based Offline Inverse Reinforcement Learning
von: Ahn, Woo-Jin, et al.
Veröffentlicht: (2025)
von: Ahn, Woo-Jin, et al.
Veröffentlicht: (2025)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
Imagination-Limited Q-Learning for Offline Reinforcement Learning
von: Liu, Wenhui, et al.
Veröffentlicht: (2025)
von: Liu, Wenhui, et al.
Veröffentlicht: (2025)
Mildly Conservative Q-Learning for Offline Reinforcement Learning
von: Lyu, Jiafei, et al.
Veröffentlicht: (2022)
von: Lyu, Jiafei, et al.
Veröffentlicht: (2022)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
von: Wibault, Clarisse, et al.
Veröffentlicht: (2026)
von: Wibault, Clarisse, et al.
Veröffentlicht: (2026)
Offline Reinforcement Learning with Universal Horizon Models
von: Chung, Hojun, et al.
Veröffentlicht: (2026)
von: Chung, Hojun, et al.
Veröffentlicht: (2026)
Flow Actor-Critic for Offline Reinforcement Learning
von: Chae, Jongseong, et al.
Veröffentlicht: (2026)
von: Chae, Jongseong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Confidence-gated training for efficient early-exit neural networks
von: Mokssit, Saad, et al.
Veröffentlicht: (2025) -
Applications of machine learning and IoT for Outdoor Air Pollution Monitoring and Prediction: A Systematic Literature Review
von: Gryech, Ihsane, et al.
Veröffentlicht: (2024) -
Mutual Information Regularized Offline Reinforcement Learning
von: Ma, Xiao, et al.
Veröffentlicht: (2022) -
Mildly Conservative Regularized Evaluation for Offline Reinforcement Learning
von: Chen, Haohui, et al.
Veröffentlicht: (2025) -
Discrete Flow Matching for Offline-to-Online Reinforcement Learning
von: Khan, Fairoz Nower, et al.
Veröffentlicht: (2026)