Local Reinforcement Learning with Action-Conditioned Root Mean Squared Q-Functions
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Frank, Ren, Mengye |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reward Learning through Ranking Mean Squared Error
por: Kharyal, Chaitanya, et al.
Publicado: (2026)
por: Kharyal, Chaitanya, et al.
Publicado: (2026)
Improving Offline-to-Online Reinforcement Learning with Q Conditioned State Entropy Exploration
por: Zhang, Ziqi, et al.
Publicado: (2023)
por: Zhang, Ziqi, et al.
Publicado: (2023)
A Unifying Framework for Action-Conditional Self-Predictive Reinforcement Learning
por: Khetarpal, Khimya, et al.
Publicado: (2024)
por: Khetarpal, Khimya, et al.
Publicado: (2024)
SAMG: Offline-to-Online Reinforcement Learning via State-Action-Conditional Offline Model Guidance
por: Zhang, Liyu, et al.
Publicado: (2024)
por: Zhang, Liyu, et al.
Publicado: (2024)
Imagination-Limited Q-Learning for Offline Reinforcement Learning
por: Liu, Wenhui, et al.
Publicado: (2025)
por: Liu, Wenhui, et al.
Publicado: (2025)
On the Statistical Efficiency of Mean-Field Reinforcement Learning with General Function Approximation
por: Huang, Jiawei, et al.
Publicado: (2023)
por: Huang, Jiawei, et al.
Publicado: (2023)
Coarse-to-fine Q-Network with Action Sequence for Data-Efficient Reinforcement Learning
por: Seo, Younggyo, et al.
Publicado: (2024)
por: Seo, Younggyo, et al.
Publicado: (2024)
Are LLMs Prescient? A Continuous Evaluation using Daily News as the Oracle
por: Dai, Hui, et al.
Publicado: (2024)
por: Dai, Hui, et al.
Publicado: (2024)
Optimization Guarantees for Square-Root Natural-Gradient Variational Inference
por: Kumar, Navish, et al.
Publicado: (2025)
por: Kumar, Navish, et al.
Publicado: (2025)
Reinforcement Learning With Sparse-Executing Actions via Sparsity Regularization
por: Pang, Jing-Cheng, et al.
Publicado: (2021)
por: Pang, Jing-Cheng, et al.
Publicado: (2021)
Adapting the Behavior of Reinforcement Learning Agents to Changing Action Spaces and Reward Functions
por: de la Rosa, Raul, et al.
Publicado: (2026)
por: de la Rosa, Raul, et al.
Publicado: (2026)
FAST-Q: Fast-track Exploration with Adversarially Balanced State Representations for Counterfactual Action Estimation in Offline Reinforcement Learning
por: Agrawal, Pulkit, et al.
Publicado: (2025)
por: Agrawal, Pulkit, et al.
Publicado: (2025)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
por: Xu, Qiushui, et al.
Publicado: (2025)
por: Xu, Qiushui, et al.
Publicado: (2025)
Mildly Conservative Q-Learning for Offline Reinforcement Learning
por: Lyu, Jiafei, et al.
Publicado: (2022)
por: Lyu, Jiafei, et al.
Publicado: (2022)
Non-Linear Reinforcement Learning in Large Action Spaces: Structural Conditions and Sample-efficiency of Posterior Sampling
por: Agarwal, Alekh, et al.
Publicado: (2022)
por: Agarwal, Alekh, et al.
Publicado: (2022)
Context Tuning for In-Context Optimization
por: Lu, Jack, et al.
Publicado: (2025)
por: Lu, Jack, et al.
Publicado: (2025)
Aligning LLMs with Human Uncertainty: A Beta-Bernoulli Calibrator for LLM Forecasting
por: Dai, Hui, et al.
Publicado: (2026)
por: Dai, Hui, et al.
Publicado: (2026)
QSIM: Mitigating Overestimation in Multi-Agent Reinforcement Learning via Action Similarity Weighted Q-Learning
por: Li, Yuanjun, et al.
Publicado: (2026)
por: Li, Yuanjun, et al.
Publicado: (2026)
Adaptive Neighborhood-Constrained Q Learning for Offline Reinforcement Learning
por: Mao, Yixiu, et al.
Publicado: (2025)
por: Mao, Yixiu, et al.
Publicado: (2025)
One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow
por: Wang, Zeyuan, et al.
Publicado: (2025)
por: Wang, Zeyuan, et al.
Publicado: (2025)
Exclusively Penalized Q-learning for Offline Reinforcement Learning
por: Yeom, Junghyuk, et al.
Publicado: (2024)
por: Yeom, Junghyuk, et al.
Publicado: (2024)
Hypergraph and Latent ODE Learning for Multimodal Root Cause Localization in Microservices
por: Liu, Xin, et al.
Publicado: (2026)
por: Liu, Xin, et al.
Publicado: (2026)
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Counterfactual Explanations for Continuous Action Reinforcement Learning
por: Dong, Shuyang, et al.
Publicado: (2025)
por: Dong, Shuyang, et al.
Publicado: (2025)
Goal Exploration via Adaptive Skill Distribution for Goal-Conditioned Reinforcement Learning
por: Wu, Lisheng, et al.
Publicado: (2024)
por: Wu, Lisheng, et al.
Publicado: (2024)
In-Context Reinforcement Learning for Variable Action Spaces
por: Sinii, Viacheslav, et al.
Publicado: (2023)
por: Sinii, Viacheslav, et al.
Publicado: (2023)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
por: Huang, Xingshuai, et al.
Publicado: (2024)
por: Huang, Xingshuai, et al.
Publicado: (2024)
Model-based Offline Reinforcement Learning with Lower Expectile Q-Learning
por: Park, Kwanyoung, et al.
Publicado: (2024)
por: Park, Kwanyoung, et al.
Publicado: (2024)
Root Cause Attribution of Delivery Risks via Causal Discovery with Reinforcement Learning
por: Xiao, Minheng
Publicado: (2024)
por: Xiao, Minheng
Publicado: (2024)
Teaching RL Agents to Act Better: VLM as Action Advisor for Online Reinforcement Learning
por: Wu, Xiefeng, et al.
Publicado: (2025)
por: Wu, Xiefeng, et al.
Publicado: (2025)
ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
por: Wu, Kun, et al.
Publicado: (2024)
por: Wu, Kun, et al.
Publicado: (2024)
Sat-EnQ: Satisficing Ensembles of Weak Q-Learners for Reliable and Compute-Efficient Reinforcement Learning
por: Çiftçi, Ünver
Publicado: (2025)
por: Çiftçi, Ünver
Publicado: (2025)
SVL: Goal-Conditioned Reinforcement Learning as Survival Learning
por: Tiofack, Franki Nguimatsia, et al.
Publicado: (2026)
por: Tiofack, Franki Nguimatsia, et al.
Publicado: (2026)
Offline Reinforcement Learning with Penalized Action Noise Injection
por: Oh, JunHyeok, et al.
Publicado: (2025)
por: Oh, JunHyeok, et al.
Publicado: (2025)
Context-Sensitive Abstractions for Reinforcement Learning with Parameterized Actions
por: Nayyar, Rashmeet Kaur, et al.
Publicado: (2025)
por: Nayyar, Rashmeet Kaur, et al.
Publicado: (2025)
Model-based Reinforcement Learning for Parameterized Action Spaces
por: Zhang, Renhao, et al.
Publicado: (2024)
por: Zhang, Renhao, et al.
Publicado: (2024)
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces
por: Tiwari, Saket, et al.
Publicado: (2022)
por: Tiwari, Saket, et al.
Publicado: (2022)
Neuro-symbolic Action Masking for Deep Reinforcement Learning
por: Han, Shuai, et al.
Publicado: (2026)
por: Han, Shuai, et al.
Publicado: (2026)
PIQL: Projective Implicit Q-Learning with Support Constraint for Offline Reinforcement Learning
por: Han, Xinchen, et al.
Publicado: (2025)
por: Han, Xinchen, et al.
Publicado: (2025)
Double Successive Over-Relaxation Q-Learning with an Extension to Deep Reinforcement Learning
por: R, Shreyas S
Publicado: (2024)
por: R, Shreyas S
Publicado: (2024)
Ejemplares similares
-
Reward Learning through Ranking Mean Squared Error
por: Kharyal, Chaitanya, et al.
Publicado: (2026) -
Improving Offline-to-Online Reinforcement Learning with Q Conditioned State Entropy Exploration
por: Zhang, Ziqi, et al.
Publicado: (2023) -
A Unifying Framework for Action-Conditional Self-Predictive Reinforcement Learning
por: Khetarpal, Khimya, et al.
Publicado: (2024) -
SAMG: Offline-to-Online Reinforcement Learning via State-Action-Conditional Offline Model Guidance
por: Zhang, Liyu, et al.
Publicado: (2024) -
Imagination-Limited Q-Learning for Offline Reinforcement Learning
por: Liu, Wenhui, et al.
Publicado: (2025)