Stochastic Actor-Critic: Mitigating Overestimation via Temporal Aleatoric Uncertainty
Fuente:
arXiv
Guardado en:
| Autor principal: | Özalp, Uğurcan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
por: Lei, Yuheng, et al.
Publicado: (2022)
por: Lei, Yuheng, et al.
Publicado: (2022)
Application of Soft Actor-Critic Algorithms in Optimizing Wastewater Treatment with Time Delays Integration
por: Mohammadi, Esmaeel, et al.
Publicado: (2024)
por: Mohammadi, Esmaeel, et al.
Publicado: (2024)
ACING: Actor-Critic for Instruction Learning in Black-Box LLMs
por: Kharrat, Salma, et al.
Publicado: (2024)
por: Kharrat, Salma, et al.
Publicado: (2024)
MSACL: Multi-Step Actor-Critic Learning with Lyapunov Certificates for Exponentially Stabilizing Control
por: Zhang, Yongwei, et al.
Publicado: (2025)
por: Zhang, Yongwei, et al.
Publicado: (2025)
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences
por: Singh, Nikhil Kumar, et al.
Publicado: (2024)
por: Singh, Nikhil Kumar, et al.
Publicado: (2024)
MAGICS: Adversarial RL with Minimax Actors Guided by Implicit Critic Stackelberg for Convergent Neural Synthesis of Robot Safety
por: Wang, Justin, et al.
Publicado: (2024)
por: Wang, Justin, et al.
Publicado: (2024)
RL for Mitigating Cascading Failures: Targeted Exploration via Sensitivity Factors
por: Dwivedi, Anmol, et al.
Publicado: (2024)
por: Dwivedi, Anmol, et al.
Publicado: (2024)
Temporal Memory for Resource-Constrained Agents: Continual Learning via Stochastic Compress-Add-Smooth
por: Chertkov, Michael
Publicado: (2026)
por: Chertkov, Michael
Publicado: (2026)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
por: Cui, Mingxuan, et al.
Publicado: (2025)
por: Cui, Mingxuan, et al.
Publicado: (2025)
QSIM: Mitigating Overestimation in Multi-Agent Reinforcement Learning via Action Similarity Weighted Q-Learning
por: Li, Yuanjun, et al.
Publicado: (2026)
por: Li, Yuanjun, et al.
Publicado: (2026)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
por: Ibrahim, Sinan, et al.
Publicado: (2026)
por: Ibrahim, Sinan, et al.
Publicado: (2026)
Virtual Smart Metering in District Heating Networks via Heterogeneous Spatial-Temporal Graph Neural Networks
por: Niresi, Keivan Faghih, et al.
Publicado: (2026)
por: Niresi, Keivan Faghih, et al.
Publicado: (2026)
STO-RL: Offline RL under Sparse Rewards via LLM-Guided Subgoal Temporal Order
por: Gu, Chengyang, et al.
Publicado: (2026)
por: Gu, Chengyang, et al.
Publicado: (2026)
Wiener Chaos in Kernel Regression: Towards Untangling Aleatoric and Epistemic Uncertainty
por: Faulwasser, T., et al.
Publicado: (2023)
por: Faulwasser, T., et al.
Publicado: (2023)
Annealing Optimization for Progressive Learning with Stochastic Approximation
por: Mavridis, Christos, et al.
Publicado: (2022)
por: Mavridis, Christos, et al.
Publicado: (2022)
LogicGuard: Improving Embodied LLM agents through Temporal Logic based Critics
por: Gokhale, Anand, et al.
Publicado: (2025)
por: Gokhale, Anand, et al.
Publicado: (2025)
PID Accelerated Temporal Difference Algorithms
por: Bedaywi, Mark, et al.
Publicado: (2024)
por: Bedaywi, Mark, et al.
Publicado: (2024)
TubeDAgger: Reducing the Number of Expert Interventions with Stochastic Reach-Tubes
por: Lemmel, Julian, et al.
Publicado: (2025)
por: Lemmel, Julian, et al.
Publicado: (2025)
Fail Fast, or Ask: Mitigating the Deficiencies of Reasoning LLMs with Human-in-the-Loop Systems Engineering
por: Zellinger, Michael J., et al.
Publicado: (2025)
por: Zellinger, Michael J., et al.
Publicado: (2025)
Zono-Conformal Prediction: Zonotope-Based Uncertainty Quantification for Regression and Classification Tasks
por: Lützow, Laura, et al.
Publicado: (2025)
por: Lützow, Laura, et al.
Publicado: (2025)
Improving Mixed-Criticality Scheduling with Reinforcement Learning
por: El-Mahdy, Muhammad, et al.
Publicado: (2025)
por: El-Mahdy, Muhammad, et al.
Publicado: (2025)
Learning from Imperfect Demonstrations via Temporal Behavior Tree-Guided Trajectory Repair
por: Puranic, Aniruddh G., et al.
Publicado: (2026)
por: Puranic, Aniruddh G., et al.
Publicado: (2026)
Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability
por: Liu, Yushen, et al.
Publicado: (2026)
por: Liu, Yushen, et al.
Publicado: (2026)
Reinforcement Learning-enabled Satellite Constellation Reconfiguration and Retasking for Mission-Critical Applications
por: Alami, Hassan El, et al.
Publicado: (2024)
por: Alami, Hassan El, et al.
Publicado: (2024)
Learning Hidden Subgoals under Temporal Ordering Constraints in Reinforcement Learning
por: Xu, Duo, et al.
Publicado: (2024)
por: Xu, Duo, et al.
Publicado: (2024)
Deep Learning for Sequential Decision Making under Uncertainty: Foundations, Frameworks, and Frontiers
por: Buyuktahtakin, I. Esra
Publicado: (2026)
por: Buyuktahtakin, I. Esra
Publicado: (2026)
EMFusion: An Uncertainty-Aware Conditional Diffusion Framework for Frequency-Selective EMF Forecasting in Wireless Networks
por: Yan, Zijiang, et al.
Publicado: (2025)
por: Yan, Zijiang, et al.
Publicado: (2025)
Stochastic Learning of Computational Resource Usage as Graph Structured Multimarginal Schrödinger Bridge
por: Bondar, Georgiy A., et al.
Publicado: (2024)
por: Bondar, Georgiy A., et al.
Publicado: (2024)
ON-Traffic: An Operator Learning Framework for Online Traffic Flow Estimation and Uncertainty Quantification from Lagrangian Sensors
por: Rap, Jake, et al.
Publicado: (2025)
por: Rap, Jake, et al.
Publicado: (2025)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
por: Adibi, Arman, et al.
Publicado: (2024)
por: Adibi, Arman, et al.
Publicado: (2024)
Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm
por: Qiao, Ting, et al.
Publicado: (2024)
por: Qiao, Ting, et al.
Publicado: (2024)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
por: Zhu, Feng, et al.
Publicado: (2025)
por: Zhu, Feng, et al.
Publicado: (2025)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
por: Mitra, Aritra, et al.
Publicado: (2023)
por: Mitra, Aritra, et al.
Publicado: (2023)
FTT-GRU: A Hybrid Fast Temporal Transformer with GRU for Remaining Useful Life Prediction
por: Chirukiri, Varun Teja, et al.
Publicado: (2025)
por: Chirukiri, Varun Teja, et al.
Publicado: (2025)
Multiscale Spatio-Temporal Enhanced Short-term Load Forecasting of Electric Vehicle Charging Stations
por: Zhang, Zongbao, et al.
Publicado: (2024)
por: Zhang, Zongbao, et al.
Publicado: (2024)
A cGAN Ensemble-based Uncertainty-aware Surrogate Model for Offline Model-based Optimization in Industrial Control Problems
por: Feng, Cheng
Publicado: (2022)
por: Feng, Cheng
Publicado: (2022)
DASA: Delay-Adaptive Multi-Agent Stochastic Approximation
por: Fabbro, Nicolò Dal, et al.
Publicado: (2024)
por: Fabbro, Nicolò Dal, et al.
Publicado: (2024)
Generalized Information Gathering Under Dynamics Uncertainty
por: Palafox, Fernando, et al.
Publicado: (2026)
por: Palafox, Fernando, et al.
Publicado: (2026)
Online Location Planning for AI-Defined Vehicles: Optimizing Joint Tasks of Order Serving and Spatio-Temporal Heterogeneous Model Fine-Tuning
por: Zheng, Bokeng, et al.
Publicado: (2025)
por: Zheng, Bokeng, et al.
Publicado: (2025)
Uncertainty-Aware DRL for Autonomous Vehicle Crowd Navigation in Shared Space
por: Golchoubian, Mahsa, et al.
Publicado: (2024)
por: Golchoubian, Mahsa, et al.
Publicado: (2024)
Ejemplares similares
-
Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
por: Lei, Yuheng, et al.
Publicado: (2022) -
Application of Soft Actor-Critic Algorithms in Optimizing Wastewater Treatment with Time Delays Integration
por: Mohammadi, Esmaeel, et al.
Publicado: (2024) -
ACING: Actor-Critic for Instruction Learning in Black-Box LLMs
por: Kharrat, Salma, et al.
Publicado: (2024) -
MSACL: Multi-Step Actor-Critic Learning with Lyapunov Certificates for Exponentially Stabilizing Control
por: Zhang, Yongwei, et al.
Publicado: (2025) -
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences
por: Singh, Nikhil Kumar, et al.
Publicado: (2024)