A Theoretical Justification for Asymmetric Actor-Critic Algorithms
Fuente:
arXiv
Saved in:
| Main Authors: | Lambrechts, Gaspard, Ernst, Damien, Mahajan, Aditya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access
by: Ebi, Daniel, et al.
Published: (2025)
by: Ebi, Daniel, et al.
Published: (2025)
Behind the Myth of Exploration in Policy Gradients
by: Bolland, Adrien, et al.
Published: (2024)
by: Bolland, Adrien, et al.
Published: (2024)
Informed POMDP: Leveraging Additional Information in Model-Based RL
by: Lambrechts, Gaspard, et al.
Published: (2023)
by: Lambrechts, Gaspard, et al.
Published: (2023)
Maximum-Entropy Exploration with Future State-Action Visitation Measures
by: Bolland, Adrien, et al.
Published: (2026)
by: Bolland, Adrien, et al.
Published: (2026)
Off-Policy Maximum Entropy RL with Future State and Action Visitation Measures
by: Bolland, Adrien, et al.
Published: (2024)
by: Bolland, Adrien, et al.
Published: (2024)
Parallelizing Autoregressive Generation with Variational State Space Models
by: Lambrechts, Gaspard, et al.
Published: (2024)
by: Lambrechts, Gaspard, et al.
Published: (2024)
Parallelizable memory recurrent units
by: De Geeter, Florent, et al.
Published: (2026)
by: De Geeter, Florent, et al.
Published: (2026)
Compatible Gradient Approximations for Actor-Critic Algorithms
by: Saglam, Baturay, et al.
Published: (2024)
by: Saglam, Baturay, et al.
Published: (2024)
Finite-Time Analysis of Three-Timescale Constrained Actor-Critic and Constrained Natural Actor-Critic Algorithms
by: Panda, Prashansa, et al.
Published: (2023)
by: Panda, Prashansa, et al.
Published: (2023)
Value Improved Actor Critic Algorithms
by: Oren, Yaniv, et al.
Published: (2024)
by: Oren, Yaniv, et al.
Published: (2024)
Actor-Critic without Actor
by: Ki, Donghyeon, et al.
Published: (2025)
by: Ki, Donghyeon, et al.
Published: (2025)
Actor-Critic Algorithm for Dynamic Expectile and CVaR
by: Luo, Yudong, et al.
Published: (2026)
by: Luo, Yudong, et al.
Published: (2026)
A Communication-Efficient Decentralized Actor-Critic Algorithm
by: Ren, Xiaoxing, et al.
Published: (2025)
by: Ren, Xiaoxing, et al.
Published: (2025)
Actor-Critic or Critic-Actor? A Tale of Two Time Scales
by: Bhatnagar, Shalabh, et al.
Published: (2022)
by: Bhatnagar, Shalabh, et al.
Published: (2022)
An Investigation of Batch Normalization in Off-Policy Actor-Critic Algorithms
by: Wang, Li, et al.
Published: (2025)
by: Wang, Li, et al.
Published: (2025)
D2 Actor Critic: Diffusion Actor Meets Distributional Critic
by: Zhang, Lunjun, et al.
Published: (2025)
by: Zhang, Lunjun, et al.
Published: (2025)
Actor-Critic Reinforcement Learning with Phased Actor
by: Wu, Ruofan, et al.
Published: (2024)
by: Wu, Ruofan, et al.
Published: (2024)
Generative Actor Critic
by: Qin, Aoyang, et al.
Published: (2025)
by: Qin, Aoyang, et al.
Published: (2025)
TAROT: Towards Essentially Domain-Invariant Robustness with Theoretical Justification
by: Yang, Dongyoon, et al.
Published: (2025)
by: Yang, Dongyoon, et al.
Published: (2025)
Efficient Exploration in Deep Reinforcement Learning: A Novel Bayesian Actor-Critic Algorithm
by: Rozanov, Nikolai
Published: (2024)
by: Rozanov, Nikolai
Published: (2024)
Primal-Only Actor Critic Algorithm for Robust Constrained Average Cost MDPs
by: Satheesh, Anirudh, et al.
Published: (2025)
by: Satheesh, Anirudh, et al.
Published: (2025)
Pseudo-Quantized Actor-Critic Algorithm for Robustness to Noisy Temporal Difference Error
by: Kobayashi, Taisuke
Published: (2026)
by: Kobayashi, Taisuke
Published: (2026)
Weak Convergence Analysis of Online Neural Actor-Critic Algorithms
by: Lam, Samuel Chun-Hei, et al.
Published: (2024)
by: Lam, Samuel Chun-Hei, et al.
Published: (2024)
Operator-Theoretic Foundations and Policy Gradient Methods for General MDPs with Unbounded Costs
by: Gupta, Abhishek, et al.
Published: (2026)
by: Gupta, Abhishek, et al.
Published: (2026)
A Single-Loop Deep Actor-Critic Algorithm for Constrained Reinforcement Learning with Provable Convergence
by: Wang, Kexuan, et al.
Published: (2023)
by: Wang, Kexuan, et al.
Published: (2023)
Limits of Actor-Critic Algorithms for Decision Tree Policies Learning in IBMDPs
by: Kohler, Hector, et al.
Published: (2023)
by: Kohler, Hector, et al.
Published: (2023)
An Actor-Critic Algorithm with Function Approximation for Risk Sensitive Cost Markov Decision Processes
by: Guin, Soumyajit, et al.
Published: (2025)
by: Guin, Soumyajit, et al.
Published: (2025)
XQCfD: Accelerating Fast Actor-Critic Algorithms with Prior Data and Prior Policies
by: Palenicek, Daniel, et al.
Published: (2026)
by: Palenicek, Daniel, et al.
Published: (2026)
Scaling Effects and Uncertainty Quantification in Neural Actor Critic Algorithms
by: Georgoudios, Nikos, et al.
Published: (2026)
by: Georgoudios, Nikos, et al.
Published: (2026)
Adaptive Ensemble Aggregation for Actor-Critics
by: Werge, Nicklas, et al.
Published: (2025)
by: Werge, Nicklas, et al.
Published: (2025)
Risk-Sensitive Exponential Actor Critic
by: Granados, Alonso, et al.
Published: (2026)
by: Granados, Alonso, et al.
Published: (2026)
Safe Langevin Soft Actor Critic
by: Keswani, Mahesh, et al.
Published: (2026)
by: Keswani, Mahesh, et al.
Published: (2026)
On the Convergence of Single-Timescale Actor-Critic
by: Kumar, Navdeep, et al.
Published: (2024)
by: Kumar, Navdeep, et al.
Published: (2024)
Actor-Critic with Active Importance Sampling
by: Molaei, Majid, et al.
Published: (2026)
by: Molaei, Majid, et al.
Published: (2026)
Agent-state based policies in POMDPs: Beyond belief-state MDPs
by: Sinha, Amit, et al.
Published: (2024)
by: Sinha, Amit, et al.
Published: (2024)
A Case for Validation Buffer in Pessimistic Actor-Critic
by: Nauman, Michal, et al.
Published: (2024)
by: Nauman, Michal, et al.
Published: (2024)
Deep Actor-Critics with Tight Risk Certificates
by: Tasdighi, Bahareh, et al.
Published: (2025)
by: Tasdighi, Bahareh, et al.
Published: (2025)
Refined Analysis of Entropy-Regularized Actor-Critic
by: Labbi, Safwan, et al.
Published: (2026)
by: Labbi, Safwan, et al.
Published: (2026)
Actor-Critic Pretraining for Proximal Policy Optimization
by: Kernbach, Andreas, et al.
Published: (2026)
by: Kernbach, Andreas, et al.
Published: (2026)
PAC-Bayesian Soft Actor-Critic Learning
by: Tasdighi, Bahareh, et al.
Published: (2023)
by: Tasdighi, Bahareh, et al.
Published: (2023)
Similar Items
-
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access
by: Ebi, Daniel, et al.
Published: (2025) -
Behind the Myth of Exploration in Policy Gradients
by: Bolland, Adrien, et al.
Published: (2024) -
Informed POMDP: Leveraging Additional Information in Model-Based RL
by: Lambrechts, Gaspard, et al.
Published: (2023) -
Maximum-Entropy Exploration with Future State-Action Visitation Measures
by: Bolland, Adrien, et al.
Published: (2026) -
Off-Policy Maximum Entropy RL with Future State and Action Visitation Measures
by: Bolland, Adrien, et al.
Published: (2024)