Towards Global Optimality for Practical Average Reward Reinforcement Learning without Mixing Time Oracles

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Patel, Bhrij, Suttle, Wesley A., Koppel, Alec, Aggarwal, Vaneet, Sadler, Brian M., Bedi, Amrit Singh, Manocha, Dinesh
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!

Similar Items