Wasserstein Policy Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pfau, David, Davies, Ian, Borsa, Diana, Araujo, Joao G. M., Tracey, Brendan, van Hasselt, Hado |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
General Uncertainty Estimation with Delta Variances
von: Schmitt, Simon, et al.
Veröffentlicht: (2025)
von: Schmitt, Simon, et al.
Veröffentlicht: (2025)
Scalable Meta-Learning via Mixed-Mode Differentiation
von: Kemaev, Iurii, et al.
Veröffentlicht: (2025)
von: Kemaev, Iurii, et al.
Veröffentlicht: (2025)
Normalization and effective learning rates in reinforcement learning
von: Lyle, Clare, et al.
Veröffentlicht: (2024)
von: Lyle, Clare, et al.
Veröffentlicht: (2024)
A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2023)
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2023)
DataRater: Meta-Learned Dataset Curation
von: Calian, Dan A., et al.
Veröffentlicht: (2025)
von: Calian, Dan A., et al.
Veröffentlicht: (2025)
Optimizing Return Distributions with Distributional Dynamic Programming
von: Pires, Bernardo Ávila, et al.
Veröffentlicht: (2025)
von: Pires, Bernardo Ávila, et al.
Veröffentlicht: (2025)
Semantic-aware Wasserstein Policy Regularization for Large Language Model Alignment
von: Na, Byeonghu, et al.
Veröffentlicht: (2026)
von: Na, Byeonghu, et al.
Veröffentlicht: (2026)
Wasserstein Distributionally Robust Bayesian Optimization with Continuous Context
von: Micheli, Francesco, et al.
Veröffentlicht: (2025)
von: Micheli, Francesco, et al.
Veröffentlicht: (2025)
Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow
von: Koo, Juil, et al.
Veröffentlicht: (2026)
von: Koo, Juil, et al.
Veröffentlicht: (2026)
On-Policy Optimization of ANFIS Policies Using Proximal Policy Optimization
von: Shankar, Kaaustaaub, et al.
Veröffentlicht: (2025)
von: Shankar, Kaaustaaub, et al.
Veröffentlicht: (2025)
Spherical Tree-Sliced Wasserstein Distance
von: Tran, Viet-Hoang, et al.
Veröffentlicht: (2025)
von: Tran, Viet-Hoang, et al.
Veröffentlicht: (2025)
Soft Policy Optimization: Online Off-Policy RL for Sequence Models
von: Cohen, Taco, et al.
Veröffentlicht: (2025)
von: Cohen, Taco, et al.
Veröffentlicht: (2025)
Private Wasserstein Distance
von: Li, Wenqian, et al.
Veröffentlicht: (2024)
von: Li, Wenqian, et al.
Veröffentlicht: (2024)
Matrix Low-Rank Trust Region Policy Optimization
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
Distance-Based Tree-Sliced Wasserstein Distance
von: Tran, Hoang V., et al.
Veröffentlicht: (2025)
von: Tran, Hoang V., et al.
Veröffentlicht: (2025)
Autoregressive Policy Optimization for Constrained Allocation Tasks
von: Winkel, David, et al.
Veröffentlicht: (2024)
von: Winkel, David, et al.
Veröffentlicht: (2024)
Orthogonalized Policy Optimization:Policy Optimization as Orthogonal Projection in Hilbert Space
von: Zixian, Wang
Veröffentlicht: (2026)
von: Zixian, Wang
Veröffentlicht: (2026)
Disentangling the Causes of Plasticity Loss in Neural Networks
von: Lyle, Clare, et al.
Veröffentlicht: (2024)
von: Lyle, Clare, et al.
Veröffentlicht: (2024)
Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2022)
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2022)
Reflective Policy Optimization
von: Gan, Yaozhong, et al.
Veröffentlicht: (2024)
von: Gan, Yaozhong, et al.
Veröffentlicht: (2024)
Tree-Sliced Wasserstein Distance with Nonlinear Projection
von: Tran, Thanh, et al.
Veröffentlicht: (2025)
von: Tran, Thanh, et al.
Veröffentlicht: (2025)
Relative Policy-Transition Optimization for Fast Policy Transfer
von: Xu, Jiawei, et al.
Veröffentlicht: (2022)
von: Xu, Jiawei, et al.
Veröffentlicht: (2022)
Group Orthogonalized Policy Optimization:Group Policy Optimization as Orthogonal Projection in Hilbert Space
von: Zixian, Wang
Veröffentlicht: (2026)
von: Zixian, Wang
Veröffentlicht: (2026)
Delightful Policy Gradient
von: Osband, Ian
Veröffentlicht: (2026)
von: Osband, Ian
Veröffentlicht: (2026)
Provable Robust Overfitting Mitigation in Wasserstein Distributionally Robust Optimization
von: Liu, Shuang, et al.
Veröffentlicht: (2025)
von: Liu, Shuang, et al.
Veröffentlicht: (2025)
Wasserstein Distances, Neuronal Entanglement, and Sparsity
von: Sawmya, Shashata, et al.
Veröffentlicht: (2024)
von: Sawmya, Shashata, et al.
Veröffentlicht: (2024)
Canonical Variates in Wasserstein Metric Space
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
Tree Search-Based Policy Optimization under Stochastic Execution Delay
von: Valensi, David, et al.
Veröffentlicht: (2024)
von: Valensi, David, et al.
Veröffentlicht: (2024)
Single-stream Policy Optimization
von: Xu, Zhongwen, et al.
Veröffentlicht: (2025)
von: Xu, Zhongwen, et al.
Veröffentlicht: (2025)
Reparameterization Proximal Policy Optimization
von: Zhong, Hai, et al.
Veröffentlicht: (2025)
von: Zhong, Hai, et al.
Veröffentlicht: (2025)
Divergence-Augmented Policy Optimization
von: Wang, Qing, et al.
Veröffentlicht: (2025)
von: Wang, Qing, et al.
Veröffentlicht: (2025)
Decision Flow Policy Optimization
von: Hu, Jifeng, et al.
Veröffentlicht: (2025)
von: Hu, Jifeng, et al.
Veröffentlicht: (2025)
Soft Sequence Policy Optimization
von: Glazyrina, Svetlana, et al.
Veröffentlicht: (2026)
von: Glazyrina, Svetlana, et al.
Veröffentlicht: (2026)
Variational Delayed Policy Optimization
von: Wu, Qingyuan, et al.
Veröffentlicht: (2024)
von: Wu, Qingyuan, et al.
Veröffentlicht: (2024)
Reparameterization Flow Policy Optimization
von: Zhong, Hai, et al.
Veröffentlicht: (2026)
von: Zhong, Hai, et al.
Veröffentlicht: (2026)
Fractal Landscapes in Policy Optimization
von: Wang, Tao, et al.
Veröffentlicht: (2023)
von: Wang, Tao, et al.
Veröffentlicht: (2023)
Hybrid Group Relative Policy Optimization: A Multi-Sample Approach to Enhancing Policy Optimization
von: Sane, Soham
Veröffentlicht: (2025)
von: Sane, Soham
Veröffentlicht: (2025)
Tree-Sliced Wasserstein Distance: A Geometric Perspective
von: Tran, Viet-Hoang, et al.
Veröffentlicht: (2024)
von: Tran, Viet-Hoang, et al.
Veröffentlicht: (2024)
Offline Imitation from Observation via Primal Wasserstein State Occupancy Matching
von: Yan, Kai, et al.
Veröffentlicht: (2023)
von: Yan, Kai, et al.
Veröffentlicht: (2023)
A Unifying Framework for Action-Conditional Self-Predictive Reinforcement Learning
von: Khetarpal, Khimya, et al.
Veröffentlicht: (2024)
von: Khetarpal, Khimya, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
General Uncertainty Estimation with Delta Variances
von: Schmitt, Simon, et al.
Veröffentlicht: (2025) -
Scalable Meta-Learning via Mixed-Mode Differentiation
von: Kemaev, Iurii, et al.
Veröffentlicht: (2025) -
Normalization and effective learning rates in reinforcement learning
von: Lyle, Clare, et al.
Veröffentlicht: (2024) -
A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2023) -
DataRater: Meta-Learned Dataset Curation
von: Calian, Dan A., et al.
Veröffentlicht: (2025)