Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Dus, Mathias |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Designing Algorithms for Entropic Optimal Transport from an Optimisation Perspective
von: Srinivasan, Vishwak, et al.
Veröffentlicht: (2025)
von: Srinivasan, Vishwak, et al.
Veröffentlicht: (2025)
Two-layers neural networks for Schr{ö}dinger eigenvalue problems
von: Dus, Mathias, et al.
Veröffentlicht: (2024)
von: Dus, Mathias, et al.
Veröffentlicht: (2024)
Generalized Wasserstein Flow Matching: Transport Plans, Everywhere, All at Once
von: Piening, Moritz, et al.
Veröffentlicht: (2026)
von: Piening, Moritz, et al.
Veröffentlicht: (2026)
The geometry of financial institutions -- Wasserstein clustering of financial data
von: Riess, Lorenz, et al.
Veröffentlicht: (2023)
von: Riess, Lorenz, et al.
Veröffentlicht: (2023)
Properties of Discrete Sliced Wasserstein Losses
von: Tanguy, Eloi, et al.
Veröffentlicht: (2023)
von: Tanguy, Eloi, et al.
Veröffentlicht: (2023)
Admission Control of Quasi-Reversible Queueing Systems: Optimization and Reinforcement Learning
von: Comte, Céline, et al.
Veröffentlicht: (2025)
von: Comte, Céline, et al.
Veröffentlicht: (2025)
Linear convergence of proximal descent schemes on the Wasserstein space
von: Lascu, Razvan-Andrei, et al.
Veröffentlicht: (2024)
von: Lascu, Razvan-Andrei, et al.
Veröffentlicht: (2024)
Convergence of SGD for Training Neural Networks with Sliced Wasserstein Losses
von: Tanguy, Eloi
Veröffentlicht: (2023)
von: Tanguy, Eloi
Veröffentlicht: (2023)
Stochastic Inverse Problem: stability, regularization and Wasserstein gradient flow
von: Li, Qin, et al.
Veröffentlicht: (2024)
von: Li, Qin, et al.
Veröffentlicht: (2024)
Neural Wasserstein Gradient Flows for Maximum Mean Discrepancies with Riesz Kernels
von: Altekrüger, Fabian, et al.
Veröffentlicht: (2023)
von: Altekrüger, Fabian, et al.
Veröffentlicht: (2023)
Wasserstein Convergence of Score-based Generative Models under Semiconvexity and Discontinuous Gradients
von: Bruno, Stefano, et al.
Veröffentlicht: (2025)
von: Bruno, Stefano, et al.
Veröffentlicht: (2025)
Value Mirror Descent for Reinforcement Learning
von: Jia, Zhichao, et al.
Veröffentlicht: (2026)
von: Jia, Zhichao, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Random Time Horizons
von: Borrell, Enric Ribera, et al.
Veröffentlicht: (2025)
von: Borrell, Enric Ribera, et al.
Veröffentlicht: (2025)
Constrained Density Estimation via Optimal Transport
von: Hu, Yinan, et al.
Veröffentlicht: (2026)
von: Hu, Yinan, et al.
Veröffentlicht: (2026)
A Generalization Result for Convergence in Learning-to-Optimize
von: Sucker, Michael, et al.
Veröffentlicht: (2024)
von: Sucker, Michael, et al.
Veröffentlicht: (2024)
Online Learning and Optimization for Queues with Unknown Demand Curve and Service Distribution
von: Chen, Xinyun, et al.
Veröffentlicht: (2023)
von: Chen, Xinyun, et al.
Veröffentlicht: (2023)
Structure Matters: Dynamic Policy Gradient
von: Klein, Sara, et al.
Veröffentlicht: (2024)
von: Klein, Sara, et al.
Veröffentlicht: (2024)
The Fundamental Theorem of Weak Optimal Transport
von: Beiglböck, Mathias, et al.
Veröffentlicht: (2025)
von: Beiglböck, Mathias, et al.
Veröffentlicht: (2025)
Model Predictive Control is Almost Optimal for Restless Bandit
von: Gast, Nicolas, et al.
Veröffentlicht: (2024)
von: Gast, Nicolas, et al.
Veröffentlicht: (2024)
Model Predictive Control is almost Optimal for Heterogeneous Restless Multi-armed Bandits
von: Narasimha, Dheeraj, et al.
Veröffentlicht: (2025)
von: Narasimha, Dheeraj, et al.
Veröffentlicht: (2025)
Grassmannian Geometry and Global Convergence of Variable Projection for Neural Networks
von: Dus, Mathias
Veröffentlicht: (2026)
von: Dus, Mathias
Veröffentlicht: (2026)
A Distributional View of High Dimensional Optimization
von: Benning, Felix
Veröffentlicht: (2025)
von: Benning, Felix
Veröffentlicht: (2025)
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
Causal Optimal Coupling for Gaussian Input-Output Distributional Data
von: Xu, Daran, et al.
Veröffentlicht: (2026)
von: Xu, Daran, et al.
Veröffentlicht: (2026)
Accelerating Distributed Stochastic Optimization via Self-Repellent Random Walks
von: Hu, Jie, et al.
Veröffentlicht: (2024)
von: Hu, Jie, et al.
Veröffentlicht: (2024)
Ito Diffusion Approximation of Universal Ito Chains for Sampling, Optimization and Boosting
von: Ustimenko, Aleksei, et al.
Veröffentlicht: (2023)
von: Ustimenko, Aleksei, et al.
Veröffentlicht: (2023)
Data-driven Multistage Distributionally Robust Linear Optimization with Nested Distance
von: Gao, Rui, et al.
Veröffentlicht: (2024)
von: Gao, Rui, et al.
Veröffentlicht: (2024)
Improved Approximation Algorithms for Orthogonally Constrained Problems Using Semidefinite Optimization
von: Cory-Wright, Ryan, et al.
Veröffentlicht: (2025)
von: Cory-Wright, Ryan, et al.
Veröffentlicht: (2025)
Wasserstein Contraction of Coordinate Ascent Variational Inference
von: Caprio, Rocco, et al.
Veröffentlicht: (2026)
von: Caprio, Rocco, et al.
Veröffentlicht: (2026)
Faster Computation of Entropic Optimal Transport via Stable Low Frequency Modes
von: Chhaibi, Reda, et al.
Veröffentlicht: (2025)
von: Chhaibi, Reda, et al.
Veröffentlicht: (2025)
Optimization Trade-offs in Asynchronous Federated Learning: A Stochastic Networks Approach
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2026)
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2026)
Convergence Error Analysis of Reflected Gradient Langevin Dynamics for Globally Optimizing Non-Convex Constrained Problems
von: Sato, Kanji, et al.
Veröffentlicht: (2022)
von: Sato, Kanji, et al.
Veröffentlicht: (2022)
The Wasserstein Space of Stochastic Processes in Continuous Time
von: Bartl, Daniel, et al.
Veröffentlicht: (2025)
von: Bartl, Daniel, et al.
Veröffentlicht: (2025)
Optimizing Asynchronous Federated Learning: A Delicate Trade-Off Between Model-Parameter Staleness and Update Frequency
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2025)
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2025)
Learning-Based Pricing and Matching for Two-Sided Queues
von: Yang, Zixian, et al.
Veröffentlicht: (2024)
von: Yang, Zixian, et al.
Veröffentlicht: (2024)
Score-Aware Policy-Gradient and Performance Guarantees using Local Lyapunov Stability
von: Comte, Céline, et al.
Veröffentlicht: (2023)
von: Comte, Céline, et al.
Veröffentlicht: (2023)
Convergence of Actor-Critic Learning for Mean Field Games and Mean Field Control in Continuous Spaces
von: Fouque, Jean-Pierre, et al.
Veröffentlicht: (2025)
von: Fouque, Jean-Pierre, et al.
Veröffentlicht: (2025)
When Machine Learning Meets Importance Sampling: A More Efficient Rare Event Estimation Approach
von: Zhao, Ruoning, et al.
Veröffentlicht: (2025)
von: Zhao, Ruoning, et al.
Veröffentlicht: (2025)
Slicing Wasserstein Over Wasserstein Via Functional Optimal Transport
von: Piening, Moritz, et al.
Veröffentlicht: (2025)
von: Piening, Moritz, et al.
Veröffentlicht: (2025)
A Provably Convergent and Practical Algorithm for Gromov--Wasserstein Optimal Transport
von: Liang, Ling, et al.
Veröffentlicht: (2026)
von: Liang, Ling, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Designing Algorithms for Entropic Optimal Transport from an Optimisation Perspective
von: Srinivasan, Vishwak, et al.
Veröffentlicht: (2025) -
Two-layers neural networks for Schr{ö}dinger eigenvalue problems
von: Dus, Mathias, et al.
Veröffentlicht: (2024) -
Generalized Wasserstein Flow Matching: Transport Plans, Everywhere, All at Once
von: Piening, Moritz, et al.
Veröffentlicht: (2026) -
The geometry of financial institutions -- Wasserstein clustering of financial data
von: Riess, Lorenz, et al.
Veröffentlicht: (2023) -
Properties of Discrete Sliced Wasserstein Losses
von: Tanguy, Eloi, et al.
Veröffentlicht: (2023)