Flow Matching Policy Optimization with Mirror Descent and Entropy Constraints
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Ting, Orfanoudakis, Stavros, Lin, Nan, Daamen, Winnie, Hoogendoorn, Serge, Isufi, Elvin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Q-Net: Queue Length Estimation via Kalman-based Neural Networks
by: Gao, Ting, et al.
Published: (2025)
by: Gao, Ting, et al.
Published: (2025)
Sparse Covariance Neural Networks
by: Cavallo, Andrea, et al.
Published: (2024)
by: Cavallo, Andrea, et al.
Published: (2024)
Online Graph Filtering Over Expanding Graphs
by: Das, Bishwadeep, et al.
Published: (2024)
by: Das, Bishwadeep, et al.
Published: (2024)
SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control
by: Orfanoudakis, Stavros, et al.
Published: (2026)
by: Orfanoudakis, Stavros, et al.
Published: (2026)
Stochastic Sequential Decision Making over Expanding Networks with Graph Filtering
by: Gao, Zhan, et al.
Published: (2026)
by: Gao, Zhan, et al.
Published: (2026)
Beyond State-Wise Mirror Descent: Offline Policy Optimization with Parametric Policies
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
One-Step Flow Policy Mirror Descent
by: Chen, Tianyi, et al.
Published: (2025)
by: Chen, Tianyi, et al.
Published: (2025)
PowerFlowNet: Power Flow Approximation Using Message Passing Graph Neural Networks
by: Lin, Nan, et al.
Published: (2023)
by: Lin, Nan, et al.
Published: (2023)
Optimizing Electric Vehicles Charging using Large Language Models and Graph Neural Networks
by: Orfanoudakis, Stavros, et al.
Published: (2025)
by: Orfanoudakis, Stavros, et al.
Published: (2025)
Matched Topological Subspace Detector
by: Liu, Chengen, et al.
Published: (2025)
by: Liu, Chengen, et al.
Published: (2025)
On the Convergence of Policy in Unregularized Policy Mirror Descent
by: Lin, Dachao, et al.
Published: (2022)
by: Lin, Dachao, et al.
Published: (2022)
GKNet: Graph Kalman Filtering and Model Inference via Model-based Deep Learning
by: Sabbaqi, Mohammad, et al.
Published: (2025)
by: Sabbaqi, Mohammad, et al.
Published: (2025)
Spatiotemporal Covariance Neural Networks
by: Cavallo, Andrea, et al.
Published: (2024)
by: Cavallo, Andrea, et al.
Published: (2024)
Hodge-Compositional Edge Gaussian Processes
by: Yang, Maosheng, et al.
Published: (2023)
by: Yang, Maosheng, et al.
Published: (2023)
CoVariance Filters and Neural Networks over Hilbert Spaces
by: Battiloro, Claudio, et al.
Published: (2025)
by: Battiloro, Claudio, et al.
Published: (2025)
Online Learning Of Expanding Graphs
by: Rey, Samuel, et al.
Published: (2024)
by: Rey, Samuel, et al.
Published: (2024)
On the Effect of Regularization in Policy Mirror Descent
by: Kleuker, Jan Felix, et al.
Published: (2025)
by: Kleuker, Jan Felix, et al.
Published: (2025)
Policy Mirror Descent with Lookahead
by: Protopapas, Kimon, et al.
Published: (2024)
by: Protopapas, Kimon, et al.
Published: (2024)
Sample Complexity of Neural Policy Mirror Descent for Policy Optimization on Low-Dimensional Manifolds
by: Xu, Zhenghao, et al.
Published: (2023)
by: Xu, Zhenghao, et al.
Published: (2023)
Fair CoVariance Neural Networks
by: Cavallo, Andrea, et al.
Published: (2024)
by: Cavallo, Andrea, et al.
Published: (2024)
Functional Acceleration for Policy Mirror Descent
by: Chelu, Veronica, et al.
Published: (2024)
by: Chelu, Veronica, et al.
Published: (2024)
GNN-DT: Graph Neural Network Enhanced Decision Transformer for Efficient Optimization in Dynamic Environments
by: Orfanoudakis, Stavros, et al.
Published: (2025)
by: Orfanoudakis, Stavros, et al.
Published: (2025)
Topological Kalman Filtering on Cell Complexes
by: Liu, Chengen, et al.
Published: (2026)
by: Liu, Chengen, et al.
Published: (2026)
Precision Neural Networks: Joint Graph And Relational Learning
by: Cavallo, Andrea, et al.
Published: (2025)
by: Cavallo, Andrea, et al.
Published: (2025)
Covariance Scattering Transforms
by: Cavallo, Andrea, et al.
Published: (2025)
by: Cavallo, Andrea, et al.
Published: (2025)
Never Saddle for Reparameterized Steepest Descent as Mirror Flow
by: Jacobs, Tom, et al.
Published: (2026)
by: Jacobs, Tom, et al.
Published: (2026)
Stochastic MeanFlow Policies: One-Step Generative Control with Entropic Mirror Descent
by: Wang, Zeyuan, et al.
Published: (2026)
by: Wang, Zeyuan, et al.
Published: (2026)
Topology-Aware Graph Reinforcement Learning for Energy Storage Systems Optimal Dispatch in Distribution Networks
by: Gao, Shuyi, et al.
Published: (2026)
by: Gao, Shuyi, et al.
Published: (2026)
Higher-Order Topological Directionality and Directed Simplicial Neural Networks
by: Lecha, Manuel, et al.
Published: (2024)
by: Lecha, Manuel, et al.
Published: (2024)
Graph Filters for Signal Processing and Machine Learning on Graphs
by: Isufi, Elvin, et al.
Published: (2022)
by: Isufi, Elvin, et al.
Published: (2022)
On the Convergence of Policy Mirror Descent with Temporal Difference Evaluation
by: Liu, Jiacai, et al.
Published: (2025)
by: Liu, Jiacai, et al.
Published: (2025)
Mirror Descent-Type Algorithms for the Variational Inequality Problem with Functional Constraints
by: Alkousa, Mohammad S., et al.
Published: (2026)
by: Alkousa, Mohammad S., et al.
Published: (2026)
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
by: Sherman, Uri, et al.
Published: (2025)
by: Sherman, Uri, et al.
Published: (2025)
StaQ it! Growing neural networks for Policy Mirror Descent
by: Shilova, Alena, et al.
Published: (2025)
by: Shilova, Alena, et al.
Published: (2025)
Reference-Sampled Boltzmann Projection for KL-Regularized RLVR: Target-Matched Weighted SFT, Finite One-Shot Gaps, and Policy Mirror Descent
by: Shu, Yao, et al.
Published: (2026)
by: Shu, Yao, et al.
Published: (2026)
Group Entropies and Mirror Duality: A Class of Flexible Mirror Descent Updates for Machine Learning
by: Cichocki, Andrzej, et al.
Published: (2026)
by: Cichocki, Andrzej, et al.
Published: (2026)
Simplicial Convolutional Filters
by: Yang, Maosheng, et al.
Published: (2022)
by: Yang, Maosheng, et al.
Published: (2022)
Path-Space Mirror Descent for On-Policy Reinforcement Learning under the Generalized Schrödinger Bridge
by: Gong, Yuehu, et al.
Published: (2026)
by: Gong, Yuehu, et al.
Published: (2026)
Mirror Descent on Riemannian Manifolds
by: Jiang, Jiaxin, et al.
Published: (2026)
by: Jiang, Jiaxin, et al.
Published: (2026)
Parameter-free Mirror Descent
by: Jacobsen, Andrew, et al.
Published: (2022)
by: Jacobsen, Andrew, et al.
Published: (2022)
Similar Items
-
Q-Net: Queue Length Estimation via Kalman-based Neural Networks
by: Gao, Ting, et al.
Published: (2025) -
Sparse Covariance Neural Networks
by: Cavallo, Andrea, et al.
Published: (2024) -
Online Graph Filtering Over Expanding Graphs
by: Das, Bishwadeep, et al.
Published: (2024) -
SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control
by: Orfanoudakis, Stavros, et al.
Published: (2026) -
Stochastic Sequential Decision Making over Expanding Networks with Graph Filtering
by: Gao, Zhan, et al.
Published: (2026)