Scalable Bi-causal Optimal Transport via KL Relaxation and Policy Gradients
Fuente:
arXiv
Guardado en:
| Autores principales: | Cao, Haoyang, Hoekstra, Jesse, Xu, Renyuan, Xu, Yumin, Zhang, Ruixun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Convergence Rate in Nonlinear Two-Time-Scale Stochastic Approximation with State (Time)-Dependence
por: Chen, Zixi, et al.
Publicado: (2025)
por: Chen, Zixi, et al.
Publicado: (2025)
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
por: Han, Yinbin, et al.
Publicado: (2023)
por: Han, Yinbin, et al.
Publicado: (2023)
Inference of Utilities and Time Preference in Sequential Decision-Making
por: Cao, Haoyang, et al.
Publicado: (2024)
por: Cao, Haoyang, et al.
Publicado: (2024)
Fast Policy Learning for Linear Quadratic Control with Entropy Regularization
por: Guo, Xin, et al.
Publicado: (2023)
por: Guo, Xin, et al.
Publicado: (2023)
Stochastic Control for Fine-tuning Diffusion Models: Optimality, Regularity, and Convergence
por: Han, Yinbin, et al.
Publicado: (2024)
por: Han, Yinbin, et al.
Publicado: (2024)
Risk-sensitive Markov Decision Process and Learning under General Utility Functions
por: Wu, Zhengqi, et al.
Publicado: (2023)
por: Wu, Zhengqi, et al.
Publicado: (2023)
Enhancing Convergence of Decentralized Gradient Tracking under the KL Property
por: Chen, Xiaokai, et al.
Publicado: (2024)
por: Chen, Xiaokai, et al.
Publicado: (2024)
Scalable Approximate Algorithms for Optimal Transport Linear Models
por: Kacprzak, Tomasz, et al.
Publicado: (2025)
por: Kacprzak, Tomasz, et al.
Publicado: (2025)
ScaleBiO: Scalable Bilevel Optimization for LLM Data Reweighting
por: Pan, Rui, et al.
Publicado: (2024)
por: Pan, Rui, et al.
Publicado: (2024)
On Unbalanced Optimal Transport: Gradient Methods, Sparsity and Approximation Error
por: Nguyen, Quang Minh, et al.
Publicado: (2022)
por: Nguyen, Quang Minh, et al.
Publicado: (2022)
Solving Sparse \& High-Dimensional-Output Regression via Compression
por: Li, Renyuan, et al.
Publicado: (2024)
por: Li, Renyuan, et al.
Publicado: (2024)
Riemannian Neural Optimal Transport
por: Micheli, Alessandro, et al.
Publicado: (2026)
por: Micheli, Alessandro, et al.
Publicado: (2026)
Fast and Large-Scale Unbalanced Optimal Transport via its Semi-Dual and Adaptive Gradient Methods
por: Genans, Ferdinand
Publicado: (2026)
por: Genans, Ferdinand
Publicado: (2026)
Sobolev Gradient Ascent for Optimal Transport: Barycenter Optimization and Convergence Analysis
por: Kim, Kaheon, et al.
Publicado: (2025)
por: Kim, Kaheon, et al.
Publicado: (2025)
Sequential Bayesian Optimal Experimental Design in Infinite Dimensions via Policy Gradient Reinforcement Learning
por: Shen, Kaichen, et al.
Publicado: (2026)
por: Shen, Kaichen, et al.
Publicado: (2026)
Convergence Properties of Natural Gradient Descent for Minimizing KL Divergence
por: Datar, Adwait, et al.
Publicado: (2025)
por: Datar, Adwait, et al.
Publicado: (2025)
Adaptive Partitioning and Learning for Stochastic Control of Diffusion Processes
por: Jin, Hanqing, et al.
Publicado: (2025)
por: Jin, Hanqing, et al.
Publicado: (2025)
Feed m Birds with One Scone: Accelerating Multi-task Gradient Balancing via Bi-level Optimization
por: Chen, Xuxing, et al.
Publicado: (2026)
por: Chen, Xuxing, et al.
Publicado: (2026)
Inclusive KL Minimization: A Wasserstein-Fisher-Rao Gradient Flow Perspective
por: Zhu, Jia-Jie
Publicado: (2024)
por: Zhu, Jia-Jie
Publicado: (2024)
Learning Over-Relaxation Policies for ADMM with Convergence Guarantees
por: Lin, Junan, et al.
Publicado: (2026)
por: Lin, Junan, et al.
Publicado: (2026)
Model-Free Output Feedback Stabilization via Policy Gradient Methods
por: Zhang, Ankang, et al.
Publicado: (2026)
por: Zhang, Ankang, et al.
Publicado: (2026)
Unifying Distributionally Robust Optimization via Optimal Transport Theory
por: Blanchet, Jose, et al.
Publicado: (2023)
por: Blanchet, Jose, et al.
Publicado: (2023)
Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
por: Dus, Mathias
Publicado: (2026)
por: Dus, Mathias
Publicado: (2026)
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
por: Basu, Debabrota, et al.
Publicado: (2025)
por: Basu, Debabrota, et al.
Publicado: (2025)
Dual Conic Proxy for Semidefinite Relaxation of AC Optimal Power Flow
por: Qiu, Guancheng, et al.
Publicado: (2025)
por: Qiu, Guancheng, et al.
Publicado: (2025)
Slicing Unbalanced Optimal Transport
por: Bonet, Clément, et al.
Publicado: (2023)
por: Bonet, Clément, et al.
Publicado: (2023)
Decentralized and Equitable Optimal Transport
por: Lau, Ivan, et al.
Publicado: (2024)
por: Lau, Ivan, et al.
Publicado: (2024)
Automatic Outlier Rectification via Optimal Transport
por: Blanchet, Jose, et al.
Publicado: (2024)
por: Blanchet, Jose, et al.
Publicado: (2024)
Towards Optimal Branching of Linear and Semidefinite Relaxations for Neural Network Robustness Certification
por: Anderson, Brendon G., et al.
Publicado: (2021)
por: Anderson, Brendon G., et al.
Publicado: (2021)
Recurrent Natural Policy Gradient for POMDPs
por: Cayci, Semih, et al.
Publicado: (2024)
por: Cayci, Semih, et al.
Publicado: (2024)
Elementary Analysis of Policy Gradient Methods
por: Liu, Jiacai, et al.
Publicado: (2024)
por: Liu, Jiacai, et al.
Publicado: (2024)
Schrödinger bridge for generative AI: Soft-constrained formulation and convergence analysis
por: Ma, Jin, et al.
Publicado: (2025)
por: Ma, Jin, et al.
Publicado: (2025)
Optimal Guarantees for Algorithmic Reproducibility and Gradient Complexity in Convex Optimization
por: Zhang, Liang, et al.
Publicado: (2023)
por: Zhang, Liang, et al.
Publicado: (2023)
Heuristic Optimal Transport in Branching Networks
por: Andrecut, M.
Publicado: (2023)
por: Andrecut, M.
Publicado: (2023)
An Optimal Transport Approach for Network Regression
por: Zalles, Alex G., et al.
Publicado: (2024)
por: Zalles, Alex G., et al.
Publicado: (2024)
Optimal Transport with Tempered Exponential Measures
por: Amid, Ehsan, et al.
Publicado: (2023)
por: Amid, Ehsan, et al.
Publicado: (2023)
MMD-Regularized Unbalanced Optimal Transport
por: Manupriya, Piyushi, et al.
Publicado: (2020)
por: Manupriya, Piyushi, et al.
Publicado: (2020)
Linear Optimal Partial Transport Embedding
por: Bai, Yikun, et al.
Publicado: (2023)
por: Bai, Yikun, et al.
Publicado: (2023)
Complexity Lower Bounds of Adaptive Gradient Algorithms for Non-convex Stochastic Optimization under Relaxed Smoothness
por: Crawshaw, Michael, et al.
Publicado: (2025)
por: Crawshaw, Michael, et al.
Publicado: (2025)
Data-Driven Density Steering via the Gromov-Wasserstein Optimal Transport Distance
por: Nakashima, Haruto, et al.
Publicado: (2025)
por: Nakashima, Haruto, et al.
Publicado: (2025)
Ejemplares similares
-
Convergence Rate in Nonlinear Two-Time-Scale Stochastic Approximation with State (Time)-Dependence
por: Chen, Zixi, et al.
Publicado: (2025) -
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
por: Han, Yinbin, et al.
Publicado: (2023) -
Inference of Utilities and Time Preference in Sequential Decision-Making
por: Cao, Haoyang, et al.
Publicado: (2024) -
Fast Policy Learning for Linear Quadratic Control with Entropy Regularization
por: Guo, Xin, et al.
Publicado: (2023) -
Stochastic Control for Fine-tuning Diffusion Models: Optimality, Regularity, and Convergence
por: Han, Yinbin, et al.
Publicado: (2024)