Variational Online Mirror Descent for Robust Learning in Schrödinger Bridge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Dong-Sig, Kim, Jaein, Yoo, Hee Bin, Zhang, Byoung-Tak |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Coordinate-based Convolutional Kernels for Continuous SE(3) Equivariant and Efficient Point Cloud Analysis
von: Kim, Jaein, et al.
Veröffentlicht: (2026)
von: Kim, Jaein, et al.
Veröffentlicht: (2026)
Voronoi-based Second-order Descriptor with Whitened Metric in LiDAR Place Recognition
von: Kim, Jaein, et al.
Veröffentlicht: (2026)
von: Kim, Jaein, et al.
Veröffentlicht: (2026)
Optimistic Online Mirror Descent for Bridging Stochastic and Adversarial Online Convex Optimization
von: Chen, Sijia, et al.
Veröffentlicht: (2023)
von: Chen, Sijia, et al.
Veröffentlicht: (2023)
DUEL: Duplicate Elimination on Active Memory for Self-Supervised Class-Imbalanced Learning
von: Choi, Won-Seok, et al.
Veröffentlicht: (2024)
von: Choi, Won-Seok, et al.
Veröffentlicht: (2024)
Path-Space Mirror Descent for On-Policy Reinforcement Learning under the Generalized Schrödinger Bridge
von: Gong, Yuehu, et al.
Veröffentlicht: (2026)
von: Gong, Yuehu, et al.
Veröffentlicht: (2026)
OBSER: Object-Based Sub-Environment Recognition for Zero-Shot Environmental Inference
von: Choi, Won-Seok, et al.
Veröffentlicht: (2025)
von: Choi, Won-Seok, et al.
Veröffentlicht: (2025)
Adaptive Online Mirror Descent for Tchebycheff Scalarization in Multi-Objective Learning
von: Liu, Meitong, et al.
Veröffentlicht: (2024)
von: Liu, Meitong, et al.
Veröffentlicht: (2024)
PGA: Personalizing Grasping Agents with Single Human-Robot Interaction
von: Kim, Junghyun, et al.
Veröffentlicht: (2023)
von: Kim, Junghyun, et al.
Veröffentlicht: (2023)
Multimodal Anomaly Detection based on Deep Auto-Encoder for Object Slip Perception of Mobile Manipulation Robots
von: Yoo, Youngjae, et al.
Veröffentlicht: (2024)
von: Yoo, Youngjae, et al.
Veröffentlicht: (2024)
Unveiling the Significance of Toddler-Inspired Reward Transition in Goal-Oriented Reinforcement Learning
von: Park, Junseok, et al.
Veröffentlicht: (2024)
von: Park, Junseok, et al.
Veröffentlicht: (2024)
The Hidden Cost of Approximation in Online Mirror Descent
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
Nonstationary Generalized Linear Bandits with Discounted Online Mirror Descent
von: Lee, Joongkyu, et al.
Veröffentlicht: (2026)
von: Lee, Joongkyu, et al.
Veröffentlicht: (2026)
Fine-Grained Causal Dynamics Learning with Quantization for Improving Robustness in Reinforcement Learning
von: Hwang, Inwoo, et al.
Veröffentlicht: (2024)
von: Hwang, Inwoo, et al.
Veröffentlicht: (2024)
Treatment Stitching with Schrödinger Bridge for Enhancing Offline Reinforcement Learning in Adaptive Treatment Strategies
von: Shin, Dong-Hee, et al.
Veröffentlicht: (2025)
von: Shin, Dong-Hee, et al.
Veröffentlicht: (2025)
Improving LLM General Preference Alignment via Optimistic Online Mirror Descent
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
Mirror Descent-Type Algorithms for the Variational Inequality Problem with Functional Constraints
von: Alkousa, Mohammad S., et al.
Veröffentlicht: (2026)
von: Alkousa, Mohammad S., et al.
Veröffentlicht: (2026)
Value Mirror Descent for Reinforcement Learning
von: Jia, Zhichao, et al.
Veröffentlicht: (2026)
von: Jia, Zhichao, et al.
Veröffentlicht: (2026)
Learnable Loss Geometries with Mirror Descent for Scalable and Convergent Meta-Learning
von: Zhang, Yilang, et al.
Veröffentlicht: (2025)
von: Zhang, Yilang, et al.
Veröffentlicht: (2025)
Policy Mirror Descent with Temporal Difference Learning: Sample Complexity under Online Markov Data
von: Li, Wenye, et al.
Veröffentlicht: (2025)
von: Li, Wenye, et al.
Veröffentlicht: (2025)
On the Effect of Regularization in Policy Mirror Descent
von: Kleuker, Jan Felix, et al.
Veröffentlicht: (2025)
von: Kleuker, Jan Felix, et al.
Veröffentlicht: (2025)
Mirror Descent Actor Critic via Bounded Advantage Learning
von: Iwaki, Ryo
Veröffentlicht: (2025)
von: Iwaki, Ryo
Veröffentlicht: (2025)
Learning Mixtures of Experts with EM: A Mirror Descent Perspective
von: Fruytier, Quentin, et al.
Veröffentlicht: (2024)
von: Fruytier, Quentin, et al.
Veröffentlicht: (2024)
Meta-Learning with Versatile Loss Geometries for Fast Adaptation Using Mirror Descent
von: Zhang, Yilang, et al.
Veröffentlicht: (2023)
von: Zhang, Yilang, et al.
Veröffentlicht: (2023)
Parameter-free Mirror Descent
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2022)
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2022)
Policy Mirror Descent with Lookahead
von: Protopapas, Kimon, et al.
Veröffentlicht: (2024)
von: Protopapas, Kimon, et al.
Veröffentlicht: (2024)
Mirror Descent on Riemannian Manifolds
von: Jiang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Jiang, Jiaxin, et al.
Veröffentlicht: (2026)
A Mirror Descent Perspective of Smoothed Sign Descent
von: Wang, Shuyang, et al.
Veröffentlicht: (2024)
von: Wang, Shuyang, et al.
Veröffentlicht: (2024)
On the Convergence of Policy in Unregularized Policy Mirror Descent
von: Lin, Dachao, et al.
Veröffentlicht: (2022)
von: Lin, Dachao, et al.
Veröffentlicht: (2022)
Stability and Robustness via Regularization: Bandit Inference via Regularized Stochastic Mirror Descent
von: Halder, Budhaditya, et al.
Veröffentlicht: (2026)
von: Halder, Budhaditya, et al.
Veröffentlicht: (2026)
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
von: Xu, Hang, et al.
Veröffentlicht: (2024)
von: Xu, Hang, et al.
Veröffentlicht: (2024)
Adaptively Perturbed Mirror Descent for Learning in Games
von: Abe, Kenshi, et al.
Veröffentlicht: (2023)
von: Abe, Kenshi, et al.
Veröffentlicht: (2023)
One-Step Flow Policy Mirror Descent
von: Chen, Tianyi, et al.
Veröffentlicht: (2025)
von: Chen, Tianyi, et al.
Veröffentlicht: (2025)
Transformers Learn Latent Mixture Models In-Context via Mirror Descent
von: D'Angelo, Francesco, et al.
Veröffentlicht: (2026)
von: D'Angelo, Francesco, et al.
Veröffentlicht: (2026)
Mirror Descent Policy Optimisation for Robust Constrained Markov Decision Processes
von: Bossens, David M., et al.
Veröffentlicht: (2025)
von: Bossens, David M., et al.
Veröffentlicht: (2025)
Functional Acceleration for Policy Mirror Descent
von: Chelu, Veronica, et al.
Veröffentlicht: (2024)
von: Chelu, Veronica, et al.
Veröffentlicht: (2024)
Limited Memory Online Gradient Descent for Kernelized Pairwise Learning with Dynamic Averaging
von: AlQuabeh, Hilal, et al.
Veröffentlicht: (2024)
von: AlQuabeh, Hilal, et al.
Veröffentlicht: (2024)
Never Saddle for Reparameterized Steepest Descent as Mirror Flow
von: Jacobs, Tom, et al.
Veröffentlicht: (2026)
von: Jacobs, Tom, et al.
Veröffentlicht: (2026)
Efficient Monte Carlo Tree Search via On-the-Fly State-Conditioned Action Abstraction
von: Kwak, Yunhyeok, et al.
Veröffentlicht: (2024)
von: Kwak, Yunhyeok, et al.
Veröffentlicht: (2024)
Population-aware Online Mirror Descent for Mean-Field Games by Deep Reinforcement Learning
von: Wu, Zida, et al.
Veröffentlicht: (2024)
von: Wu, Zida, et al.
Veröffentlicht: (2024)
Population-aware Online Mirror Descent for Mean-Field Games with Common Noise by Deep Reinforcement Learning
von: Wu, Zida, et al.
Veröffentlicht: (2025)
von: Wu, Zida, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning Coordinate-based Convolutional Kernels for Continuous SE(3) Equivariant and Efficient Point Cloud Analysis
von: Kim, Jaein, et al.
Veröffentlicht: (2026) -
Voronoi-based Second-order Descriptor with Whitened Metric in LiDAR Place Recognition
von: Kim, Jaein, et al.
Veröffentlicht: (2026) -
Optimistic Online Mirror Descent for Bridging Stochastic and Adversarial Online Convex Optimization
von: Chen, Sijia, et al.
Veröffentlicht: (2023) -
DUEL: Duplicate Elimination on Active Memory for Self-Supervised Class-Imbalanced Learning
von: Choi, Won-Seok, et al.
Veröffentlicht: (2024) -
Path-Space Mirror Descent for On-Policy Reinforcement Learning under the Generalized Schrödinger Bridge
von: Gong, Yuehu, et al.
Veröffentlicht: (2026)