Saved in:
| Main Authors: | Lim, Joonyoung, Yoo, Younghwan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.00880 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Boosted Distributional Reinforcement Learning: Analysis and Healthcare Applications
by: Chen, Zequn, et al.
Published: (2026)
by: Chen, Zequn, et al.
Published: (2026)
Convergence of Neural Network Policies for Risk--Reward Optimization
by: Chen, Chang, et al.
Published: (2026)
by: Chen, Chang, et al.
Published: (2026)
Neural Actor-Critic Methods for Hamilton-Jacobi-Bellman PDEs: Asymptotic Analysis and Numerical Studies
by: Cohen, Samuel N., et al.
Published: (2025)
by: Cohen, Samuel N., et al.
Published: (2025)
System stabilization with policy optimization on unstable latent manifolds
by: Werner, Steffen W. R., et al.
Published: (2024)
by: Werner, Steffen W. R., et al.
Published: (2024)
Communicating Plans, Not Percepts: Scalable Multi-Agent Coordination with Embodied World Models
by: Hill, Brennen A., et al.
Published: (2025)
by: Hill, Brennen A., et al.
Published: (2025)
The Hard-Constraint PINNs for Interface Optimal Control Problems
by: Lai, Ming-Chih, et al.
Published: (2023)
by: Lai, Ming-Chih, et al.
Published: (2023)
Edge-Wise Graph-Instructed Neural Networks
by: Della Santa, Francesco, et al.
Published: (2024)
by: Della Santa, Francesco, et al.
Published: (2024)
Sample Complexity of Policy Gradient for Log-Growth Control
by: Pan, Qiuhua, et al.
Published: (2026)
by: Pan, Qiuhua, et al.
Published: (2026)
Two-Steps Diffusion Policy for Robotic Manipulation via Genetic Denoising
by: Clemente, Mateo, et al.
Published: (2025)
by: Clemente, Mateo, et al.
Published: (2025)
Integrating Vision Foundation Models with Reinforcement Learning for Enhanced Object Interaction
by: Farooq, Ahmad, et al.
Published: (2025)
by: Farooq, Ahmad, et al.
Published: (2025)
An Efficient Conditional Score-based Filter for High Dimensional Nonlinear Filtering Problems
by: Zeng, Zhijun, et al.
Published: (2025)
by: Zeng, Zhijun, et al.
Published: (2025)
Deep Relaxation of Controlled Stochastic Gradient Descent via Singular Perturbations
by: Bardi, Martino, et al.
Published: (2022)
by: Bardi, Martino, et al.
Published: (2022)
How to beat a Bayesian adversary
by: Ding, Zihan, et al.
Published: (2024)
by: Ding, Zihan, et al.
Published: (2024)
Global Convergence of Adjoint-Optimized Neural PDEs
by: Riedl, Konstantin, et al.
Published: (2025)
by: Riedl, Konstantin, et al.
Published: (2025)
Modeling Vehicle-Type-Specific Pedestrian Crash Avoidance Behavior in Safety-Critical Interactions Using Smooth-Mamba Deep Reinforcement Learning
by: Pu, Qingwen, et al.
Published: (2026)
by: Pu, Qingwen, et al.
Published: (2026)
BROS: Bias-Corrected Randomized Subspaces for Memory-Efficient Single-Loop Bilevel Optimization
by: Zhang, Hengrui, et al.
Published: (2026)
by: Zhang, Hengrui, et al.
Published: (2026)
Bayesian Conservative Policy Optimization (BCPO): A Novel Uncertainty-Calibrated Offline Reinforcement Learning with Credible Lower Bounds
by: Chatterjee, Debashis
Published: (2026)
by: Chatterjee, Debashis
Published: (2026)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
by: Zheng, Yaowei, et al.
Published: (2026)
by: Zheng, Yaowei, et al.
Published: (2026)
Improving LLM Agent Planning with In-Context Learning via Atomic Fact Augmentation and Lookahead Search
by: Holt, Samuel, et al.
Published: (2025)
by: Holt, Samuel, et al.
Published: (2025)
Neural Policy Iteration for Stochastic Optimal Control: A Physics-Informed Approach
by: Kim, Yeongjong, et al.
Published: (2025)
by: Kim, Yeongjong, et al.
Published: (2025)
Self-Organizing Dual-Buffer Adaptive Clustering Experience Replay (SODACER) for Safe Reinforcement Learning in Optimal Control
by: Amirabadi, Roya Khalili, et al.
Published: (2026)
by: Amirabadi, Roya Khalili, et al.
Published: (2026)
Eq.Bot: Enhance Robotic Manipulation Learning via Group Equivariant Canonicalization
by: Deng, Jian, et al.
Published: (2025)
by: Deng, Jian, et al.
Published: (2025)
Knowledge Integration in Differentiable Models: A Comparative Study of Data-Driven, Soft-Constrained, and Hard-Constrained Paradigms for Identification and Control of the Single Machine Infinite Bus System
by: Kang, Shinhoo, et al.
Published: (2026)
by: Kang, Shinhoo, et al.
Published: (2026)
Optimal sampling for stochastic and natural gradient descent
by: Gruhlke, Robert, et al.
Published: (2024)
by: Gruhlke, Robert, et al.
Published: (2024)
Objective Value Change and Shape-Based Accelerated Optimization for the Neural Network Approximation
by: Xie, Pengcheng, et al.
Published: (2025)
by: Xie, Pengcheng, et al.
Published: (2025)
Deep neural networks can provably solve Bellman equations for Markov decision processes without the curse of dimensionality
by: Jentzen, Arnulf, et al.
Published: (2025)
by: Jentzen, Arnulf, et al.
Published: (2025)
Physics-informed approach for exploratory Hamilton--Jacobi--Bellman equations via policy iterations
by: Kim, Yeongjong, et al.
Published: (2025)
by: Kim, Yeongjong, et al.
Published: (2025)
Deep Hilbert--Galerkin Methods for Infinite-Dimensional PDEs and Optimal Control
by: Cohen, Samuel N., et al.
Published: (2026)
by: Cohen, Samuel N., et al.
Published: (2026)
Machine Learning Algorithms for Improving Black Box Optimization Solvers
by: Kimiaei, Morteza, et al.
Published: (2025)
by: Kimiaei, Morteza, et al.
Published: (2025)
Wake-Informed 3D Path Planning for Autonomous Underwater Vehicles Using A* and Neural Network Approximations
by: Cooper-Baldock, Zachary, et al.
Published: (2025)
by: Cooper-Baldock, Zachary, et al.
Published: (2025)
Can Computational Reducibility Lead to Transferable Models for Graph Combinatorial Optimization?
by: Cantürk, Semih, et al.
Published: (2026)
by: Cantürk, Semih, et al.
Published: (2026)
Operator-Theoretic Foundations and Policy Gradient Methods for General MDPs with Unbounded Costs
by: Gupta, Abhishek, et al.
Published: (2026)
by: Gupta, Abhishek, et al.
Published: (2026)
A Robust Model-Based Approach for Continuous-Time Policy Evaluation with Unknown Lévy Process Dynamics
by: Ye, Qihao, et al.
Published: (2025)
by: Ye, Qihao, et al.
Published: (2025)
Policy Optimization over General State and Action Spaces
by: Ju, Caleb, et al.
Published: (2022)
by: Ju, Caleb, et al.
Published: (2022)
Monotone and Conservative Policy Iteration Beyond the Tabular Case
by: Eshwar, S. R., et al.
Published: (2025)
by: Eshwar, S. R., et al.
Published: (2025)
A Learning Stability Profile for Finite-Dimensional Learning Dynamics
by: Katende, Ronald
Published: (2025)
by: Katende, Ronald
Published: (2025)
Towards a General Recipe for Combinatorial Optimization with Multi-Filter GNNs
by: Wenkel, Frederik, et al.
Published: (2024)
by: Wenkel, Frederik, et al.
Published: (2024)
From Semi-Infinite Constraints to Structured Robust Policies: Optimal Gain Selection for Financial Systems
by: Hsieh, Chung-Han
Published: (2022)
by: Hsieh, Chung-Han
Published: (2022)
Convex Regularization and Convergence of Policy Gradient Flows under Safety Constraints
by: Malo, Pekka, et al.
Published: (2024)
by: Malo, Pekka, et al.
Published: (2024)
SAGE: Sign-Adaptive Gradient for Memory-Efficient LLM Optimization
by: Lee, Wooin, et al.
Published: (2026)
by: Lee, Wooin, et al.
Published: (2026)
Similar Items
-
Boosted Distributional Reinforcement Learning: Analysis and Healthcare Applications
by: Chen, Zequn, et al.
Published: (2026) -
Convergence of Neural Network Policies for Risk--Reward Optimization
by: Chen, Chang, et al.
Published: (2026) -
Neural Actor-Critic Methods for Hamilton-Jacobi-Bellman PDEs: Asymptotic Analysis and Numerical Studies
by: Cohen, Samuel N., et al.
Published: (2025) -
System stabilization with policy optimization on unstable latent manifolds
by: Werner, Steffen W. R., et al.
Published: (2024) -
Communicating Plans, Not Percepts: Scalable Multi-Agent Coordination with Embodied World Models
by: Hill, Brennen A., et al.
Published: (2025)