A Differential and Pointwise Control Approach to Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Nguyen, Minh, Bajaj, Chandrajit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Decoupled Continuous-Time Reinforcement Learning via Hamiltonian Flow
por: Nguyen, Minh
Publicado: (2026)
por: Nguyen, Minh
Publicado: (2026)
Reinforcement Learning for Molecular Dynamics Optimization: A Stochastic Pontryagin Maximum Principle Approach
por: Bajaj, Chandrajit, et al.
Publicado: (2022)
por: Bajaj, Chandrajit, et al.
Publicado: (2022)
Statistical and Algorithmic Foundations of Reinforcement Learning
por: Chi, Yuejie, et al.
Publicado: (2025)
por: Chi, Yuejie, et al.
Publicado: (2025)
Physics-informed neural networks via stochastic Hamiltonian dynamics learning
por: Bajaj, Chandrajit, et al.
Publicado: (2021)
por: Bajaj, Chandrajit, et al.
Publicado: (2021)
Learning to Fuse Temporal Proximity Networks: A Case Study in Chimpanzee Social Interactions
por: He, Yixuan, et al.
Publicado: (2025)
por: He, Yixuan, et al.
Publicado: (2025)
When Descent Is Too Stable: Event-Triggered Hamiltonian Learning to Optimize
por: Wang, Yi, et al.
Publicado: (2026)
por: Wang, Yi, et al.
Publicado: (2026)
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
por: Bareilles, Gilles, et al.
Publicado: (2026)
por: Bareilles, Gilles, et al.
Publicado: (2026)
Inverse Mixed-Integer Programming: Learning Constraints then Objective Functions
por: Kitaoka, Akira
Publicado: (2025)
por: Kitaoka, Akira
Publicado: (2025)
Training Dynamics of Multi-Head Softmax Attention for In-Context Learning: Emergence, Convergence, and Optimality
por: Chen, Siyu, et al.
Publicado: (2024)
por: Chen, Siyu, et al.
Publicado: (2024)
Motion Code: Robust Time Series Classification and Forecasting via Sparse Variational Multi-Stochastic Processes Learning
por: Bajaj, Chandrajit, et al.
Publicado: (2024)
por: Bajaj, Chandrajit, et al.
Publicado: (2024)
Sail into the Headwind: Alignment via Robust Rewards and Dynamic Labels against Reward Hacking
por: Rashidinejad, Paria, et al.
Publicado: (2024)
por: Rashidinejad, Paria, et al.
Publicado: (2024)
Sinkhorn Based Associative Memory Retrieval Using Spherical Hellinger Kantorovich Dynamics
por: Mustafi, Aratrika, et al.
Publicado: (2026)
por: Mustafi, Aratrika, et al.
Publicado: (2026)
Straight-Through meets Sparse Recovery: the Support Exploration Algorithm
por: Mohamed, Mimoun, et al.
Publicado: (2023)
por: Mohamed, Mimoun, et al.
Publicado: (2023)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
por: Han, Qiyang, et al.
Publicado: (2025)
por: Han, Qiyang, et al.
Publicado: (2025)
FraPPE: Fast and Efficient Preference-based Pure Exploration
por: Das, Udvas, et al.
Publicado: (2025)
por: Das, Udvas, et al.
Publicado: (2025)
Optimism Stabilizes Thompson Sampling for Adaptive Inference
por: Yan, Shunxing, et al.
Publicado: (2026)
por: Yan, Shunxing, et al.
Publicado: (2026)
Smooth Non-Stationary Bandits
por: Jia, Su, et al.
Publicado: (2023)
por: Jia, Su, et al.
Publicado: (2023)
Piecewise Polynomial Regression of Tame Functions via Integer Programming
por: Bareilles, Gilles, et al.
Publicado: (2023)
por: Bareilles, Gilles, et al.
Publicado: (2023)
Pessimism Meets Risk: Risk-Sensitive Offline Reinforcement Learning
por: Zhang, Dake, et al.
Publicado: (2024)
por: Zhang, Dake, et al.
Publicado: (2024)
Learning the Uncertainty Sets for Control Dynamics via Set Membership: A Non-Asymptotic Analysis
por: Li, Yingying, et al.
Publicado: (2023)
por: Li, Yingying, et al.
Publicado: (2023)
Analytic Bridge Diffusions for Controlled Path Generation
por: Chertkov, Michael
Publicado: (2026)
por: Chertkov, Michael
Publicado: (2026)
Geometry-induced Regularization in Deep ReLU Neural Networks
por: Bona-Pellissier, Joachim, et al.
Publicado: (2024)
por: Bona-Pellissier, Joachim, et al.
Publicado: (2024)
Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
por: Li, Gen, et al.
Publicado: (2020)
por: Li, Gen, et al.
Publicado: (2020)
A Theory of Feature Learning in Kernel Models
por: Chen, Yunlu, et al.
Publicado: (2023)
por: Chen, Yunlu, et al.
Publicado: (2023)
The High Line: Exact Risk and Learning Rate Curves of Stochastic Adaptive Learning Rate Algorithms
por: Collins-Woodfin, Elizabeth, et al.
Publicado: (2024)
por: Collins-Woodfin, Elizabeth, et al.
Publicado: (2024)
On the Uniform Convergence of Subdifferentials in Stochastic Optimization and Learning
por: Ruan, Feng
Publicado: (2024)
por: Ruan, Feng
Publicado: (2024)
Gradient Equilibrium in Online Learning: Theory and Applications
por: Angelopoulos, Anastasios N., et al.
Publicado: (2025)
por: Angelopoulos, Anastasios N., et al.
Publicado: (2025)
Bayesian Design Principles for Frequentist Sequential Learning
por: Xu, Yunbei, et al.
Publicado: (2023)
por: Xu, Yunbei, et al.
Publicado: (2023)
Learning an Optimal Assortment Policy under Observational Data
por: Han, Yuxuan, et al.
Publicado: (2025)
por: Han, Yuxuan, et al.
Publicado: (2025)
Learning and Decision-Making with Data: Optimal Formulations and Phase Transitions
por: Bennouna, Amine, et al.
Publicado: (2021)
por: Bennouna, Amine, et al.
Publicado: (2021)
PHAST: Port-Hamiltonian Architecture for Structured Temporal Dynamics Forecasting
por: Bhardwaj, Shubham, et al.
Publicado: (2026)
por: Bhardwaj, Shubham, et al.
Publicado: (2026)
ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule
por: Huang, Yilie, et al.
Publicado: (2026)
por: Huang, Yilie, et al.
Publicado: (2026)
Robustly Learning Monotone Generalized Linear Models via Data Augmentation
por: Zarifis, Nikos, et al.
Publicado: (2025)
por: Zarifis, Nikos, et al.
Publicado: (2025)
Blessings and Curses of Covariate Shifts: Adversarial Learning Dynamics, Directional Convergence, and Equilibria
por: Liang, Tengyuan
Publicado: (2022)
por: Liang, Tengyuan
Publicado: (2022)
A Graphical Global Optimization Framework for Parameter Estimation of Statistical Models with Nonconvex Regularization Functions
por: Davarnia, Danial, et al.
Publicado: (2025)
por: Davarnia, Danial, et al.
Publicado: (2025)
Function Gradient Approximation with Random Shallow ReLU Networks with Control Applications
por: Lamperski, Andrew, et al.
Publicado: (2024)
por: Lamperski, Andrew, et al.
Publicado: (2024)
Low-cost Robust Night-time Aerial Material Segmentation through Hyperspectral Data and Sparse Spatio-Temporal Learning
por: Bajaj, Chandrajit, et al.
Publicado: (2024)
por: Bajaj, Chandrajit, et al.
Publicado: (2024)
Shifted Interpolation for Differential Privacy
por: Bok, Jinho, et al.
Publicado: (2024)
por: Bok, Jinho, et al.
Publicado: (2024)
Differentiable Nonlinear Model Predictive Control
por: Frey, Jonathan, et al.
Publicado: (2025)
por: Frey, Jonathan, et al.
Publicado: (2025)
Learning linear dynamical systems under convex constraints
por: Tyagi, Hemant, et al.
Publicado: (2023)
por: Tyagi, Hemant, et al.
Publicado: (2023)
Ejemplares similares
-
Decoupled Continuous-Time Reinforcement Learning via Hamiltonian Flow
por: Nguyen, Minh
Publicado: (2026) -
Reinforcement Learning for Molecular Dynamics Optimization: A Stochastic Pontryagin Maximum Principle Approach
por: Bajaj, Chandrajit, et al.
Publicado: (2022) -
Statistical and Algorithmic Foundations of Reinforcement Learning
por: Chi, Yuejie, et al.
Publicado: (2025) -
Physics-informed neural networks via stochastic Hamiltonian dynamics learning
por: Bajaj, Chandrajit, et al.
Publicado: (2021) -
Learning to Fuse Temporal Proximity Networks: A Case Study in Chimpanzee Social Interactions
por: He, Yixuan, et al.
Publicado: (2025)