Alignment of large language models with constrained learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Botong, Li, Shuo, Hounie, Ignacio, Bastani, Osbert, Ding, Dongsheng, Ribeiro, Alejandro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Resilient Constrained Reinforcement Learning
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
One-Shot Safety Alignment for Large Language Models via Optimal Dualization
by: Huang, Xinmeng, et al.
Published: (2024)
by: Huang, Xinmeng, et al.
Published: (2024)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2022)
by: Ding, Dongsheng, et al.
Published: (2022)
Distributed Thompson sampling under constrained communication
by: Zerefa, Saba, et al.
Published: (2024)
by: Zerefa, Saba, et al.
Published: (2024)
Scalable spectral representations for multi-agent reinforcement learning in network MDPs
by: Ren, Zhaolin, et al.
Published: (2024)
by: Ren, Zhaolin, et al.
Published: (2024)
Independent policy gradient-based reinforcement learning for economic and reliable energy management of multi-microgrid systems
by: Hu, Junkai, et al.
Published: (2025)
by: Hu, Junkai, et al.
Published: (2025)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
by: Manenti, Massimiliano, et al.
Published: (2025)
by: Manenti, Massimiliano, et al.
Published: (2025)
Variance-Reduced Cascade Q-learning: Algorithms and Sample Complexity
by: Boveiri, Mohammad, et al.
Published: (2024)
by: Boveiri, Mohammad, et al.
Published: (2024)
Safe learning-based control via function-based uncertainty quantification
by: Tokmak, Abdullah, et al.
Published: (2026)
by: Tokmak, Abdullah, et al.
Published: (2026)
Linear quadratic control of nonlinear systems with Koopman operator learning and the Nyström method
by: Caldarelli, Edoardo, et al.
Published: (2024)
by: Caldarelli, Edoardo, et al.
Published: (2024)
PolyFormer: learning efficient reformulations for scalable optimization under complex physical constraints
by: Wen, Yilin, et al.
Published: (2026)
by: Wen, Yilin, et al.
Published: (2026)
GreenLight-Gym: Reinforcement learning benchmark environment for control of greenhouse production systems
by: van Laatum, Bart, et al.
Published: (2024)
by: van Laatum, Bart, et al.
Published: (2024)
Efficient model predictive control for nonlinear systems modelled by deep neural networks
by: Lan, Jianglin
Published: (2024)
by: Lan, Jianglin
Published: (2024)
Revealing design archetypes and flexibility in e-molecule import pathways using Modeling to Generate Alternatives and interpretable machine learning
by: Kchaou, Mahdi, et al.
Published: (2025)
by: Kchaou, Mahdi, et al.
Published: (2025)
A robust and adaptive MPC formulation for Gaussian process models
by: Dubied, Mathieu, et al.
Published: (2025)
by: Dubied, Mathieu, et al.
Published: (2025)
Efficient identification of linear, parameter-varying, and nonlinear systems with noise models
by: Bemporad, Alberto, et al.
Published: (2025)
by: Bemporad, Alberto, et al.
Published: (2025)
Safe Bayesian optimization across noise models via scenario programming
by: Tokmak, Abdullah, et al.
Published: (2025)
by: Tokmak, Abdullah, et al.
Published: (2025)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
by: Muehlebach, Michael, et al.
Published: (2025)
by: Muehlebach, Michael, et al.
Published: (2025)
Neural Contraction Metrics with Formal Guarantees for Discrete-Time Nonlinear Dynamical Systems
by: Li, Haoyu, et al.
Published: (2025)
by: Li, Haoyu, et al.
Published: (2025)
Temporal-Aware Deep Reinforcement Learning for Energy Storage Bidding in Energy and Contingency Reserve Markets
by: Li, Jinhao, et al.
Published: (2024)
by: Li, Jinhao, et al.
Published: (2024)
Exploiting inter-agent coupling information for efficient reinforcement learning of cooperative LQR
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
Approximate non-linear model predictive control with safety-augmented neural networks
by: Hose, Henrik, et al.
Published: (2023)
by: Hose, Henrik, et al.
Published: (2023)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
by: Zhang, Xiaole, et al.
Published: (2025)
by: Zhang, Xiaole, et al.
Published: (2025)
An active learning method for solving competitive multi-agent decision-making and control problems
by: Fabiani, Filippo, et al.
Published: (2022)
by: Fabiani, Filippo, et al.
Published: (2022)
Optimism as Risk-Seeking in Multi-Agent Reinforcement Learning
by: Zhang, Runyu, et al.
Published: (2025)
by: Zhang, Runyu, et al.
Published: (2025)
Bayesian dynamic scheduling of multipurpose batch processes under incomplete look-ahead information
by: Zheng, Taicheng, et al.
Published: (2025)
by: Zheng, Taicheng, et al.
Published: (2025)
Slack More, Predict Better: Proximal Relaxation for Probabilistic Latent Variable Model-based Soft Sensors
by: Zou, Zehua, et al.
Published: (2026)
by: Zou, Zehua, et al.
Published: (2026)
Efficient Reachability Analysis for Convolutional Neural Networks Using Hybrid Zonotopes
by: Zhang, Yuhao, et al.
Published: (2025)
by: Zhang, Yuhao, et al.
Published: (2025)
Neural Approximators for Low-Thrust Trajectory Transfer Cost and Reachability
by: Zhang, Zhong, et al.
Published: (2025)
by: Zhang, Zhong, et al.
Published: (2025)
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
by: Boone, Victor, et al.
Published: (2024)
by: Boone, Victor, et al.
Published: (2024)
ReLU Networks for Model Predictive Control: Network Complexity and Performance Guarantees
by: Li, Xingchen, et al.
Published: (2026)
by: Li, Xingchen, et al.
Published: (2026)
Bayesian Ambiguity Contraction-based Adaptive Robust Markov Decision Processes for Adversarial Surveillance Missions
by: Choi, Jimin, et al.
Published: (2025)
by: Choi, Jimin, et al.
Published: (2025)
Global Search of Optimal Spacecraft Trajectories using Amortization and Deep Generative Models
by: Beeson, Ryne, et al.
Published: (2024)
by: Beeson, Ryne, et al.
Published: (2024)
Attentive Convolutional Deep Reinforcement Learning for Optimizing Solar-Storage Systems in Real-Time Electricity Markets
by: Li, Jinhao, et al.
Published: (2024)
by: Li, Jinhao, et al.
Published: (2024)
Observability conditions for neural state-space models with eigenvalues and their roots of unity
by: Gracyk, Andrew
Published: (2025)
by: Gracyk, Andrew
Published: (2025)
Provably-Stable Neural Network-Based Control of Nonlinear Systems
by: Li, Anran, et al.
Published: (2025)
by: Li, Anran, et al.
Published: (2025)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part I
by: Tian, Yi, et al.
Published: (2022)
by: Tian, Yi, et al.
Published: (2022)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Model-Free Output Feedback Stabilization via Policy Gradient Methods
by: Zhang, Ankang, et al.
Published: (2026)
by: Zhang, Ankang, et al.
Published: (2026)
Similar Items
-
Resilient Constrained Reinforcement Learning
by: Ding, Dongsheng, et al.
Published: (2023) -
One-Shot Safety Alignment for Large Language Models via Optimal Dualization
by: Huang, Xinmeng, et al.
Published: (2024) -
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023) -
Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2022) -
Distributed Thompson sampling under constrained communication
by: Zerefa, Saba, et al.
Published: (2024)