Imitation Learning from Observations: An Autoregressive Mixture of Experts Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Renzi, Acerbo, Flavia Sofia, Son, Tong Duy, Patrinos, Panagiotis |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Parametric Nonconvex Optimization via Convex Surrogates
by: Wang, Renzi, et al.
Published: (2026)
by: Wang, Renzi, et al.
Published: (2026)
Rethinking PCA Through Duality
by: Quan, Jan, et al.
Published: (2025)
by: Quan, Jan, et al.
Published: (2025)
Real-time MPC with Control Barrier Functions for Autonomous Driving using Safety Enhanced Collocation
by: Allamaa, Jean Pierre, et al.
Published: (2024)
by: Allamaa, Jean Pierre, et al.
Published: (2024)
Risk-Sensitive Model Predictive Control for Interaction-Aware Planning -- A Sequential Convexification Algorithm
by: Wang, Renzi, et al.
Published: (2025)
by: Wang, Renzi, et al.
Published: (2025)
The inexact power augmented Lagrangian method for constrained nonconvex optimization
by: Bodard, Alexander, et al.
Published: (2024)
by: Bodard, Alexander, et al.
Published: (2024)
Asynchronous Message-Passing and Zeroth-Order Optimization Based Distributed Learning with a Use-Case in Resource Allocation in Communication Networks
by: Behmandpoor, Pourya, et al.
Published: (2023)
by: Behmandpoor, Pourya, et al.
Published: (2023)
Lasry-Lions Envelopes and Nonconvex Optimization: A Homotopy Approach
by: Simões, Miguel, et al.
Published: (2021)
by: Simões, Miguel, et al.
Published: (2021)
Guided by the Experts: Provable Feature Learning Dynamic of Soft-Routed Mixture-of-Experts
by: Liao, Fangshuo, et al.
Published: (2025)
by: Liao, Fangshuo, et al.
Published: (2025)
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
EM++: A parameter learning framework for stochastic switching systems
by: Wang, Renzi, et al.
Published: (2024)
by: Wang, Renzi, et al.
Published: (2024)
$ϕ$-Balancing for Mixture-of-Experts Training
by: Chen, Lizhang, et al.
Published: (2026)
by: Chen, Lizhang, et al.
Published: (2026)
Constrained Stochastic Spectral Preconditioning Converges for Nonconvex Objectives
by: Oikonomidis, Konstantinos, et al.
Published: (2026)
by: Oikonomidis, Konstantinos, et al.
Published: (2026)
A Deep Learning Based Resource Allocator for Communication Networks with Dynamic User Utility Demands
by: Behmandpoor, Pourya, et al.
Published: (2023)
by: Behmandpoor, Pourya, et al.
Published: (2023)
Imitation Learning for Combinatorial Optimisation under Uncertainty
by: Gawas, Prakash, et al.
Published: (2026)
by: Gawas, Prakash, et al.
Published: (2026)
Hierarchical Mixture-of-Experts with Two-Stage Optimization
by: Molodtsov, Gleb, et al.
Published: (2026)
by: Molodtsov, Gleb, et al.
Published: (2026)
Population-Aware Imitation Learning in Mean-field Games with Common Noise
by: Lambrecht, Grégoire, et al.
Published: (2026)
by: Lambrecht, Grégoire, et al.
Published: (2026)
Escaping saddle points without Lipschitz smoothness: the power of nonlinear preconditioning
by: Bodard, Alexander, et al.
Published: (2025)
by: Bodard, Alexander, et al.
Published: (2025)
Anisotropic Proximal Point Algorithm
by: Laude, Emanuel, et al.
Published: (2023)
by: Laude, Emanuel, et al.
Published: (2023)
Anisotropic Proximal Gradient
by: Laude, Emanuel, et al.
Published: (2022)
by: Laude, Emanuel, et al.
Published: (2022)
Newton methods beyond Hessian Lipschitz continuity: A nonlinear preconditioning approach
by: Bodard, Alexander, et al.
Published: (2026)
by: Bodard, Alexander, et al.
Published: (2026)
Mixtures Closest to a Given Measure: A Semidefinite Programming Approach
by: Đurašinović, Srećko, et al.
Published: (2025)
by: Đurašinović, Srećko, et al.
Published: (2025)
Online Residual Learning from Offline Experts for Pedestrian Tracking
by: Vlachos, Anastasios, et al.
Published: (2024)
by: Vlachos, Anastasios, et al.
Published: (2024)
In-Context Learning for Data-Driven Censored Inventory Control
by: Mukherjee, Sohom, et al.
Published: (2026)
by: Mukherjee, Sohom, et al.
Published: (2026)
Learning Linear Dynamics from Bilinear Observations
by: Sattar, Yahya, et al.
Published: (2024)
by: Sattar, Yahya, et al.
Published: (2024)
Reinforcement Learning from Partial Observation: Linear Function Approximation with Provable Sample Efficiency
by: Cai, Qi, et al.
Published: (2022)
by: Cai, Qi, et al.
Published: (2022)
Offline-Online Reinforcement Learning for Linear Mixture MDPs
by: Zhang, Zhongjun, et al.
Published: (2026)
by: Zhang, Zhongjun, et al.
Published: (2026)
Safe Imitation Learning of Nonlinear Model Predictive Control for Flexible Robots
by: Mamedov, Shamil, et al.
Published: (2022)
by: Mamedov, Shamil, et al.
Published: (2022)
Learning Infinite-Horizon Average-Reward Linear Mixture MDPs of Bounded Span
by: Chae, Woojin, et al.
Published: (2024)
by: Chae, Woojin, et al.
Published: (2024)
Near-Optimal Primal-Dual Algorithm for Learning Linear Mixture CMDPs with Adversarial Rewards
by: Yu, Kihyun, et al.
Published: (2026)
by: Yu, Kihyun, et al.
Published: (2026)
Dynamic Estimation of Learning Rates Using a Non-Linear Autoregressive Model
by: Okhrati, Ramin
Published: (2024)
by: Okhrati, Ramin
Published: (2024)
Identifying Systems with Symmetries using Equivariant Autoregressive Reservoir Computers
by: Vides, Fredy, et al.
Published: (2023)
by: Vides, Fredy, et al.
Published: (2023)
A Relaxed Wasserstein Distance Formulation for Mixtures of Radially Contoured Distributions
by: Chen, Keyu, et al.
Published: (2025)
by: Chen, Keyu, et al.
Published: (2025)
Probabilistic Safety under Arbitrary Disturbance Distributions using Piecewise-Affine Control Barrier Functions
by: Teuwen, Matisse, et al.
Published: (2025)
by: Teuwen, Matisse, et al.
Published: (2025)
Exact worst-case convergence rates of gradient descent: a complete analysis for all constant stepsizes over nonconvex and convex functions
by: Rotaru, Teodor, et al.
Published: (2024)
by: Rotaru, Teodor, et al.
Published: (2024)
Improved convergence rates for the Difference-of-Convex algorithm
by: Rotaru, Teodor, et al.
Published: (2024)
by: Rotaru, Teodor, et al.
Published: (2024)
Nonlinearly Preconditioned Gradient Methods: Momentum and Stochastic Analysis
by: Oikonomidis, Konstantinos, et al.
Published: (2025)
by: Oikonomidis, Konstantinos, et al.
Published: (2025)
Tight Analysis of Difference-of-Convex Algorithm (DCA) Improves Convergence Rates for Proximal Gradient Descent
by: Rotaru, Teodor, et al.
Published: (2025)
by: Rotaru, Teodor, et al.
Published: (2025)
Dualities for non-Euclidean smoothness and strong convexity under the light of generalized conjugacy
by: Laude, Emanuel, et al.
Published: (2021)
by: Laude, Emanuel, et al.
Published: (2021)
On the Regularity of Generalized Conjugate Functions
by: Oikonomidis, Konstantinos, et al.
Published: (2026)
by: Oikonomidis, Konstantinos, et al.
Published: (2026)
Forward-backward splitting under the light of generalized convexity
by: Oikonomidis, Konstantinos, et al.
Published: (2025)
by: Oikonomidis, Konstantinos, et al.
Published: (2025)
Similar Items
-
Parametric Nonconvex Optimization via Convex Surrogates
by: Wang, Renzi, et al.
Published: (2026) -
Rethinking PCA Through Duality
by: Quan, Jan, et al.
Published: (2025) -
Real-time MPC with Control Barrier Functions for Autonomous Driving using Safety Enhanced Collocation
by: Allamaa, Jean Pierre, et al.
Published: (2024) -
Risk-Sensitive Model Predictive Control for Interaction-Aware Planning -- A Sequential Convexification Algorithm
by: Wang, Renzi, et al.
Published: (2025) -
The inexact power augmented Lagrangian method for constrained nonconvex optimization
by: Bodard, Alexander, et al.
Published: (2024)