Online Decision-Focused Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Capitaine, Aymeric, Haddouche, Maxime, Moulines, Eric, Jordan, Michael I., Boursier, Etienne, Durmus, Alain |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality
por: Scheid, Antoine, et al.
Publicado: (2024)
por: Scheid, Antoine, et al.
Publicado: (2024)
Prediction-Aware Learning in Multi-Agent Systems
por: Capitaine, Aymeric, et al.
Publicado: (2025)
por: Capitaine, Aymeric, et al.
Publicado: (2025)
Incentivized Learning in Principal-Agent Bandit Games
por: Scheid, Antoine, et al.
Publicado: (2024)
por: Scheid, Antoine, et al.
Publicado: (2024)
Online Decision-Making in Tree-Like Multi-Agent Games with Transfers
por: Scheid, Antoine, et al.
Publicado: (2025)
por: Scheid, Antoine, et al.
Publicado: (2025)
Optimal Design for Reward Modeling in RLHF
por: Scheid, Antoine, et al.
Publicado: (2024)
por: Scheid, Antoine, et al.
Publicado: (2024)
Unravelling in Collaborative Learning
por: Capitaine, Aymeric, et al.
Publicado: (2024)
por: Capitaine, Aymeric, et al.
Publicado: (2024)
Test-then-Punish: A Statistical Approach to Repeated Games
por: Capitaine, Aymeric, et al.
Publicado: (2026)
por: Capitaine, Aymeric, et al.
Publicado: (2026)
Scaffold with Stochastic Gradients: New Analysis with Linear Speed-Up
por: Mangold, Paul, et al.
Publicado: (2025)
por: Mangold, Paul, et al.
Publicado: (2025)
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
por: Mangold, Paul, et al.
Publicado: (2024)
por: Mangold, Paul, et al.
Publicado: (2024)
Algorithm- and Data-Dependent Generalization Bounds for Diffusion Models
por: Dupuis, Benjamin, et al.
Publicado: (2025)
por: Dupuis, Benjamin, et al.
Publicado: (2025)
Online (Non-)Convex Learning via Tempered Optimism
por: Haddouche, Maxime, et al.
Publicado: (2023)
por: Haddouche, Maxime, et al.
Publicado: (2023)
Sequential Off-Policy Learning with Logarithmic Smoothing
por: Haddouche, Maxime, et al.
Publicado: (2025)
por: Haddouche, Maxime, et al.
Publicado: (2025)
Briding Diffusion Posterior Sampling and Monte Carlo methods: a survey
por: Janati, Yazid, et al.
Publicado: (2025)
por: Janati, Yazid, et al.
Publicado: (2025)
Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
por: Huix, Tom, et al.
Publicado: (2024)
por: Huix, Tom, et al.
Publicado: (2024)
On Sampling with Approximate Transport Maps
por: Grenioux, Louis, et al.
Publicado: (2023)
por: Grenioux, Louis, et al.
Publicado: (2023)
Federated Learning with Nonvacuous Generalisation Bounds
por: Jobic, Pierre, et al.
Publicado: (2023)
por: Jobic, Pierre, et al.
Publicado: (2023)
Categorical Reparameterization with Denoising Diffusion models
por: Gourevitch, Samson, et al.
Publicado: (2026)
por: Gourevitch, Samson, et al.
Publicado: (2026)
Piecewise deterministic generative models
por: Bertazzi, Andrea, et al.
Publicado: (2024)
por: Bertazzi, Andrea, et al.
Publicado: (2024)
Divide-and-Conquer Posterior Sampling for Denoising Diffusion Priors
por: Janati, Yazid, et al.
Publicado: (2024)
por: Janati, Yazid, et al.
Publicado: (2024)
Softmax as Linear Attention in the Large-Prompt Regime: a Measure-based Perspective
por: Boursier, Etienne, et al.
Publicado: (2025)
por: Boursier, Etienne, et al.
Publicado: (2025)
Early alignment in two-layer networks training is a two-edged sword
por: Boursier, Etienne, et al.
Publicado: (2024)
por: Boursier, Etienne, et al.
Publicado: (2024)
Simplicity bias and optimization threshold in two-layer ReLU networks
por: Boursier, Etienne, et al.
Publicado: (2024)
por: Boursier, Etienne, et al.
Publicado: (2024)
Penalising the biases in norm regularisation enforces sparsity
por: Boursier, Etienne, et al.
Publicado: (2023)
por: Boursier, Etienne, et al.
Publicado: (2023)
Conditional Diffusion Models with Classifier-Free Gibbs-like Guidance
por: Moufad, Badr, et al.
Publicado: (2025)
por: Moufad, Badr, et al.
Publicado: (2025)
Implicit Bias in Noisy-SGD: With Applications to Differentially Private Training
por: Sander, Tom, et al.
Publicado: (2024)
por: Sander, Tom, et al.
Publicado: (2024)
Uniform Diffusion Models Revisited: Leave-One-Out Denoiser and Absorbing State Reformulation
por: Gourevitch, Samson, et al.
Publicado: (2026)
por: Gourevitch, Samson, et al.
Publicado: (2026)
A survey on multi-player bandits
por: Boursier, Etienne, et al.
Publicado: (2022)
por: Boursier, Etienne, et al.
Publicado: (2022)
Generalization Bounds for Markov Algorithms through Entropy Flow Computations
por: Dupuis, Benjamin, et al.
Publicado: (2025)
por: Dupuis, Benjamin, et al.
Publicado: (2025)
A PAC-Bayesian Link Between Generalisation and Flat Minima
por: Haddouche, Maxime, et al.
Publicado: (2024)
por: Haddouche, Maxime, et al.
Publicado: (2024)
Tighter Generalisation Bounds via Interpolation
por: Viallard, Paul, et al.
Publicado: (2024)
por: Viallard, Paul, et al.
Publicado: (2024)
A Mixture-Based Framework for Guiding Diffusion Models
por: Janati, Yazid, et al.
Publicado: (2025)
por: Janati, Yazid, et al.
Publicado: (2025)
Variational Diffusion Posterior Sampling with Midpoint Guidance
por: Moufad, Badr, et al.
Publicado: (2024)
por: Moufad, Badr, et al.
Publicado: (2024)
Nonasymptotic Analysis of Stochastic Gradient Descent with the Richardson-Romberg Extrapolation
por: Sheshukova, Marina, et al.
Publicado: (2024)
por: Sheshukova, Marina, et al.
Publicado: (2024)
Rosenthal-type inequalities for linear statistics of Markov chains
por: Durmus, Alain, et al.
Publicado: (2023)
por: Durmus, Alain, et al.
Publicado: (2023)
Gradient flow dynamics of shallow ReLU networks for square loss and orthogonal inputs
por: Boursier, Etienne, et al.
Publicado: (2022)
por: Boursier, Etienne, et al.
Publicado: (2022)
First-order ANIL provably learns representations despite overparametrization
por: Yüksel, Oğuz Kaan, et al.
Publicado: (2023)
por: Yüksel, Oğuz Kaan, et al.
Publicado: (2023)
Joint Channel Selection using FedDRL in V2X
por: Mancini, Lorenzo, et al.
Publicado: (2024)
por: Mancini, Lorenzo, et al.
Publicado: (2024)
Mildly Overparameterized ReLU Networks on Orthogonal Data: Incremental Learning and Implicit Bias
por: Town, James, et al.
Publicado: (2026)
por: Town, James, et al.
Publicado: (2026)
Benignity of loss landscape with weight decay requires both large overparametrization and initialization
por: Boursier, Etienne, et al.
Publicado: (2025)
por: Boursier, Etienne, et al.
Publicado: (2025)
Regret and Sample Complexity of Online Q-Learning via Concentration of Stochastic Approximation with Time-Inhomogeneous Markov Chains
por: Singh, Rahul, et al.
Publicado: (2026)
por: Singh, Rahul, et al.
Publicado: (2026)
Ejemplares similares
-
Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality
por: Scheid, Antoine, et al.
Publicado: (2024) -
Prediction-Aware Learning in Multi-Agent Systems
por: Capitaine, Aymeric, et al.
Publicado: (2025) -
Incentivized Learning in Principal-Agent Bandit Games
por: Scheid, Antoine, et al.
Publicado: (2024) -
Online Decision-Making in Tree-Like Multi-Agent Games with Transfers
por: Scheid, Antoine, et al.
Publicado: (2025) -
Optimal Design for Reward Modeling in RLHF
por: Scheid, Antoine, et al.
Publicado: (2024)