Online Decision-Focused Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Capitaine, Aymeric, Haddouche, Maxime, Moulines, Eric, Jordan, Michael I., Boursier, Etienne, Durmus, Alain |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality
di: Scheid, Antoine, et al.
Pubblicazione: (2024)
di: Scheid, Antoine, et al.
Pubblicazione: (2024)
Prediction-Aware Learning in Multi-Agent Systems
di: Capitaine, Aymeric, et al.
Pubblicazione: (2025)
di: Capitaine, Aymeric, et al.
Pubblicazione: (2025)
Incentivized Learning in Principal-Agent Bandit Games
di: Scheid, Antoine, et al.
Pubblicazione: (2024)
di: Scheid, Antoine, et al.
Pubblicazione: (2024)
Online Decision-Making in Tree-Like Multi-Agent Games with Transfers
di: Scheid, Antoine, et al.
Pubblicazione: (2025)
di: Scheid, Antoine, et al.
Pubblicazione: (2025)
Optimal Design for Reward Modeling in RLHF
di: Scheid, Antoine, et al.
Pubblicazione: (2024)
di: Scheid, Antoine, et al.
Pubblicazione: (2024)
Unravelling in Collaborative Learning
di: Capitaine, Aymeric, et al.
Pubblicazione: (2024)
di: Capitaine, Aymeric, et al.
Pubblicazione: (2024)
Test-then-Punish: A Statistical Approach to Repeated Games
di: Capitaine, Aymeric, et al.
Pubblicazione: (2026)
di: Capitaine, Aymeric, et al.
Pubblicazione: (2026)
Scaffold with Stochastic Gradients: New Analysis with Linear Speed-Up
di: Mangold, Paul, et al.
Pubblicazione: (2025)
di: Mangold, Paul, et al.
Pubblicazione: (2025)
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
di: Mangold, Paul, et al.
Pubblicazione: (2024)
di: Mangold, Paul, et al.
Pubblicazione: (2024)
Algorithm- and Data-Dependent Generalization Bounds for Diffusion Models
di: Dupuis, Benjamin, et al.
Pubblicazione: (2025)
di: Dupuis, Benjamin, et al.
Pubblicazione: (2025)
Online (Non-)Convex Learning via Tempered Optimism
di: Haddouche, Maxime, et al.
Pubblicazione: (2023)
di: Haddouche, Maxime, et al.
Pubblicazione: (2023)
Sequential Off-Policy Learning with Logarithmic Smoothing
di: Haddouche, Maxime, et al.
Pubblicazione: (2025)
di: Haddouche, Maxime, et al.
Pubblicazione: (2025)
Briding Diffusion Posterior Sampling and Monte Carlo methods: a survey
di: Janati, Yazid, et al.
Pubblicazione: (2025)
di: Janati, Yazid, et al.
Pubblicazione: (2025)
Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
di: Huix, Tom, et al.
Pubblicazione: (2024)
di: Huix, Tom, et al.
Pubblicazione: (2024)
On Sampling with Approximate Transport Maps
di: Grenioux, Louis, et al.
Pubblicazione: (2023)
di: Grenioux, Louis, et al.
Pubblicazione: (2023)
Federated Learning with Nonvacuous Generalisation Bounds
di: Jobic, Pierre, et al.
Pubblicazione: (2023)
di: Jobic, Pierre, et al.
Pubblicazione: (2023)
Categorical Reparameterization with Denoising Diffusion models
di: Gourevitch, Samson, et al.
Pubblicazione: (2026)
di: Gourevitch, Samson, et al.
Pubblicazione: (2026)
Piecewise deterministic generative models
di: Bertazzi, Andrea, et al.
Pubblicazione: (2024)
di: Bertazzi, Andrea, et al.
Pubblicazione: (2024)
Divide-and-Conquer Posterior Sampling for Denoising Diffusion Priors
di: Janati, Yazid, et al.
Pubblicazione: (2024)
di: Janati, Yazid, et al.
Pubblicazione: (2024)
Softmax as Linear Attention in the Large-Prompt Regime: a Measure-based Perspective
di: Boursier, Etienne, et al.
Pubblicazione: (2025)
di: Boursier, Etienne, et al.
Pubblicazione: (2025)
Early alignment in two-layer networks training is a two-edged sword
di: Boursier, Etienne, et al.
Pubblicazione: (2024)
di: Boursier, Etienne, et al.
Pubblicazione: (2024)
Simplicity bias and optimization threshold in two-layer ReLU networks
di: Boursier, Etienne, et al.
Pubblicazione: (2024)
di: Boursier, Etienne, et al.
Pubblicazione: (2024)
Penalising the biases in norm regularisation enforces sparsity
di: Boursier, Etienne, et al.
Pubblicazione: (2023)
di: Boursier, Etienne, et al.
Pubblicazione: (2023)
Conditional Diffusion Models with Classifier-Free Gibbs-like Guidance
di: Moufad, Badr, et al.
Pubblicazione: (2025)
di: Moufad, Badr, et al.
Pubblicazione: (2025)
Implicit Bias in Noisy-SGD: With Applications to Differentially Private Training
di: Sander, Tom, et al.
Pubblicazione: (2024)
di: Sander, Tom, et al.
Pubblicazione: (2024)
Uniform Diffusion Models Revisited: Leave-One-Out Denoiser and Absorbing State Reformulation
di: Gourevitch, Samson, et al.
Pubblicazione: (2026)
di: Gourevitch, Samson, et al.
Pubblicazione: (2026)
A survey on multi-player bandits
di: Boursier, Etienne, et al.
Pubblicazione: (2022)
di: Boursier, Etienne, et al.
Pubblicazione: (2022)
Generalization Bounds for Markov Algorithms through Entropy Flow Computations
di: Dupuis, Benjamin, et al.
Pubblicazione: (2025)
di: Dupuis, Benjamin, et al.
Pubblicazione: (2025)
A PAC-Bayesian Link Between Generalisation and Flat Minima
di: Haddouche, Maxime, et al.
Pubblicazione: (2024)
di: Haddouche, Maxime, et al.
Pubblicazione: (2024)
Tighter Generalisation Bounds via Interpolation
di: Viallard, Paul, et al.
Pubblicazione: (2024)
di: Viallard, Paul, et al.
Pubblicazione: (2024)
A Mixture-Based Framework for Guiding Diffusion Models
di: Janati, Yazid, et al.
Pubblicazione: (2025)
di: Janati, Yazid, et al.
Pubblicazione: (2025)
Variational Diffusion Posterior Sampling with Midpoint Guidance
di: Moufad, Badr, et al.
Pubblicazione: (2024)
di: Moufad, Badr, et al.
Pubblicazione: (2024)
Nonasymptotic Analysis of Stochastic Gradient Descent with the Richardson-Romberg Extrapolation
di: Sheshukova, Marina, et al.
Pubblicazione: (2024)
di: Sheshukova, Marina, et al.
Pubblicazione: (2024)
Rosenthal-type inequalities for linear statistics of Markov chains
di: Durmus, Alain, et al.
Pubblicazione: (2023)
di: Durmus, Alain, et al.
Pubblicazione: (2023)
Gradient flow dynamics of shallow ReLU networks for square loss and orthogonal inputs
di: Boursier, Etienne, et al.
Pubblicazione: (2022)
di: Boursier, Etienne, et al.
Pubblicazione: (2022)
First-order ANIL provably learns representations despite overparametrization
di: Yüksel, Oğuz Kaan, et al.
Pubblicazione: (2023)
di: Yüksel, Oğuz Kaan, et al.
Pubblicazione: (2023)
Joint Channel Selection using FedDRL in V2X
di: Mancini, Lorenzo, et al.
Pubblicazione: (2024)
di: Mancini, Lorenzo, et al.
Pubblicazione: (2024)
Mildly Overparameterized ReLU Networks on Orthogonal Data: Incremental Learning and Implicit Bias
di: Town, James, et al.
Pubblicazione: (2026)
di: Town, James, et al.
Pubblicazione: (2026)
Benignity of loss landscape with weight decay requires both large overparametrization and initialization
di: Boursier, Etienne, et al.
Pubblicazione: (2025)
di: Boursier, Etienne, et al.
Pubblicazione: (2025)
Regret and Sample Complexity of Online Q-Learning via Concentration of Stochastic Approximation with Time-Inhomogeneous Markov Chains
di: Singh, Rahul, et al.
Pubblicazione: (2026)
di: Singh, Rahul, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality
di: Scheid, Antoine, et al.
Pubblicazione: (2024) -
Prediction-Aware Learning in Multi-Agent Systems
di: Capitaine, Aymeric, et al.
Pubblicazione: (2025) -
Incentivized Learning in Principal-Agent Bandit Games
di: Scheid, Antoine, et al.
Pubblicazione: (2024) -
Online Decision-Making in Tree-Like Multi-Agent Games with Transfers
di: Scheid, Antoine, et al.
Pubblicazione: (2025) -
Optimal Design for Reward Modeling in RLHF
di: Scheid, Antoine, et al.
Pubblicazione: (2024)