Isoperimetry is All We Need: Langevin Posterior Sampling for RL with Sublinear Regret
Fuente:
arXiv
Saved in:
| Main Authors: | Jorge, Emilio, Dimitrakakis, Christos, Basu, Debabrota |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sublinear Algorithms for Wasserstein and Total Variation Distances: Applications to Fairness and Privacy Auditing
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
Stochastic Online Instrumental Variable Regression: Regrets for Endogeneity and Bandit Feedback
by: Della Vecchia, Riccardo, et al.
Published: (2023)
by: Della Vecchia, Riccardo, et al.
Published: (2023)
Discrete Langevin-Inspired Posterior Sampling
by: Amballa, Chaitanya, et al.
Published: (2026)
by: Amballa, Chaitanya, et al.
Published: (2026)
Posterior Sampling Reinforcement Learning with Gaussian Processes for Continuous Control: Sublinear Regret Bounds for Unbounded State Spaces
by: Flynn, Hamish, et al.
Published: (2026)
by: Flynn, Hamish, et al.
Published: (2026)
Learning Equilibria in Matching Games with Bandit Feedback
by: Athanasopoulos, Andreas, et al.
Published: (2025)
by: Athanasopoulos, Andreas, et al.
Published: (2025)
Preference-based Pure Exploration
by: Shukla, Apurv, et al.
Published: (2024)
by: Shukla, Apurv, et al.
Published: (2024)
Two-Player Zero-Sum Games with Bandit Feedback
by: Yılmaz, Elif, et al.
Published: (2025)
by: Yılmaz, Elif, et al.
Published: (2025)
How does Your RL Agent Explore? An Optimal Transport Analysis of Occupancy Measure Trajectories
by: Nkhumise, Reabetswe M., et al.
Published: (2024)
by: Nkhumise, Reabetswe M., et al.
Published: (2024)
Rawlsian many-to-one matching with non-linear utility
by: Nana, Hortence, et al.
Published: (2025)
by: Nana, Hortence, et al.
Published: (2025)
Learning to Explore with Lagrangians for Bandits under Unknown Linear Constraints
by: Das, Udvas, et al.
Published: (2024)
by: Das, Udvas, et al.
Published: (2024)
Some Targets Are Harder to Identify than Others: Quantifying the Target-dependent Membership Leakage
by: Azize, Achraf, et al.
Published: (2024)
by: Azize, Achraf, et al.
Published: (2024)
Auditing Fairness under Model Updates: Fundamental Complexity and Property-Preserving Updates
by: Ajarra, Ayoub, et al.
Published: (2026)
by: Ajarra, Ayoub, et al.
Published: (2026)
Probably Correct Optimal Stable Matching for Two-Sided Markets Under Uncertainty
by: Athanasopoulos, Andreas, et al.
Published: (2025)
by: Athanasopoulos, Andreas, et al.
Published: (2025)
Optimal Regret of Bernoulli Bandits under Global Differential Privacy
by: Azize, Achraf, et al.
Published: (2025)
by: Azize, Achraf, et al.
Published: (2025)
Concentrated Differential Privacy for Bandits
by: Azize, Achraf, et al.
Published: (2023)
by: Azize, Achraf, et al.
Published: (2023)
Sequential Cohort Selection
by: Nana, Hortence Phalonne, et al.
Published: (2025)
by: Nana, Hortence Phalonne, et al.
Published: (2025)
Optimal Scalarizations for Sublinear Hypervolume Regret
by: Zhang, Qiuyi
Published: (2023)
by: Zhang, Qiuyi
Published: (2023)
FLIPHAT: Joint Differential Privacy for High Dimensional Sparse Linear Bandits
by: Chakraborty, Sunrit, et al.
Published: (2024)
by: Chakraborty, Sunrit, et al.
Published: (2024)
Environment Design for Inverse Reinforcement Learning
by: Buening, Thomas Kleine, et al.
Published: (2022)
by: Buening, Thomas Kleine, et al.
Published: (2022)
Posterior Sampling-Based Bayesian Optimization with Tighter Bayesian Regret Bounds
by: Takeno, Shion, et al.
Published: (2023)
by: Takeno, Shion, et al.
Published: (2023)
Regret Analysis of Posterior Sampling-Based Expected Improvement for Bayesian Optimization
by: Takeno, Shion, et al.
Published: (2025)
by: Takeno, Shion, et al.
Published: (2025)
Efficient Approximate Posterior Sampling with Annealed Langevin Monte Carlo
by: Parulekar, Advait, et al.
Published: (2025)
by: Parulekar, Advait, et al.
Published: (2025)
DOT: Dynamic Knob Selection and Online Sampling for Automated Database Tuning
by: Wang, Yifan, et al.
Published: (2026)
by: Wang, Yifan, et al.
Published: (2026)
Fair Contracts in Principal-Agent Games with Heterogeneous Types
by: Tłuczek, Jakub, et al.
Published: (2025)
by: Tłuczek, Jakub, et al.
Published: (2025)
Active Fourier Auditor for Estimating Distributional Properties of ML Models
by: Ajarra, Ayoub, et al.
Published: (2024)
by: Ajarra, Ayoub, et al.
Published: (2024)
Sequential Membership Inference Attacks
by: Michel, Thomas, et al.
Published: (2026)
by: Michel, Thomas, et al.
Published: (2026)
DP-SPRT: Differentially Private Sequential Probability Ratio Tests
by: Michel, Thomas, et al.
Published: (2025)
by: Michel, Thomas, et al.
Published: (2025)
Lagrangian-based Equilibrium Propagation: generalisation to arbitrary boundary conditions & equivalence with Hamiltonian Echo Learning
by: Pourcel, Guillaume, et al.
Published: (2025)
by: Pourcel, Guillaume, et al.
Published: (2025)
Faster Sampling without Isoperimetry via Diffusion-based Monte Carlo
by: Huang, Xunpeng, et al.
Published: (2024)
by: Huang, Xunpeng, et al.
Published: (2024)
FraPPE: Fast and Efficient Preference-based Pure Exploration
by: Das, Udvas, et al.
Published: (2025)
by: Das, Udvas, et al.
Published: (2025)
Were RNNs All We Needed?
by: Feng, Leo, et al.
Published: (2024)
by: Feng, Leo, et al.
Published: (2024)
Temperature is All You Need for Generalization in Langevin Dynamics and other Markov Processes
by: Harel, Itamar, et al.
Published: (2025)
by: Harel, Itamar, et al.
Published: (2025)
When Witnesses Defend: A Witness Graph Topological Layer for Adversarial Graph Learning
by: Arafat, Naheed Anjum, et al.
Published: (2024)
by: Arafat, Naheed Anjum, et al.
Published: (2024)
Augmented Bayesian Policy Search
by: Kallel, Mahdi, et al.
Published: (2024)
by: Kallel, Mahdi, et al.
Published: (2024)
Pure Exploration in Bandits with Linear Constraints
by: Carlsson, Emil, et al.
Published: (2023)
by: Carlsson, Emil, et al.
Published: (2023)
Posterior Sampling by Combining Diffusion Models with Annealed Langevin Dynamics
by: Xun, Zhiyang, et al.
Published: (2025)
by: Xun, Zhiyang, et al.
Published: (2025)
Revisiting Matrix Sketching in Linear Bandits: Achieving Sublinear Regret via Dyadic Block Sketching
by: Wen, Dongxie, et al.
Published: (2024)
by: Wen, Dongxie, et al.
Published: (2024)
Better Regret Rates in Bilateral Trade via Sublinear Budget Violation
by: Lunghi, Anna, et al.
Published: (2025)
by: Lunghi, Anna, et al.
Published: (2025)
Dynamical-VAE-based Hindsight to Learn the Causal Dynamics of Factored-POMDPs
by: Han, Chao, et al.
Published: (2024)
by: Han, Chao, et al.
Published: (2024)
Efficient Exploration in Average-Reward Constrained Reinforcement Learning: Achieving Near-Optimal Regret With Posterior Sampling
by: Provodin, Danil, et al.
Published: (2024)
by: Provodin, Danil, et al.
Published: (2024)
Similar Items
-
Sublinear Algorithms for Wasserstein and Total Variation Distances: Applications to Fairness and Privacy Auditing
by: Basu, Debabrota, et al.
Published: (2025) -
Stochastic Online Instrumental Variable Regression: Regrets for Endogeneity and Bandit Feedback
by: Della Vecchia, Riccardo, et al.
Published: (2023) -
Discrete Langevin-Inspired Posterior Sampling
by: Amballa, Chaitanya, et al.
Published: (2026) -
Posterior Sampling Reinforcement Learning with Gaussian Processes for Continuous Control: Sublinear Regret Bounds for Unbounded State Spaces
by: Flynn, Hamish, et al.
Published: (2026) -
Learning Equilibria in Matching Games with Bandit Feedback
by: Athanasopoulos, Andreas, et al.
Published: (2025)