Generator-Mediated Bandits: Thompson Sampling for GenAI-Powered Adaptive Interventions
Fuente:
arXiv
Saved in:
| Main Authors: | Brooks, Marc, Durham, Gabriel, Hong, Kihyuk, Tewari, Ambuj |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GenAI-Powered Inference
by: Imai, Kosuke, et al.
Published: (2025)
by: Imai, Kosuke, et al.
Published: (2025)
A Primal-Dual Algorithm for Offline Constrained Reinforcement Learning with Linear MDPs
by: Hong, Kihyuk, et al.
Published: (2024)
by: Hong, Kihyuk, et al.
Published: (2024)
A Computationally Efficient Algorithm for Infinite-Horizon Average-Reward Linear MDPs
by: Hong, Kihyuk, et al.
Published: (2025)
by: Hong, Kihyuk, et al.
Published: (2025)
Learning to Partially Defer for Sequences
by: Rayan, Sahana, et al.
Published: (2025)
by: Rayan, Sahana, et al.
Published: (2025)
Distribution-Free Robust Predict-Then-Optimize in Function Spaces
by: Patel, Yash, et al.
Published: (2026)
by: Patel, Yash, et al.
Published: (2026)
A Greedy PDE Router for Blending Neural Operators and Classical Methods
by: Rayan, Sahana, et al.
Published: (2025)
by: Rayan, Sahana, et al.
Published: (2025)
Offline Constrained Reinforcement Learning under Partial Data Coverage
by: Ko, Seokmin, et al.
Published: (2025)
by: Ko, Seokmin, et al.
Published: (2025)
Conformal Prediction for Ensembles: Improving Efficiency via Score-Based Aggregation
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
On the Computational Complexity of Private High-dimensional Model Selection
by: Roy, Saptarshi, et al.
Published: (2023)
by: Roy, Saptarshi, et al.
Published: (2023)
Industrializing Prediction-Powered Inference: The GLIDE Library for Reliable GenAI and Agentic Systems Evaluation
by: Martinon, Grégoire, et al.
Published: (2026)
by: Martinon, Grégoire, et al.
Published: (2026)
Variational Inference with Coverage Guarantees in Simulation-Based Inference
by: Patel, Yash, et al.
Published: (2023)
by: Patel, Yash, et al.
Published: (2023)
BFTS: Thompson Sampling with Bayesian Additive Regression Trees
by: Deng, Ruizhe, et al.
Published: (2026)
by: Deng, Ruizhe, et al.
Published: (2026)
If generative AI is the answer, what is the question?
by: Tewari, Ambuj
Published: (2025)
by: Tewari, Ambuj
Published: (2025)
Counterfactual Inference under Thompson Sampling
by: Jeunen, Olivier
Published: (2025)
by: Jeunen, Olivier
Published: (2025)
Optimal Thresholding Linear Bandit
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
Reinforcement Learning for Infinite-Horizon Average-Reward Linear MDPs via Approximation by Discounted-Reward MDPs
by: Hong, Kihyuk, et al.
Published: (2024)
by: Hong, Kihyuk, et al.
Published: (2024)
Near Optimal Pure Exploration in Logistic Bandits
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
by: Rivera, Eduardo Ochoa, et al.
Published: (2024)
Learning Infinite-Horizon Average-Reward Linear Mixture MDPs of Bounded Span
by: Chae, Woojin, et al.
Published: (2024)
by: Chae, Woojin, et al.
Published: (2024)
Tree Bandits for Generative Bayes
by: O'Hagan, Sean, et al.
Published: (2024)
by: O'Hagan, Sean, et al.
Published: (2024)
Sampling as Bandits: Evaluation-Efficient Design for Black-Box Densities
by: Matsubara, Takuo, et al.
Published: (2025)
by: Matsubara, Takuo, et al.
Published: (2025)
DoubleGen: Debiased Generative Modeling of Counterfactuals
by: Luedtke, Alex, et al.
Published: (2025)
by: Luedtke, Alex, et al.
Published: (2025)
Causality-Encoded Diffusion Models for Interventional Sampling and Edge Inference
by: Chen, Li, et al.
Published: (2026)
by: Chen, Li, et al.
Published: (2026)
M-learner:A Flexible And Powerful Framework To Study Heterogeneous Treatment Effect In Mediation Model
by: Li, Xingyu, et al.
Published: (2025)
by: Li, Xingyu, et al.
Published: (2025)
Leveraging Offline Data in Linear Latent Contextual Bandits
by: Kausik, Chinmaya, et al.
Published: (2024)
by: Kausik, Chinmaya, et al.
Published: (2024)
Off-Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits
by: Zhan, Ruohan, et al.
Published: (2021)
by: Zhan, Ruohan, et al.
Published: (2021)
GenAI Powered Dynamic Causal Inference with Unstructured Data
by: Nakamura, Kentaro, et al.
Published: (2026)
by: Nakamura, Kentaro, et al.
Published: (2026)
Prediction-Powered Adaptive Shrinkage Estimation
by: Li, Sida, et al.
Published: (2025)
by: Li, Sida, et al.
Published: (2025)
Linear Contextual Bandits with Interference
by: Xu, Yang, et al.
Published: (2024)
by: Xu, Yang, et al.
Published: (2024)
Deployment of AI-Assisted Interventions: Capacity Constraints and Noisy Compliance
by: Chan, Carri W., et al.
Published: (2026)
by: Chan, Carri W., et al.
Published: (2026)
Latency-Aware Contextual Bandit: Application to Cryo-EM Data Collection
by: Wei, Lai, et al.
Published: (2024)
by: Wei, Lai, et al.
Published: (2024)
Sample Efficient Bayesian Learning of Causal Graphs from Interventions
by: Zhou, Zihan, et al.
Published: (2024)
by: Zhou, Zihan, et al.
Published: (2024)
Multi-Armed Bandits with Network Interference
by: Agarwal, Abhineet, et al.
Published: (2024)
by: Agarwal, Abhineet, et al.
Published: (2024)
Sample size planning for conditional counterfactual mean estimation with a K-armed randomized experiment
by: Ruiz, Gabriel
Published: (2024)
by: Ruiz, Gabriel
Published: (2024)
Reinforcement Learning in Modern Biostatistics: Constructing Optimal Adaptive Interventions
by: Deliu, Nina, et al.
Published: (2022)
by: Deliu, Nina, et al.
Published: (2022)
High-dimensional Nonparametric Contextual Bandit Problem
by: Iwazaki, Shogo, et al.
Published: (2025)
by: Iwazaki, Shogo, et al.
Published: (2025)
Q-Learning with Clustered-SMART (cSMART) Data: Examining Moderators in the Construction of Clustered Adaptive Interventions
by: Song, Yao, et al.
Published: (2025)
by: Song, Yao, et al.
Published: (2025)
Locally Private Nonparametric Contextual Multi-armed Bandits
by: Ma, Yuheng, et al.
Published: (2025)
by: Ma, Yuheng, et al.
Published: (2025)
Nearly Optimal Best Arm Identification for Semiparametric Bandits
by: Kim, Seok-Jin
Published: (2026)
by: Kim, Seok-Jin
Published: (2026)
PPI is the Difference Estimator: Recognizing the Survey Sampling Roots of Prediction-Powered Inference
by: Mozer, Reagan
Published: (2026)
by: Mozer, Reagan
Published: (2026)
High-Dimensional Linear Bandits under Stochastic Latent Heterogeneity
by: Chen, Elynn, et al.
Published: (2025)
by: Chen, Elynn, et al.
Published: (2025)
Similar Items
-
GenAI-Powered Inference
by: Imai, Kosuke, et al.
Published: (2025) -
A Primal-Dual Algorithm for Offline Constrained Reinforcement Learning with Linear MDPs
by: Hong, Kihyuk, et al.
Published: (2024) -
A Computationally Efficient Algorithm for Infinite-Horizon Average-Reward Linear MDPs
by: Hong, Kihyuk, et al.
Published: (2025) -
Learning to Partially Defer for Sequences
by: Rayan, Sahana, et al.
Published: (2025) -
Distribution-Free Robust Predict-Then-Optimize in Function Spaces
by: Patel, Yash, et al.
Published: (2026)