Statistical Inference for Misspecified Contextual Bandits
Fuente:
arXiv
Guardado en:
| Autores principales: | Guo, Yongyi, Xu, Ziping |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Identifying All ε-Best Arms in (Misspecified) Linear Bandits
por: Li, Zhekai, et al.
Publicado: (2025)
por: Li, Zhekai, et al.
Publicado: (2025)
Towards a Sharp Analysis of Offline Policy Learning for $f$-Divergence-Regularized Contextual Bandits
por: Zhao, Qingyue, et al.
Publicado: (2025)
por: Zhao, Qingyue, et al.
Publicado: (2025)
Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability
por: Zhao, Qingyue, et al.
Publicado: (2026)
por: Zhao, Qingyue, et al.
Publicado: (2026)
Statistical Inference for Optimal Transport Maps: Recent Advances and Perspectives
por: Balakrishnan, Sivaraman, et al.
Publicado: (2025)
por: Balakrishnan, Sivaraman, et al.
Publicado: (2025)
CITE: Anytime-Valid Statistical Inference in LLM Self-Consistency
por: Ota, Hirofumi, et al.
Publicado: (2026)
por: Ota, Hirofumi, et al.
Publicado: (2026)
Retraining as Approximate Bayesian Inference
por: Katz, Harrison
Publicado: (2026)
por: Katz, Harrison
Publicado: (2026)
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
por: Lattimore, Tor
Publicado: (2026)
por: Lattimore, Tor
Publicado: (2026)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
por: Ji, Kaixuan, et al.
Publicado: (2026)
por: Ji, Kaixuan, et al.
Publicado: (2026)
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
por: Ji, Kaixuan, et al.
Publicado: (2026)
por: Ji, Kaixuan, et al.
Publicado: (2026)
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
por: Boudart, Pierre, et al.
Publicado: (2025)
por: Boudart, Pierre, et al.
Publicado: (2025)
LIBRA: Language Model Informed Bandit Recourse Algorithm for Personalized Treatment Planning
por: Cao, Junyu, et al.
Publicado: (2026)
por: Cao, Junyu, et al.
Publicado: (2026)
Smooth Non-Stationary Bandits
por: Jia, Su, et al.
Publicado: (2023)
por: Jia, Su, et al.
Publicado: (2023)
Early Stopping in Contextual Bandits and Inferences
por: Cui, Zihan
Publicado: (2025)
por: Cui, Zihan
Publicado: (2025)
On the Statistical Capacity of Deep Generative Models
por: Tam, Edric, et al.
Publicado: (2025)
por: Tam, Edric, et al.
Publicado: (2025)
Statistical inference with belief functions: A survey
por: Cuzzolin, Fabio
Publicado: (2026)
por: Cuzzolin, Fabio
Publicado: (2026)
Multiple-Prediction-Powered Inference
por: Cowen-Breen, Charlie, et al.
Publicado: (2026)
por: Cowen-Breen, Charlie, et al.
Publicado: (2026)
Enhancing Conformal Prediction Using E-Test Statistics
por: Balinsky, A. A., et al.
Publicado: (2024)
por: Balinsky, A. A., et al.
Publicado: (2024)
Avoiding the Price of Adaptivity: Inference in Linear Contextual Bandits via Stability
por: Praharaj, Samya, et al.
Publicado: (2025)
por: Praharaj, Samya, et al.
Publicado: (2025)
Solving a Research Problem in Mathematical Statistics with AI Assistance
por: Dobriban, Edgar
Publicado: (2025)
por: Dobriban, Edgar
Publicado: (2025)
Training Implicit Generative Models via an Invariant Statistical Loss
por: de Frutos, José Manuel, et al.
Publicado: (2024)
por: de Frutos, José Manuel, et al.
Publicado: (2024)
Counterfactual Generative Modeling with Variational Causal Inference
por: Wu, Yulun, et al.
Publicado: (2024)
por: Wu, Yulun, et al.
Publicado: (2024)
On the Statistical Properties of Generative Adversarial Models for Low Intrinsic Data Dimension
por: Chakraborty, Saptarshi, et al.
Publicado: (2024)
por: Chakraborty, Saptarshi, et al.
Publicado: (2024)
Bayesian Modular Inference for Copula Models with Potentially Misspecified Marginals
por: Kock, Lucas, et al.
Publicado: (2026)
por: Kock, Lucas, et al.
Publicado: (2026)
A Statistical Analysis of Deep Federated Learning for Intrinsically Low-dimensional Data
por: Chakraborty, Saptarshi, et al.
Publicado: (2024)
por: Chakraborty, Saptarshi, et al.
Publicado: (2024)
Debiased Prediction Inference with Non-sparse Loadings in Misspecified High-dimensional Regression Models
por: Liang, Libin, et al.
Publicado: (2025)
por: Liang, Libin, et al.
Publicado: (2025)
Variational Causal Inference
por: Wu, Yulun, et al.
Publicado: (2022)
por: Wu, Yulun, et al.
Publicado: (2022)
A Graphical Global Optimization Framework for Parameter Estimation of Statistical Models with Nonconvex Regularization Functions
por: Davarnia, Danial, et al.
Publicado: (2025)
por: Davarnia, Danial, et al.
Publicado: (2025)
Navigating the Exploration-Exploitation Tradeoff in Inference-Time Scaling of Diffusion Models
por: Su, Xun, et al.
Publicado: (2025)
por: Su, Xun, et al.
Publicado: (2025)
Federated Causal Inference from Multi-Site Observational Data via Propensity Score Aggregation
por: Khellaf, Rémi, et al.
Publicado: (2025)
por: Khellaf, Rémi, et al.
Publicado: (2025)
Bayesian Inference of Minimally Complex Models with Interactions of Arbitrary Order
por: de Mulatier, Clélia, et al.
Publicado: (2020)
por: de Mulatier, Clélia, et al.
Publicado: (2020)
Batched Nonparametric Contextual Bandits
por: Jiang, Rong, et al.
Publicado: (2024)
por: Jiang, Rong, et al.
Publicado: (2024)
Theory of Evolutionary Spectra for Heteroskedasticity and Autocorrelation Robust Inference in Possibly Misspecified and Nonstationary Models
por: Casini, Alessandro
Publicado: (2021)
por: Casini, Alessandro
Publicado: (2021)
Statistical and Algorithmic Foundations of Reinforcement Learning
por: Chi, Yuejie, et al.
Publicado: (2025)
por: Chi, Yuejie, et al.
Publicado: (2025)
Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals
por: Dong, Zihan, et al.
Publicado: (2026)
por: Dong, Zihan, et al.
Publicado: (2026)
Unveiling the Statistical Foundations of Chain-of-Thought Prompting Methods
por: Hu, Xinyang, et al.
Publicado: (2024)
por: Hu, Xinyang, et al.
Publicado: (2024)
Multitask Learning and Bandits via Robust Statistics
por: Xu, Kan, et al.
Publicado: (2021)
por: Xu, Kan, et al.
Publicado: (2021)
Inference with the Upper Confidence Bound Algorithm
por: Khamaru, Koulik, et al.
Publicado: (2024)
por: Khamaru, Koulik, et al.
Publicado: (2024)
Beyond Covariance Matrix: The Statistical Complexity of Private Linear Regression
por: Chen, Fan, et al.
Publicado: (2025)
por: Chen, Fan, et al.
Publicado: (2025)
A Reflection on the Impact of Misspecifying Unidentifiable Causal Inference Models in Surrogate Endpoint Evaluation
por: Deliorman, Gokce, et al.
Publicado: (2024)
por: Deliorman, Gokce, et al.
Publicado: (2024)
On the Optimality of Misspecified Spectral Algorithms
por: Zhang, Haobo, et al.
Publicado: (2023)
por: Zhang, Haobo, et al.
Publicado: (2023)
Ejemplares similares
-
Identifying All ε-Best Arms in (Misspecified) Linear Bandits
por: Li, Zhekai, et al.
Publicado: (2025) -
Towards a Sharp Analysis of Offline Policy Learning for $f$-Divergence-Regularized Contextual Bandits
por: Zhao, Qingyue, et al.
Publicado: (2025) -
Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability
por: Zhao, Qingyue, et al.
Publicado: (2026) -
Statistical Inference for Optimal Transport Maps: Recent Advances and Perspectives
por: Balakrishnan, Sivaraman, et al.
Publicado: (2025) -
CITE: Anytime-Valid Statistical Inference in LLM Self-Consistency
por: Ota, Hirofumi, et al.
Publicado: (2026)