Stochastic Online Conformal Prediction with Semi-Bandit Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Ge, Haosen, Bastani, Hamsa, Bastani, Osbert |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Algorithmic Fairness for Human-AI Collaboration
by: Ge, Haosen, et al.
Published: (2023)
by: Ge, Haosen, et al.
Published: (2023)
Stochastic Bandits with ReLU Neural Networks
by: Xu, Kan, et al.
Published: (2024)
by: Xu, Kan, et al.
Published: (2024)
Are AI Capabilities Increasing Exponentially? A Competing Hypothesis
by: Ge, Haosen, et al.
Published: (2026)
by: Ge, Haosen, et al.
Published: (2026)
Beating the Winner's Curse via Inference-Aware Policy Optimization
by: Bastani, Hamsa, et al.
Published: (2025)
by: Bastani, Hamsa, et al.
Published: (2025)
Winner's Curse Drives False Promises in Data-Driven Decisions: A Case Study in Refugee Matching
by: Bastani, Hamsa, et al.
Published: (2026)
by: Bastani, Hamsa, et al.
Published: (2026)
Improving Human Sequential Decision-Making with Reinforcement Learning
by: Bastani, Hamsa, et al.
Published: (2021)
by: Bastani, Hamsa, et al.
Published: (2021)
Group-Sparse Matrix Factorization for Transfer Learning of Word Embeddings
by: Xu, Kan, et al.
Published: (2021)
by: Xu, Kan, et al.
Published: (2021)
Multitask Learning and Bandits via Robust Statistics
by: Xu, Kan, et al.
Published: (2021)
by: Xu, Kan, et al.
Published: (2021)
Conformal Structured Prediction
by: Zhang, Botong, et al.
Published: (2024)
by: Zhang, Botong, et al.
Published: (2024)
Asymptotic Normality of Generalized Low-Rank Matrix Sensing via Riemannian Geometry
by: Bastani, Osbert
Published: (2024)
by: Bastani, Osbert
Published: (2024)
Uncertainty Quantification for Neurosymbolic Programs via Compositional Conformal Prediction
by: Ramalingam, Ramya, et al.
Published: (2024)
by: Ramalingam, Ramya, et al.
Published: (2024)
Generative Adversarial Model-Based Optimization via Source Critic Regularization
by: Yao, Michael S., et al.
Published: (2024)
by: Yao, Michael S., et al.
Published: (2024)
Conformal Constrained Policy Optimization for Cost-Effective LLM Agents
by: Si, Wenwen, et al.
Published: (2025)
by: Si, Wenwen, et al.
Published: (2025)
LLM Program Optimization via Retrieval Augmented Search
by: Anupam, Sagnik, et al.
Published: (2025)
by: Anupam, Sagnik, et al.
Published: (2025)
Strategic Hiring under Algorithmic Monoculture
by: Baek, Jackie, et al.
Published: (2025)
by: Baek, Jackie, et al.
Published: (2025)
SPARLING: Learning Latent Representations with Extremely Sparse Activations
by: Gupta, Kavi, et al.
Published: (2023)
by: Gupta, Kavi, et al.
Published: (2023)
Prior-Agnostic Incentive-Compatible Exploration
by: Ramalingam, Ramya, et al.
Published: (2026)
by: Ramalingam, Ramya, et al.
Published: (2026)
Optimal Multitask Linear Regression and Contextual Bandits under Sparse Heterogeneity
by: Huang, Xinmeng, et al.
Published: (2023)
by: Huang, Xinmeng, et al.
Published: (2023)
Diversity By Design: Leveraging Distribution Matching for Offline Model-Based Optimization
by: Yao, Michael S., et al.
Published: (2025)
by: Yao, Michael S., et al.
Published: (2025)
SeekerGym: A Benchmark for Reliable Information Seeking
by: Kim, Remy, et al.
Published: (2026)
by: Kim, Remy, et al.
Published: (2026)
Improving Structural Diversity of Blackbox LLMs via Chain-of-Specification Prompting
by: Young, Halley, et al.
Published: (2024)
by: Young, Halley, et al.
Published: (2024)
RAPID: An Efficient Reinforcement Learning Algorithm for Small Language Models
by: Huang, Lianghuan, et al.
Published: (2025)
by: Huang, Lianghuan, et al.
Published: (2025)
Decaf: Improving Neural Decompilation with Automatic Feedback and Search
by: Shypula, Alexander, et al.
Published: (2026)
by: Shypula, Alexander, et al.
Published: (2026)
TRAQ: Trustworthy Retrieval Augmented Question Answering via Conformal Prediction
by: Li, Shuo, et al.
Published: (2023)
by: Li, Shuo, et al.
Published: (2023)
BrowserArena: Evaluating LLM Agents on Real-World Web Navigation Tasks
by: Anupam, Sagnik, et al.
Published: (2025)
by: Anupam, Sagnik, et al.
Published: (2025)
Knowledgeable Language Models as Black-Box Optimizers for Personalized Medicine
by: Yao, Michael S., et al.
Published: (2025)
by: Yao, Michael S., et al.
Published: (2025)
Synthesizing Trajectory Queries from Examples
by: Mell, Stephen, et al.
Published: (2026)
by: Mell, Stephen, et al.
Published: (2026)
Alignment of large language models with constrained learning
by: Zhang, Botong, et al.
Published: (2025)
by: Zhang, Botong, et al.
Published: (2025)
One-Shot Safety Alignment for Large Language Models via Optimal Dualization
by: Huang, Xinmeng, et al.
Published: (2024)
by: Huang, Xinmeng, et al.
Published: (2024)
Online Conformal Prediction with Adversarial Semi-bandit Feedback via Regret Minimization
by: Yang, Junyoung, et al.
Published: (2026)
by: Yang, Junyoung, et al.
Published: (2026)
Eurekaverse: Environment Curriculum Generation via Large Language Models
by: Liang, William, et al.
Published: (2024)
by: Liang, William, et al.
Published: (2024)
Online Conformal Abstention for Factuality Control Under Adversarial Bandit Feedback
by: Lee, Minjae, et al.
Published: (2025)
by: Lee, Minjae, et al.
Published: (2025)
Bandit and Delayed Feedback in Online Structured Prediction
by: Shibukawa, Yuki, et al.
Published: (2025)
by: Shibukawa, Yuki, et al.
Published: (2025)
Online Budget Allocation with Censored Semi-Bandit Feedback
by: Bachoc, François, et al.
Published: (2025)
by: Bachoc, François, et al.
Published: (2025)
Adversarial Query Synthesis via Bayesian Optimization
by: Tao, Jeffrey, et al.
Published: (2026)
by: Tao, Jeffrey, et al.
Published: (2026)
Stochastic Online Instrumental Variable Regression: Regrets for Endogeneity and Bandit Feedback
by: Della Vecchia, Riccardo, et al.
Published: (2023)
by: Della Vecchia, Riccardo, et al.
Published: (2023)
Online Conformal Prediction with Corrupted Feedback
by: Wang, Bowen, et al.
Published: (2026)
by: Wang, Bowen, et al.
Published: (2026)
Policy-Guided Stepwise Model Routing for Cost-Effective Reasoning
by: Si, Wenwen, et al.
Published: (2026)
by: Si, Wenwen, et al.
Published: (2026)
Optimal Program Synthesis via Abstract Interpretation
by: Mell, Stephen, et al.
Published: (2026)
by: Mell, Stephen, et al.
Published: (2026)
Uncertainty in Language Models: Assessment through Rank-Calibration
by: Huang, Xinmeng, et al.
Published: (2024)
by: Huang, Xinmeng, et al.
Published: (2024)
Similar Items
-
Rethinking Algorithmic Fairness for Human-AI Collaboration
by: Ge, Haosen, et al.
Published: (2023) -
Stochastic Bandits with ReLU Neural Networks
by: Xu, Kan, et al.
Published: (2024) -
Are AI Capabilities Increasing Exponentially? A Competing Hypothesis
by: Ge, Haosen, et al.
Published: (2026) -
Beating the Winner's Curse via Inference-Aware Policy Optimization
by: Bastani, Hamsa, et al.
Published: (2025) -
Winner's Curse Drives False Promises in Data-Driven Decisions: A Case Study in Refugee Matching
by: Bastani, Hamsa, et al.
Published: (2026)