Cost-optimal Sequential Testing via Doubly Robust Q-learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Doudou, Zhang, Yiran, Jin, Dian, Zheng, Yingye, Tian, Lu, Cai, Tianxi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sail into the Headwind: Alignment via Robust Rewards and Dynamic Labels against Reward Hacking
by: Rashidinejad, Paria, et al.
Published: (2024)
by: Rashidinejad, Paria, et al.
Published: (2024)
Provable Robust Overfitting Mitigation in Wasserstein Distributionally Robust Optimization
by: Liu, Shuang, et al.
Published: (2025)
by: Liu, Shuang, et al.
Published: (2025)
Psychometric Tests for AI Agents and Their Moduli Space
by: Chojecki, Przemyslaw
Published: (2025)
by: Chojecki, Przemyslaw
Published: (2025)
Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals
by: Dong, Zihan, et al.
Published: (2026)
by: Dong, Zihan, et al.
Published: (2026)
Enhancing Conformal Prediction Using E-Test Statistics
by: Balinsky, A. A., et al.
Published: (2024)
by: Balinsky, A. A., et al.
Published: (2024)
Influence functions and regularity tangents for efficient active learning
by: Eaton, Frederik
Published: (2024)
by: Eaton, Frederik
Published: (2024)
Generalising realisability in statistical learning theory under epistemic uncertainty
by: Cuzzolin, Fabio
Published: (2024)
by: Cuzzolin, Fabio
Published: (2024)
Interaction Testing in Variation Analysis
by: Plecko, Drago
Published: (2024)
by: Plecko, Drago
Published: (2024)
A comparative study of conformal prediction methods for valid uncertainty quantification in machine learning
by: Dewolf, Nicolas
Published: (2024)
by: Dewolf, Nicolas
Published: (2024)
Conditional Distributional Treatment Effects: Doubly Robust Estimation and Testing
by: Jain, Saksham, et al.
Published: (2026)
by: Jain, Saksham, et al.
Published: (2026)
Representation learning with a transformer by contrastive learning for money laundering detection
by: Guéneau, Harold, et al.
Published: (2025)
by: Guéneau, Harold, et al.
Published: (2025)
Machine learning and optimization-based approaches to duality in statistical physics
by: Ferrari, Andrea E. V., et al.
Published: (2024)
by: Ferrari, Andrea E. V., et al.
Published: (2024)
Knowledge-Embedded Latent Projection for Robust Representation Learning
by: Tang, Weijing, et al.
Published: (2026)
by: Tang, Weijing, et al.
Published: (2026)
Test-time Alignment of Diffusion Models without Reward Over-optimization
by: Kim, Sunwoo, et al.
Published: (2025)
by: Kim, Sunwoo, et al.
Published: (2025)
Prior-dependent analysis of posterior sampling reinforcement learning with function approximation
by: Li, Yingru, et al.
Published: (2024)
by: Li, Yingru, et al.
Published: (2024)
Beyond identifiability: Learning causal representations with few environments and finite samples
by: Lee, Inbeom, et al.
Published: (2026)
by: Lee, Inbeom, et al.
Published: (2026)
Label Noise Robustness of Conformal Prediction
by: Einbinder, Bat-Sheva, et al.
Published: (2022)
by: Einbinder, Bat-Sheva, et al.
Published: (2022)
Decision Making in Changing Environments: Robustness, Query-Based Learning, and Differential Privacy
by: Chen, Fan, et al.
Published: (2025)
by: Chen, Fan, et al.
Published: (2025)
How Particle-System Random Batch Methods Enhance Graph Transformer: Memory Efficiency and Parallel Computing Strategy
by: Liu, Hanwen, et al.
Published: (2025)
by: Liu, Hanwen, et al.
Published: (2025)
Efficient Knowledge Distillation via Curriculum Extraction
by: Gupta, Shivam, et al.
Published: (2025)
by: Gupta, Shivam, et al.
Published: (2025)
MATES: Multi-view Aggregated Two-Sample Test
by: Cai, Zexi, et al.
Published: (2024)
by: Cai, Zexi, et al.
Published: (2024)
Doubly Robust Interval Estimation for Optimal Policy Evaluation in Online Learning
by: Shen, Ye, et al.
Published: (2021)
by: Shen, Ye, et al.
Published: (2021)
Training Implicit Generative Models via an Invariant Statistical Loss
by: de Frutos, José Manuel, et al.
Published: (2024)
by: de Frutos, José Manuel, et al.
Published: (2024)
Beyond Demand Estimation: Consumer Surplus Evaluation via Cumulative Propensity Weights
by: Bian, Zeyu, et al.
Published: (2026)
by: Bian, Zeyu, et al.
Published: (2026)
Le Cam Distortion: A Decision-Theoretic Framework for Robust Transfer Learning
by: Akdemir, Deniz
Published: (2025)
by: Akdemir, Deniz
Published: (2025)
A Statistical Hypothesis Testing Framework for Data Misappropriation Detection in Large Language Models
by: Cai, Yinpeng, et al.
Published: (2025)
by: Cai, Yinpeng, et al.
Published: (2025)
Adaptive auditing of AI systems with anytime-valid guarantees
by: Zhou, Siyu, et al.
Published: (2026)
by: Zhou, Siyu, et al.
Published: (2026)
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
by: Bareilles, Gilles, et al.
Published: (2026)
by: Bareilles, Gilles, et al.
Published: (2026)
Tail-Aware Information-Theoretic Generalization for RLHF and SGLD
by: Zhang, Huiming, et al.
Published: (2026)
by: Zhang, Huiming, et al.
Published: (2026)
About the Cost of Central Privacy in Density Estimation
by: Lalanne, Clément, et al.
Published: (2023)
by: Lalanne, Clément, et al.
Published: (2023)
Hierarchical Contrastive Learning for Multimodal Data
by: Li, Huichao, et al.
Published: (2026)
by: Li, Huichao, et al.
Published: (2026)
Entropy, concentration, and learning: a statistical mechanics primer
by: Balsubramani, Akshay
Published: (2024)
by: Balsubramani, Akshay
Published: (2024)
Differentially Private Two-Stage Gradient Descent for Instrumental Variable Regression
by: Liang, Haodong, et al.
Published: (2025)
by: Liang, Haodong, et al.
Published: (2025)
Improving Efficiency and Robustness of the Prognostic Accuracy of Biomarkers With Partial Incomplete Failure‐Time Data and Auxiliary Outcome: Application to Prostate Cancer Active Surveillance Study
by: Yunro Chung, et al.
Published: (2025)
by: Yunro Chung, et al.
Published: (2025)
Structure-agnostic Optimality of Doubly Robust Learning for Treatment Effect Estimation
by: Jin, Jikai, et al.
Published: (2024)
by: Jin, Jikai, et al.
Published: (2024)
Pessimism in the Face of Confounders: Provably Efficient Offline Reinforcement Learning in Partially Observable Markov Decision Processes
by: Lu, Miao, et al.
Published: (2022)
by: Lu, Miao, et al.
Published: (2022)
Automatic Doubly Robust Forests
by: Chen, Zhaomeng, et al.
Published: (2024)
by: Chen, Zhaomeng, et al.
Published: (2024)
Representation-Enhanced Neural Knowledge Integration with Application to Large-Scale Medical Ontology Learning
by: Liu, Suqi, et al.
Published: (2024)
by: Liu, Suqi, et al.
Published: (2024)
Wasserstein Transfer Learning
by: Zhang, Kaicheng, et al.
Published: (2025)
by: Zhang, Kaicheng, et al.
Published: (2025)
Mitigating loss of variance in ensemble data assimilation: machine learning-based and distance-free localization
by: Silva, Vinicius L. S., et al.
Published: (2025)
by: Silva, Vinicius L. S., et al.
Published: (2025)
Similar Items
-
Sail into the Headwind: Alignment via Robust Rewards and Dynamic Labels against Reward Hacking
by: Rashidinejad, Paria, et al.
Published: (2024) -
Provable Robust Overfitting Mitigation in Wasserstein Distributionally Robust Optimization
by: Liu, Shuang, et al.
Published: (2025) -
Psychometric Tests for AI Agents and Their Moduli Space
by: Chojecki, Przemyslaw
Published: (2025) -
Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals
by: Dong, Zihan, et al.
Published: (2026) -
Enhancing Conformal Prediction Using E-Test Statistics
by: Balinsky, A. A., et al.
Published: (2024)