Breaking Determinism: Stochastic Modeling for Reliable Off-Policy Evaluation in Ad Auctions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yeom, Hongseon, Shin, Jaeyoul, Min, Soojin, Yoon, Jeongmin, Yu, Seunghak, Kang, Dongyeop |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Off-Policy Evaluation and Counterfactual Methods in Dynamic Auction Environments
von: Guha, Ritam, et al.
Veröffentlicht: (2025)
von: Guha, Ritam, et al.
Veröffentlicht: (2025)
Learning a High-quality Robotic Wiping Policy Using Systematic Reward Analysis and Visual-Language Model Based Curriculum
von: Liu, Yihong, et al.
Veröffentlicht: (2025)
von: Liu, Yihong, et al.
Veröffentlicht: (2025)
Ad Auctions for LLMs via Retrieval Augmented Generation
von: Hajiaghayi, MohammadTaghi, et al.
Veröffentlicht: (2024)
von: Hajiaghayi, MohammadTaghi, et al.
Veröffentlicht: (2024)
Cross-Validated Off-Policy Evaluation
von: Cief, Matej, et al.
Veröffentlicht: (2024)
von: Cief, Matej, et al.
Veröffentlicht: (2024)
$K$-Nearest-Neighbor Resampling for Off-Policy Evaluation in Stochastic Control
von: Giegrich, Michael, et al.
Veröffentlicht: (2023)
von: Giegrich, Michael, et al.
Veröffentlicht: (2023)
Off-Policy Evaluation for Ranking Policies under Deterministic Logging Policies
von: Tanaka, Koichi, et al.
Veröffentlicht: (2026)
von: Tanaka, Koichi, et al.
Veröffentlicht: (2026)
Off-Policy Evaluation of Ranking Policies via Embedding-Space User Behavior Modeling
von: Takahashi, Tatsuki, et al.
Veröffentlicht: (2025)
von: Takahashi, Tatsuki, et al.
Veröffentlicht: (2025)
Doubly-Robust Off-Policy Evaluation with Estimated Logging Policy
von: Lee, Kyungbok, et al.
Veröffentlicht: (2024)
von: Lee, Kyungbok, et al.
Veröffentlicht: (2024)
Off-Policy Actor-Critic for Adversarial Observation Robustness: Virtual Alternative Training via Symmetric Policy Evaluation
von: Nakanishi, Kosuke, et al.
Veröffentlicht: (2025)
von: Nakanishi, Kosuke, et al.
Veröffentlicht: (2025)
Long-term Off-Policy Evaluation and Learning
von: Saito, Yuta, et al.
Veröffentlicht: (2024)
von: Saito, Yuta, et al.
Veröffentlicht: (2024)
Breaking the Curse of Repulsion: Optimistic Distributionally Robust Policy Optimization for Off-Policy Generative Recommendation
von: Jiang, Jie, et al.
Veröffentlicht: (2026)
von: Jiang, Jie, et al.
Veröffentlicht: (2026)
From Weighting to Modeling: A Nonparametric Estimator for Off-Policy Evaluation
von: Zhu, Rong J. B.
Veröffentlicht: (2026)
von: Zhu, Rong J. B.
Veröffentlicht: (2026)
Off-Policy Evaluation of Slate Bandit Policies via Optimizing Abstraction
von: Kiyohara, Haruka, et al.
Veröffentlicht: (2024)
von: Kiyohara, Haruka, et al.
Veröffentlicht: (2024)
Clustering Context in Off-Policy Evaluation
von: Guzman-Olivares, Daniel, et al.
Veröffentlicht: (2025)
von: Guzman-Olivares, Daniel, et al.
Veröffentlicht: (2025)
Concept-driven Off Policy Evaluation
von: Majumdar, Ritam, et al.
Veröffentlicht: (2024)
von: Majumdar, Ritam, et al.
Veröffentlicht: (2024)
Advancing Ad Auction Realism: Practical Insights & Modeling Implications
von: Chen, Ming, et al.
Veröffentlicht: (2023)
von: Chen, Ming, et al.
Veröffentlicht: (2023)
Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards
von: Takahashi, Tatsuki, et al.
Veröffentlicht: (2025)
von: Takahashi, Tatsuki, et al.
Veröffentlicht: (2025)
Off-Policy Evaluation Under Nonignorable Missing Data
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
Off-Policy Evaluation from Logged Human Feedback
von: Bhargava, Aniruddha, et al.
Veröffentlicht: (2024)
von: Bhargava, Aniruddha, et al.
Veröffentlicht: (2024)
Simulation-Based Sensitivity Analysis in Optimal Treatment Regimes and Causal Decomposition with Individualized Interventions
von: Park, Soojin, et al.
Veröffentlicht: (2025)
von: Park, Soojin, et al.
Veröffentlicht: (2025)
Auction-Based Online Policy Adaptation for Evolving Objectives
von: Shabadi, Guruprerana, et al.
Veröffentlicht: (2026)
von: Shabadi, Guruprerana, et al.
Veröffentlicht: (2026)
SASSHA: Sharpness-aware Adaptive Second-order Optimization with Stable Hessian Approximation
von: Shin, Dahun, et al.
Veröffentlicht: (2025)
von: Shin, Dahun, et al.
Veröffentlicht: (2025)
Critical Influence of Overparameterization on Sharpness-aware Minimization
von: Shin, Sungbin, et al.
Veröffentlicht: (2023)
von: Shin, Sungbin, et al.
Veröffentlicht: (2023)
EdiText: Controllable Coarse-to-Fine Text Editing with Diffusion Language Models
von: Lee, Che Hyun, et al.
Veröffentlicht: (2025)
von: Lee, Che Hyun, et al.
Veröffentlicht: (2025)
Iterated Energy-based Flow Matching for Sampling from Boltzmann Densities
von: Woo, Dongyeop, et al.
Veröffentlicht: (2024)
von: Woo, Dongyeop, et al.
Veröffentlicht: (2024)
Off-Policy Evaluation and Learning for Matching Markets
von: Hayashi, Yudai, et al.
Veröffentlicht: (2025)
von: Hayashi, Yudai, et al.
Veröffentlicht: (2025)
Learning Action Embeddings for Off-Policy Evaluation
von: Cief, Matej, et al.
Veröffentlicht: (2023)
von: Cief, Matej, et al.
Veröffentlicht: (2023)
Log-Sum-Exponential Estimator for Off-Policy Evaluation and Learning
von: Behnamnia, Armin, et al.
Veröffentlicht: (2025)
von: Behnamnia, Armin, et al.
Veröffentlicht: (2025)
Logarithmic Smoothing for Pessimistic Off-Policy Evaluation, Selection and Learning
von: Sakhi, Otmane, et al.
Veröffentlicht: (2024)
von: Sakhi, Otmane, et al.
Veröffentlicht: (2024)
CANDOR: Counterfactual ANnotated DOubly Robust Off-Policy Evaluation
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2024)
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2024)
Becoming Experienced Judges: Selective Test-Time Learning for Evaluators
von: Jwa, Seungyeon, et al.
Veröffentlicht: (2025)
von: Jwa, Seungyeon, et al.
Veröffentlicht: (2025)
Context-Action Embedding Learning for Off-Policy Evaluation in Contextual Bandits
von: Chandak, Kushagra, et al.
Veröffentlicht: (2025)
von: Chandak, Kushagra, et al.
Veröffentlicht: (2025)
DOLCE: Decomposing Off-Policy Evaluation/Learning into Lagged and Current Effects
von: Tamano, Shu
Veröffentlicht: (2025)
von: Tamano, Shu
Veröffentlicht: (2025)
Off-Policy Evaluation Using Information Borrowing and Context-Based Switching
von: Dasgupta, Sutanoy, et al.
Veröffentlicht: (2021)
von: Dasgupta, Sutanoy, et al.
Veröffentlicht: (2021)
IntOPE: Off-Policy Evaluation in the Presence of Interference
von: Bai, Yuqi, et al.
Veröffentlicht: (2024)
von: Bai, Yuqi, et al.
Veröffentlicht: (2024)
Improved Online Learning Algorithms for CTR Prediction in Ad Auctions
von: Feng, Zhe, et al.
Veröffentlicht: (2024)
von: Feng, Zhe, et al.
Veröffentlicht: (2024)
Symmetry-Breaking Augmentations for Ad Hoc Teamwork
von: Hammond, Ravi, et al.
Veröffentlicht: (2024)
von: Hammond, Ravi, et al.
Veröffentlicht: (2024)
Distributional Off-Policy Evaluation with Deep Quantile Process Regression
von: Kuang, Qi, et al.
Veröffentlicht: (2026)
von: Kuang, Qi, et al.
Veröffentlicht: (2026)
Data Poisoning Attacks on Off-Policy Policy Evaluation Methods
von: Lobo, Elita, et al.
Veröffentlicht: (2024)
von: Lobo, Elita, et al.
Veröffentlicht: (2024)
The Privacy-Utility Trade-Off of Location Tracking in Ad Personalization
von: Mosaffa, Mohammad, et al.
Veröffentlicht: (2026)
von: Mosaffa, Mohammad, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Off-Policy Evaluation and Counterfactual Methods in Dynamic Auction Environments
von: Guha, Ritam, et al.
Veröffentlicht: (2025) -
Learning a High-quality Robotic Wiping Policy Using Systematic Reward Analysis and Visual-Language Model Based Curriculum
von: Liu, Yihong, et al.
Veröffentlicht: (2025) -
Ad Auctions for LLMs via Retrieval Augmented Generation
von: Hajiaghayi, MohammadTaghi, et al.
Veröffentlicht: (2024) -
Cross-Validated Off-Policy Evaluation
von: Cief, Matej, et al.
Veröffentlicht: (2024) -
$K$-Nearest-Neighbor Resampling for Off-Policy Evaluation in Stochastic Control
von: Giegrich, Michael, et al.
Veröffentlicht: (2023)