Saved in:
| Main Authors: | Bakker, Hua Chang, Gupta, Shashank, Oosterhuis, Harrie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2409.09819 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Practical and Robust Safety Guarantees for Advanced Counterfactual Learning to Rank
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Proximal Ranking Policy Optimization for Practical Safety in Counterfactual Learning to Rank
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Learning to Rank with Variable Result Presentation Lengths
by: Knyazev, Norman, et al.
Published: (2025)
by: Knyazev, Norman, et al.
Published: (2025)
Local Feature Selection without Label or Feature Leakage for Interpretable Machine Learning Predictions
by: Oosterhuis, Harrie, et al.
Published: (2024)
by: Oosterhuis, Harrie, et al.
Published: (2024)
Estimating the Hessian Matrix of Ranking Objectives for Stochastic Learning to Rank with Gradient Boosted Trees
by: Kang, Jingwei, et al.
Published: (2024)
by: Kang, Jingwei, et al.
Published: (2024)
A First Look at Selection Bias in Preference Elicitation for Recommendation
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Optimizing Compound Retrieval Systems
by: Oosterhuis, Harrie, et al.
Published: (2025)
by: Oosterhuis, Harrie, et al.
Published: (2025)
Reliable Confidence Intervals for Information Retrieval Evaluation Using Generative A.I
by: Oosterhuis, Harrie, et al.
Published: (2024)
by: Oosterhuis, Harrie, et al.
Published: (2024)
A Simple and Effective Reinforcement Learning Method for Text-to-Image Diffusion Fine-tuning
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
A Non-Parametric Choice Model That Learns How Users Choose Between Recommended Options
by: Krause, Thorsten, et al.
Published: (2025)
by: Krause, Thorsten, et al.
Published: (2025)
Additive Control Variates Dominate Self-Normalisation in Off-Policy Evaluation
by: Jeunen, Olivier, et al.
Published: (2026)
by: Jeunen, Olivier, et al.
Published: (2026)
Safe, Efficient, and Robust Reinforcement Learning for Ranking and Diffusion Models
by: Gupta, Shashank
Published: (2025)
by: Gupta, Shashank
Published: (2025)
Empirical Risk Minimization with $f$-Divergence Regularization
by: Daunas, Francisco, et al.
Published: (2026)
by: Daunas, Francisco, et al.
Published: (2026)
Invariant Risk Minimization Is A Total Variation Model
by: Lai, Zhao-Rong, et al.
Published: (2024)
by: Lai, Zhao-Rong, et al.
Published: (2024)
Keep your distance: learning dispersed embeddings on $\mathbb{S}_m$
by: Tokarchuk, Evgeniia, et al.
Published: (2025)
by: Tokarchuk, Evgeniia, et al.
Published: (2025)
Counterfactual Risk Minimization with IPS-Weighted BPR and Self-Normalized Evaluation in Recommender Systems
by: Raja, Rahul, et al.
Published: (2025)
by: Raja, Rahul, et al.
Published: (2025)
Hierarchical Deep Counterfactual Regret Minimization
by: Chen, Jiayu, et al.
Published: (2023)
by: Chen, Jiayu, et al.
Published: (2023)
Asymmetry of the Relative Entropy in the Regularization of Empirical Risk Minimization
by: Daunas, Francisco, et al.
Published: (2024)
by: Daunas, Francisco, et al.
Published: (2024)
Empirical Risk Minimization with Relative Entropy Regularization
by: Perlaza, Samir M., et al.
Published: (2022)
by: Perlaza, Samir M., et al.
Published: (2022)
A Dual Optimization View to Empirical Risk Minimization with f-Divergence Regularization
by: Daunas, Francisco, et al.
Published: (2025)
by: Daunas, Francisco, et al.
Published: (2025)
Equivalence of the Empirical Risk Minimization to Regularization on the Family of f-Divergences
by: Daunas, Francisco, et al.
Published: (2024)
by: Daunas, Francisco, et al.
Published: (2024)
Meta-reinforcement learning with minimum attention
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
Multi-Turn Jailbreaks Are Simpler Than They Seem
by: Yang, Xiaoxue, et al.
Published: (2025)
by: Yang, Xiaoxue, et al.
Published: (2025)
Out-of-distribution Generalization for Total Variation based Invariant Risk Minimization
by: Wang, Yuanchao, et al.
Published: (2025)
by: Wang, Yuanchao, et al.
Published: (2025)
Offline Imitation Learning with Variational Counterfactual Reasoning
by: He, Bowei, et al.
Published: (2023)
by: He, Bowei, et al.
Published: (2023)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
by: Belaire, Roman, et al.
Published: (2024)
by: Belaire, Roman, et al.
Published: (2024)
Following the Eye-Tracking Evidence: Established Web-Search Assumptions Fail in Carousel Interfaces
by: Kang, Jingwei, et al.
Published: (2026)
by: Kang, Jingwei, et al.
Published: (2026)
Orthogonal Gradient Boosting for Simpler Additive Rule Ensembles
by: Yang, Fan, et al.
Published: (2024)
by: Yang, Fan, et al.
Published: (2024)
GPU-Accelerated Counterfactual Regret Minimization
by: Kim, Juho
Published: (2024)
by: Kim, Juho
Published: (2024)
Adversary-Free Counterfactual Prediction via Information-Regularized Representations
by: Tang, Shiqin, et al.
Published: (2025)
by: Tang, Shiqin, et al.
Published: (2025)
Environment-Conditioned Tail Reweighting for Total Variation Invariant Risk Minimization
by: Wang, Yuanchao, et al.
Published: (2026)
by: Wang, Yuanchao, et al.
Published: (2026)
A Unified Empirical Risk Minimization Framework for Flexible N-Tuples Weak Supervision
by: Huang, Shuying, et al.
Published: (2025)
by: Huang, Shuying, et al.
Published: (2025)
OccamNets: Mitigating Dataset Bias by Favoring Simpler Hypotheses
by: Shrestha, Robik, et al.
Published: (2022)
by: Shrestha, Robik, et al.
Published: (2022)
Causal Dynamic Variational Autoencoder for Counterfactual Regression in Longitudinal Data
by: Bouchattaoui, Mouad El, et al.
Published: (2023)
by: Bouchattaoui, Mouad El, et al.
Published: (2023)
Variational f-divergence Minimization
by: Zhang, Mingtian, et al.
Published: (2019)
by: Zhang, Mingtian, et al.
Published: (2019)
CounterFlowNet: From Minimal Changes to Meaningful Counterfactual Explanations
by: Furman, Oleksii, et al.
Published: (2026)
by: Furman, Oleksii, et al.
Published: (2026)
$L_2$-Regularized Empirical Risk Minimization Guarantees Small Smooth Calibration Error
by: Fujisawa, Masahiro, et al.
Published: (2025)
by: Fujisawa, Masahiro, et al.
Published: (2025)
Adversarial Alignment for LLMs Requires Simpler, Reproducible, and More Measurable Objectives
by: Schwinn, Leo, et al.
Published: (2025)
by: Schwinn, Leo, et al.
Published: (2025)
Revisiting Generative Policies: A Simpler Reinforcement Learning Algorithmic Perspective
by: Zhang, Jinouwen, et al.
Published: (2024)
by: Zhang, Jinouwen, et al.
Published: (2024)
Similar Items
-
Practical and Robust Safety Guarantees for Advanced Counterfactual Learning to Rank
by: Gupta, Shashank, et al.
Published: (2024) -
Proximal Ranking Policy Optimization for Practical Safety in Counterfactual Learning to Rank
by: Gupta, Shashank, et al.
Published: (2024) -
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024) -
Learning to Rank with Variable Result Presentation Lengths
by: Knyazev, Norman, et al.
Published: (2025) -
Local Feature Selection without Label or Feature Leakage for Interpretable Machine Learning Predictions
by: Oosterhuis, Harrie, et al.
Published: (2024)