Contextual Bandits in Payment Processing: Non-uniform Exploration and Supervised Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Vangara, Akhila, Egg, Alex |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Off-policy Evaluation for Payments at Adyen
by: Egg, Alex
Published: (2025)
by: Egg, Alex
Published: (2025)
Online Learning for Recommendations at Grubhub
by: Egg, Alex
Published: (2021)
by: Egg, Alex
Published: (2021)
Calibrated Recommendations with Contextual Bandits
by: Feijer, Diego, et al.
Published: (2025)
by: Feijer, Diego, et al.
Published: (2025)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Causal Feature Selection Method for Contextual Multi-Armed Bandits in Recommender System
by: Zhao, Zhenyu, et al.
Published: (2024)
by: Zhao, Zhenyu, et al.
Published: (2024)
Exploitation Over Exploration: Unmasking the Bias in Linear Bandit Recommender Offline Evaluation
by: Pires, Pedro R., et al.
Published: (2025)
by: Pires, Pedro R., et al.
Published: (2025)
Contextual Bandit with Herding Effects: Algorithms and Recommendation Applications
by: Xu, Luyue, et al.
Published: (2024)
by: Xu, Luyue, et al.
Published: (2024)
AdaptEx: A Self-Service Contextual Bandit Platform
by: Black, William, et al.
Published: (2023)
by: Black, William, et al.
Published: (2023)
Meta Clustering of Neural Bandits
by: Ban, Yikun, et al.
Published: (2024)
by: Ban, Yikun, et al.
Published: (2024)
Modeling Attrition in Recommender Systems with Departing Bandits
by: Ben-Porat, Omer, et al.
Published: (2022)
by: Ben-Porat, Omer, et al.
Published: (2022)
Mixed Supervised Graph Contrastive Learning for Recommendation
by: Zhang, Weizhi, et al.
Published: (2024)
by: Zhang, Weizhi, et al.
Published: (2024)
Individualized non-uniform quantization for vector search
by: Tepper, Mariano, et al.
Published: (2025)
by: Tepper, Mariano, et al.
Published: (2025)
The Bandit's Blind Spot: The Critical Role of User State Representation in Recommender Systems
by: Pires, Pedro R., et al.
Published: (2026)
by: Pires, Pedro R., et al.
Published: (2026)
Low-Rank Online Dynamic Assortment with Dual Contextual Information
by: Lee, Seong Jin, et al.
Published: (2024)
by: Lee, Seong Jin, et al.
Published: (2024)
ActionPiece: Contextually Tokenizing Action Sequences for Generative Recommendation
by: Hou, Yupeng, et al.
Published: (2025)
by: Hou, Yupeng, et al.
Published: (2025)
Contextual Scalarisation Thompson Sampling for multi-objective decisions in public media
by: Maëtz, Théo, et al.
Published: (2026)
by: Maëtz, Théo, et al.
Published: (2026)
SemStruct: Contextualizing Semantic Embeddings with Structural Information for Schema Matching
by: Kang, Inwon, et al.
Published: (2026)
by: Kang, Inwon, et al.
Published: (2026)
Enhancing User Sequence Modeling through Barlow Twins-based Self-Supervised Learning
by: Liu, Yuhan, et al.
Published: (2025)
by: Liu, Yuhan, et al.
Published: (2025)
SLSREC: Self-Supervised Contrastive Learning for Adaptive Fusion of Long- and Short-Term User Interests
by: Zhou, Wei, et al.
Published: (2026)
by: Zhou, Wei, et al.
Published: (2026)
Generative Auto-Bidding with Value-Guided Explorations
by: Gao, Jingtong, et al.
Published: (2025)
by: Gao, Jingtong, et al.
Published: (2025)
ContextWIN: Whittle Index Based Mixture-of-Experts Neural Model For Restless Bandits Via Deep RL
by: Guo, Zhanqiu, et al.
Published: (2024)
by: Guo, Zhanqiu, et al.
Published: (2024)
Ranking In Generalized Linear Bandits
by: Shidani, Amitis, et al.
Published: (2022)
by: Shidani, Amitis, et al.
Published: (2022)
Scaled Supervision is an Implicit Lipschitz Regularizer
by: Ouyang, Zhongyu, et al.
Published: (2025)
by: Ouyang, Zhongyu, et al.
Published: (2025)
Cost-Aware Query Policies in Active Learning for Efficient Autonomous Robotic Exploration
by: Akins, Sapphira, et al.
Published: (2024)
by: Akins, Sapphira, et al.
Published: (2024)
Prompt Optimization with Logged Bandit Data
by: Kiyohara, Haruka, et al.
Published: (2025)
by: Kiyohara, Haruka, et al.
Published: (2025)
The Nah Bandit: Modeling User Non-compliance in Recommendation Systems
by: Zhou, Tianyue, et al.
Published: (2024)
by: Zhou, Tianyue, et al.
Published: (2024)
Error Bounds of Supervised Classification from Information-Theoretic Perspective
by: Qi, Binchuan
Published: (2024)
by: Qi, Binchuan
Published: (2024)
Understanding and Guiding Weakly Supervised Entity Alignment with Potential Isomorphism Propagation
by: Wang, Yuanyi, et al.
Published: (2024)
by: Wang, Yuanyi, et al.
Published: (2024)
Lipschitz Dueling Bandits over Continuous Action Spaces
by: Sharma, Mudit, et al.
Published: (2026)
by: Sharma, Mudit, et al.
Published: (2026)
Leave No One Behind: Online Self-Supervised Self-Distillation for Sequential Recommendation
by: Wei, Shaowei, et al.
Published: (2024)
by: Wei, Shaowei, et al.
Published: (2024)
Contextual Multilingual Spellchecker for User Queries
by: Sharma, Sanat, et al.
Published: (2023)
by: Sharma, Sanat, et al.
Published: (2023)
Diffusion-Pretrained Dense and Contextual Embeddings
by: Eslami, Sedigheh, et al.
Published: (2026)
by: Eslami, Sedigheh, et al.
Published: (2026)
Abacus: Self-Supervised Event Counting-Aligned Distributional Pretraining for Sequential User Modeling
by: Castro, Sullivan, et al.
Published: (2025)
by: Castro, Sullivan, et al.
Published: (2025)
A Human-Centered Approach for Improving Supervised Learning
by: Bansal, Shubhi, et al.
Published: (2024)
by: Bansal, Shubhi, et al.
Published: (2024)
Goal-Conditioned Supervised Learning for Multi-Objective Recommendation
by: Li, Shijun, et al.
Published: (2024)
by: Li, Shijun, et al.
Published: (2024)
OmniSage: Large Scale, Multi-Entity Heterogeneous Graph Representation Learning
by: Badrinath, Anirudhan, et al.
Published: (2025)
by: Badrinath, Anirudhan, et al.
Published: (2025)
Towards Efficient Quantity Retrieval from Text:An Approach via Description Parsing and Weak Supervision
by: Cao, Yixuan, et al.
Published: (2025)
by: Cao, Yixuan, et al.
Published: (2025)
Judgement Citation Retrieval using Contextual Similarity
by: Dasula, Akshat Mohan, et al.
Published: (2024)
by: Dasula, Akshat Mohan, et al.
Published: (2024)
CoCoB: Adaptive Collaborative Combinatorial Bandits for Online Recommendation
by: Yan, Cairong, et al.
Published: (2025)
by: Yan, Cairong, et al.
Published: (2025)
Non-autoregressive Personalized Bundle Generation
by: Yang, Wenchuan, et al.
Published: (2024)
by: Yang, Wenchuan, et al.
Published: (2024)
Similar Items
-
Off-policy Evaluation for Payments at Adyen
by: Egg, Alex
Published: (2025) -
Online Learning for Recommendations at Grubhub
by: Egg, Alex
Published: (2021) -
Calibrated Recommendations with Contextual Bandits
by: Feijer, Diego, et al.
Published: (2025) -
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024) -
Causal Feature Selection Method for Contextual Multi-Armed Bandits in Recommender System
by: Zhao, Zhenyu, et al.
Published: (2024)