Budgeted Recommendation with Delayed Feedback
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Kweiguu, Maghsudi, Setareh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robust Optimization Approach and Learning Based Hide-and-Seek Game for Resilient Network Design
von: Khosravi, Mohammad, et al.
Veröffentlicht: (2026)
von: Khosravi, Mohammad, et al.
Veröffentlicht: (2026)
Anomaly Detection in Networked Bandits
von: Cheng, Xiaotong, et al.
Veröffentlicht: (2025)
von: Cheng, Xiaotong, et al.
Veröffentlicht: (2025)
Stochastic Multi-Objective Multi-Armed Bandits: Regret Definition and Algorithm
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025)
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025)
Emergence of Fair Leaders via Mediators in Multi-Agent Reinforcement Learning
von: Dodwadmath, Akshay, et al.
Veröffentlicht: (2025)
von: Dodwadmath, Akshay, et al.
Veröffentlicht: (2025)
Pareto Multi-Objective Alignment for Language Models
von: He, Qiang, et al.
Veröffentlicht: (2025)
von: He, Qiang, et al.
Veröffentlicht: (2025)
Meta Learning in Bandits within Shared Affine Subspaces
von: Bilaj, Steven, et al.
Veröffentlicht: (2024)
von: Bilaj, Steven, et al.
Veröffentlicht: (2024)
Distributed Management of Fluctuating Energy Resources in Dynamic Networked Systems
von: Cheng, Xiaotong, et al.
Veröffentlicht: (2024)
von: Cheng, Xiaotong, et al.
Veröffentlicht: (2024)
Decentralized Task Offloading and Load-Balancing for Mobile Edge Computing in Dense Networks
von: Yahya, Mariam, et al.
Veröffentlicht: (2024)
von: Yahya, Mariam, et al.
Veröffentlicht: (2024)
Service Placement in Small Cell Networks Using Distributed Best Arm Identification in Linear Bandits
von: Yahya, Mariam, et al.
Veröffentlicht: (2025)
von: Yahya, Mariam, et al.
Veröffentlicht: (2025)
Adaptive Regularization of Representation Rank as an Implicit Constraint of Bellman Equation
von: He, Qiang, et al.
Veröffentlicht: (2024)
von: He, Qiang, et al.
Veröffentlicht: (2024)
Meta-Learning Multi-armed Bandits for Beam Tracking in 5G and 6G Networks
von: Mattick, Alexander, et al.
Veröffentlicht: (2025)
von: Mattick, Alexander, et al.
Veröffentlicht: (2025)
Unveiling the Decision-Making Process in Reinforcement Learning with Genetic Programming
von: Eberhardinger, Manuel, et al.
Veröffentlicht: (2024)
von: Eberhardinger, Manuel, et al.
Veröffentlicht: (2024)
Quantum-Inspired Reinforcement Learning in the Presence of Epistemic Ambivalence
von: Habibi, Alireza, et al.
Veröffentlicht: (2025)
von: Habibi, Alireza, et al.
Veröffentlicht: (2025)
One Model for All: Multi-Objective Controllable Language Models
von: He, Qiang, et al.
Veröffentlicht: (2026)
von: He, Qiang, et al.
Veröffentlicht: (2026)
Lipschitz Bandits with Stochastic Delayed Feedback
von: Liu, Zhongxuan, et al.
Veröffentlicht: (2025)
von: Liu, Zhongxuan, et al.
Veröffentlicht: (2025)
Online Influence Maximization with Semi-Bandit Feedback under Corruptions
von: Cheng, Xiaotong, et al.
Veröffentlicht: (2024)
von: Cheng, Xiaotong, et al.
Veröffentlicht: (2024)
Safe and Efficient Online Convex Optimization with Linear Budget Constraints and Partial Feedback
von: Liu, Shanqi, et al.
Veröffentlicht: (2024)
von: Liu, Shanqi, et al.
Veröffentlicht: (2024)
Feedback Control for Small Budget Pacing
von: Apparaju, Sreeja, et al.
Veröffentlicht: (2025)
von: Apparaju, Sreeja, et al.
Veröffentlicht: (2025)
Biased Dueling Bandits with Stochastic Delayed Feedback
von: Yi, Bongsoo, et al.
Veröffentlicht: (2024)
von: Yi, Bongsoo, et al.
Veröffentlicht: (2024)
A Robust Optimization Approach for Regenerator Placement in Fault-Tolerant Networks Under Discrete Cost Uncertainty
von: Khosravi, Mohammad, et al.
Veröffentlicht: (2026)
von: Khosravi, Mohammad, et al.
Veröffentlicht: (2026)
Efficient Resource Allocation under Adversary Attacks: A Decomposition-Based Approach
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025)
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025)
Bandit and Delayed Feedback in Online Structured Prediction
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2025)
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2025)
A Best-of-both-worlds Algorithm for Bandits with Delayed Feedback with Robustness to Excessive Delays
von: Masoudian, Saeed, et al.
Veröffentlicht: (2023)
von: Masoudian, Saeed, et al.
Veröffentlicht: (2023)
A Reduction-based Framework for Sequential Decision Making with Delayed Feedback
von: Yang, Yunchang, et al.
Veröffentlicht: (2023)
von: Yang, Yunchang, et al.
Veröffentlicht: (2023)
Improved Regret for Bandit Convex Optimization with Delayed Feedback
von: Wan, Yuanyu, et al.
Veröffentlicht: (2024)
von: Wan, Yuanyu, et al.
Veröffentlicht: (2024)
Exploiting Curvature in Online Convex Optimization with Delayed Feedback
von: Qiu, Hao, et al.
Veröffentlicht: (2025)
von: Qiu, Hao, et al.
Veröffentlicht: (2025)
Online Nonsubmodular Optimization with Delayed Feedback in the Bandit Setting
von: Yang, Sifan, et al.
Veröffentlicht: (2025)
von: Yang, Sifan, et al.
Veröffentlicht: (2025)
Neural Contextual Bandits Under Delayed Feedback Constraints
von: Moghimi, Mohammadali, et al.
Veröffentlicht: (2025)
von: Moghimi, Mohammadali, et al.
Veröffentlicht: (2025)
Differentiable Attenuation Filters for Feedback Delay Networks
von: Ibnyahya, Ilias, et al.
Veröffentlicht: (2025)
von: Ibnyahya, Ilias, et al.
Veröffentlicht: (2025)
Improved Best-of-Both-Worlds Regret for Bandits with Delayed Feedback
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
Adversarial Bandits with Multi-User Delayed Feedback: Theory and Application
von: Li, Yandi, et al.
Veröffentlicht: (2023)
von: Li, Yandi, et al.
Veröffentlicht: (2023)
Modeling Attention during Dimensional Shifts with Counterfactual and Delayed Feedback
von: Malloy, Tailia, et al.
Veröffentlicht: (2025)
von: Malloy, Tailia, et al.
Veröffentlicht: (2025)
Linear and Neural Dueling Bandits with Delayed Feedback
von: Wang, Xiangyi, et al.
Veröffentlicht: (2026)
von: Wang, Xiangyi, et al.
Veröffentlicht: (2026)
Debiased Recommendation with Noisy Feedback
von: Li, Haoxuan, et al.
Veröffentlicht: (2024)
von: Li, Haoxuan, et al.
Veröffentlicht: (2024)
Online Budget Allocation with Censored Semi-Bandit Feedback
von: Bachoc, François, et al.
Veröffentlicht: (2025)
von: Bachoc, François, et al.
Veröffentlicht: (2025)
Merit-based Fair Combinatorial Semi-Bandit with Unrestricted Feedback Delays
von: Chen, Ziqun, et al.
Veröffentlicht: (2024)
von: Chen, Ziqun, et al.
Veröffentlicht: (2024)
Choice-Model-Assisted Q-learning for Delayed-Feedback Revenue Management
von: Shen, Owen, et al.
Veröffentlicht: (2026)
von: Shen, Owen, et al.
Veröffentlicht: (2026)
Delayed Feedback Modeling with Influence Functions
von: Ding, Chenlu, et al.
Veröffentlicht: (2025)
von: Ding, Chenlu, et al.
Veröffentlicht: (2025)
End-to-End Cost-Effective Incentive Recommendation under Budget Constraint with Uplift Modeling
von: Sun, Zexu, et al.
Veröffentlicht: (2024)
von: Sun, Zexu, et al.
Veröffentlicht: (2024)
Decentralized Online Convex Optimization with Unknown Feedback Delays
von: Qiu, Hao, et al.
Veröffentlicht: (2026)
von: Qiu, Hao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Robust Optimization Approach and Learning Based Hide-and-Seek Game for Resilient Network Design
von: Khosravi, Mohammad, et al.
Veröffentlicht: (2026) -
Anomaly Detection in Networked Bandits
von: Cheng, Xiaotong, et al.
Veröffentlicht: (2025) -
Stochastic Multi-Objective Multi-Armed Bandits: Regret Definition and Algorithm
von: Davoodi, Mansoor, et al.
Veröffentlicht: (2025) -
Emergence of Fair Leaders via Mediators in Multi-Agent Reinforcement Learning
von: Dodwadmath, Akshay, et al.
Veröffentlicht: (2025) -
Pareto Multi-Objective Alignment for Language Models
von: He, Qiang, et al.
Veröffentlicht: (2025)