Gespeichert in:
| Hauptverfasser: | Shen, Owen, Jaillet, Patrick |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.02283 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Single-Sample Polylogarithmic Regret Bound for Nonstationary Online Linear Programming
von: Xu, Haoran, et al.
Veröffentlicht: (2026)
von: Xu, Haoran, et al.
Veröffentlicht: (2026)
Distribution-Dependent Rates for Multi-Distribution Learning
von: Hanashiro, Rafael, et al.
Veröffentlicht: (2023)
von: Hanashiro, Rafael, et al.
Veröffentlicht: (2023)
Multi-Timescale Primal Dual Hybrid Gradient with Application to Distributed Optimization
von: Zhang, Junhui, et al.
Veröffentlicht: (2025)
von: Zhang, Junhui, et al.
Veröffentlicht: (2025)
Grace Period is All You Need: Individual Fairness without Revenue Loss in Revenue Management
von: Jaillet, Patrick, et al.
Veröffentlicht: (2024)
von: Jaillet, Patrick, et al.
Veröffentlicht: (2024)
Is Multi-Distribution Learning as Easy as PAC Learning: Sharp Rates with Bounded Label Noise
von: Hanashiro, Rafael, et al.
Veröffentlicht: (2026)
von: Hanashiro, Rafael, et al.
Veröffentlicht: (2026)
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
von: Verma, Arun, et al.
Veröffentlicht: (2024)
von: Verma, Arun, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management
von: Meng, Huiling, et al.
Veröffentlicht: (2024)
von: Meng, Huiling, et al.
Veröffentlicht: (2024)
Prompt Optimization with Human Feedback
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2024)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2024)
Incentive-Aware Dynamic Resource Allocation under Long-Term Cost Constraints
von: Dai, Yan, et al.
Veröffentlicht: (2025)
von: Dai, Yan, et al.
Veröffentlicht: (2025)
Online Resource Allocation with Convex-set Machine-Learned Advice
von: Golrezaei, Negin, et al.
Veröffentlicht: (2023)
von: Golrezaei, Negin, et al.
Veröffentlicht: (2023)
Dynamic Retail Pricing via Q-Learning -- A Reinforcement Learning Framework for Enhanced Revenue Management
von: Apte, Mohit, et al.
Veröffentlicht: (2024)
von: Apte, Mohit, et al.
Veröffentlicht: (2024)
Learning with Exact Invariances in Polynomial Time
von: Soleymani, Ashkan, et al.
Veröffentlicht: (2025)
von: Soleymani, Ashkan, et al.
Veröffentlicht: (2025)
A Universal Class of Sharpness-Aware Minimization Algorithms
von: Tahmasebi, Behrooz, et al.
Veröffentlicht: (2024)
von: Tahmasebi, Behrooz, et al.
Veröffentlicht: (2024)
Double Machine Learning Based Structure Identification from Temporal Data
von: Angelis, Emmanouil, et al.
Veröffentlicht: (2023)
von: Angelis, Emmanouil, et al.
Veröffentlicht: (2023)
Data-Driven Revenue Management for Air Cargo
von: Eren, Ezgi, et al.
Veröffentlicht: (2024)
von: Eren, Ezgi, et al.
Veröffentlicht: (2024)
Budgeted Recommendation with Delayed Feedback
von: Liu, Kweiguu, et al.
Veröffentlicht: (2024)
von: Liu, Kweiguu, et al.
Veröffentlicht: (2024)
Demand Balancing in Primal-Dual Optimization for Blind Network Revenue Management
von: Miao, Sentao, et al.
Veröffentlicht: (2024)
von: Miao, Sentao, et al.
Veröffentlicht: (2024)
Learning with Posterior Sampling for Revenue Management under Time-varying Demand
von: Shimizu, Kazuma, et al.
Veröffentlicht: (2024)
von: Shimizu, Kazuma, et al.
Veröffentlicht: (2024)
DelayPTC-LLM: Metro Passenger Travel Choice Prediction under Train Delays with Large Language Models
von: Chen, Chen, et al.
Veröffentlicht: (2024)
von: Chen, Chen, et al.
Veröffentlicht: (2024)
Delayed Feedback Modeling with Influence Functions
von: Ding, Chenlu, et al.
Veröffentlicht: (2025)
von: Ding, Chenlu, et al.
Veröffentlicht: (2025)
Degeneracy is OK: Logarithmic Regret for Network Revenue Management with Indiscrete Distributions
von: Jiang, Jiashuo, et al.
Veröffentlicht: (2022)
von: Jiang, Jiashuo, et al.
Veröffentlicht: (2022)
Lipschitz Bandits with Stochastic Delayed Feedback
von: Liu, Zhongxuan, et al.
Veröffentlicht: (2025)
von: Liu, Zhongxuan, et al.
Veröffentlicht: (2025)
Modeling Attention during Dimensional Shifts with Counterfactual and Delayed Feedback
von: Malloy, Tailia, et al.
Veröffentlicht: (2025)
von: Malloy, Tailia, et al.
Veröffentlicht: (2025)
Movie Revenue Prediction using Machine Learning Models
von: Udandarao, Vikranth, et al.
Veröffentlicht: (2024)
von: Udandarao, Vikranth, et al.
Veröffentlicht: (2024)
Online Scheduling for LLM Inference with KV Cache Constraints
von: Jaillet, Patrick, et al.
Veröffentlicht: (2025)
von: Jaillet, Patrick, et al.
Veröffentlicht: (2025)
Bandit and Delayed Feedback in Online Structured Prediction
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2025)
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2025)
Biased Dueling Bandits with Stochastic Delayed Feedback
von: Yi, Bongsoo, et al.
Veröffentlicht: (2024)
von: Yi, Bongsoo, et al.
Veröffentlicht: (2024)
A Best-of-both-worlds Algorithm for Bandits with Delayed Feedback with Robustness to Excessive Delays
von: Masoudian, Saeed, et al.
Veröffentlicht: (2023)
von: Masoudian, Saeed, et al.
Veröffentlicht: (2023)
Rankability-enhanced Revenue Uplift Modeling Framework for Online Marketing
von: He, Bowei, et al.
Veröffentlicht: (2024)
von: He, Bowei, et al.
Veröffentlicht: (2024)
Revenue Maximization and Learning in Products Ranking
von: Chen, Ningyuan, et al.
Veröffentlicht: (2020)
von: Chen, Ningyuan, et al.
Veröffentlicht: (2020)
Differentiable Attenuation Filters for Feedback Delay Networks
von: Ibnyahya, Ilias, et al.
Veröffentlicht: (2025)
von: Ibnyahya, Ilias, et al.
Veröffentlicht: (2025)
Exploiting Curvature in Online Convex Optimization with Delayed Feedback
von: Qiu, Hao, et al.
Veröffentlicht: (2025)
von: Qiu, Hao, et al.
Veröffentlicht: (2025)
Improved Regret for Bandit Convex Optimization with Delayed Feedback
von: Wan, Yuanyu, et al.
Veröffentlicht: (2024)
von: Wan, Yuanyu, et al.
Veröffentlicht: (2024)
Online Nonsubmodular Optimization with Delayed Feedback in the Bandit Setting
von: Yang, Sifan, et al.
Veröffentlicht: (2025)
von: Yang, Sifan, et al.
Veröffentlicht: (2025)
Neural Contextual Bandits Under Delayed Feedback Constraints
von: Moghimi, Mohammadali, et al.
Veröffentlicht: (2025)
von: Moghimi, Mohammadali, et al.
Veröffentlicht: (2025)
Incentives in Private Collaborative Machine Learning
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2024)
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2024)
Improved Q-learning based Multi-hop Routing for UAV-Assisted Communication
von: Sharvari, N P, et al.
Veröffentlicht: (2024)
von: Sharvari, N P, et al.
Veröffentlicht: (2024)
Regularized Q-learning
von: Lim, Han-Dong, et al.
Veröffentlicht: (2022)
von: Lim, Han-Dong, et al.
Veröffentlicht: (2022)
Linear and Neural Dueling Bandits with Delayed Feedback
von: Wang, Xiangyi, et al.
Veröffentlicht: (2026)
von: Wang, Xiangyi, et al.
Veröffentlicht: (2026)
Improved Best-of-Both-Worlds Regret for Bandits with Delayed Feedback
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Single-Sample Polylogarithmic Regret Bound for Nonstationary Online Linear Programming
von: Xu, Haoran, et al.
Veröffentlicht: (2026) -
Distribution-Dependent Rates for Multi-Distribution Learning
von: Hanashiro, Rafael, et al.
Veröffentlicht: (2023) -
Multi-Timescale Primal Dual Hybrid Gradient with Application to Distributed Optimization
von: Zhang, Junhui, et al.
Veröffentlicht: (2025) -
Grace Period is All You Need: Individual Fairness without Revenue Loss in Revenue Management
von: Jaillet, Patrick, et al.
Veröffentlicht: (2024) -
Is Multi-Distribution Learning as Easy as PAC Learning: Sharp Rates with Bounded Label Noise
von: Hanashiro, Rafael, et al.
Veröffentlicht: (2026)