Doubly-Bounded Queue for Constrained Online Learning: Keeping Pace with Dynamics of Both Loss and Constraint
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Juncheng, Yan, Bingjie, Liu, Yituo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Constrained Over-the-Air Model Updating for Wireless Online Federated Learning with Delayed Information
por: Wang, Juncheng, et al.
Publicado: (2025)
por: Wang, Juncheng, et al.
Publicado: (2025)
Universal Dynamic Regret and Constraint Violation Bounds for Constrained Online Convex Optimization
por: Supantha, Subhamon, et al.
Publicado: (2025)
por: Supantha, Subhamon, et al.
Publicado: (2025)
Queue Length Regret Bounds for Contextual Queueing Bandits
por: Bae, Seoungbin, et al.
Publicado: (2026)
por: Bae, Seoungbin, et al.
Publicado: (2026)
A Best-of-Both-Worlds Algorithm for Constrained MDPs with Long-Term Constraints
por: Germano, Jacopo, et al.
Publicado: (2023)
por: Germano, Jacopo, et al.
Publicado: (2023)
Structure-Dependent Regret and Constraint Violation Bounds for Online Convex Optimization with Time-Varying Constraints
por: Liu, Xiufeng, et al.
Publicado: (2026)
por: Liu, Xiufeng, et al.
Publicado: (2026)
Hierarchical Upper Confidence Bounds for Constrained Online Learning
por: Baheri, Ali
Publicado: (2024)
por: Baheri, Ali
Publicado: (2024)
Online Learning and Optimization for Queues with Unknown Demand Curve and Service Distribution
por: Chen, Xinyun, et al.
Publicado: (2023)
por: Chen, Xinyun, et al.
Publicado: (2023)
Autobidders with Budget and ROI Constraints: Efficiency, Regret, and Pacing Dynamics
por: Lucier, Brendan, et al.
Publicado: (2023)
por: Lucier, Brendan, et al.
Publicado: (2023)
Dynamic Regret Bounds for Online Omniprediction with Long Term Constraints
por: Bechavod, Yahav, et al.
Publicado: (2025)
por: Bechavod, Yahav, et al.
Publicado: (2025)
Small Loss Bounds for Online Learning Separated Function Classes: A Gaussian Process Perspective
por: Block, Adam, et al.
Publicado: (2025)
por: Block, Adam, et al.
Publicado: (2025)
Tight Bounds for Online Convex Optimization with Adversarial Constraints
por: Sinha, Abhishek, et al.
Publicado: (2024)
por: Sinha, Abhishek, et al.
Publicado: (2024)
Posterior Probability Matters: Doubly-Adaptive Calibration for Neural Predictions in Online Advertising
por: Wei, Penghui, et al.
Publicado: (2022)
por: Wei, Penghui, et al.
Publicado: (2022)
From Static Constraints to Dynamic Adaptation: Sample-Level Constraint Relaxation for Offline-to-Online Reinforcement Learning
por: Zu, Lipeng, et al.
Publicado: (2025)
por: Zu, Lipeng, et al.
Publicado: (2025)
Doubly Outlier-Robust Online Infinite Hidden Markov Model
por: Yiu, Horace, et al.
Publicado: (2026)
por: Yiu, Horace, et al.
Publicado: (2026)
Doubly Optimal Policy Evaluation for Reinforcement Learning
por: Liu, Shuze Daniel, et al.
Publicado: (2024)
por: Liu, Shuze Daniel, et al.
Publicado: (2024)
Finite-Time Minimax Bounds and an Optimal Lyapunov Policy in Queueing Control
por: Liu, Yujie, et al.
Publicado: (2025)
por: Liu, Yujie, et al.
Publicado: (2025)
Doubly Robust Interval Estimation for Optimal Policy Evaluation in Online Learning
por: Shen, Ye, et al.
Publicado: (2021)
por: Shen, Ye, et al.
Publicado: (2021)
Doubly Adaptive Social Learning
por: Carpentiero, Marco, et al.
Publicado: (2025)
por: Carpentiero, Marco, et al.
Publicado: (2025)
Variation-Bounded Loss for Noise-Tolerant Learning
por: Wang, Jialiang, et al.
Publicado: (2025)
por: Wang, Jialiang, et al.
Publicado: (2025)
Non-stationary Online Learning for Curved Losses: Improved Dynamic Regret via Mixability
por: Zhang, Yu-Jie, et al.
Publicado: (2025)
por: Zhang, Yu-Jie, et al.
Publicado: (2025)
Optimal Bounds for Adversarial Constrained Online Convex Optimization
por: Ferreira, Ricardo N., et al.
Publicado: (2025)
por: Ferreira, Ricardo N., et al.
Publicado: (2025)
Revisiting Projection-Free Online Learning with Time-Varying Constraints
por: Wang, Yibo, et al.
Publicado: (2025)
por: Wang, Yibo, et al.
Publicado: (2025)
Model-Free Robust $ϕ$-Divergence Reinforcement Learning Using Both Offline and Online Data
por: Panaganti, Kishan, et al.
Publicado: (2024)
por: Panaganti, Kishan, et al.
Publicado: (2024)
Efficient Reinforcement Learning for Routing Jobs in Heterogeneous Queueing Systems
por: Jali, Neharika, et al.
Publicado: (2024)
por: Jali, Neharika, et al.
Publicado: (2024)
Online Learning in MDPs with Partially Adversarial Transitions and Losses
por: Schlisselberg, Ofir, et al.
Publicado: (2026)
por: Schlisselberg, Ofir, et al.
Publicado: (2026)
Queue-based Eco-Driving at Roundabouts with Reinforcement Learning
por: Schlamp, Anna-Lena, et al.
Publicado: (2024)
por: Schlamp, Anna-Lena, et al.
Publicado: (2024)
Distributionally-Constrained Adversaries in Online Learning
por: Blanchard, Moïse, et al.
Publicado: (2025)
por: Blanchard, Moïse, et al.
Publicado: (2025)
Causal-Paced Deep Reinforcement Learning
por: Cho, Geonwoo, et al.
Publicado: (2025)
por: Cho, Geonwoo, et al.
Publicado: (2025)
Distributionally Robust Self Paced Curriculum Reinforcement Learning
por: Satheesh, Anirudh, et al.
Publicado: (2025)
por: Satheesh, Anirudh, et al.
Publicado: (2025)
Doubly Inhomogeneous Reinforcement Learning
por: Hu, Liyuan, et al.
Publicado: (2022)
por: Hu, Liyuan, et al.
Publicado: (2022)
Doubly Robust Proximal Causal Learning for Continuous Treatments
por: Wu, Yong, et al.
Publicado: (2023)
por: Wu, Yong, et al.
Publicado: (2023)
Angular Constraint Embedding via SpherePair Loss for Constrained Clustering
por: Zhang, Shaojie, et al.
Publicado: (2025)
por: Zhang, Shaojie, et al.
Publicado: (2025)
FedQueue: Queue-Aware Federated Learning for Cross-Facility HPC Training
por: Li, Yijiang, et al.
Publicado: (2026)
por: Li, Yijiang, et al.
Publicado: (2026)
Forming Auxiliary High-confident Instance-level Loss to Promote Learning from Label Proportions
por: Ma, Tianhao, et al.
Publicado: (2024)
por: Ma, Tianhao, et al.
Publicado: (2024)
Gradient-Variation Regret Bounds for Unconstrained Online Learning
por: Zhao, Yuheng, et al.
Publicado: (2026)
por: Zhao, Yuheng, et al.
Publicado: (2026)
Doubly Mild Generalization for Offline Reinforcement Learning
por: Mao, Yixiu, et al.
Publicado: (2024)
por: Mao, Yixiu, et al.
Publicado: (2024)
Optimal Mistake Bounds for Transductive Online Learning
por: Chase, Zachary, et al.
Publicado: (2025)
por: Chase, Zachary, et al.
Publicado: (2025)
Constrained Optimal Fuel Consumption of HEV: A Constrained Reinforcement Learning Approach
por: Yan, Shuchang
Publicado: (2024)
por: Yan, Shuchang
Publicado: (2024)
Few Batches or Little Memory, But Not Both: Simultaneous Space and Adaptivity Constraints in Stochastic Bandits
por: Huang, Ruiyuan, et al.
Publicado: (2026)
por: Huang, Ruiyuan, et al.
Publicado: (2026)
Learning Constrained Markov Decision Processes With Non-stationary Rewards and Constraints
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
Ejemplares similares
-
Constrained Over-the-Air Model Updating for Wireless Online Federated Learning with Delayed Information
por: Wang, Juncheng, et al.
Publicado: (2025) -
Universal Dynamic Regret and Constraint Violation Bounds for Constrained Online Convex Optimization
por: Supantha, Subhamon, et al.
Publicado: (2025) -
Queue Length Regret Bounds for Contextual Queueing Bandits
por: Bae, Seoungbin, et al.
Publicado: (2026) -
A Best-of-Both-Worlds Algorithm for Constrained MDPs with Long-Term Constraints
por: Germano, Jacopo, et al.
Publicado: (2023) -
Structure-Dependent Regret and Constraint Violation Bounds for Online Convex Optimization with Time-Varying Constraints
por: Liu, Xiufeng, et al.
Publicado: (2026)