Pairwise Elimination with Instance-Dependent Guarantees for Bandits with Cost Subsidy
Fuente:
arXiv
Saved in:
| Main Authors: | Juneja, Ishank, Joe-Wong, Carlee, Yağan, Osman |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
by: Juneja, Ishank, et al.
Published: (2026)
by: Juneja, Ishank, et al.
Published: (2026)
FedSPD: A Soft-clustering Approach for Personalized Decentralized Federated Learning
by: Lin, I-Cheng, et al.
Published: (2024)
by: Lin, I-Cheng, et al.
Published: (2024)
Neural Combinatorial Clustered Bandits for Recommendation Systems
by: Atalar, Baran, et al.
Published: (2024)
by: Atalar, Baran, et al.
Published: (2024)
Bandits with Anytime Knapsacks
by: Elumar, Eray Can, et al.
Published: (2025)
by: Elumar, Eray Can, et al.
Published: (2025)
Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks
by: Atalar, Baran, et al.
Published: (2025)
by: Atalar, Baran, et al.
Published: (2025)
Cost-aware LLM-based Online Dataset Annotation
by: Elumar, Eray Can, et al.
Published: (2025)
by: Elumar, Eray Can, et al.
Published: (2025)
MIRA: Memory-Integrated Reinforcement Learning Agent with Limited LLM Guidance
by: Nourzad, Narjes, et al.
Published: (2026)
by: Nourzad, Narjes, et al.
Published: (2026)
Memory-Based Advantage Shaping for LLM-Guided Reinforcement Learning
by: Nourzad, Narjes, et al.
Published: (2026)
by: Nourzad, Narjes, et al.
Published: (2026)
M3Net: A Multi-Metric Mixture of Experts Network Digital Twin with Graph Neural Networks
by: Guda, Blessed, et al.
Published: (2025)
by: Guda, Blessed, et al.
Published: (2025)
Offline Learning for Combinatorial Multi-armed Bandits
by: Liu, Xutong, et al.
Published: (2025)
by: Liu, Xutong, et al.
Published: (2025)
Offline Clustering of Linear Bandits: The Power of Clusters under Limited Data
by: Liu, Jingyuan, et al.
Published: (2025)
by: Liu, Jingyuan, et al.
Published: (2025)
Batch-Size Independent Regret Bounds for Combinatorial Semi-Bandits with Probabilistically Triggered Arms or Independent Arms
by: Liu, Xutong, et al.
Published: (2022)
by: Liu, Xutong, et al.
Published: (2022)
Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Continuous Semantic Caching for Low-Cost LLM Serving
by: Atalar, Baran, et al.
Published: (2026)
by: Atalar, Baran, et al.
Published: (2026)
FedTLU: Federated Learning with Targeted Layer Updates
by: Park, Jong-Ik, et al.
Published: (2024)
by: Park, Jong-Ik, et al.
Published: (2024)
Federated Learning with Flexible Architectures
by: Park, Jong-Ik, et al.
Published: (2024)
by: Park, Jong-Ik, et al.
Published: (2024)
An LLM-Based Digital Twin for Optimizing Human-in-the Loop Systems
by: Yang, Hanqing, et al.
Published: (2024)
by: Yang, Hanqing, et al.
Published: (2024)
Evaluating Selective Encryption Against Gradient Inversion Attacks
by: Gu, Jiajun, et al.
Published: (2025)
by: Gu, Jiajun, et al.
Published: (2025)
The Five Ws of Multi-Agent Communication: Who Talks to Whom, When, What, and Why -- A Survey from MARL to Emergent Language and LLMs
by: Chen, Jingdi, et al.
Published: (2026)
by: Chen, Jingdi, et al.
Published: (2026)
Semantic Caching for Low-Cost LLM Serving: From Offline Learning to Online Adaptation
by: Liu, Xutong, et al.
Published: (2025)
by: Liu, Xutong, et al.
Published: (2025)
BanditQ: Fair Bandits with Guaranteed Rewards
by: Sinha, Abhishek
Published: (2023)
by: Sinha, Abhishek
Published: (2023)
Tin-Tin: Towards Tiny Learning on Tiny Devices with Integer-based Neural Network Training
by: Hu, Yi, et al.
Published: (2025)
by: Hu, Yi, et al.
Published: (2025)
PubSwap: Public-Data Off-Policy Coordination for Federated RLVR
by: Nayak, Anupam, et al.
Published: (2026)
by: Nayak, Anupam, et al.
Published: (2026)
On Balancing Sparsity with Reliable Connectivity in Distributed Network Design with Random K-out Graphs
by: Sood, Mansi, et al.
Published: (2025)
by: Sood, Mansi, et al.
Published: (2025)
Federated Communication-Efficient Multi-Objective Optimization
by: Askin, Baris, et al.
Published: (2024)
by: Askin, Baris, et al.
Published: (2024)
FedGAT: A Privacy-Preserving Federated Approximation Algorithm for Graph Attention Networks
by: Ambekar, Siddharth, et al.
Published: (2024)
by: Ambekar, Siddharth, et al.
Published: (2024)
Intelligent Communication Planning for Constrained Environmental IoT Sensing with Reinforcement Learning
by: Hu, Yi, et al.
Published: (2023)
by: Hu, Yi, et al.
Published: (2023)
FedAST: Federated Asynchronous Simultaneous Training
by: Askin, Baris, et al.
Published: (2024)
by: Askin, Baris, et al.
Published: (2024)
Continuum-armed Bandit Optimization with Batch Pairwise Comparison Oracles
by: Chang, Xiangyu, et al.
Published: (2025)
by: Chang, Xiangyu, et al.
Published: (2025)
Adversarial Robustness Unhardening via Backdoor Attacks in Federated Learning
by: Kim, Taejin, et al.
Published: (2023)
by: Kim, Taejin, et al.
Published: (2023)
CoRAST: Towards Foundation Model-Powered Correlated Data Analysis in Resource-Constrained CPS and IoT
by: Hu, Yi, et al.
Published: (2024)
by: Hu, Yi, et al.
Published: (2024)
DR. WELL: Dynamic Reasoning and Learning with Symbolic World Model for Embodied LLM-Based Multi-Agent Collaboration
by: Nourzad, Narjes, et al.
Published: (2025)
by: Nourzad, Narjes, et al.
Published: (2025)
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling
by: Guda, Blessed, et al.
Published: (2024)
by: Guda, Blessed, et al.
Published: (2024)
Offline Clustering of Preference Learning with Active-data Augmentation
by: Liu, Jingyuan, et al.
Published: (2025)
by: Liu, Jingyuan, et al.
Published: (2025)
Interference-Aware Edge Runtime Prediction with Conformal Matrix Completion
by: Huang, Tianshu, et al.
Published: (2025)
by: Huang, Tianshu, et al.
Published: (2025)
Optimal Arm Elimination Algorithms for Combinatorial Bandits
by: Wen, Yuxiao, et al.
Published: (2025)
by: Wen, Yuxiao, et al.
Published: (2025)
Fairness and Privacy Guarantees in Federated Contextual Bandits
by: Solanki, Sambhav, et al.
Published: (2024)
by: Solanki, Sambhav, et al.
Published: (2024)
Achieving Logarithmic Regret in KL-Regularized Zero-Sum Markov Games
by: Nayak, Anupam, et al.
Published: (2025)
by: Nayak, Anupam, et al.
Published: (2025)
GLUE: Gradient-free Learning to Unify Experts
by: Park, Jong-Ik, et al.
Published: (2025)
by: Park, Jong-Ik, et al.
Published: (2025)
Reviving Stale Updates: Data-Free Knowledge Distillation for Asynchronous Federated Learning
by: Askin, Baris, et al.
Published: (2025)
by: Askin, Baris, et al.
Published: (2025)
Similar Items
-
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
by: Juneja, Ishank, et al.
Published: (2026) -
FedSPD: A Soft-clustering Approach for Personalized Decentralized Federated Learning
by: Lin, I-Cheng, et al.
Published: (2024) -
Neural Combinatorial Clustered Bandits for Recommendation Systems
by: Atalar, Baran, et al.
Published: (2024) -
Bandits with Anytime Knapsacks
by: Elumar, Eray Can, et al.
Published: (2025) -
Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks
by: Atalar, Baran, et al.
Published: (2025)