Pairwise Elimination with Instance-Dependent Guarantees for Bandits with Cost Subsidy
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Juneja, Ishank, Joe-Wong, Carlee, Yağan, Osman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
von: Juneja, Ishank, et al.
Veröffentlicht: (2026)
von: Juneja, Ishank, et al.
Veröffentlicht: (2026)
FedSPD: A Soft-clustering Approach for Personalized Decentralized Federated Learning
von: Lin, I-Cheng, et al.
Veröffentlicht: (2024)
von: Lin, I-Cheng, et al.
Veröffentlicht: (2024)
Neural Combinatorial Clustered Bandits for Recommendation Systems
von: Atalar, Baran, et al.
Veröffentlicht: (2024)
von: Atalar, Baran, et al.
Veröffentlicht: (2024)
Bandits with Anytime Knapsacks
von: Elumar, Eray Can, et al.
Veröffentlicht: (2025)
von: Elumar, Eray Can, et al.
Veröffentlicht: (2025)
Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks
von: Atalar, Baran, et al.
Veröffentlicht: (2025)
von: Atalar, Baran, et al.
Veröffentlicht: (2025)
Cost-aware LLM-based Online Dataset Annotation
von: Elumar, Eray Can, et al.
Veröffentlicht: (2025)
von: Elumar, Eray Can, et al.
Veröffentlicht: (2025)
MIRA: Memory-Integrated Reinforcement Learning Agent with Limited LLM Guidance
von: Nourzad, Narjes, et al.
Veröffentlicht: (2026)
von: Nourzad, Narjes, et al.
Veröffentlicht: (2026)
Memory-Based Advantage Shaping for LLM-Guided Reinforcement Learning
von: Nourzad, Narjes, et al.
Veröffentlicht: (2026)
von: Nourzad, Narjes, et al.
Veröffentlicht: (2026)
M3Net: A Multi-Metric Mixture of Experts Network Digital Twin with Graph Neural Networks
von: Guda, Blessed, et al.
Veröffentlicht: (2025)
von: Guda, Blessed, et al.
Veröffentlicht: (2025)
Offline Learning for Combinatorial Multi-armed Bandits
von: Liu, Xutong, et al.
Veröffentlicht: (2025)
von: Liu, Xutong, et al.
Veröffentlicht: (2025)
Offline Clustering of Linear Bandits: The Power of Clusters under Limited Data
von: Liu, Jingyuan, et al.
Veröffentlicht: (2025)
von: Liu, Jingyuan, et al.
Veröffentlicht: (2025)
Batch-Size Independent Regret Bounds for Combinatorial Semi-Bandits with Probabilistically Triggered Arms or Independent Arms
von: Liu, Xutong, et al.
Veröffentlicht: (2022)
von: Liu, Xutong, et al.
Veröffentlicht: (2022)
Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2026)
Continuous Semantic Caching for Low-Cost LLM Serving
von: Atalar, Baran, et al.
Veröffentlicht: (2026)
von: Atalar, Baran, et al.
Veröffentlicht: (2026)
FedTLU: Federated Learning with Targeted Layer Updates
von: Park, Jong-Ik, et al.
Veröffentlicht: (2024)
von: Park, Jong-Ik, et al.
Veröffentlicht: (2024)
Federated Learning with Flexible Architectures
von: Park, Jong-Ik, et al.
Veröffentlicht: (2024)
von: Park, Jong-Ik, et al.
Veröffentlicht: (2024)
An LLM-Based Digital Twin for Optimizing Human-in-the Loop Systems
von: Yang, Hanqing, et al.
Veröffentlicht: (2024)
von: Yang, Hanqing, et al.
Veröffentlicht: (2024)
Evaluating Selective Encryption Against Gradient Inversion Attacks
von: Gu, Jiajun, et al.
Veröffentlicht: (2025)
von: Gu, Jiajun, et al.
Veröffentlicht: (2025)
The Five Ws of Multi-Agent Communication: Who Talks to Whom, When, What, and Why -- A Survey from MARL to Emergent Language and LLMs
von: Chen, Jingdi, et al.
Veröffentlicht: (2026)
von: Chen, Jingdi, et al.
Veröffentlicht: (2026)
Semantic Caching for Low-Cost LLM Serving: From Offline Learning to Online Adaptation
von: Liu, Xutong, et al.
Veröffentlicht: (2025)
von: Liu, Xutong, et al.
Veröffentlicht: (2025)
BanditQ: Fair Bandits with Guaranteed Rewards
von: Sinha, Abhishek
Veröffentlicht: (2023)
von: Sinha, Abhishek
Veröffentlicht: (2023)
Tin-Tin: Towards Tiny Learning on Tiny Devices with Integer-based Neural Network Training
von: Hu, Yi, et al.
Veröffentlicht: (2025)
von: Hu, Yi, et al.
Veröffentlicht: (2025)
PubSwap: Public-Data Off-Policy Coordination for Federated RLVR
von: Nayak, Anupam, et al.
Veröffentlicht: (2026)
von: Nayak, Anupam, et al.
Veröffentlicht: (2026)
On Balancing Sparsity with Reliable Connectivity in Distributed Network Design with Random K-out Graphs
von: Sood, Mansi, et al.
Veröffentlicht: (2025)
von: Sood, Mansi, et al.
Veröffentlicht: (2025)
Federated Communication-Efficient Multi-Objective Optimization
von: Askin, Baris, et al.
Veröffentlicht: (2024)
von: Askin, Baris, et al.
Veröffentlicht: (2024)
FedGAT: A Privacy-Preserving Federated Approximation Algorithm for Graph Attention Networks
von: Ambekar, Siddharth, et al.
Veröffentlicht: (2024)
von: Ambekar, Siddharth, et al.
Veröffentlicht: (2024)
Intelligent Communication Planning for Constrained Environmental IoT Sensing with Reinforcement Learning
von: Hu, Yi, et al.
Veröffentlicht: (2023)
von: Hu, Yi, et al.
Veröffentlicht: (2023)
FedAST: Federated Asynchronous Simultaneous Training
von: Askin, Baris, et al.
Veröffentlicht: (2024)
von: Askin, Baris, et al.
Veröffentlicht: (2024)
Continuum-armed Bandit Optimization with Batch Pairwise Comparison Oracles
von: Chang, Xiangyu, et al.
Veröffentlicht: (2025)
von: Chang, Xiangyu, et al.
Veröffentlicht: (2025)
Adversarial Robustness Unhardening via Backdoor Attacks in Federated Learning
von: Kim, Taejin, et al.
Veröffentlicht: (2023)
von: Kim, Taejin, et al.
Veröffentlicht: (2023)
CoRAST: Towards Foundation Model-Powered Correlated Data Analysis in Resource-Constrained CPS and IoT
von: Hu, Yi, et al.
Veröffentlicht: (2024)
von: Hu, Yi, et al.
Veröffentlicht: (2024)
DR. WELL: Dynamic Reasoning and Learning with Symbolic World Model for Embodied LLM-Based Multi-Agent Collaboration
von: Nourzad, Narjes, et al.
Veröffentlicht: (2025)
von: Nourzad, Narjes, et al.
Veröffentlicht: (2025)
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling
von: Guda, Blessed, et al.
Veröffentlicht: (2024)
von: Guda, Blessed, et al.
Veröffentlicht: (2024)
Offline Clustering of Preference Learning with Active-data Augmentation
von: Liu, Jingyuan, et al.
Veröffentlicht: (2025)
von: Liu, Jingyuan, et al.
Veröffentlicht: (2025)
Interference-Aware Edge Runtime Prediction with Conformal Matrix Completion
von: Huang, Tianshu, et al.
Veröffentlicht: (2025)
von: Huang, Tianshu, et al.
Veröffentlicht: (2025)
Optimal Arm Elimination Algorithms for Combinatorial Bandits
von: Wen, Yuxiao, et al.
Veröffentlicht: (2025)
von: Wen, Yuxiao, et al.
Veröffentlicht: (2025)
Fairness and Privacy Guarantees in Federated Contextual Bandits
von: Solanki, Sambhav, et al.
Veröffentlicht: (2024)
von: Solanki, Sambhav, et al.
Veröffentlicht: (2024)
Achieving Logarithmic Regret in KL-Regularized Zero-Sum Markov Games
von: Nayak, Anupam, et al.
Veröffentlicht: (2025)
von: Nayak, Anupam, et al.
Veröffentlicht: (2025)
GLUE: Gradient-free Learning to Unify Experts
von: Park, Jong-Ik, et al.
Veröffentlicht: (2025)
von: Park, Jong-Ik, et al.
Veröffentlicht: (2025)
Reviving Stale Updates: Data-Free Knowledge Distillation for Asynchronous Federated Learning
von: Askin, Baris, et al.
Veröffentlicht: (2025)
von: Askin, Baris, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
von: Juneja, Ishank, et al.
Veröffentlicht: (2026) -
FedSPD: A Soft-clustering Approach for Personalized Decentralized Federated Learning
von: Lin, I-Cheng, et al.
Veröffentlicht: (2024) -
Neural Combinatorial Clustered Bandits for Recommendation Systems
von: Atalar, Baran, et al.
Veröffentlicht: (2024) -
Bandits with Anytime Knapsacks
von: Elumar, Eray Can, et al.
Veröffentlicht: (2025) -
Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks
von: Atalar, Baran, et al.
Veröffentlicht: (2025)