Online Prompt Pricing based on Combinatorial Multi-Armed Bandit and Hierarchical Stackelberg Game
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Meiling, Ren, Hongrun, Xiong, Haixu, Qian, Zhenxing, Zhang, Xinpeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Federated Combinatorial Multi-Agent Multi-Armed Bandits
von: Fourati, Fares, et al.
Veröffentlicht: (2024)
von: Fourati, Fares, et al.
Veröffentlicht: (2024)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Multi-Armed Bandits With Best-Action Queries
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2026)
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2026)
Introduction to Multi-Armed Bandits
von: Slivkins, Aleksandrs
Veröffentlicht: (2019)
von: Slivkins, Aleksandrs
Veröffentlicht: (2019)
Fair Algorithms with Probing for Multi-Agent Multi-Armed Bandits
von: Xu, Tianyi, et al.
Veröffentlicht: (2025)
von: Xu, Tianyi, et al.
Veröffentlicht: (2025)
Large Language Model-Enhanced Multi-Armed Bandits
von: Sun, Jiahang, et al.
Veröffentlicht: (2025)
von: Sun, Jiahang, et al.
Veröffentlicht: (2025)
Multi-Armed Bandits-Based Optimization of Decision Trees
von: Shanto, Hasibul Karim, et al.
Veröffentlicht: (2025)
von: Shanto, Hasibul Karim, et al.
Veröffentlicht: (2025)
Bandits with Preference Feedback: A Stackelberg Game Perspective
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
ParBalans: Parallel Multi-Armed Bandits-based Adaptive Large Neighborhood Search
von: Yilmaz, Alican, et al.
Veröffentlicht: (2025)
von: Yilmaz, Alican, et al.
Veröffentlicht: (2025)
Global Rewards in Restless Multi-Armed Bandits
von: Raman, Naveen, et al.
Veröffentlicht: (2024)
von: Raman, Naveen, et al.
Veröffentlicht: (2024)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
Thresholding Data Shapley for Data Cleansing Using Multi-Armed Bandits
von: Namba, Hiroyuki, et al.
Veröffentlicht: (2024)
von: Namba, Hiroyuki, et al.
Veröffentlicht: (2024)
KernelBand: Steering LLM-based Kernel Optimization via Hardware-Aware Multi-Armed Bandits
von: Ran, Dezhi, et al.
Veröffentlicht: (2025)
von: Ran, Dezhi, et al.
Veröffentlicht: (2025)
Mobility-Aware Federated Learning: Multi-Armed Bandit Based Selection in Vehicular Network
von: Tu, Haoyu, et al.
Veröffentlicht: (2024)
von: Tu, Haoyu, et al.
Veröffentlicht: (2024)
Regeneration Based Training-free Attribution of Fake Images Generated by Text-to-Image Generative Models
von: Li, Meiling, et al.
Veröffentlicht: (2024)
von: Li, Meiling, et al.
Veröffentlicht: (2024)
TripleWin: Fixed-Point Equilibrium Pricing for Data-Model Coupled Markets
von: Ren, Hongrun, et al.
Veröffentlicht: (2025)
von: Ren, Hongrun, et al.
Veröffentlicht: (2025)
Annealed Softmax Greedy in Many-Armed Bayesian Bandits
von: Overman, William, et al.
Veröffentlicht: (2026)
von: Overman, William, et al.
Veröffentlicht: (2026)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
CoCoB: Adaptive Collaborative Combinatorial Bandits for Online Recommendation
von: Yan, Cairong, et al.
Veröffentlicht: (2025)
von: Yan, Cairong, et al.
Veröffentlicht: (2025)
Effective Off-Policy Evaluation and Learning in Contextual Combinatorial Bandits
von: Shimizu, Tatsuhiro, et al.
Veröffentlicht: (2024)
von: Shimizu, Tatsuhiro, et al.
Veröffentlicht: (2024)
A Contextual Combinatorial Bandit Approach to Negotiation
von: Li, Yexin, et al.
Veröffentlicht: (2024)
von: Li, Yexin, et al.
Veröffentlicht: (2024)
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
Stackelberg Coupling of Online Representation Learning and Reinforcement Learning
von: Martinez, Fernando, et al.
Veröffentlicht: (2025)
von: Martinez, Fernando, et al.
Veröffentlicht: (2025)
On the Regularity and Fairness of Combinatorial Multi-Armed Bandit
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyi, et al.
Veröffentlicht: (2025)
MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks
von: Hong, Zhi, et al.
Veröffentlicht: (2026)
von: Hong, Zhi, et al.
Veröffentlicht: (2026)
Balans: Multi-Armed Bandits-based Adaptive Large Neighborhood Search for Mixed-Integer Programming Problem
von: Cai, Junyang, et al.
Veröffentlicht: (2024)
von: Cai, Junyang, et al.
Veröffentlicht: (2024)
Contextual Combinatorial Bandits with Probabilistically Triggered Arms
von: Liu, Xutong, et al.
Veröffentlicht: (2023)
von: Liu, Xutong, et al.
Veröffentlicht: (2023)
Neural Combinatorial Clustered Bandits for Recommendation Systems
von: Atalar, Baran, et al.
Veröffentlicht: (2024)
von: Atalar, Baran, et al.
Veröffentlicht: (2024)
Bayesian Analysis of Combinatorial Gaussian Process Bandits
von: Sandberg, Jack, et al.
Veröffentlicht: (2023)
von: Sandberg, Jack, et al.
Veröffentlicht: (2023)
Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning
von: Feng, Xinsong, et al.
Veröffentlicht: (2025)
von: Feng, Xinsong, et al.
Veröffentlicht: (2025)
Adaptive Budgeted Multi-Armed Bandits for IoT with Dynamic Resource Constraints
von: Vaishnav, Shubham, et al.
Veröffentlicht: (2025)
von: Vaishnav, Shubham, et al.
Veröffentlicht: (2025)
Reward Shaping for Inference-Time Alignment: A Stackelberg Game Perspective
von: Wang, Haichuan, et al.
Veröffentlicht: (2026)
von: Wang, Haichuan, et al.
Veröffentlicht: (2026)
Efficient Multi-objective Prompt Optimization via Pure-exploration Bandits
von: Li, Donghao, et al.
Veröffentlicht: (2026)
von: Li, Donghao, et al.
Veröffentlicht: (2026)
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)
Put CASH on Bandits: A Max K-Armed Problem for Automated Machine Learning
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2025)
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2025)
Hierarchical Multi-Armed Bandits for the Concurrent Intelligent Tutoring of Concepts and Problems of Varying Difficulty Levels
von: Castleman, Blake, et al.
Veröffentlicht: (2024)
von: Castleman, Blake, et al.
Veröffentlicht: (2024)
Search, Examine and Early-Termination: Fake News Detection with Annotation-Free Evidences
von: Yang, Yuzhou, et al.
Veröffentlicht: (2024)
von: Yang, Yuzhou, et al.
Veröffentlicht: (2024)
Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health
von: Verma, Shresth, et al.
Veröffentlicht: (2026)
von: Verma, Shresth, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Federated Combinatorial Multi-Agent Multi-Armed Bandits
von: Fourati, Fares, et al.
Veröffentlicht: (2024) -
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
von: Xiong, Guojun, et al.
Veröffentlicht: (2024) -
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026) -
Multi-Armed Bandits With Best-Action Queries
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2026) -
Introduction to Multi-Armed Bandits
von: Slivkins, Aleksandrs
Veröffentlicht: (2019)