Bandits with Single-Peaked Preferences and Limited Resources
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ben-Porat, Omer, Keinan, Gur, Torkan, Rotem |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Modeling Churn in Recommender Systems with Aggregated Preferences
von: Keinan, Gur, et al.
Veröffentlicht: (2025)
von: Keinan, Gur, et al.
Veröffentlicht: (2025)
Churn-Aware Recommendation Planning under Aggregated Preference Feedback
von: Keinan, Gur, et al.
Veröffentlicht: (2025)
von: Keinan, Gur, et al.
Veröffentlicht: (2025)
Strategic Content Creation with Age of GenAI: To Share or Not to Share?
von: Keinan, Gur, et al.
Veröffentlicht: (2025)
von: Keinan, Gur, et al.
Veröffentlicht: (2025)
Envious Explore and Exploit
von: Ben-Porat, Omer, et al.
Veröffentlicht: (2025)
von: Ben-Porat, Omer, et al.
Veröffentlicht: (2025)
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
von: Verma, Arun, et al.
Veröffentlicht: (2024)
von: Verma, Arun, et al.
Veröffentlicht: (2024)
Geometric-Averaged Preference Optimization for Soft Preference Labels
von: Furuta, Hiroki, et al.
Veröffentlicht: (2024)
von: Furuta, Hiroki, et al.
Veröffentlicht: (2024)
Exposing Limitations of Language Model Agents in Sequential-Task Compositions on the Web
von: Furuta, Hiroki, et al.
Veröffentlicht: (2023)
von: Furuta, Hiroki, et al.
Veröffentlicht: (2023)
Semi-Supervised Preference Optimization with Limited Feedback
von: Lee, Seonggyun, et al.
Veröffentlicht: (2025)
von: Lee, Seonggyun, et al.
Veröffentlicht: (2025)
Hummer: Towards Limited Competitive Preference Dataset
von: Jiang, Li, et al.
Veröffentlicht: (2024)
von: Jiang, Li, et al.
Veröffentlicht: (2024)
Bi-Level Contextual Bandits for Individualized Resource Allocation under Delayed Feedback
von: Almasi, Mohammadsina, et al.
Veröffentlicht: (2025)
von: Almasi, Mohammadsina, et al.
Veröffentlicht: (2025)
Preferences Evolve And So Should Your Bandits: Bandits with Evolving States for Online Platforms
von: Khosravi, Khashayar, et al.
Veröffentlicht: (2023)
von: Khosravi, Khashayar, et al.
Veröffentlicht: (2023)
Diffusion-Driven Inertial Generated Data for Smartphone Location Classification
von: Cohen, Noa, et al.
Veröffentlicht: (2025)
von: Cohen, Noa, et al.
Veröffentlicht: (2025)
INSIGHTS: Demonstration-Based Summaries of Time Series Predictors
von: Porat, Bar Eini, et al.
Veröffentlicht: (2026)
von: Porat, Bar Eini, et al.
Veröffentlicht: (2026)
Peak-Detector: Explainable Peak Detection via Instruction-Tuned Large Language Models in Physiological Sign
von: Li, Jiahui, et al.
Veröffentlicht: (2026)
von: Li, Jiahui, et al.
Veröffentlicht: (2026)
Personalized Reinforcement Learning with a Budget of Policies
von: Ivanov, Dmitry, et al.
Veröffentlicht: (2024)
von: Ivanov, Dmitry, et al.
Veröffentlicht: (2024)
Bandits with Preference Feedback: A Stackelberg Game Perspective
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
BanditSpec: Adaptive Speculative Decoding via Bandit Algorithms
von: Hou, Yunlong, et al.
Veröffentlicht: (2025)
von: Hou, Yunlong, et al.
Veröffentlicht: (2025)
Incentivized Lipschitz Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2025)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2025)
NestQuant: Nested Lattice Quantization for Matrix Products and LLMs
von: Savkin, Semyon, et al.
Veröffentlicht: (2025)
von: Savkin, Semyon, et al.
Veröffentlicht: (2025)
Modeling Attrition in Recommender Systems with Departing Bandits
von: Ben-Porat, Omer, et al.
Veröffentlicht: (2022)
von: Ben-Porat, Omer, et al.
Veröffentlicht: (2022)
Peak-Controlled Logits Poisoning Attack in Federated Distillation
von: Tang, Yuhan, et al.
Veröffentlicht: (2024)
von: Tang, Yuhan, et al.
Veröffentlicht: (2024)
Online Clustering of Dueling Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)
Tree Ensembles for Contextual Bandits
von: Nilsson, Hannes, et al.
Veröffentlicht: (2024)
von: Nilsson, Hannes, et al.
Veröffentlicht: (2024)
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
EviTrack: Selection over Sampling for Delayed Disambiguation
von: Haq, Omer
Veröffentlicht: (2026)
von: Haq, Omer
Veröffentlicht: (2026)
CollaFuse: Navigating Limited Resources and Privacy in Collaborative Generative AI
von: Zipperling, Domenique, et al.
Veröffentlicht: (2024)
von: Zipperling, Domenique, et al.
Veröffentlicht: (2024)
Collaborating with GenAI: Incentives and Replacements
von: Taitler, Boaz, et al.
Veröffentlicht: (2025)
von: Taitler, Boaz, et al.
Veröffentlicht: (2025)
Enhancing Preference-based Linear Bandits via Human Response Time
von: Li, Shen, et al.
Veröffentlicht: (2024)
von: Li, Shen, et al.
Veröffentlicht: (2024)
Universal One-third Time Scaling in Learning Peaked Distributions
von: Liu, Yizhou, et al.
Veröffentlicht: (2026)
von: Liu, Yizhou, et al.
Veröffentlicht: (2026)
From Contextual Combinatorial Semi-Bandits to Bandit List Classification: Improved Sample Complexity with Sparse Rewards
von: Erez, Liad, et al.
Veröffentlicht: (2025)
von: Erez, Liad, et al.
Veröffentlicht: (2025)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
Tokenized Bandit for LLM Decoding and Alignment
von: Shin, Suho, et al.
Veröffentlicht: (2025)
von: Shin, Suho, et al.
Veröffentlicht: (2025)
Deceptive Exploration in Multi-armed Bandits
von: Vurankaya, I. Arda, et al.
Veröffentlicht: (2025)
von: Vurankaya, I. Arda, et al.
Veröffentlicht: (2025)
Causal Contextual Bandits with Adaptive Context
von: Madhavan, Rahul, et al.
Veröffentlicht: (2024)
von: Madhavan, Rahul, et al.
Veröffentlicht: (2024)
Causally Abstracted Multi-armed Bandits
von: Zennaro, Fabio Massimo, et al.
Veröffentlicht: (2024)
von: Zennaro, Fabio Massimo, et al.
Veröffentlicht: (2024)
Neural Active Learning Beyond Bandits
von: Ban, Yikun, et al.
Veröffentlicht: (2024)
von: Ban, Yikun, et al.
Veröffentlicht: (2024)
Diffusion Models Meet Contextual Bandits
von: Aouali, Imad
Veröffentlicht: (2024)
von: Aouali, Imad
Veröffentlicht: (2024)
Best-Arm Identification in Unimodal Bandits
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
Learning When to Trust in Contextual Bandits
von: Ghasemi, Majid, et al.
Veröffentlicht: (2026)
von: Ghasemi, Majid, et al.
Veröffentlicht: (2026)
Operationalizing Fairness: Post-Hoc Threshold Optimization Under Hard Resource Limits
von: Singh, Moirangthem Tiken, et al.
Veröffentlicht: (2026)
von: Singh, Moirangthem Tiken, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Modeling Churn in Recommender Systems with Aggregated Preferences
von: Keinan, Gur, et al.
Veröffentlicht: (2025) -
Churn-Aware Recommendation Planning under Aggregated Preference Feedback
von: Keinan, Gur, et al.
Veröffentlicht: (2025) -
Strategic Content Creation with Age of GenAI: To Share or Not to Share?
von: Keinan, Gur, et al.
Veröffentlicht: (2025) -
Envious Explore and Exploit
von: Ben-Porat, Omer, et al.
Veröffentlicht: (2025) -
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
von: Verma, Arun, et al.
Veröffentlicht: (2024)