Incentivized Lipschitz Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Chakraborty, Sourav, Rege, Amit Kiran, Monteleoni, Claire, Chen, Lijun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Flickering Multi-Armed Bandits
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
Multi-Agent Lipschitz Bandits
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
A Unified Framework for Locality in Scalable MARL
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
Incentivized Exploration of Non-Stationary Stochastic Bandits
by: Chakraborty, Sourav, et al.
Published: (2024)
by: Chakraborty, Sourav, et al.
Published: (2024)
ArchesWeather: An efficient AI weather forecasting model at 1.5° resolution
by: Couairon, Guillaume, et al.
Published: (2024)
by: Couairon, Guillaume, et al.
Published: (2024)
The Role of Generator Access in Autoregressive Post-Training
by: Rege, Amit Kiran
Published: (2026)
by: Rege, Amit Kiran
Published: (2026)
Data Attribution in Adaptive Learning
by: Rege, Amit Kiran
Published: (2026)
by: Rege, Amit Kiran
Published: (2026)
Where Did Your Model Learn That? Label-free Influence for Self-supervised Learning
by: Harilal, Nidhin, et al.
Published: (2024)
by: Harilal, Nidhin, et al.
Published: (2024)
Put CASH on Bandits: A Max K-Armed Problem for Automated Machine Learning
by: Balef, Amir Rezaei, et al.
Published: (2025)
by: Balef, Amir Rezaei, et al.
Published: (2025)
BanditSpec: Adaptive Speculative Decoding via Bandit Algorithms
by: Hou, Yunlong, et al.
Published: (2025)
by: Hou, Yunlong, et al.
Published: (2025)
Variational Sampling of Temporal Trajectories
by: Nazarovs, Jurijs, et al.
Published: (2024)
by: Nazarovs, Jurijs, et al.
Published: (2024)
Application-Driven Innovation in Machine Learning
by: Rolnick, David, et al.
Published: (2024)
by: Rolnick, David, et al.
Published: (2024)
Incentivizing LLMs to Self-Verify Their Answers
by: Zhang, Fuxiang, et al.
Published: (2025)
by: Zhang, Fuxiang, et al.
Published: (2025)
AI Alignment via Incentives and Correction
by: Agarwal, Rohit, et al.
Published: (2026)
by: Agarwal, Rohit, et al.
Published: (2026)
Principles of Lipschitz continuity in neural networks
by: Luo, Róisín
Published: (2026)
by: Luo, Róisín
Published: (2026)
The Sample Complexity of Multiclass and Sparse Contextual Bandits
by: Erez, Liad, et al.
Published: (2026)
by: Erez, Liad, et al.
Published: (2026)
Contextual Combinatorial Bandits with Probabilistically Triggered Arms
by: Liu, Xutong, et al.
Published: (2023)
by: Liu, Xutong, et al.
Published: (2023)
Non-Stationary Lipschitz Bandits
by: Nguyen, Nicolas, et al.
Published: (2025)
by: Nguyen, Nicolas, et al.
Published: (2025)
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
by: Sim, Rachael Hwee Ling, et al.
Published: (2026)
by: Sim, Rachael Hwee Ling, et al.
Published: (2026)
IP-FL: Incentivized and Personalized Federated Learning
by: Khan, Ahmad Faraz, et al.
Published: (2023)
by: Khan, Ahmad Faraz, et al.
Published: (2023)
Online Clustering of Dueling Bandits
by: Wang, Zhiyong, et al.
Published: (2025)
by: Wang, Zhiyong, et al.
Published: (2025)
Tree Ensembles for Contextual Bandits
by: Nilsson, Hannes, et al.
Published: (2024)
by: Nilsson, Hannes, et al.
Published: (2024)
Data-Driven Lipschitz Continuity: A Cost-Effective Approach to Improve Adversarial Robustness
by: Chen, Erh-Chung, et al.
Published: (2024)
by: Chen, Erh-Chung, et al.
Published: (2024)
DQNC2S: DQN-based Cross-stream Crisis event Summarizer
by: Cambrin, Daniele Rege, et al.
Published: (2024)
by: Cambrin, Daniele Rege, et al.
Published: (2024)
Lipschitz-aware Linearity Grafting for Certified Robustness
by: Han, Yongjin, et al.
Published: (2025)
by: Han, Yongjin, et al.
Published: (2025)
L-Lipschitz Gershgorin ResNet Network
by: Juston, Marius F. R., et al.
Published: (2025)
by: Juston, Marius F. R., et al.
Published: (2025)
Certificate-Guided Pruning for Stochastic Lipschitz Optimization
by: Shihab, Ibne Farabi, et al.
Published: (2026)
by: Shihab, Ibne Farabi, et al.
Published: (2026)
VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
by: Wang, Haozhe, et al.
Published: (2025)
by: Wang, Haozhe, et al.
Published: (2025)
BanditLP: Large-Scale Stochastic Optimization for Personalized Recommendations
by: Nguyen, Phuc, et al.
Published: (2026)
by: Nguyen, Phuc, et al.
Published: (2026)
From Contextual Combinatorial Semi-Bandits to Bandit List Classification: Improved Sample Complexity with Sparse Rewards
by: Erez, Liad, et al.
Published: (2025)
by: Erez, Liad, et al.
Published: (2025)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
by: Xiong, Guojun, et al.
Published: (2024)
by: Xiong, Guojun, et al.
Published: (2024)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
by: Wang, Zhiyong, et al.
Published: (2024)
by: Wang, Zhiyong, et al.
Published: (2024)
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
by: Lu, Xiaodong, et al.
Published: (2026)
by: Lu, Xiaodong, et al.
Published: (2026)
Tokenized Bandit for LLM Decoding and Alignment
by: Shin, Suho, et al.
Published: (2025)
by: Shin, Suho, et al.
Published: (2025)
Deceptive Exploration in Multi-armed Bandits
by: Vurankaya, I. Arda, et al.
Published: (2025)
by: Vurankaya, I. Arda, et al.
Published: (2025)
Causal Contextual Bandits with Adaptive Context
by: Madhavan, Rahul, et al.
Published: (2024)
by: Madhavan, Rahul, et al.
Published: (2024)
Causally Abstracted Multi-armed Bandits
by: Zennaro, Fabio Massimo, et al.
Published: (2024)
by: Zennaro, Fabio Massimo, et al.
Published: (2024)
Neural Active Learning Beyond Bandits
by: Ban, Yikun, et al.
Published: (2024)
by: Ban, Yikun, et al.
Published: (2024)
Diffusion Models Meet Contextual Bandits
by: Aouali, Imad
Published: (2024)
by: Aouali, Imad
Published: (2024)
Best-Arm Identification in Unimodal Bandits
by: Poiani, Riccardo, et al.
Published: (2024)
by: Poiani, Riccardo, et al.
Published: (2024)
Similar Items
-
Flickering Multi-Armed Bandits
by: Chakraborty, Sourav, et al.
Published: (2026) -
Multi-Agent Lipschitz Bandits
by: Chakraborty, Sourav, et al.
Published: (2026) -
A Unified Framework for Locality in Scalable MARL
by: Chakraborty, Sourav, et al.
Published: (2026) -
Incentivized Exploration of Non-Stationary Stochastic Bandits
by: Chakraborty, Sourav, et al.
Published: (2024) -
ArchesWeather: An efficient AI weather forecasting model at 1.5° resolution
by: Couairon, Guillaume, et al.
Published: (2024)