Incentivized Lipschitz Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chakraborty, Sourav, Rege, Amit Kiran, Monteleoni, Claire, Chen, Lijun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Multi-Agent Lipschitz Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
A Unified Framework for Locality in Scalable MARL
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Incentivized Exploration of Non-Stationary Stochastic Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2024)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2024)
ArchesWeather: An efficient AI weather forecasting model at 1.5° resolution
von: Couairon, Guillaume, et al.
Veröffentlicht: (2024)
von: Couairon, Guillaume, et al.
Veröffentlicht: (2024)
The Role of Generator Access in Autoregressive Post-Training
von: Rege, Amit Kiran
Veröffentlicht: (2026)
von: Rege, Amit Kiran
Veröffentlicht: (2026)
Data Attribution in Adaptive Learning
von: Rege, Amit Kiran
Veröffentlicht: (2026)
von: Rege, Amit Kiran
Veröffentlicht: (2026)
Where Did Your Model Learn That? Label-free Influence for Self-supervised Learning
von: Harilal, Nidhin, et al.
Veröffentlicht: (2024)
von: Harilal, Nidhin, et al.
Veröffentlicht: (2024)
Put CASH on Bandits: A Max K-Armed Problem for Automated Machine Learning
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2025)
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2025)
BanditSpec: Adaptive Speculative Decoding via Bandit Algorithms
von: Hou, Yunlong, et al.
Veröffentlicht: (2025)
von: Hou, Yunlong, et al.
Veröffentlicht: (2025)
Variational Sampling of Temporal Trajectories
von: Nazarovs, Jurijs, et al.
Veröffentlicht: (2024)
von: Nazarovs, Jurijs, et al.
Veröffentlicht: (2024)
Application-Driven Innovation in Machine Learning
von: Rolnick, David, et al.
Veröffentlicht: (2024)
von: Rolnick, David, et al.
Veröffentlicht: (2024)
Incentivizing LLMs to Self-Verify Their Answers
von: Zhang, Fuxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Fuxiang, et al.
Veröffentlicht: (2025)
AI Alignment via Incentives and Correction
von: Agarwal, Rohit, et al.
Veröffentlicht: (2026)
von: Agarwal, Rohit, et al.
Veröffentlicht: (2026)
Principles of Lipschitz continuity in neural networks
von: Luo, Róisín
Veröffentlicht: (2026)
von: Luo, Róisín
Veröffentlicht: (2026)
The Sample Complexity of Multiclass and Sparse Contextual Bandits
von: Erez, Liad, et al.
Veröffentlicht: (2026)
von: Erez, Liad, et al.
Veröffentlicht: (2026)
Contextual Combinatorial Bandits with Probabilistically Triggered Arms
von: Liu, Xutong, et al.
Veröffentlicht: (2023)
von: Liu, Xutong, et al.
Veröffentlicht: (2023)
Non-Stationary Lipschitz Bandits
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2025)
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2025)
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2026)
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2026)
IP-FL: Incentivized and Personalized Federated Learning
von: Khan, Ahmad Faraz, et al.
Veröffentlicht: (2023)
von: Khan, Ahmad Faraz, et al.
Veröffentlicht: (2023)
Online Clustering of Dueling Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)
Tree Ensembles for Contextual Bandits
von: Nilsson, Hannes, et al.
Veröffentlicht: (2024)
von: Nilsson, Hannes, et al.
Veröffentlicht: (2024)
Data-Driven Lipschitz Continuity: A Cost-Effective Approach to Improve Adversarial Robustness
von: Chen, Erh-Chung, et al.
Veröffentlicht: (2024)
von: Chen, Erh-Chung, et al.
Veröffentlicht: (2024)
DQNC2S: DQN-based Cross-stream Crisis event Summarizer
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
Lipschitz-aware Linearity Grafting for Certified Robustness
von: Han, Yongjin, et al.
Veröffentlicht: (2025)
von: Han, Yongjin, et al.
Veröffentlicht: (2025)
L-Lipschitz Gershgorin ResNet Network
von: Juston, Marius F. R., et al.
Veröffentlicht: (2025)
von: Juston, Marius F. R., et al.
Veröffentlicht: (2025)
Certificate-Guided Pruning for Stochastic Lipschitz Optimization
von: Shihab, Ibne Farabi, et al.
Veröffentlicht: (2026)
von: Shihab, Ibne Farabi, et al.
Veröffentlicht: (2026)
VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
BanditLP: Large-Scale Stochastic Optimization for Personalized Recommendations
von: Nguyen, Phuc, et al.
Veröffentlicht: (2026)
von: Nguyen, Phuc, et al.
Veröffentlicht: (2026)
From Contextual Combinatorial Semi-Bandits to Bandit List Classification: Improved Sample Complexity with Sparse Rewards
von: Erez, Liad, et al.
Veröffentlicht: (2025)
von: Erez, Liad, et al.
Veröffentlicht: (2025)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
von: Lu, Xiaodong, et al.
Veröffentlicht: (2026)
von: Lu, Xiaodong, et al.
Veröffentlicht: (2026)
Tokenized Bandit for LLM Decoding and Alignment
von: Shin, Suho, et al.
Veröffentlicht: (2025)
von: Shin, Suho, et al.
Veröffentlicht: (2025)
Deceptive Exploration in Multi-armed Bandits
von: Vurankaya, I. Arda, et al.
Veröffentlicht: (2025)
von: Vurankaya, I. Arda, et al.
Veröffentlicht: (2025)
Causal Contextual Bandits with Adaptive Context
von: Madhavan, Rahul, et al.
Veröffentlicht: (2024)
von: Madhavan, Rahul, et al.
Veröffentlicht: (2024)
Causally Abstracted Multi-armed Bandits
von: Zennaro, Fabio Massimo, et al.
Veröffentlicht: (2024)
von: Zennaro, Fabio Massimo, et al.
Veröffentlicht: (2024)
Neural Active Learning Beyond Bandits
von: Ban, Yikun, et al.
Veröffentlicht: (2024)
von: Ban, Yikun, et al.
Veröffentlicht: (2024)
Diffusion Models Meet Contextual Bandits
von: Aouali, Imad
Veröffentlicht: (2024)
von: Aouali, Imad
Veröffentlicht: (2024)
Best-Arm Identification in Unimodal Bandits
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Flickering Multi-Armed Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026) -
Multi-Agent Lipschitz Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026) -
A Unified Framework for Locality in Scalable MARL
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026) -
Incentivized Exploration of Non-Stationary Stochastic Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2024) -
ArchesWeather: An efficient AI weather forecasting model at 1.5° resolution
von: Couairon, Guillaume, et al.
Veröffentlicht: (2024)