Incentivized Lipschitz Bandits
Fuente:
arXiv
Guardado en:
| Autores principales: | Chakraborty, Sourav, Rege, Amit Kiran, Monteleoni, Claire, Chen, Lijun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Flickering Multi-Armed Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2026)
por: Chakraborty, Sourav, et al.
Publicado: (2026)
Multi-Agent Lipschitz Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2026)
por: Chakraborty, Sourav, et al.
Publicado: (2026)
A Unified Framework for Locality in Scalable MARL
por: Chakraborty, Sourav, et al.
Publicado: (2026)
por: Chakraborty, Sourav, et al.
Publicado: (2026)
Incentivized Exploration of Non-Stationary Stochastic Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2024)
por: Chakraborty, Sourav, et al.
Publicado: (2024)
ArchesWeather: An efficient AI weather forecasting model at 1.5° resolution
por: Couairon, Guillaume, et al.
Publicado: (2024)
por: Couairon, Guillaume, et al.
Publicado: (2024)
The Role of Generator Access in Autoregressive Post-Training
por: Rege, Amit Kiran
Publicado: (2026)
por: Rege, Amit Kiran
Publicado: (2026)
Data Attribution in Adaptive Learning
por: Rege, Amit Kiran
Publicado: (2026)
por: Rege, Amit Kiran
Publicado: (2026)
Where Did Your Model Learn That? Label-free Influence for Self-supervised Learning
por: Harilal, Nidhin, et al.
Publicado: (2024)
por: Harilal, Nidhin, et al.
Publicado: (2024)
Put CASH on Bandits: A Max K-Armed Problem for Automated Machine Learning
por: Balef, Amir Rezaei, et al.
Publicado: (2025)
por: Balef, Amir Rezaei, et al.
Publicado: (2025)
BanditSpec: Adaptive Speculative Decoding via Bandit Algorithms
por: Hou, Yunlong, et al.
Publicado: (2025)
por: Hou, Yunlong, et al.
Publicado: (2025)
Variational Sampling of Temporal Trajectories
por: Nazarovs, Jurijs, et al.
Publicado: (2024)
por: Nazarovs, Jurijs, et al.
Publicado: (2024)
Application-Driven Innovation in Machine Learning
por: Rolnick, David, et al.
Publicado: (2024)
por: Rolnick, David, et al.
Publicado: (2024)
Incentivizing LLMs to Self-Verify Their Answers
por: Zhang, Fuxiang, et al.
Publicado: (2025)
por: Zhang, Fuxiang, et al.
Publicado: (2025)
AI Alignment via Incentives and Correction
por: Agarwal, Rohit, et al.
Publicado: (2026)
por: Agarwal, Rohit, et al.
Publicado: (2026)
Principles of Lipschitz continuity in neural networks
por: Luo, Róisín
Publicado: (2026)
por: Luo, Róisín
Publicado: (2026)
The Sample Complexity of Multiclass and Sparse Contextual Bandits
por: Erez, Liad, et al.
Publicado: (2026)
por: Erez, Liad, et al.
Publicado: (2026)
Contextual Combinatorial Bandits with Probabilistically Triggered Arms
por: Liu, Xutong, et al.
Publicado: (2023)
por: Liu, Xutong, et al.
Publicado: (2023)
Non-Stationary Lipschitz Bandits
por: Nguyen, Nicolas, et al.
Publicado: (2025)
por: Nguyen, Nicolas, et al.
Publicado: (2025)
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
por: Sim, Rachael Hwee Ling, et al.
Publicado: (2026)
por: Sim, Rachael Hwee Ling, et al.
Publicado: (2026)
IP-FL: Incentivized and Personalized Federated Learning
por: Khan, Ahmad Faraz, et al.
Publicado: (2023)
por: Khan, Ahmad Faraz, et al.
Publicado: (2023)
Online Clustering of Dueling Bandits
por: Wang, Zhiyong, et al.
Publicado: (2025)
por: Wang, Zhiyong, et al.
Publicado: (2025)
Tree Ensembles for Contextual Bandits
por: Nilsson, Hannes, et al.
Publicado: (2024)
por: Nilsson, Hannes, et al.
Publicado: (2024)
Data-Driven Lipschitz Continuity: A Cost-Effective Approach to Improve Adversarial Robustness
por: Chen, Erh-Chung, et al.
Publicado: (2024)
por: Chen, Erh-Chung, et al.
Publicado: (2024)
DQNC2S: DQN-based Cross-stream Crisis event Summarizer
por: Cambrin, Daniele Rege, et al.
Publicado: (2024)
por: Cambrin, Daniele Rege, et al.
Publicado: (2024)
Lipschitz-aware Linearity Grafting for Certified Robustness
por: Han, Yongjin, et al.
Publicado: (2025)
por: Han, Yongjin, et al.
Publicado: (2025)
L-Lipschitz Gershgorin ResNet Network
por: Juston, Marius F. R., et al.
Publicado: (2025)
por: Juston, Marius F. R., et al.
Publicado: (2025)
Certificate-Guided Pruning for Stochastic Lipschitz Optimization
por: Shihab, Ibne Farabi, et al.
Publicado: (2026)
por: Shihab, Ibne Farabi, et al.
Publicado: (2026)
VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
por: Wang, Haozhe, et al.
Publicado: (2025)
por: Wang, Haozhe, et al.
Publicado: (2025)
BanditLP: Large-Scale Stochastic Optimization for Personalized Recommendations
por: Nguyen, Phuc, et al.
Publicado: (2026)
por: Nguyen, Phuc, et al.
Publicado: (2026)
From Contextual Combinatorial Semi-Bandits to Bandit List Classification: Improved Sample Complexity with Sparse Rewards
por: Erez, Liad, et al.
Publicado: (2025)
por: Erez, Liad, et al.
Publicado: (2025)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
por: Xiong, Guojun, et al.
Publicado: (2024)
por: Xiong, Guojun, et al.
Publicado: (2024)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
por: Wang, Zhiyong, et al.
Publicado: (2024)
por: Wang, Zhiyong, et al.
Publicado: (2024)
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
por: Lu, Xiaodong, et al.
Publicado: (2026)
por: Lu, Xiaodong, et al.
Publicado: (2026)
Tokenized Bandit for LLM Decoding and Alignment
por: Shin, Suho, et al.
Publicado: (2025)
por: Shin, Suho, et al.
Publicado: (2025)
Deceptive Exploration in Multi-armed Bandits
por: Vurankaya, I. Arda, et al.
Publicado: (2025)
por: Vurankaya, I. Arda, et al.
Publicado: (2025)
Causal Contextual Bandits with Adaptive Context
por: Madhavan, Rahul, et al.
Publicado: (2024)
por: Madhavan, Rahul, et al.
Publicado: (2024)
Causally Abstracted Multi-armed Bandits
por: Zennaro, Fabio Massimo, et al.
Publicado: (2024)
por: Zennaro, Fabio Massimo, et al.
Publicado: (2024)
Neural Active Learning Beyond Bandits
por: Ban, Yikun, et al.
Publicado: (2024)
por: Ban, Yikun, et al.
Publicado: (2024)
Diffusion Models Meet Contextual Bandits
por: Aouali, Imad
Publicado: (2024)
por: Aouali, Imad
Publicado: (2024)
Best-Arm Identification in Unimodal Bandits
por: Poiani, Riccardo, et al.
Publicado: (2024)
por: Poiani, Riccardo, et al.
Publicado: (2024)
Ejemplares similares
-
Flickering Multi-Armed Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2026) -
Multi-Agent Lipschitz Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2026) -
A Unified Framework for Locality in Scalable MARL
por: Chakraborty, Sourav, et al.
Publicado: (2026) -
Incentivized Exploration of Non-Stationary Stochastic Bandits
por: Chakraborty, Sourav, et al.
Publicado: (2024) -
ArchesWeather: An efficient AI weather forecasting model at 1.5° resolution
por: Couairon, Guillaume, et al.
Publicado: (2024)