UCB-type Algorithm for Budget-Constrained Expert Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Latypov, Ilgam, Suvorikova, Alexandra, Kroshnin, Alexey, Gasnikov, Alexander, Dorn, Yuriy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast UCB-type algorithms for stochastic bandits with heavy and super heavy symmetric noise
by: Dorn, Yuriy, et al.
Published: (2024)
by: Dorn, Yuriy, et al.
Published: (2024)
Cooperative Multi-Agent Graph Bandits: UCB Algorithm and Regret Analysis
by: Paschalidis, Phevos, et al.
Published: (2024)
by: Paschalidis, Phevos, et al.
Published: (2024)
SLDP: Semi-Local Differential Privacy for Density-Adaptive Analytics
by: Kroshnin, Alexey, et al.
Published: (2026)
by: Kroshnin, Alexey, et al.
Published: (2026)
Functional multi-armed bandit and the best function identification problems
by: Dorn, Yuriy, et al.
Published: (2025)
by: Dorn, Yuriy, et al.
Published: (2025)
$γ$-Competitiveness: An Approach to Multi-Objective Optimization with High Computation Costs in Lipschitz Functions
by: Latypov, Ilgam, et al.
Published: (2024)
by: Latypov, Ilgam, et al.
Published: (2024)
Constrained Black-Box Attacks Against Cooperative Multi-Agent Reinforcement Learning
by: Andam, Amine, et al.
Published: (2025)
by: Andam, Amine, et al.
Published: (2025)
Mean-Field Approximation of Cooperative Constrained Multi-Agent Reinforcement Learning (CMARL)
by: Mondal, Washim Uddin, et al.
Published: (2022)
by: Mondal, Washim Uddin, et al.
Published: (2022)
Bernstein-type and Bennett-type inequalities for unbounded matrix martingales
by: Kroshnin, Alexey, et al.
Published: (2024)
by: Kroshnin, Alexey, et al.
Published: (2024)
A Constrained Multi-Agent Reinforcement Learning Approach to Autonomous Traffic Signal Control
by: Satheesh, Anirudh, et al.
Published: (2025)
by: Satheesh, Anirudh, et al.
Published: (2025)
LEED: A Highly Efficient and Scalable LLM-Empowered Expert Demonstrations Framework for Multi-Agent Reinforcement Learning
by: Duan, Tianyang, et al.
Published: (2025)
by: Duan, Tianyang, et al.
Published: (2025)
Intelligent Communication Planning for Constrained Environmental IoT Sensing with Reinforcement Learning
by: Hu, Yi, et al.
Published: (2023)
by: Hu, Yi, et al.
Published: (2023)
PokéChamp: an Expert-level Minimax Language Agent
by: Karten, Seth, et al.
Published: (2025)
by: Karten, Seth, et al.
Published: (2025)
EcoFair-CH-MARL: Scalable Constrained Hierarchical Multi-Agent RL with Real-Time Emission Budgets and Fairness Guarantees
by: Alqithami, Saad
Published: (2026)
by: Alqithami, Saad
Published: (2026)
Multi-Agent Systems are Mixtures of Experts: Who Becomes an Influencer?
by: Bause, Franka, et al.
Published: (2026)
by: Bause, Franka, et al.
Published: (2026)
Interpretable Cascading Mixture-of-Experts for Urban Traffic Congestion Prediction
by: Jiang, Wenzhao, et al.
Published: (2024)
by: Jiang, Wenzhao, et al.
Published: (2024)
Cooperative Multi-Agent Constrained Stochastic Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2024)
by: Afsharrad, Amirhossein, et al.
Published: (2024)
LEARN: Learning End-to-End Aerial Resource-Constrained Multi-Robot Navigation
by: Chiu, Darren, et al.
Published: (2025)
by: Chiu, Darren, et al.
Published: (2025)
Learning safe, constrained policies via imitation learning: Connection to Probabilistic Inference and a Naive Algorithm
by: Papadopoulos, George, et al.
Published: (2025)
by: Papadopoulos, George, et al.
Published: (2025)
MixTTE: Multi-Level Mixture-of-Experts for Scalable and Adaptive Travel Time Estimation
by: Jiang, Wenzhao, et al.
Published: (2026)
by: Jiang, Wenzhao, et al.
Published: (2026)
DePAint: A Decentralized Safe Multi-Agent Reinforcement Learning Algorithm considering Peak and Average Constraints
by: Hassan, Raheeb, et al.
Published: (2023)
by: Hassan, Raheeb, et al.
Published: (2023)
YOTOnet: Zero-Shot Cross-Domain Fault Diagnosis via Domain-Conditioned Mixture of Experts
by: Wang, Zesen, et al.
Published: (2026)
by: Wang, Zesen, et al.
Published: (2026)
MermaidFlow: Redefining Agentic Workflow Generation via Safety-Constrained Evolutionary Programming
by: Zheng, Chengqi, et al.
Published: (2025)
by: Zheng, Chengqi, et al.
Published: (2025)
Expert-Free Online Transfer Learning in Multi-Agent Reinforcement Learning
by: Castagna, Alberto
Published: (2025)
by: Castagna, Alberto
Published: (2025)
Ready from Day 1: Population-Aware Coordination for Large-Scale Constrained Multi-Agent Systems
by: Wang, Angel, et al.
Published: (2026)
by: Wang, Angel, et al.
Published: (2026)
OFA-MAS: One-for-All Multi-Agent System Topology Design based on Mixture-of-Experts Graph Generative Models
by: Li, Shiyuan, et al.
Published: (2026)
by: Li, Shiyuan, et al.
Published: (2026)
Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale
by: Omari, Bassel Al, et al.
Published: (2025)
by: Omari, Bassel Al, et al.
Published: (2025)
MAPF-GPT: Imitation Learning for Multi-Agent Pathfinding at Scale
by: Andreychuk, Anton, et al.
Published: (2024)
by: Andreychuk, Anton, et al.
Published: (2024)
A Reinforcement Learning Inspired Latent Yield Based Adaptive Algorithm Switching Mechanism
by: Nair, Jayprakash S., et al.
Published: (2026)
by: Nair, Jayprakash S., et al.
Published: (2026)
Meritocratic Fairness in Budgeted Combinatorial Multi-armed Bandits via Shapley Values
by: Sharma, Shradha, et al.
Published: (2026)
by: Sharma, Shradha, et al.
Published: (2026)
MAPFAST: A Deep Algorithm Selector for Multi Agent Path Finding using Shortest Path Embeddings
by: Ren, Jingyao, et al.
Published: (2021)
by: Ren, Jingyao, et al.
Published: (2021)
Bayesian Decision Making around Experts
by: Ornia, Daniel Jarne, et al.
Published: (2025)
by: Ornia, Daniel Jarne, et al.
Published: (2025)
POGEMA: A Benchmark Platform for Cooperative Multi-Agent Pathfinding
by: Skrynnik, Alexey, et al.
Published: (2024)
by: Skrynnik, Alexey, et al.
Published: (2024)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
by: Rutherford, Alexander, et al.
Published: (2023)
by: Rutherford, Alexander, et al.
Published: (2023)
Transfer in Reinforcement Learning via Regret Bounds for Learning Agents
by: Tuynman, Adrienne, et al.
Published: (2022)
by: Tuynman, Adrienne, et al.
Published: (2022)
BIPPO: Budget-Aware Independent PPO for Energy-Efficient Federated Learning Services
by: Lackinger, Anna, et al.
Published: (2025)
by: Lackinger, Anna, et al.
Published: (2025)
Preference-Guided Learning for Sparse-Reward Multi-Agent Reinforcement Learning
by: Bui, The Viet, et al.
Published: (2025)
by: Bui, The Viet, et al.
Published: (2025)
Learning to Coordinate via Quantum Entanglement in Multi-Agent Reinforcement Learning
by: Gardiner, John, et al.
Published: (2026)
by: Gardiner, John, et al.
Published: (2026)
Distributed Continual Learning
by: Le, Long, et al.
Published: (2024)
by: Le, Long, et al.
Published: (2024)
CAMAR: Continuous Actions Multi-Agent Routing
by: Pshenitsyn, Artem, et al.
Published: (2025)
by: Pshenitsyn, Artem, et al.
Published: (2025)
Multi-Agent Model-Based Reinforcement Learning with Joint State-Action Learned Embeddings
by: Wang, Zhizun, et al.
Published: (2026)
by: Wang, Zhizun, et al.
Published: (2026)
Similar Items
-
Fast UCB-type algorithms for stochastic bandits with heavy and super heavy symmetric noise
by: Dorn, Yuriy, et al.
Published: (2024) -
Cooperative Multi-Agent Graph Bandits: UCB Algorithm and Regret Analysis
by: Paschalidis, Phevos, et al.
Published: (2024) -
SLDP: Semi-Local Differential Privacy for Density-Adaptive Analytics
by: Kroshnin, Alexey, et al.
Published: (2026) -
Functional multi-armed bandit and the best function identification problems
by: Dorn, Yuriy, et al.
Published: (2025) -
$γ$-Competitiveness: An Approach to Multi-Objective Optimization with High Computation Costs in Lipschitz Functions
by: Latypov, Ilgam, et al.
Published: (2024)