Restless Multi-Process Multi-Armed Bandits with Applications to Self-Driving Microscopies
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peris, Jaume Anguera, Cheng, Songtao, Zhang, Hanzhao, Ouyang, Wei, Jaldén, Joakim |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Stochastic Modeling and Resource Dimensioning of Multi-Cellular Edge Intelligent Systems
von: Peris, Jaume Anguera, et al.
Veröffentlicht: (2026)
von: Peris, Jaume Anguera, et al.
Veröffentlicht: (2026)
Extreme Distance Distributions of Poisson Voronoi Cells
von: Peris, Jaume Anguera, et al.
Veröffentlicht: (2024)
von: Peris, Jaume Anguera, et al.
Veröffentlicht: (2024)
Risk-Based Prognostics and Health Management
von: Sheppard, John W.
Veröffentlicht: (2025)
von: Sheppard, John W.
Veröffentlicht: (2025)
A Hybrid Artificial Intelligence Method for Estimating Flicker in Power Systems
von: Enayati, Javad, et al.
Veröffentlicht: (2025)
von: Enayati, Javad, et al.
Veröffentlicht: (2025)
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)
Global Rewards in Restless Multi-Armed Bandits
von: Raman, Naveen, et al.
Veröffentlicht: (2024)
von: Raman, Naveen, et al.
Veröffentlicht: (2024)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
An N-of-1 Artificial Intelligence Ecosystem for Precision Medicine
von: Fard, Pedram, et al.
Veröffentlicht: (2025)
von: Fard, Pedram, et al.
Veröffentlicht: (2025)
Contextual Restless Multi-Armed Bandits with Application to Demand Response Decision-Making
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
Can the Waymo Open Motion Dataset Support Realistic Behavioral Modeling? A Validation Study with Naturalistic Trajectories
von: Zhang, Yanlin, et al.
Veröffentlicht: (2025)
von: Zhang, Yanlin, et al.
Veröffentlicht: (2025)
Cooling Channel Design Optimization for High Power Multi-chip Packages
von: Acquah, Michael, et al.
Veröffentlicht: (2026)
von: Acquah, Michael, et al.
Veröffentlicht: (2026)
Adapting Probabilistic Risk Assessment for AI
von: Wisakanto, Anna Katariina, et al.
Veröffentlicht: (2025)
von: Wisakanto, Anna Katariina, et al.
Veröffentlicht: (2025)
DynaMark: A Reinforcement Learning Framework for Dynamic Watermarking in Industrial Machine Tool Controllers
von: Aftabi, Navid, et al.
Veröffentlicht: (2025)
von: Aftabi, Navid, et al.
Veröffentlicht: (2025)
Exploration of Multi-Element Collaborative Research and Application for Modern Power System Based on Generative Large Models
von: Cheng, Lu, et al.
Veröffentlicht: (2025)
von: Cheng, Lu, et al.
Veröffentlicht: (2025)
Establishment and Solution of a Multi-Stage Decision Model Based on Hypothesis Testing and Dynamic Programming Algorithm
von: Liu, Ziyang, et al.
Veröffentlicht: (2025)
von: Liu, Ziyang, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Delivery Delay Duration Prediction via Multi-Task Deep Learning
von: Faulkner, Stefan, et al.
Veröffentlicht: (2026)
von: Faulkner, Stefan, et al.
Veröffentlicht: (2026)
Federated Combinatorial Multi-Agent Multi-Armed Bandits
von: Fourati, Fares, et al.
Veröffentlicht: (2024)
von: Fourati, Fares, et al.
Veröffentlicht: (2024)
Gaussian process-based online health monitoring and fault analysis of lithium-ion battery systems from field data
von: Schaeffer, Joachim, et al.
Veröffentlicht: (2024)
von: Schaeffer, Joachim, et al.
Veröffentlicht: (2024)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
von: Xiong, Guojun, et al.
Veröffentlicht: (2024)
Byzantine-Resilient Decentralized Multi-Armed Bandits
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2023)
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2023)
Towards Interpretable Renal Health Decline Forecasting via Multi-LMM Collaborative Reasoning Framework
von: Wu, Peng-Yi, et al.
Veröffentlicht: (2025)
von: Wu, Peng-Yi, et al.
Veröffentlicht: (2025)
Continuous-Time Distributed Dynamic Programming for Networked Multi-Agent Markov Decision Processes
von: Lee, Donghwan, et al.
Veröffentlicht: (2023)
von: Lee, Donghwan, et al.
Veröffentlicht: (2023)
Stochastic Redistribution of Indistinguishable Items in Shared Habitation: A Multi-Agent Simulation Framework
von: Shah, Syed Haseeb
Veröffentlicht: (2025)
von: Shah, Syed Haseeb
Veröffentlicht: (2025)
TSS GAZ PTP: Towards Improving Gumbel AlphaZero with Two-stage Self-play for Multi-constrained Electric Vehicle Routing Problems
von: Wang, Hui, et al.
Veröffentlicht: (2025)
von: Wang, Hui, et al.
Veröffentlicht: (2025)
Integration of Multi-Mode Preference into Home Energy Management System Using Deep Reinforcement Learning
von: Sumayli, Mohammed, et al.
Veröffentlicht: (2025)
von: Sumayli, Mohammed, et al.
Veröffentlicht: (2025)
Lagrangian Index Policy for Restless Bandits with Average Reward
von: Avrachenkov, Konstantin, et al.
Veröffentlicht: (2024)
von: Avrachenkov, Konstantin, et al.
Veröffentlicht: (2024)
Multi-Mode Process Control Using Multi-Task Inverse Reinforcement Learning
von: Lin, Runze, et al.
Veröffentlicht: (2025)
von: Lin, Runze, et al.
Veröffentlicht: (2025)
Multi-task Neural Diffusion Processes
von: Rawson, Joseph, et al.
Veröffentlicht: (2025)
von: Rawson, Joseph, et al.
Veröffentlicht: (2025)
Data-driven Power Flow Linearization: Simulation
von: Jia, Mengshuo, et al.
Veröffentlicht: (2024)
von: Jia, Mengshuo, et al.
Veröffentlicht: (2024)
A Review of Stop-and-Go Traffic Wave Suppression Strategies: Variable Speed Limit vs. Jam-Absorption Driving
von: He, Zhengbing, et al.
Veröffentlicht: (2025)
von: He, Zhengbing, et al.
Veröffentlicht: (2025)
Calibrated Adversarial Sampling: Multi-Armed Bandit-Guided Generalization Against Unforeseen Attacks
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
Agentic AI and Occupational Displacement: A Multi-Regional Task Exposure Analysis of Emerging Labor Market Disruption
von: Gupta, Ravish, et al.
Veröffentlicht: (2026)
von: Gupta, Ravish, et al.
Veröffentlicht: (2026)
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
von: Juneja, Ishank, et al.
Veröffentlicht: (2026)
von: Juneja, Ishank, et al.
Veröffentlicht: (2026)
Beyond Conservative Automated Driving in Multi-Agent Scenarios via Coupled Model Predictive Control and Deep Reinforcement Learning
von: Rahmani, Saeed, et al.
Veröffentlicht: (2026)
von: Rahmani, Saeed, et al.
Veröffentlicht: (2026)
Balans: Multi-Armed Bandits-based Adaptive Large Neighborhood Search for Mixed-Integer Programming Problem
von: Cai, Junyang, et al.
Veröffentlicht: (2024)
von: Cai, Junyang, et al.
Veröffentlicht: (2024)
Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning
von: Merati, Mohammad, et al.
Veröffentlicht: (2026)
von: Merati, Mohammad, et al.
Veröffentlicht: (2026)
Lagrangian Relaxation for Multi-Action Partially Observable Restless Bandits: Heuristic Policies and Indexability
von: Meshram, Rahul, et al.
Veröffentlicht: (2025)
von: Meshram, Rahul, et al.
Veröffentlicht: (2025)
Performance-Aware Self-Configurable Multi-Agent Networks: A Distributed Submodular Approach for Simultaneous Coordination and Network Design
von: Xu, Zirui, et al.
Veröffentlicht: (2024)
von: Xu, Zirui, et al.
Veröffentlicht: (2024)
Smart and Efficient IoT-Based Irrigation System Design: Utilizing a Hybrid Agent-Based and System Dynamics Approach
von: Pargo, Taha Ahmadi, et al.
Veröffentlicht: (2025)
von: Pargo, Taha Ahmadi, et al.
Veröffentlicht: (2025)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2021)
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Stochastic Modeling and Resource Dimensioning of Multi-Cellular Edge Intelligent Systems
von: Peris, Jaume Anguera, et al.
Veröffentlicht: (2026) -
Extreme Distance Distributions of Poisson Voronoi Cells
von: Peris, Jaume Anguera, et al.
Veröffentlicht: (2024) -
Risk-Based Prognostics and Health Management
von: Sheppard, John W.
Veröffentlicht: (2025) -
A Hybrid Artificial Intelligence Method for Estimating Flicker in Power Systems
von: Enayati, Javad, et al.
Veröffentlicht: (2025) -
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)