Online MDP with Transition Prototypes: A Robust Adaptive Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Shuo, Qi, Meng, Shen, Zuo-Jun Max |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Decision-Focused Sequential Experimental Design: A Directional Uncertainty-Guided Approach
von: Wan, Beichen, et al.
Veröffentlicht: (2026)
von: Wan, Beichen, et al.
Veröffentlicht: (2026)
Active Learning For Contextual Linear Optimization: A Margin-Based Approach
von: Liu, Mo, et al.
Veröffentlicht: (2023)
von: Liu, Mo, et al.
Veröffentlicht: (2023)
Spatial Supply Repositioning with Censored Demand Data
von: Jiang, Hansheng, et al.
Veröffentlicht: (2025)
von: Jiang, Hansheng, et al.
Veröffentlicht: (2025)
MDP Planning as Policy Inference
von: Tolpin, David
Veröffentlicht: (2026)
von: Tolpin, David
Veröffentlicht: (2026)
MDP: Multidimensional Vision Model Pruning with Latency Constraint
von: Sun, Xinglong, et al.
Veröffentlicht: (2025)
von: Sun, Xinglong, et al.
Veröffentlicht: (2025)
MDP3: A Training-free Approach for List-wise Frame Selection in Video-LLMs
von: Sun, Hui, et al.
Veröffentlicht: (2025)
von: Sun, Hui, et al.
Veröffentlicht: (2025)
Daily Physical Activity Monitoring -- Adaptive Learning from Multi-source Motion Sensor Data
von: Zhang, Haoting, et al.
Veröffentlicht: (2024)
von: Zhang, Haoting, et al.
Veröffentlicht: (2024)
Geometric Re-Analysis of Classical MDP Solving Algorithms
von: Mustafin, Arsenii, et al.
Veröffentlicht: (2025)
von: Mustafin, Arsenii, et al.
Veröffentlicht: (2025)
Dual Prototypes for Adaptive Pre-Trained Model in Class-Incremental Learning
von: Xu, Zhiming, et al.
Veröffentlicht: (2024)
von: Xu, Zhiming, et al.
Veröffentlicht: (2024)
PRAGA: Prototype-aware Graph Adaptive Aggregation for Spatial Multi-modal Omics Analysis
von: Huang, Xinlei, et al.
Veröffentlicht: (2024)
von: Huang, Xinlei, et al.
Veröffentlicht: (2024)
Adaptive and Robust DBSCAN with Multi-agent Reinforcement Learning
von: Peng, Hao, et al.
Veröffentlicht: (2025)
von: Peng, Hao, et al.
Veröffentlicht: (2025)
MDP Geometry, Normalization and Reward Balancing Solvers
von: Mustafin, Arsenii, et al.
Veröffentlicht: (2024)
von: Mustafin, Arsenii, et al.
Veröffentlicht: (2024)
None To Optima in Few Shots: Bayesian Optimization with MDP Priors
von: Li, Diantong, et al.
Veröffentlicht: (2025)
von: Li, Diantong, et al.
Veröffentlicht: (2025)
Adaptive, Robust and Scalable Bayesian Filtering for Online Learning
von: Duran-Martin, Gerardo
Veröffentlicht: (2025)
von: Duran-Martin, Gerardo
Veröffentlicht: (2025)
ICU-Sepsis: A Benchmark MDP Built from Real Medical Data
von: Choudhary, Kartik, et al.
Veröffentlicht: (2024)
von: Choudhary, Kartik, et al.
Veröffentlicht: (2024)
A-PETE: Adaptive Prototype Explanations of Tree Ensembles
von: Karolczak, Jacek, et al.
Veröffentlicht: (2024)
von: Karolczak, Jacek, et al.
Veröffentlicht: (2024)
Vertical Federated Continual Learning via Evolving Prototype Knowledge
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
Harnessing the Continuous Structure: Utilizing the First-order Approach in Online Contract Design
von: Zuo, Shiliang
Veröffentlicht: (2024)
von: Zuo, Shiliang
Veröffentlicht: (2024)
A Factored MDP Approach To Moving Target Defense With Dynamic Threat Modeling and Cost Efficiency
von: Bose, Megha, et al.
Veröffentlicht: (2024)
von: Bose, Megha, et al.
Veröffentlicht: (2024)
Federated Learning With Energy Harvesting Devices: An MDP Framework
von: Zhang, Kai, et al.
Veröffentlicht: (2024)
von: Zhang, Kai, et al.
Veröffentlicht: (2024)
Using Forwards-Backwards Models to Approximate MDP Homomorphisms
von: Mavor-Parker, Augustine N., et al.
Veröffentlicht: (2022)
von: Mavor-Parker, Augustine N., et al.
Veröffentlicht: (2022)
CHAM-net: A Contrastive Hierarchical Adaptive Meta-network for Robust Global Methane Flux Prediction
von: Dong, Rongchao, et al.
Veröffentlicht: (2026)
von: Dong, Rongchao, et al.
Veröffentlicht: (2026)
Enhancing Job Salary Prediction with Disentangled Composition Effect Modeling: A Neural Prototyping Approach
von: Ji, Yang, et al.
Veröffentlicht: (2025)
von: Ji, Yang, et al.
Veröffentlicht: (2025)
Predictive Control and Regret Analysis of Non-Stationary MDP with Look-ahead Information
von: Zhang, Ziyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ziyi, et al.
Veröffentlicht: (2024)
A Deep Generative Learning Approach for Two-stage Adaptive Robust Optimization
von: Brenner, Aron, et al.
Veröffentlicht: (2024)
von: Brenner, Aron, et al.
Veröffentlicht: (2024)
Track-MDP: Reinforcement Learning for Target Tracking with Controlled Sensing
von: Subramaniam, Adarsh M., et al.
Veröffentlicht: (2024)
von: Subramaniam, Adarsh M., et al.
Veröffentlicht: (2024)
Corruption-Robust Lipschitz Contextual Search
von: Zuo, Shiliang
Veröffentlicht: (2023)
von: Zuo, Shiliang
Veröffentlicht: (2023)
A Minimax-MDP Framework with Future-imposed Conditions for Learning-augmented Problems
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Learning a Fast Mixing Exogenous Block MDP using a Single Trajectory
von: Levine, Alexander, et al.
Veröffentlicht: (2024)
von: Levine, Alexander, et al.
Veröffentlicht: (2024)
Online Learning to Rank under Corruption: A Robust Cascading Bandits Approach
von: Ghaffari, Fatemeh, et al.
Veröffentlicht: (2025)
von: Ghaffari, Fatemeh, et al.
Veröffentlicht: (2025)
A Unified Online-Offline Framework for Co-Branding Campaign Recommendations
von: Dai, Xiangxiang, et al.
Veröffentlicht: (2025)
von: Dai, Xiangxiang, et al.
Veröffentlicht: (2025)
Robust, Online, and Adaptive Decentralized Gaussian Processes
von: Llorente, Fernando, et al.
Veröffentlicht: (2025)
von: Llorente, Fernando, et al.
Veröffentlicht: (2025)
SaVeR: Optimal Data Collection Strategy for Safe Policy Evaluation in Tabular MDP
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
Group Resonance Network: Learnable Prototypes and Multi-Subject Resonance for EEG Emotion Recognition
von: Meng, Renwei
Veröffentlicht: (2026)
von: Meng, Renwei
Veröffentlicht: (2026)
Pre-training Epidemic Time Series Forecasters with Compartmental Prototypes
von: Liu, Zewen, et al.
Veröffentlicht: (2025)
von: Liu, Zewen, et al.
Veröffentlicht: (2025)
Exploring the Noise Robustness of Online Conformal Prediction
von: Xi, Huajun, et al.
Veröffentlicht: (2025)
von: Xi, Huajun, et al.
Veröffentlicht: (2025)
Adversarially Robust Detection of Harmful Online Content: A Computational Design Science Approach
von: Chai, Yidong, et al.
Veröffentlicht: (2025)
von: Chai, Yidong, et al.
Veröffentlicht: (2025)
TimeHF: Billion-Scale Time Series Models Guided by Human Feedback
von: Qi, Yongzhi, et al.
Veröffentlicht: (2025)
von: Qi, Yongzhi, et al.
Veröffentlicht: (2025)
Contribution Evaluation of Heterogeneous Participants in Federated Learning via Prototypical Representations
von: Guo, Qi, et al.
Veröffentlicht: (2024)
von: Guo, Qi, et al.
Veröffentlicht: (2024)
Deep reinforcement learning for weakly coupled MDP's with continuous actions
von: Robledo, Francisco, et al.
Veröffentlicht: (2024)
von: Robledo, Francisco, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Decision-Focused Sequential Experimental Design: A Directional Uncertainty-Guided Approach
von: Wan, Beichen, et al.
Veröffentlicht: (2026) -
Active Learning For Contextual Linear Optimization: A Margin-Based Approach
von: Liu, Mo, et al.
Veröffentlicht: (2023) -
Spatial Supply Repositioning with Censored Demand Data
von: Jiang, Hansheng, et al.
Veröffentlicht: (2025) -
MDP Planning as Policy Inference
von: Tolpin, David
Veröffentlicht: (2026) -
MDP: Multidimensional Vision Model Pruning with Latency Constraint
von: Sun, Xinglong, et al.
Veröffentlicht: (2025)