Saved in:
| Main Author: | Kim, Mintae |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.03132 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Computing the Exact Pareto Front in Average-Cost Multi-Objective Markov Decision Processes
by: Luo, Jiping, et al.
Published: (2026)
by: Luo, Jiping, et al.
Published: (2026)
Optimal Control of Logically Constrained Partially Observable and Multi-Agent Markov Decision Processes
by: Kalagarla, Krishna C., et al.
Published: (2023)
by: Kalagarla, Krishna C., et al.
Published: (2023)
Feature-Based Belief Aggregation for Partially Observable Markov Decision Problems
by: Li, Yuchao, et al.
Published: (2025)
by: Li, Yuchao, et al.
Published: (2025)
GNN-Empowered Effective Partial Observation MARL Method for AoI Management in Multi-UAV Network
by: Pan, Yuhao, et al.
Published: (2024)
by: Pan, Yuhao, et al.
Published: (2024)
Minimax-Optimal Reward-Agnostic Exploration in Reinforcement Learning
by: Li, Gen, et al.
Published: (2023)
by: Li, Gen, et al.
Published: (2023)
Finite-Time Minimax Bounds and an Optimal Lyapunov Policy in Queueing Control
by: Liu, Yujie, et al.
Published: (2025)
by: Liu, Yujie, et al.
Published: (2025)
On Convex Data-Driven Inverse Optimal Control for Nonlinear, Non-stationary and Stochastic Systems
by: Garrabe, Emiland, et al.
Published: (2023)
by: Garrabe, Emiland, et al.
Published: (2023)
Semantic-Aware Remote Estimation of Multiple Markov Sources Under Constraints
by: Luo, Jiping, et al.
Published: (2024)
by: Luo, Jiping, et al.
Published: (2024)
Decision Transformer as a Foundation Model for Partially Observable Continuous Control
by: Zhang, Xiangyuan, et al.
Published: (2024)
by: Zhang, Xiangyuan, et al.
Published: (2024)
Communication-Control Codesign for Large-Scale Wireless Networked Control Systems
by: Pang, Gaoyang, et al.
Published: (2024)
by: Pang, Gaoyang, et al.
Published: (2024)
Semi-Markov Decision Process Framework for Age of Incorrect Information Minimization
by: Cosandal, Ismail, et al.
Published: (2025)
by: Cosandal, Ismail, et al.
Published: (2025)
Bayesian Learning of Optimal Policies in Markov Decision Processes with Countably Infinite State-Space
by: Adler, Saghar, et al.
Published: (2023)
by: Adler, Saghar, et al.
Published: (2023)
Learning Markov Processes as Sum-of-Square Forms for Analytical Belief Propagation
by: Amorese, Peter, et al.
Published: (2026)
by: Amorese, Peter, et al.
Published: (2026)
Markov Kernels, Distances and Optimal Control: A Parable of Linear Quadratic Non-Gaussian Distribution Steering
by: Teter, Alexis M. H., et al.
Published: (2025)
by: Teter, Alexis M. H., et al.
Published: (2025)
Mitigating Partial Observability in Adaptive Traffic Signal Control with Transformers
by: Wang, Xiaoyu, et al.
Published: (2024)
by: Wang, Xiaoyu, et al.
Published: (2024)
Achieving Hiding and Smart Anti-Jamming Communication: A Parallel DRL Approach against Moving Reactive Jammer
by: Li, Yangyang, et al.
Published: (2025)
by: Li, Yangyang, et al.
Published: (2025)
Learning Low-dimensional Latent Dynamics from High-dimensional Observations: Non-asymptotics and Lower Bounds
by: Zhang, Yuyang, et al.
Published: (2024)
by: Zhang, Yuyang, et al.
Published: (2024)
Latent Diffusion Model Based Denoising Receiver for 6G Semantic Communication: From Stochastic Differential Theory to Application
by: Wang, Xiucheng, et al.
Published: (2025)
by: Wang, Xiucheng, et al.
Published: (2025)
Newton-Flow Particle Filters based on Generalized Cramér Distance
by: Hanebeck, Uwe D.
Published: (2025)
by: Hanebeck, Uwe D.
Published: (2025)
Taming "data-hungry" reinforcement learning? Stability in continuous state-action spaces
by: Duan, Yaqi, et al.
Published: (2024)
by: Duan, Yaqi, et al.
Published: (2024)
Wireless Resource Allocation with Collaborative Distributed and Centralized DRL under Control Channel Attacks
by: Wang, Ke, et al.
Published: (2024)
by: Wang, Ke, et al.
Published: (2024)
Concentration of Cumulative Reward in Markov Decision Processes
by: Sayedana, Borna, et al.
Published: (2024)
by: Sayedana, Borna, et al.
Published: (2024)
Joint Age-State Belief is All You Need: Minimizing AoII via Pull-Based Remote Estimation
by: Cosandal, Ismail, et al.
Published: (2024)
by: Cosandal, Ismail, et al.
Published: (2024)
Belief Samples Are All You Need For Social Learning
by: JafariNodeh, Mahyar, et al.
Published: (2024)
by: JafariNodeh, Mahyar, et al.
Published: (2024)
Structure-Enhanced DRL for Optimal Transmission Scheduling
by: Chen, Jiazheng, et al.
Published: (2022)
by: Chen, Jiazheng, et al.
Published: (2022)
Settling the Sample Complexity of Model-Based Offline Reinforcement Learning
by: Li, Gen, et al.
Published: (2022)
by: Li, Gen, et al.
Published: (2022)
Structure-Enhanced Deep Reinforcement Learning for Optimal Transmission Scheduling
by: Chen, Jiazheng, et al.
Published: (2022)
by: Chen, Jiazheng, et al.
Published: (2022)
Approximate Linear Programming for Decentralized Policy Iteration in Cooperative Multi-agent Markov Decision Processes
by: Mandal, Lakshmi, et al.
Published: (2023)
by: Mandal, Lakshmi, et al.
Published: (2023)
Low-Complexity AoI-Optimal Status Update Control with Partial Battery State Information in Energy Harvesting IoT Networks
by: Wu, Hao, et al.
Published: (2025)
by: Wu, Hao, et al.
Published: (2025)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
by: Leon, Vincent, et al.
Published: (2023)
by: Leon, Vincent, et al.
Published: (2023)
A New Finite-Horizon Dynamic Programming Analysis of Nonanticipative Rate-Distortion Function for Markov Sources
by: He, Zixuan, et al.
Published: (2024)
by: He, Zixuan, et al.
Published: (2024)
On the Global Convergence of Policy Gradient in Average Reward Markov Decision Processes
by: Kumar, Navdeep, et al.
Published: (2024)
by: Kumar, Navdeep, et al.
Published: (2024)
OCMDP: Observation-Constrained Markov Decision Process
by: Wang, Taiyi, et al.
Published: (2024)
by: Wang, Taiyi, et al.
Published: (2024)
Function Gradient Approximation with Random Shallow ReLU Networks with Control Applications
by: Lamperski, Andrew, et al.
Published: (2024)
by: Lamperski, Andrew, et al.
Published: (2024)
Learning How to Strategically Disclose Information
by: Velicheti, Raj Kiriti, et al.
Published: (2024)
by: Velicheti, Raj Kiriti, et al.
Published: (2024)
A Soft Inducement Framework for Incentive-Aided Steering of No-Regret Players
by: Yorulmaz, Asrin Efe, et al.
Published: (2025)
by: Yorulmaz, Asrin Efe, et al.
Published: (2025)
Deep Learning for Wireless Networked Systems: a joint Estimation-Control-Scheduling Approach
by: Zhao, Zihuai, et al.
Published: (2022)
by: Zhao, Zihuai, et al.
Published: (2022)
BeamAgent: LLM-Aided MIMO Beamforming with Decoupled Intent Parsing and Alternating Optimization for Joint Site Selection and Precoding
by: Wang, Xiucheng, et al.
Published: (2026)
by: Wang, Xiucheng, et al.
Published: (2026)
Sensor Design for Accuracy-Bounded Estimation via Maximum-Entropy Likelihood Synthesis
by: Bhattacharya, Raktim
Published: (2026)
by: Bhattacharya, Raktim
Published: (2026)
Effective Communication with Dynamic Feature Compression
by: Talli, Pietro, et al.
Published: (2024)
by: Talli, Pietro, et al.
Published: (2024)
Similar Items
-
Computing the Exact Pareto Front in Average-Cost Multi-Objective Markov Decision Processes
by: Luo, Jiping, et al.
Published: (2026) -
Optimal Control of Logically Constrained Partially Observable and Multi-Agent Markov Decision Processes
by: Kalagarla, Krishna C., et al.
Published: (2023) -
Feature-Based Belief Aggregation for Partially Observable Markov Decision Problems
by: Li, Yuchao, et al.
Published: (2025) -
GNN-Empowered Effective Partial Observation MARL Method for AoI Management in Multi-UAV Network
by: Pan, Yuhao, et al.
Published: (2024) -
Minimax-Optimal Reward-Agnostic Exploration in Reinforcement Learning
by: Li, Gen, et al.
Published: (2023)