Constructing Non-Markovian Decision Process via History Aggregator
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yongyi, Li, Wenxin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Decoupling Return-to-Go for Efficient Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026)
by: Wang, Yongyi, et al.
Published: (2026)
Style-Preserving Policy Optimization for Game Agents
by: Li, Lingfeng, et al.
Published: (2025)
by: Li, Lingfeng, et al.
Published: (2025)
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026)
by: Wang, Yongyi, et al.
Published: (2026)
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
by: Li, Lingfeng, et al.
Published: (2025)
by: Li, Lingfeng, et al.
Published: (2025)
Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes
by: Trapasso, Alessandro, et al.
Published: (2025)
by: Trapasso, Alessandro, et al.
Published: (2025)
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
by: Alamdari, Parand A., et al.
Published: (2023)
by: Alamdari, Parand A., et al.
Published: (2023)
Homomorphic Mappings for Value-Preserving State Aggregation in Markov Decision Processes
by: Zhao, Shuo, et al.
Published: (2025)
by: Zhao, Shuo, et al.
Published: (2025)
Non-Markovian Long-Horizon Robot Manipulation via Keyframe Chaining
by: Chen, Yipeng, et al.
Published: (2026)
by: Chen, Yipeng, et al.
Published: (2026)
Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
by: Wang, Yongyi, et al.
Published: (2025)
by: Wang, Yongyi, et al.
Published: (2025)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
by: Li, Lingfeng, et al.
Published: (2026)
by: Li, Lingfeng, et al.
Published: (2026)
Process-Tensor Tomography of SGD: Measuring Non-Markovian Memory via Back-Flow of Distinguishability
by: Sevetlidis, Vasileios, et al.
Published: (2026)
by: Sevetlidis, Vasileios, et al.
Published: (2026)
ShuttleEnv: An Interactive Data-Driven RL Environment for Badminton Strategy Modeling
by: Li, Ang, et al.
Published: (2026)
by: Li, Ang, et al.
Published: (2026)
Policy Dispersion in Non-Markovian Environment
by: Qu, Bohao, et al.
Published: (2023)
by: Qu, Bohao, et al.
Published: (2023)
Quantile Markov Decision Process
by: Li, Xiaocheng, et al.
Published: (2017)
by: Li, Xiaocheng, et al.
Published: (2017)
Act as You Learn: Adaptive Decision-Making in Non-Stationary Markov Decision Processes
by: Luo, Baiting, et al.
Published: (2024)
by: Luo, Baiting, et al.
Published: (2024)
Robust Decision Aggregation with Adversarial Experts
by: Guo, Yongkang, et al.
Published: (2024)
by: Guo, Yongkang, et al.
Published: (2024)
From Construction to Injection: Edit-Based Fingerprints for Large Language Models
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
Uncertainty-Based Methods for Automated Process Reward Data Construction and Output Aggregation in Mathematical Reasoning
by: Han, Jiuzhou, et al.
Published: (2025)
by: Han, Jiuzhou, et al.
Published: (2025)
HAEPO: History-Aggregated Exploratory Policy Optimization
by: Trivedi, Gaurish, et al.
Published: (2025)
by: Trivedi, Gaurish, et al.
Published: (2025)
Spatial-Temporal Reinforcement Learning for Network Routing with Non-Markovian Traffic
by: Wang, Molly, et al.
Published: (2025)
by: Wang, Molly, et al.
Published: (2025)
Sharpe Ratio Optimization in Markov Decision Processes
by: Ma, Shuai, et al.
Published: (2025)
by: Ma, Shuai, et al.
Published: (2025)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
by: Kar, Avik, et al.
Published: (2026)
by: Kar, Avik, et al.
Published: (2026)
Dual Formulation for Non-Rectangular Lp Robust Markov Decision Processes
by: Kumar, Navdeep, et al.
Published: (2025)
by: Kumar, Navdeep, et al.
Published: (2025)
GuardAD: Safeguarding Autonomous Driving MLLMs via Markovian Safety Logic
by: Zhang, Tianyuan, et al.
Published: (2026)
by: Zhang, Tianyuan, et al.
Published: (2026)
NS-Gym: Open-Source Simulation Environments and Benchmarks for Non-Stationary Markov Decision Processes
by: Keplinger, Nathaniel S., et al.
Published: (2025)
by: Keplinger, Nathaniel S., et al.
Published: (2025)
MPD$^2$-Router: Mask-aware Multi-expert Prior-regularized Dual-head Deferral Router in Glaucoma Screening and Diagnosis
by: Zhan, Wenxin
Published: (2026)
by: Zhan, Wenxin
Published: (2026)
Structuring Value Representations via Geometric Coherence in Markov Decision Processes
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
ABC: Any-Subset Autoregression via Non-Markovian Diffusion Bridges in Continuous Time and Space
by: Guo, Gabe, et al.
Published: (2026)
by: Guo, Gabe, et al.
Published: (2026)
State-Centric Decision Process
by: Jeong, Sungheon, et al.
Published: (2026)
by: Jeong, Sungheon, et al.
Published: (2026)
Creativity and Markov Decision Processes
by: Lahikainen, Joonas, et al.
Published: (2024)
by: Lahikainen, Joonas, et al.
Published: (2024)
Statistical Inference for Misspecified Contextual Bandits
by: Guo, Yongyi, et al.
Published: (2025)
by: Guo, Yongyi, et al.
Published: (2025)
mHC-lite: You Don't Need 20 Sinkhorn-Knopp Iterations
by: Yang, Yongyi, et al.
Published: (2026)
by: Yang, Yongyi, et al.
Published: (2026)
Safe Reinforcement Learning with Learned Non-Markovian Safety Constraints
by: Low, Siow Meng, et al.
Published: (2024)
by: Low, Siow Meng, et al.
Published: (2024)
Counterfactual Influence in Markov Decision Processes
by: Kazemi, Milad, et al.
Published: (2024)
by: Kazemi, Milad, et al.
Published: (2024)
Adapting Rules of Official International Mahjong for Online Players
by: Wang, Chucai, et al.
Published: (2026)
by: Wang, Chucai, et al.
Published: (2026)
Reinforcement Learning in Switching Non-Stationary Markov Decision Processes: Algorithms and Convergence Analysis
by: Amiri, Mohsen, et al.
Published: (2025)
by: Amiri, Mohsen, et al.
Published: (2025)
Policy Gradient Algorithms with Monte Carlo Tree Learning for Non-Markov Decision Processes
by: Morimura, Tetsuro, et al.
Published: (2022)
by: Morimura, Tetsuro, et al.
Published: (2022)
An Interpretable Client Decision Tree Aggregation process for Federated Learning
by: Argente-Garrido, Alberto, et al.
Published: (2024)
by: Argente-Garrido, Alberto, et al.
Published: (2024)
Beyond Sliding Windows: Learning to Manage Memory in Non-Markovian Environments
by: Tasse, Geraud Nangue, et al.
Published: (2025)
by: Tasse, Geraud Nangue, et al.
Published: (2025)
Efficient Test-Time Scaling via Temporal Reasoning Aggregation
by: Li, Jiakun, et al.
Published: (2026)
by: Li, Jiakun, et al.
Published: (2026)
Similar Items
-
Decoupling Return-to-Go for Efficient Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026) -
Style-Preserving Policy Optimization for Game Agents
by: Li, Lingfeng, et al.
Published: (2025) -
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026) -
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
by: Li, Lingfeng, et al.
Published: (2025) -
Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes
by: Trapasso, Alessandro, et al.
Published: (2025)