Interpreting systems as solving POMDPs: a step towards a formal understanding of agency
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Biehl, Martin, Virgo, Nathaniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A "good regulator theorem" for embodied agents
von: Virgo, Nathaniel, et al.
Veröffentlicht: (2025)
von: Virgo, Nathaniel, et al.
Veröffentlicht: (2025)
A Bayesian Interpretation of the Internal Model Principle
von: Baltieri, Manuel, et al.
Veröffentlicht: (2025)
von: Baltieri, Manuel, et al.
Veröffentlicht: (2025)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
von: Lanier, Michael, et al.
Veröffentlicht: (2024)
von: Lanier, Michael, et al.
Veröffentlicht: (2024)
Tighter Value-Function Approximations for POMDPs
von: Krale, Merlijn, et al.
Veröffentlicht: (2025)
von: Krale, Merlijn, et al.
Veröffentlicht: (2025)
Finite-State Controllers for (Hidden-Model) POMDPs using Deep Reinforcement Learning
von: Hudák, David, et al.
Veröffentlicht: (2026)
von: Hudák, David, et al.
Veröffentlicht: (2026)
Rethinking Transformers in Solving POMDPs
von: Lu, Chenhao, et al.
Veröffentlicht: (2024)
von: Lu, Chenhao, et al.
Veröffentlicht: (2024)
Multi-Environment POMDPs with Finite-Horizon Objectives
von: Brice, Léonard, et al.
Veröffentlicht: (2026)
von: Brice, Léonard, et al.
Veröffentlicht: (2026)
Simplification of Risk Averse POMDPs with Performance Guarantees
von: Pariente, Yaacov, et al.
Veröffentlicht: (2024)
von: Pariente, Yaacov, et al.
Veröffentlicht: (2024)
What should be observed for optimal reward in POMDPs?
von: Konsta, Alyzia-Maria, et al.
Veröffentlicht: (2024)
von: Konsta, Alyzia-Maria, et al.
Veröffentlicht: (2024)
Planning under Distribution Shifts with Causal POMDPs
von: Ceriscioli, Matteo, et al.
Veröffentlicht: (2026)
von: Ceriscioli, Matteo, et al.
Veröffentlicht: (2026)
Interpretable Models Capable of Handling Systematic Missingness in Imbalanced Classes and Heterogeneous Datasets
von: Ghosh, Sreejita, et al.
Veröffentlicht: (2022)
von: Ghosh, Sreejita, et al.
Veröffentlicht: (2022)
Scalable Solution Methods for Dec-POMDPs with Deterministic Dynamics
von: You, Yang, et al.
Veröffentlicht: (2025)
von: You, Yang, et al.
Veröffentlicht: (2025)
Belief-State Query Policies for User-Aligned POMDPs
von: Bramblett, Daniel, et al.
Veröffentlicht: (2024)
von: Bramblett, Daniel, et al.
Veröffentlicht: (2024)
Online Planning in POMDPs with State-Requests
von: Avalos, Raphael, et al.
Veröffentlicht: (2024)
von: Avalos, Raphael, et al.
Veröffentlicht: (2024)
Teacher agency in the age of generative AI: towards a framework of hybrid intelligence for learning design
von: Frøsig, Thomas B, et al.
Veröffentlicht: (2024)
von: Frøsig, Thomas B, et al.
Veröffentlicht: (2024)
Inducing Individual Students' Learning Strategies through Homomorphic POMDPs
von: Gao, Huifan, et al.
Veröffentlicht: (2024)
von: Gao, Huifan, et al.
Veröffentlicht: (2024)
Factored Online Planning in Many-Agent POMDPs
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2023)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2023)
Pessimistic Iterative Planning with RNNs for Robust POMDPs
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2024)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2024)
Scalable Policy-Based RL Algorithms for POMDPs
von: Anjarlekar, Ameya, et al.
Veröffentlicht: (2025)
von: Anjarlekar, Ameya, et al.
Veröffentlicht: (2025)
Multi-Environment POMDPs: Discrete Model Uncertainty Under Partial Observability
von: Bovy, Eline M., et al.
Veröffentlicht: (2025)
von: Bovy, Eline M., et al.
Veröffentlicht: (2025)
Missingness-MDPs: Bridging the Theory of Missing Data and POMDPs
von: Wendland, Joshua, et al.
Veröffentlicht: (2026)
von: Wendland, Joshua, et al.
Veröffentlicht: (2026)
Optimizing Task Completion Time Updates Using POMDPs
von: Eddy, Duncan, et al.
Veröffentlicht: (2026)
von: Eddy, Duncan, et al.
Veröffentlicht: (2026)
Value of Information and Reward Specification in Active Inference and POMDPs
von: Wei, Ran
Veröffentlicht: (2024)
von: Wei, Ran
Veröffentlicht: (2024)
Sequential Monte Carlo for Policy Optimization in Continuous POMDPs
von: Abdulsamad, Hany, et al.
Veröffentlicht: (2025)
von: Abdulsamad, Hany, et al.
Veröffentlicht: (2025)
How to Explore with Belief: State Entropy Maximization in POMDPs
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2024)
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2024)
Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
von: Wang, Yongyi, et al.
Veröffentlicht: (2025)
von: Wang, Yongyi, et al.
Veröffentlicht: (2025)
Computing the Reachability Value of Posterior-Deterministic POMDPs
von: Fijalkow, Nathanaël, et al.
Veröffentlicht: (2026)
von: Fijalkow, Nathanaël, et al.
Veröffentlicht: (2026)
Point-Based Value Iteration for POMDPs with Neural Perception Mechanisms
von: Yan, Rui, et al.
Veröffentlicht: (2023)
von: Yan, Rui, et al.
Veröffentlicht: (2023)
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2025)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2025)
Online Risk-Averse Planning in POMDPs Using Iterated CVaR Value Function
von: Pariente, Yaacov, et al.
Veröffentlicht: (2026)
von: Pariente, Yaacov, et al.
Veröffentlicht: (2026)
BetaZero: Belief-State Planning for Long-Horizon POMDPs using Learned Approximations
von: Moss, Robert J., et al.
Veröffentlicht: (2023)
von: Moss, Robert J., et al.
Veröffentlicht: (2023)
A step toward a reinforcement learning de novo genome assembler
von: Padovani, Kleber, et al.
Veröffentlicht: (2021)
von: Padovani, Kleber, et al.
Veröffentlicht: (2021)
Statistical Tractability of Off-policy Evaluation of History-dependent Policies in POMDPs
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
Welfare Maximization Algorithm for Solving Budget-Constrained Multi-Component POMDPs
von: Vora, Manav, et al.
Veröffentlicht: (2023)
von: Vora, Manav, et al.
Veröffentlicht: (2023)
Sound Heuristic Search Value Iteration for Undiscounted POMDPs with Reachability Objectives
von: Ho, Qi Heng, et al.
Veröffentlicht: (2024)
von: Ho, Qi Heng, et al.
Veröffentlicht: (2024)
Generalizing Multi-Step Inverse Models for Representation Learning to Finite-Memory POMDPs
von: Wu, Lili, et al.
Veröffentlicht: (2024)
von: Wu, Lili, et al.
Veröffentlicht: (2024)
Toward Learning POMDPs Beyond Full-Rank Actions and State Observability
von: Shaw, Seiji, et al.
Veröffentlicht: (2026)
von: Shaw, Seiji, et al.
Veröffentlicht: (2026)
Solving Collaborative Dec-POMDPs with Deep Reinforcement Learning Heuristics
von: Soffair, Nitsan
Veröffentlicht: (2022)
von: Soffair, Nitsan
Veröffentlicht: (2022)
Tru-POMDP: Task Planning Under Uncertainty via Tree of Hypotheses and Open-Ended POMDPs
von: Tang, Wenjing, et al.
Veröffentlicht: (2025)
von: Tang, Wenjing, et al.
Veröffentlicht: (2025)
No Compromise in Solution Quality: Speeding Up Belief-dependent Continuous POMDPs via Adaptive Multilevel Simplification
von: Zhitnikov, Andrey, et al.
Veröffentlicht: (2023)
von: Zhitnikov, Andrey, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
A "good regulator theorem" for embodied agents
von: Virgo, Nathaniel, et al.
Veröffentlicht: (2025) -
A Bayesian Interpretation of the Internal Model Principle
von: Baltieri, Manuel, et al.
Veröffentlicht: (2025) -
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
von: Lanier, Michael, et al.
Veröffentlicht: (2024) -
Tighter Value-Function Approximations for POMDPs
von: Krale, Merlijn, et al.
Veröffentlicht: (2025) -
Finite-State Controllers for (Hidden-Model) POMDPs using Deep Reinforcement Learning
von: Hudák, David, et al.
Veröffentlicht: (2026)