Multi-Environment POMDPs with Finite-Horizon Objectives
Fuente:
arXiv
Guardado en:
| Autores principales: | Brice, Léonard, Cano, Filip, Chatterjee, Krishnendu, Henzinger, Thomas A., Muroya, Stefanie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Energy Shields for Fairness
por: Cano, Filip, et al.
Publicado: (2026)
por: Cano, Filip, et al.
Publicado: (2026)
Algorithmic Fairness: A Runtime Perspective
por: Cano, Filip, et al.
Publicado: (2025)
por: Cano, Filip, et al.
Publicado: (2025)
Fairness Shields: Safeguarding against Biased Decision Makers
por: Cano, Filip, et al.
Publicado: (2024)
por: Cano, Filip, et al.
Publicado: (2024)
Multi-Environment POMDPs: Discrete Model Uncertainty Under Partial Observability
por: Bovy, Eline M., et al.
Publicado: (2025)
por: Bovy, Eline M., et al.
Publicado: (2025)
Qualitative Analysis of $ω$-Regular Objectives on Robust MDPs
por: Asadi, Ali, et al.
Publicado: (2025)
por: Asadi, Ali, et al.
Publicado: (2025)
Generalizing Multi-Step Inverse Models for Representation Learning to Finite-Memory POMDPs
por: Wu, Lili, et al.
Publicado: (2024)
por: Wu, Lili, et al.
Publicado: (2024)
BetaZero: Belief-State Planning for Long-Horizon POMDPs using Learned Approximations
por: Moss, Robert J., et al.
Publicado: (2023)
por: Moss, Robert J., et al.
Publicado: (2023)
Revealing POMDPs: Qualitative and Quantitative Analysis for Parity Objectives
por: Asadi, Ali, et al.
Publicado: (2025)
por: Asadi, Ali, et al.
Publicado: (2025)
Revelations: A Decidable Class of POMDPs with Omega-Regular Objectives
por: Belly, Marius, et al.
Publicado: (2024)
por: Belly, Marius, et al.
Publicado: (2024)
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
por: Galesloot, Maris F. L., et al.
Publicado: (2025)
por: Galesloot, Maris F. L., et al.
Publicado: (2025)
Finite-State Controllers for (Hidden-Model) POMDPs using Deep Reinforcement Learning
por: Hudák, David, et al.
Publicado: (2026)
por: Hudák, David, et al.
Publicado: (2026)
Towards Responsible AI: Advances in Safety, Fairness, and Accountability of Autonomous Systems
por: Cano, Filip
Publicado: (2025)
por: Cano, Filip
Publicado: (2025)
Sound Heuristic Search Value Iteration for Undiscounted POMDPs with Reachability Objectives
por: Ho, Qi Heng, et al.
Publicado: (2024)
por: Ho, Qi Heng, et al.
Publicado: (2024)
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
por: Azeem, Muqsit, et al.
Publicado: (2024)
por: Azeem, Muqsit, et al.
Publicado: (2024)
Tighter Value-Function Approximations for POMDPs
por: Krale, Merlijn, et al.
Publicado: (2025)
por: Krale, Merlijn, et al.
Publicado: (2025)
Online Planning in POMDPs with State-Requests
por: Avalos, Raphael, et al.
Publicado: (2024)
por: Avalos, Raphael, et al.
Publicado: (2024)
Rethinking Transformers in Solving POMDPs
por: Lu, Chenhao, et al.
Publicado: (2024)
por: Lu, Chenhao, et al.
Publicado: (2024)
Lower Bound on Howard Policy Iteration for Deterministic Markov Decision Processes
por: Asadi, Ali, et al.
Publicado: (2025)
por: Asadi, Ali, et al.
Publicado: (2025)
Planning under Distribution Shifts with Causal POMDPs
por: Ceriscioli, Matteo, et al.
Publicado: (2026)
por: Ceriscioli, Matteo, et al.
Publicado: (2026)
Simplification of Risk Averse POMDPs with Performance Guarantees
por: Pariente, Yaacov, et al.
Publicado: (2024)
por: Pariente, Yaacov, et al.
Publicado: (2024)
What should be observed for optimal reward in POMDPs?
por: Konsta, Alyzia-Maria, et al.
Publicado: (2024)
por: Konsta, Alyzia-Maria, et al.
Publicado: (2024)
PepTune: De Novo Generation of Therapeutic Peptides with Multi-Objective-Guided Discrete Diffusion
por: Tang, Sophia, et al.
Publicado: (2024)
por: Tang, Sophia, et al.
Publicado: (2024)
Formal Verification of Neural Certificates Done Dynamically
por: Henzinger, Thomas A., et al.
Publicado: (2025)
por: Henzinger, Thomas A., et al.
Publicado: (2025)
Neural Control and Certificate Repair via Runtime Monitoring
por: Yu, Emily, et al.
Publicado: (2024)
por: Yu, Emily, et al.
Publicado: (2024)
Welfare Maximization Algorithm for Solving Budget-Constrained Multi-Component POMDPs
por: Vora, Manav, et al.
Publicado: (2023)
por: Vora, Manav, et al.
Publicado: (2023)
ε-Stationary Nash Equilibria in Multi-player Stochastic Graph Games
por: Asadi, Ali, et al.
Publicado: (2025)
por: Asadi, Ali, et al.
Publicado: (2025)
Certified Policy Verification and Synthesis for MDPs under Distributional Reach-avoidance Properties
por: Akshay, S., et al.
Publicado: (2024)
por: Akshay, S., et al.
Publicado: (2024)
Scalable Solution Methods for Dec-POMDPs with Deterministic Dynamics
por: You, Yang, et al.
Publicado: (2025)
por: You, Yang, et al.
Publicado: (2025)
Belief-State Query Policies for User-Aligned POMDPs
por: Bramblett, Daniel, et al.
Publicado: (2024)
por: Bramblett, Daniel, et al.
Publicado: (2024)
Beyond Simple Graphs: Neural Multi-Objective Routing on Multigraphs
por: Rydin, Filip, et al.
Publicado: (2025)
por: Rydin, Filip, et al.
Publicado: (2025)
A Finite-State Controller Based Offline Solver for Deterministic POMDPs
por: Schutz, Alex, et al.
Publicado: (2025)
por: Schutz, Alex, et al.
Publicado: (2025)
Finding equilibria: simpler for pessimists, simplest for optimists
por: Brice, Léonard, et al.
Publicado: (2025)
por: Brice, Léonard, et al.
Publicado: (2025)
Mixing Any Cocktail with Limited Ingredients: On the Structure of Payoff Sets in Multi-Objective POMDPs and its Impact on Randomised Strategies
por: Main, James C. A., et al.
Publicado: (2025)
por: Main, James C. A., et al.
Publicado: (2025)
Inducing Individual Students' Learning Strategies through Homomorphic POMDPs
por: Gao, Huifan, et al.
Publicado: (2024)
por: Gao, Huifan, et al.
Publicado: (2024)
Factored Online Planning in Many-Agent POMDPs
por: Galesloot, Maris F. L., et al.
Publicado: (2023)
por: Galesloot, Maris F. L., et al.
Publicado: (2023)
Pessimistic Iterative Planning with RNNs for Robust POMDPs
por: Galesloot, Maris F. L., et al.
Publicado: (2024)
por: Galesloot, Maris F. L., et al.
Publicado: (2024)
Scalable Policy-Based RL Algorithms for POMDPs
por: Anjarlekar, Ameya, et al.
Publicado: (2025)
por: Anjarlekar, Ameya, et al.
Publicado: (2025)
Value Iteration with Guessing for Markov Chains and Markov Decision Processes
por: Chatterjee, Krishnendu, et al.
Publicado: (2025)
por: Chatterjee, Krishnendu, et al.
Publicado: (2025)
ALE-Bench: A Benchmark for Long-Horizon Objective-Driven Algorithm Engineering
por: Imajuku, Yuki, et al.
Publicado: (2025)
por: Imajuku, Yuki, et al.
Publicado: (2025)
Solving Long-run Average Reward Robust MDPs via Stochastic Games
por: Chatterjee, Krishnendu, et al.
Publicado: (2023)
por: Chatterjee, Krishnendu, et al.
Publicado: (2023)
Ejemplares similares
-
Energy Shields for Fairness
por: Cano, Filip, et al.
Publicado: (2026) -
Algorithmic Fairness: A Runtime Perspective
por: Cano, Filip, et al.
Publicado: (2025) -
Fairness Shields: Safeguarding against Biased Decision Makers
por: Cano, Filip, et al.
Publicado: (2024) -
Multi-Environment POMDPs: Discrete Model Uncertainty Under Partial Observability
por: Bovy, Eline M., et al.
Publicado: (2025) -
Qualitative Analysis of $ω$-Regular Objectives on Robust MDPs
por: Asadi, Ali, et al.
Publicado: (2025)