Selecting Belief-State Approximations in Simulators with Latent States
Fuente:
arXiv
Guardado en:
| Autor principal: | Jiang, Nan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Belief State Transformer
por: Hu, Edward S., et al.
Publicado: (2024)
por: Hu, Edward S., et al.
Publicado: (2024)
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
por: Jiang, Nan, et al.
Publicado: (2025)
por: Jiang, Nan, et al.
Publicado: (2025)
How to Explore with Belief: State Entropy Maximization in POMDPs
por: Zamboni, Riccardo, et al.
Publicado: (2024)
por: Zamboni, Riccardo, et al.
Publicado: (2024)
Circular Belief Propagation for Approximate Probabilistic Inference
por: Bouttier, Vincent, et al.
Publicado: (2024)
por: Bouttier, Vincent, et al.
Publicado: (2024)
Beyond State-Wise Mirror Descent: Offline Policy Optimization with Parametric Policies
por: Li, Xiang, et al.
Publicado: (2026)
por: Li, Xiang, et al.
Publicado: (2026)
Belief States for Cooperative Multi-Agent Reinforcement Learning under Partial Observability
por: Pritz, Paul J., et al.
Publicado: (2025)
por: Pritz, Paul J., et al.
Publicado: (2025)
Latent State Estimation Helps UI Agents to Reason
por: Bishop, William E, et al.
Publicado: (2024)
por: Bishop, William E, et al.
Publicado: (2024)
Adversarial Latent-State Training for Robust Policies in Partially Observable Domains
por: Ahuja, Angad Singh
Publicado: (2026)
por: Ahuja, Angad Singh
Publicado: (2026)
Adapting World Models with Latent-State Dynamics Residuals
por: Lanier, JB, et al.
Publicado: (2025)
por: Lanier, JB, et al.
Publicado: (2025)
PcLast: Discovering Plannable Continuous Latent States
por: Koul, Anurag, et al.
Publicado: (2023)
por: Koul, Anurag, et al.
Publicado: (2023)
Conceptual Belief-Informed Reinforcement Learning
por: Gu, Xingrui, et al.
Publicado: (2024)
por: Gu, Xingrui, et al.
Publicado: (2024)
Beyond Dense States: Elevating Sparse Transcoders to Active Operators for Latent Reasoning
por: Wang, Yadong, et al.
Publicado: (2026)
por: Wang, Yadong, et al.
Publicado: (2026)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
State Stream Transformer (SST) : Emergent Metacognitive Behaviours Through Latent State Persistence
por: Aviss, Thea
Publicado: (2025)
por: Aviss, Thea
Publicado: (2025)
Fairness Begins with State: Purifying Latent Preferences for Hierarchical Reinforcement Learning in Interactive Recommendation
por: Lu, Yun, et al.
Publicado: (2026)
por: Lu, Yun, et al.
Publicado: (2026)
Exploring the Edges of Latent State Clusters for Goal-Conditioned Reinforcement Learning
por: Duan, Yuanlin, et al.
Publicado: (2024)
por: Duan, Yuanlin, et al.
Publicado: (2024)
ss-Mamba: Semantic-Spline Selective State-Space Model
por: Ye, Zuochen
Publicado: (2025)
por: Ye, Zuochen
Publicado: (2025)
Mamba: Linear-Time Sequence Modeling with Selective State Spaces
por: Gu, Albert, et al.
Publicado: (2023)
por: Gu, Albert, et al.
Publicado: (2023)
MambaLRP: Explaining Selective State Space Sequence Models
por: Jafari, Farnoush Rezaei, et al.
Publicado: (2024)
por: Jafari, Farnoush Rezaei, et al.
Publicado: (2024)
TIDES: Implicit Time-Awareness in Selective State Space Models
por: Soydan, Taylan, et al.
Publicado: (2026)
por: Soydan, Taylan, et al.
Publicado: (2026)
Context-Selective State Space Models: Feedback is All You Need
por: Zattra, Riccardo, et al.
Publicado: (2025)
por: Zattra, Riccardo, et al.
Publicado: (2025)
Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Spaces
por: Emmanouilidis, Konstantinos, et al.
Publicado: (2026)
por: Emmanouilidis, Konstantinos, et al.
Publicado: (2026)
Sessa: Selective State Space Attention
por: Horbatko, Liubomyr
Publicado: (2026)
por: Horbatko, Liubomyr
Publicado: (2026)
A Note on Loss Functions and Error Compounding in Model-based Reinforcement Learning
por: Jiang, Nan
Publicado: (2024)
por: Jiang, Nan
Publicado: (2024)
Quamba: A Post-Training Quantization Recipe for Selective State Space Models
por: Chiang, Hung-Yueh, et al.
Publicado: (2024)
por: Chiang, Hung-Yueh, et al.
Publicado: (2024)
Decision Mamba: Reinforcement Learning via Sequence Modeling with Selective State Spaces
por: Ota, Toshihiro
Publicado: (2024)
por: Ota, Toshihiro
Publicado: (2024)
On the Laplace Approximation as Model Selection Criterion for Gaussian Processes
por: Besginow, Andreas, et al.
Publicado: (2024)
por: Besginow, Andreas, et al.
Publicado: (2024)
Latent Generative Solvers for Generalizable Long-Term Physics Simulation
por: Chen, Zituo, et al.
Publicado: (2026)
por: Chen, Zituo, et al.
Publicado: (2026)
Stochastic Multivariate Universal-Radix Finite-State Machine: a Theoretically and Practically Elegant Nonlinear Function Approximator
por: Feng, Xincheng, et al.
Publicado: (2024)
por: Feng, Xincheng, et al.
Publicado: (2024)
Variational OOD State Correction for Offline Reinforcement Learning
por: Jiang, Ke, et al.
Publicado: (2025)
por: Jiang, Ke, et al.
Publicado: (2025)
LUCoS: Latent Unsupervised Context Selection for Tabular Foundation Models
por: Ipas, Oroel, et al.
Publicado: (2026)
por: Ipas, Oroel, et al.
Publicado: (2026)
GLADMamba: Unsupervised Graph-Level Anomaly Detection Powered by Selective State Space Model
por: Fu, Yali, et al.
Publicado: (2025)
por: Fu, Yali, et al.
Publicado: (2025)
STG-Mamba: Spatial-Temporal Graph Learning via Selective State Space Model
por: Li, Lincan, et al.
Publicado: (2024)
por: Li, Lincan, et al.
Publicado: (2024)
Joint Selective State Space Model and Detrending for Robust Time Series Anomaly Detection
por: Chen, Junqi, et al.
Publicado: (2024)
por: Chen, Junqi, et al.
Publicado: (2024)
Graph-Mamba: Towards Long-Range Graph Sequence Modeling with Selective State Spaces
por: Wang, Chloe, et al.
Publicado: (2024)
por: Wang, Chloe, et al.
Publicado: (2024)
Taming Momentum: Rethinking Optimizer States Through Low-Rank Approximation
por: Wang, Zhengbo, et al.
Publicado: (2026)
por: Wang, Zhengbo, et al.
Publicado: (2026)
Approximating Pareto Frontiers in Stochastic Multi-Objective Optimization via Hashing and Randomization
por: Li, Jinzhao, et al.
Publicado: (2026)
por: Li, Jinzhao, et al.
Publicado: (2026)
Probing Latent Subspaces in LLM for AI Security: Identifying and Manipulating Adversarial States
por: Chia, Xin Wei, et al.
Publicado: (2025)
por: Chia, Xin Wei, et al.
Publicado: (2025)
Meanings and Feelings of Large Language Models: Observability of Latent States in Generative AI
por: Liu, Tian Yu, et al.
Publicado: (2024)
por: Liu, Tian Yu, et al.
Publicado: (2024)
Atrial Fibrillation Prediction Using a Lightweight Temporal Convolutional and Selective State Space Architecture
por: Lee, Yongbin, et al.
Publicado: (2025)
por: Lee, Yongbin, et al.
Publicado: (2025)
Ejemplares similares
-
The Belief State Transformer
por: Hu, Edward S., et al.
Publicado: (2024) -
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
por: Jiang, Nan, et al.
Publicado: (2025) -
How to Explore with Belief: State Entropy Maximization in POMDPs
por: Zamboni, Riccardo, et al.
Publicado: (2024) -
Circular Belief Propagation for Approximate Probabilistic Inference
por: Bouttier, Vincent, et al.
Publicado: (2024) -
Beyond State-Wise Mirror Descent: Offline Policy Optimization with Parametric Policies
por: Li, Xiang, et al.
Publicado: (2026)