Active Inference Tree Search in Large POMDPs
Fuente:
arXiv
Guardado en:
| Autores principales: | Maisto, Domenico, Gregoretti, Francesco, Friston, Karl, Pezzulo, Giovanni |
|---|---|
| Formato: | Preprint |
| Publicado: |
2021
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
por: Mitchell, Rupert, et al.
Publicado: (2025)
por: Mitchell, Rupert, et al.
Publicado: (2025)
Identifying Autism-Related Neurobiomarkers Using Hybrid Deep Learning Models
por: Chen, Ashley
Publicado: (2025)
por: Chen, Ashley
Publicado: (2025)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
por: Danieli, Federico, et al.
Publicado: (2025)
por: Danieli, Federico, et al.
Publicado: (2025)
Stealth edits to large language models
por: Sutton, Oliver J., et al.
Publicado: (2024)
por: Sutton, Oliver J., et al.
Publicado: (2024)
A Mathematical Framework for the Problem of Security for Cognition in Neurotechnology
por: Bagley, Bryce Allen, et al.
Publicado: (2024)
por: Bagley, Bryce Allen, et al.
Publicado: (2024)
Contrastive Consolidation of Top-Down Modulations Achieves Sparsely Supervised Continual Learning
por: Tran, Viet Anh Khoa, et al.
Publicado: (2025)
por: Tran, Viet Anh Khoa, et al.
Publicado: (2025)
Shortest Paths without a Map, but with an Entropic Regularizer
por: Bubeck, Sébastien, et al.
Publicado: (2022)
por: Bubeck, Sébastien, et al.
Publicado: (2022)
Towards Single Exponential Time for Temporal and Spatial Reasoning: A Study via Redundancy and Dynamic Programming
por: Lagerkvist, Victor, et al.
Publicado: (2026)
por: Lagerkvist, Victor, et al.
Publicado: (2026)
Boosting Test Performance with Importance Sampling--a Subpopulation Perspective
por: Shen, Hongyu, et al.
Publicado: (2024)
por: Shen, Hongyu, et al.
Publicado: (2024)
Platonic Representations in the Human Brain: Unsupervised Recovery of Universal Geometry
por: Marcos-Manchón, Pablo, et al.
Publicado: (2026)
por: Marcos-Manchón, Pablo, et al.
Publicado: (2026)
Embedding Explainable AI in NHS Clinical Safety: The Explainability-Enabled Clinical Safety Framework (ECSF)
por: Gigiu, Robert
Publicado: (2025)
por: Gigiu, Robert
Publicado: (2025)
Safe Distributed Control of Multi-Robot Systems with Communication Delays
por: Ballotta, Luca, et al.
Publicado: (2024)
por: Ballotta, Luca, et al.
Publicado: (2024)
Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data
por: Pasini, Massimiliano Lupo, et al.
Publicado: (2026)
por: Pasini, Massimiliano Lupo, et al.
Publicado: (2026)
Untrained CNNs Match Backpropagation at V1: A Systematic RSA Comparison of Four Learning Rules Against Human fMRI
por: Leutenegger, Nils
Publicado: (2026)
por: Leutenegger, Nils
Publicado: (2026)
BrainMAE: A Region-aware Self-supervised Learning Framework for Brain Signals
por: Yang, Yifan, et al.
Publicado: (2024)
por: Yang, Yifan, et al.
Publicado: (2024)
Enhancing Diversity in Multi-objective Feature Selection
por: Miyandoab, Sevil Zanjani, et al.
Publicado: (2024)
por: Miyandoab, Sevil Zanjani, et al.
Publicado: (2024)
Parameter-Efficient Neuroevolution for Diverse LLM Generation: Quality-Diversity Optimization via Prompt Embedding Evolution
por: Guo, Dongxin, et al.
Publicado: (2026)
por: Guo, Dongxin, et al.
Publicado: (2026)
Generating DDPM-based Samples from Tilted Distributions
por: Mandal, Himadri, et al.
Publicado: (2026)
por: Mandal, Himadri, et al.
Publicado: (2026)
SAGE: Sign-Adaptive Gradient for Memory-Efficient LLM Optimization
por: Lee, Wooin, et al.
Publicado: (2026)
por: Lee, Wooin, et al.
Publicado: (2026)
Approaching I/O-optimality for Approximate Attention
por: Papp, Pál András, et al.
Publicado: (2026)
por: Papp, Pál András, et al.
Publicado: (2026)
Automatic Depression Assessment using Machine Learning: A Comprehensive Survey
por: Song, Siyang, et al.
Publicado: (2025)
por: Song, Siyang, et al.
Publicado: (2025)
Feature Selection Based on Reinforcement Learning and Hazard State Classification for Magnetic Adhesion Wall-Climbing Robots
por: Ma, Zhen, et al.
Publicado: (2025)
por: Ma, Zhen, et al.
Publicado: (2025)
Functional Similarity Metric for Neural Networks: Overcoming Parametric Ambiguity via Activation Region Analysis
por: Hennadii, Kutomanov
Publicado: (2026)
por: Hennadii, Kutomanov
Publicado: (2026)
Understanding Reinforcement Learning for Model Training, and future directions with GRAPE
por: Patel, Rohit
Publicado: (2025)
por: Patel, Rohit
Publicado: (2025)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
por: Yang, Yibo
Publicado: (2025)
por: Yang, Yibo
Publicado: (2025)
Think Thrice Before You Speak: Dual knowledge-enhanced Theory-of-Mind Reasoning for Persuasive Agents
por: Ma, Minghui, et al.
Publicado: (2026)
por: Ma, Minghui, et al.
Publicado: (2026)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
por: Moradbeiki, Pardis, et al.
Publicado: (2024)
por: Moradbeiki, Pardis, et al.
Publicado: (2024)
Semantically-correlated memories in a dense associative model
por: Burns, Thomas F
Publicado: (2024)
por: Burns, Thomas F
Publicado: (2024)
Inference acceleration for large language models using "stairs" assisted greedy generation
por: Grigaliūnas, Domas, et al.
Publicado: (2024)
por: Grigaliūnas, Domas, et al.
Publicado: (2024)
Harnessing non-adversarial robustness in large language models
por: Zhou, Qinghua, et al.
Publicado: (2026)
por: Zhou, Qinghua, et al.
Publicado: (2026)
DISC: Dynamic Decomposition Improves LLM Inference Scaling
por: Light, Jonathan, et al.
Publicado: (2025)
por: Light, Jonathan, et al.
Publicado: (2025)
Many Objective Problems Where Crossover is Provably Essential
por: Opris, Andre
Publicado: (2024)
por: Opris, Andre
Publicado: (2024)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
por: Menon, Anjali R., et al.
Publicado: (2025)
por: Menon, Anjali R., et al.
Publicado: (2025)
Singularity-aware Optimization via Randomized Geometric Probing: Towards Stable Non-smooth Optimization
por: Xu, Ruoran, et al.
Publicado: (2026)
por: Xu, Ruoran, et al.
Publicado: (2026)
Graded Transformers
por: Shaska Sr, Tony
Publicado: (2025)
por: Shaska Sr, Tony
Publicado: (2025)
Random feature-based double Vovk-Azoury-Warmuth algorithm for online multi-kernel learning
por: Rokhlin, Dmitry B., et al.
Publicado: (2025)
por: Rokhlin, Dmitry B., et al.
Publicado: (2025)
A hierarchical Vovk-Azoury-Warmuth forecaster with discounting for online regression in RKHS
por: Rokhlin, Dmitry B.
Publicado: (2025)
por: Rokhlin, Dmitry B.
Publicado: (2025)
Quantum-Inspired Evolutionary Algorithms for Feature Subset Selection: A Comprehensive Survey
por: Vivek, Yelleti, et al.
Publicado: (2024)
por: Vivek, Yelleti, et al.
Publicado: (2024)
Improved Differential Evolution based Feature Selection through Quantum, Chaos, and Lasso
por: Vivek, Yelleti, et al.
Publicado: (2024)
por: Vivek, Yelleti, et al.
Publicado: (2024)
Mechanistic Analysis of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
por: Imanov, Olaf Yunus Laitinen
Publicado: (2026)
por: Imanov, Olaf Yunus Laitinen
Publicado: (2026)
Ejemplares similares
-
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
por: Mitchell, Rupert, et al.
Publicado: (2025) -
Identifying Autism-Related Neurobiomarkers Using Hybrid Deep Learning Models
por: Chen, Ashley
Publicado: (2025) -
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
por: Danieli, Federico, et al.
Publicado: (2025) -
Stealth edits to large language models
por: Sutton, Oliver J., et al.
Publicado: (2024) -
A Mathematical Framework for the Problem of Security for Cognition in Neurotechnology
por: Bagley, Bryce Allen, et al.
Publicado: (2024)