Active Inference Tree Search in Large POMDPs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Maisto, Domenico, Gregoretti, Francesco, Friston, Karl, Pezzulo, Giovanni |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
von: Mitchell, Rupert, et al.
Veröffentlicht: (2025)
von: Mitchell, Rupert, et al.
Veröffentlicht: (2025)
Identifying Autism-Related Neurobiomarkers Using Hybrid Deep Learning Models
von: Chen, Ashley
Veröffentlicht: (2025)
von: Chen, Ashley
Veröffentlicht: (2025)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
von: Danieli, Federico, et al.
Veröffentlicht: (2025)
von: Danieli, Federico, et al.
Veröffentlicht: (2025)
Stealth edits to large language models
von: Sutton, Oliver J., et al.
Veröffentlicht: (2024)
von: Sutton, Oliver J., et al.
Veröffentlicht: (2024)
A Mathematical Framework for the Problem of Security for Cognition in Neurotechnology
von: Bagley, Bryce Allen, et al.
Veröffentlicht: (2024)
von: Bagley, Bryce Allen, et al.
Veröffentlicht: (2024)
Contrastive Consolidation of Top-Down Modulations Achieves Sparsely Supervised Continual Learning
von: Tran, Viet Anh Khoa, et al.
Veröffentlicht: (2025)
von: Tran, Viet Anh Khoa, et al.
Veröffentlicht: (2025)
Shortest Paths without a Map, but with an Entropic Regularizer
von: Bubeck, Sébastien, et al.
Veröffentlicht: (2022)
von: Bubeck, Sébastien, et al.
Veröffentlicht: (2022)
Towards Single Exponential Time for Temporal and Spatial Reasoning: A Study via Redundancy and Dynamic Programming
von: Lagerkvist, Victor, et al.
Veröffentlicht: (2026)
von: Lagerkvist, Victor, et al.
Veröffentlicht: (2026)
Boosting Test Performance with Importance Sampling--a Subpopulation Perspective
von: Shen, Hongyu, et al.
Veröffentlicht: (2024)
von: Shen, Hongyu, et al.
Veröffentlicht: (2024)
Platonic Representations in the Human Brain: Unsupervised Recovery of Universal Geometry
von: Marcos-Manchón, Pablo, et al.
Veröffentlicht: (2026)
von: Marcos-Manchón, Pablo, et al.
Veröffentlicht: (2026)
Embedding Explainable AI in NHS Clinical Safety: The Explainability-Enabled Clinical Safety Framework (ECSF)
von: Gigiu, Robert
Veröffentlicht: (2025)
von: Gigiu, Robert
Veröffentlicht: (2025)
Safe Distributed Control of Multi-Robot Systems with Communication Delays
von: Ballotta, Luca, et al.
Veröffentlicht: (2024)
von: Ballotta, Luca, et al.
Veröffentlicht: (2024)
Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data
von: Pasini, Massimiliano Lupo, et al.
Veröffentlicht: (2026)
von: Pasini, Massimiliano Lupo, et al.
Veröffentlicht: (2026)
Untrained CNNs Match Backpropagation at V1: A Systematic RSA Comparison of Four Learning Rules Against Human fMRI
von: Leutenegger, Nils
Veröffentlicht: (2026)
von: Leutenegger, Nils
Veröffentlicht: (2026)
BrainMAE: A Region-aware Self-supervised Learning Framework for Brain Signals
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
Enhancing Diversity in Multi-objective Feature Selection
von: Miyandoab, Sevil Zanjani, et al.
Veröffentlicht: (2024)
von: Miyandoab, Sevil Zanjani, et al.
Veröffentlicht: (2024)
Parameter-Efficient Neuroevolution for Diverse LLM Generation: Quality-Diversity Optimization via Prompt Embedding Evolution
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Generating DDPM-based Samples from Tilted Distributions
von: Mandal, Himadri, et al.
Veröffentlicht: (2026)
von: Mandal, Himadri, et al.
Veröffentlicht: (2026)
SAGE: Sign-Adaptive Gradient for Memory-Efficient LLM Optimization
von: Lee, Wooin, et al.
Veröffentlicht: (2026)
von: Lee, Wooin, et al.
Veröffentlicht: (2026)
Approaching I/O-optimality for Approximate Attention
von: Papp, Pál András, et al.
Veröffentlicht: (2026)
von: Papp, Pál András, et al.
Veröffentlicht: (2026)
Automatic Depression Assessment using Machine Learning: A Comprehensive Survey
von: Song, Siyang, et al.
Veröffentlicht: (2025)
von: Song, Siyang, et al.
Veröffentlicht: (2025)
Feature Selection Based on Reinforcement Learning and Hazard State Classification for Magnetic Adhesion Wall-Climbing Robots
von: Ma, Zhen, et al.
Veröffentlicht: (2025)
von: Ma, Zhen, et al.
Veröffentlicht: (2025)
Functional Similarity Metric for Neural Networks: Overcoming Parametric Ambiguity via Activation Region Analysis
von: Hennadii, Kutomanov
Veröffentlicht: (2026)
von: Hennadii, Kutomanov
Veröffentlicht: (2026)
Understanding Reinforcement Learning for Model Training, and future directions with GRAPE
von: Patel, Rohit
Veröffentlicht: (2025)
von: Patel, Rohit
Veröffentlicht: (2025)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
von: Yang, Yibo
Veröffentlicht: (2025)
von: Yang, Yibo
Veröffentlicht: (2025)
Think Thrice Before You Speak: Dual knowledge-enhanced Theory-of-Mind Reasoning for Persuasive Agents
von: Ma, Minghui, et al.
Veröffentlicht: (2026)
von: Ma, Minghui, et al.
Veröffentlicht: (2026)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
von: Moradbeiki, Pardis, et al.
Veröffentlicht: (2024)
von: Moradbeiki, Pardis, et al.
Veröffentlicht: (2024)
Semantically-correlated memories in a dense associative model
von: Burns, Thomas F
Veröffentlicht: (2024)
von: Burns, Thomas F
Veröffentlicht: (2024)
Inference acceleration for large language models using "stairs" assisted greedy generation
von: Grigaliūnas, Domas, et al.
Veröffentlicht: (2024)
von: Grigaliūnas, Domas, et al.
Veröffentlicht: (2024)
Harnessing non-adversarial robustness in large language models
von: Zhou, Qinghua, et al.
Veröffentlicht: (2026)
von: Zhou, Qinghua, et al.
Veröffentlicht: (2026)
DISC: Dynamic Decomposition Improves LLM Inference Scaling
von: Light, Jonathan, et al.
Veröffentlicht: (2025)
von: Light, Jonathan, et al.
Veröffentlicht: (2025)
Many Objective Problems Where Crossover is Provably Essential
von: Opris, Andre
Veröffentlicht: (2024)
von: Opris, Andre
Veröffentlicht: (2024)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
von: Menon, Anjali R., et al.
Veröffentlicht: (2025)
von: Menon, Anjali R., et al.
Veröffentlicht: (2025)
Singularity-aware Optimization via Randomized Geometric Probing: Towards Stable Non-smooth Optimization
von: Xu, Ruoran, et al.
Veröffentlicht: (2026)
von: Xu, Ruoran, et al.
Veröffentlicht: (2026)
Graded Transformers
von: Shaska Sr, Tony
Veröffentlicht: (2025)
von: Shaska Sr, Tony
Veröffentlicht: (2025)
Random feature-based double Vovk-Azoury-Warmuth algorithm for online multi-kernel learning
von: Rokhlin, Dmitry B., et al.
Veröffentlicht: (2025)
von: Rokhlin, Dmitry B., et al.
Veröffentlicht: (2025)
A hierarchical Vovk-Azoury-Warmuth forecaster with discounting for online regression in RKHS
von: Rokhlin, Dmitry B.
Veröffentlicht: (2025)
von: Rokhlin, Dmitry B.
Veröffentlicht: (2025)
Quantum-Inspired Evolutionary Algorithms for Feature Subset Selection: A Comprehensive Survey
von: Vivek, Yelleti, et al.
Veröffentlicht: (2024)
von: Vivek, Yelleti, et al.
Veröffentlicht: (2024)
Improved Differential Evolution based Feature Selection through Quantum, Chaos, and Lasso
von: Vivek, Yelleti, et al.
Veröffentlicht: (2024)
von: Vivek, Yelleti, et al.
Veröffentlicht: (2024)
Mechanistic Analysis of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
von: Imanov, Olaf Yunus Laitinen
Veröffentlicht: (2026)
von: Imanov, Olaf Yunus Laitinen
Veröffentlicht: (2026)
Ähnliche Einträge
-
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
von: Mitchell, Rupert, et al.
Veröffentlicht: (2025) -
Identifying Autism-Related Neurobiomarkers Using Hybrid Deep Learning Models
von: Chen, Ashley
Veröffentlicht: (2025) -
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
von: Danieli, Federico, et al.
Veröffentlicht: (2025) -
Stealth edits to large language models
von: Sutton, Oliver J., et al.
Veröffentlicht: (2024) -
A Mathematical Framework for the Problem of Security for Cognition in Neurotechnology
von: Bagley, Bryce Allen, et al.
Veröffentlicht: (2024)