Position: An Inner Interpretability Framework for AI Inspired by Lessons from Cognitive Neuroscience
Fuente:
arXiv
Saved in:
| Main Authors: | Vilas, Martina G., Adolfi, Federico, Poeppel, David, Roig, Gemma |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Computational Complexity of Circuit Discovery for Inner Interpretability
by: Adolfi, Federico, et al.
Published: (2024)
by: Adolfi, Federico, et al.
Published: (2024)
Multilevel Interpretability Of Artificial Neural Networks: Leveraging Framework And Methods From Neuroscience
by: He, Zhonghao, et al.
Published: (2024)
by: He, Zhonghao, et al.
Published: (2024)
Lessons from Neuroscience for AI: How integrating Actions, Compositional Structure and Episodic Memory could enable Safe, Interpretable and Human-Like AI
by: Rao, Rajesh P. N., et al.
Published: (2025)
by: Rao, Rajesh P. N., et al.
Published: (2025)
Disentangled Representations for Causal Cognition
by: Torresan, Filippo, et al.
Published: (2024)
by: Torresan, Filippo, et al.
Published: (2024)
Cortex-Inspired Continual Learning: Unsupervised Instantiation and Recovery of Functional Task Networks
by: McKee, Kevin, et al.
Published: (2026)
by: McKee, Kevin, et al.
Published: (2026)
Net2Brain: A Toolbox to compare artificial vision models with human brain responses
by: Bersch, Domenic, et al.
Published: (2022)
by: Bersch, Domenic, et al.
Published: (2022)
BrainNet-MoE: Brain-Inspired Mixture-of-Experts Learning for Neurological Disease Identification
by: Zhang, Jing, et al.
Published: (2025)
by: Zhang, Jing, et al.
Published: (2025)
Stimulus-Voltage-Based Prediction of Action Potential Onset Timing: Classical vs. Quantum-Inspired Approaches
by: Johnson, Stevens, et al.
Published: (2025)
by: Johnson, Stevens, et al.
Published: (2025)
Explore Brain-Inspired Machine Intelligence for Connecting Dots on Graphs Through Holographic Blueprint of Oscillatory Synchronization
by: Dan, Tingting, et al.
Published: (2026)
by: Dan, Tingting, et al.
Published: (2026)
Triple Phase Transitions: Understanding the Learning Dynamics of Large Language Models from a Neuroscience Perspective
by: Nakagi, Yuko, et al.
Published: (2025)
by: Nakagi, Yuko, et al.
Published: (2025)
Mirror-Neuron Patterns in AI Alignment
by: Wyrick, Robyn
Published: (2025)
by: Wyrick, Robyn
Published: (2025)
Bridging Neuroscience and AI: Environmental Enrichment as a Model for Forward Knowledge Transfer
by: Saxena, Rajat, et al.
Published: (2024)
by: Saxena, Rajat, et al.
Published: (2024)
Towards Neurocognitive-Inspired Intelligence: From AI's Structural Mimicry to Human-Like Functional Cognition
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2025)
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2025)
Can Brain Signals Reveal Inner Alignment with Human Languages?
by: Han, William, et al.
Published: (2022)
by: Han, William, et al.
Published: (2022)
Systems Explaining Systems: A Framework for Intelligence and Consciousness
by: Semmler, Sean Niklas
Published: (2026)
by: Semmler, Sean Niklas
Published: (2026)
QF: Quick Feedforward AI Model Training without Gradient Back Propagation
by: Qi, Feng
Published: (2025)
by: Qi, Feng
Published: (2025)
DSAM: A Deep Learning Framework for Analyzing Temporal and Spatial Dynamics in Brain Networks
by: Thapaliya, Bishal, et al.
Published: (2024)
by: Thapaliya, Bishal, et al.
Published: (2024)
Bridging Foundation Models and Efficient Architectures: A Modular Brain Imaging Framework with Local Masking and Pretrained Representation Learning
by: Wang, Yanwen, et al.
Published: (2025)
by: Wang, Yanwen, et al.
Published: (2025)
Brian Intensify: An Adaptive Machine Learning Framework for Auditory EEG Stimulation and Cognitive Enhancement in FXS
by: ElSayed, Zag, et al.
Published: (2025)
by: ElSayed, Zag, et al.
Published: (2025)
Contrastive Self-Supervised Learning As Neural Manifold Packing
by: Zhang, Guanming, et al.
Published: (2025)
by: Zhang, Guanming, et al.
Published: (2025)
Shifting Attention to You: Personalized Brain-Inspired AI Models
by: Zhao, Stephen Chong, et al.
Published: (2025)
by: Zhao, Stephen Chong, et al.
Published: (2025)
Playing With Neuroscience: Past, Present and Future of Neuroimaging and Games
by: Burelli, Paolo, et al.
Published: (2024)
by: Burelli, Paolo, et al.
Published: (2024)
Crafting Interpretable Embeddings by Asking LLMs Questions
by: Benara, Vinamra, et al.
Published: (2024)
by: Benara, Vinamra, et al.
Published: (2024)
BRAID: Input-Driven Nonlinear Dynamical Modeling of Neural-Behavioral Data
by: Vahidi, Parsa, et al.
Published: (2025)
by: Vahidi, Parsa, et al.
Published: (2025)
Leakage and Second-Order Dynamics Improve Hippocampal RNN Replay
by: Casco-Rodriguez, Josue, et al.
Published: (2026)
by: Casco-Rodriguez, Josue, et al.
Published: (2026)
A Biologically Interpretable Cognitive Architecture for Online Structuring of Episodic Memories into Cognitive Maps
by: Dzhivelikian, E. A., et al.
Published: (2025)
by: Dzhivelikian, E. A., et al.
Published: (2025)
JEDI: Jointly Embedded Inference of Neural Dynamics
by: Jamkhandi, Anirudh, et al.
Published: (2026)
by: Jamkhandi, Anirudh, et al.
Published: (2026)
Interpretable Fusion Analytics Framework for fMRI Connectivity: Self-Attention Mechanism and Latent Space Item-Response Model
by: Kim, Jeong-Jae, et al.
Published: (2022)
by: Kim, Jeong-Jae, et al.
Published: (2022)
The Geometry of Concepts: Sparse Autoencoder Feature Structure
by: Li, Yuxiao, et al.
Published: (2024)
by: Li, Yuxiao, et al.
Published: (2024)
Nature's Insight: A Novel Framework and Comprehensive Analysis of Agentic Reasoning Through the Lens of Neuroscience
by: Liu, Zinan, et al.
Published: (2025)
by: Liu, Zinan, et al.
Published: (2025)
Bayesian Network Modeling of Causal Influence within Cognitive Domains and Clinical Dementia Severity Ratings for Western and Indian Cohorts
by: Kumar, Wupadrasta Santosh, et al.
Published: (2024)
by: Kumar, Wupadrasta Santosh, et al.
Published: (2024)
Defining and Benchmarking a Data-Centric Design Space for Brain Graph Construction
by: Ge, Qinwen, et al.
Published: (2025)
by: Ge, Qinwen, et al.
Published: (2025)
NeuroGraph: Benchmarks for Graph Machine Learning in Brain Connectomics
by: Said, Anwar, et al.
Published: (2023)
by: Said, Anwar, et al.
Published: (2023)
CTM-AI: A Blueprint for General AI Inspired by a Model of Consciousness
by: Yu, Haofei, et al.
Published: (2026)
by: Yu, Haofei, et al.
Published: (2026)
Fast weight programming and linear transformers: from machine learning to neurobiology
by: Irie, Kazuki, et al.
Published: (2025)
by: Irie, Kazuki, et al.
Published: (2025)
Complex behavior from intrinsic motivation to occupy action-state path space
by: Ramírez-Ruiz, Jorge, et al.
Published: (2022)
by: Ramírez-Ruiz, Jorge, et al.
Published: (2022)
Revealing Neurocognitive and Behavioral Patterns by Unsupervised Manifold Learning from Dynamic Brain Data
by: Zhou, Zixia, et al.
Published: (2025)
by: Zhou, Zixia, et al.
Published: (2025)
Rethinking Over-Smoothing in Graph Neural Networks: A Perspective from Anderson Localization
by: Ouyang, Kaichen
Published: (2025)
by: Ouyang, Kaichen
Published: (2025)
Brain alignment of reasoning and action representations from vision-language and action models during naturalistic gameplay
by: Oota, Subba Reddy, et al.
Published: (2026)
by: Oota, Subba Reddy, et al.
Published: (2026)
Most discriminative stimuli for functional cell type clustering
by: Burg, Max F., et al.
Published: (2023)
by: Burg, Max F., et al.
Published: (2023)
Similar Items
-
The Computational Complexity of Circuit Discovery for Inner Interpretability
by: Adolfi, Federico, et al.
Published: (2024) -
Multilevel Interpretability Of Artificial Neural Networks: Leveraging Framework And Methods From Neuroscience
by: He, Zhonghao, et al.
Published: (2024) -
Lessons from Neuroscience for AI: How integrating Actions, Compositional Structure and Episodic Memory could enable Safe, Interpretable and Human-Like AI
by: Rao, Rajesh P. N., et al.
Published: (2025) -
Disentangled Representations for Causal Cognition
by: Torresan, Filippo, et al.
Published: (2024) -
Cortex-Inspired Continual Learning: Unsupervised Instantiation and Recovery of Functional Task Networks
by: McKee, Kevin, et al.
Published: (2026)