The Computational Complexity of Circuit Discovery for Inner Interpretability
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Adolfi, Federico, Vilas, Martina G., Wareham, Todd |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Position: An Inner Interpretability Framework for AI Inspired by Lessons from Cognitive Neuroscience
von: Vilas, Martina G., et al.
Veröffentlicht: (2024)
von: Vilas, Martina G., et al.
Veröffentlicht: (2024)
The Algorithmic Regulator
von: Ruffini, Giulio
Veröffentlicht: (2025)
von: Ruffini, Giulio
Veröffentlicht: (2025)
From Clever Hans to Scientific Discovery: Interpreting EEG Foundational Transformers with LRP
von: Bexten, Justus Meyer zu, et al.
Veröffentlicht: (2026)
von: Bexten, Justus Meyer zu, et al.
Veröffentlicht: (2026)
Net2Brain: A Toolbox to compare artificial vision models with human brain responses
von: Bersch, Domenic, et al.
Veröffentlicht: (2022)
von: Bersch, Domenic, et al.
Veröffentlicht: (2022)
An Affective-Taxis Hypothesis for Alignment and Interpretability
von: Sennesh, Eli, et al.
Veröffentlicht: (2025)
von: Sennesh, Eli, et al.
Veröffentlicht: (2025)
Latent-Space Causal Discovery from Indirect Neuroimaging Observations
von: Bae, Sangyoon, et al.
Veröffentlicht: (2026)
von: Bae, Sangyoon, et al.
Veröffentlicht: (2026)
Think-Aloud Reshapes Automated Cognitive Model Discovery Beyond Behavior
von: Xie, Hanbo, et al.
Veröffentlicht: (2026)
von: Xie, Hanbo, et al.
Veröffentlicht: (2026)
Multilevel Interpretability Of Artificial Neural Networks: Leveraging Framework And Methods From Neuroscience
von: He, Zhonghao, et al.
Veröffentlicht: (2024)
von: He, Zhonghao, et al.
Veröffentlicht: (2024)
A Biologically Interpretable Cognitive Architecture for Online Structuring of Episodic Memories into Cognitive Maps
von: Dzhivelikian, E. A., et al.
Veröffentlicht: (2025)
von: Dzhivelikian, E. A., et al.
Veröffentlicht: (2025)
The Computational Mechanisms of Detached Mindfulness
von: Conway-Smith, Brendan, et al.
Veröffentlicht: (2024)
von: Conway-Smith, Brendan, et al.
Veröffentlicht: (2024)
Consciousness qua Mortal Computation
von: Kleiner, Johannes
Veröffentlicht: (2024)
von: Kleiner, Johannes
Veröffentlicht: (2024)
Can Brain Signals Reveal Inner Alignment with Human Languages?
von: Han, William, et al.
Veröffentlicht: (2022)
von: Han, William, et al.
Veröffentlicht: (2022)
DecNefSimulator: A Modular, Interpretable Framework for Decoded Neurofeedback Simulation Using Generative Models
von: Olza, Alexander, et al.
Veröffentlicht: (2025)
von: Olza, Alexander, et al.
Veröffentlicht: (2025)
Predictive Coding Enhances Meta-RL To Achieve Interpretable Bayes-Optimal Belief Representation Under Partial Observability
von: Kuo, Po-Chen, et al.
Veröffentlicht: (2025)
von: Kuo, Po-Chen, et al.
Veröffentlicht: (2025)
Computational Thought Experiments for a More Rigorous Philosophy and Science of the Mind
von: Oved, Iris, et al.
Veröffentlicht: (2024)
von: Oved, Iris, et al.
Veröffentlicht: (2024)
Report on Candidate Computational Indicators for Conscious Valenced Experience
von: Campero, Andres
Veröffentlicht: (2024)
von: Campero, Andres
Veröffentlicht: (2024)
A Reservoir-based Model for Human-like Perception of Complex Rhythm Pattern
von: Yuan, Zhongju, et al.
Veröffentlicht: (2025)
von: Yuan, Zhongju, et al.
Veröffentlicht: (2025)
Disentangling the Factors of Convergence between Brains and Computer Vision Models
von: Raugel, Joséphine, et al.
Veröffentlicht: (2025)
von: Raugel, Joséphine, et al.
Veröffentlicht: (2025)
Impact of Neuron Models on Spiking Neural Networks performance. A Complexity Based Classification Approach
von: Rudnicka, Zofia, et al.
Veröffentlicht: (2025)
von: Rudnicka, Zofia, et al.
Veröffentlicht: (2025)
Grounded Computation & Consciousness: A Framework for Exploring Consciousness in Machines & Other Organisms
von: Williams, Ryan
Veröffentlicht: (2024)
von: Williams, Ryan
Veröffentlicht: (2024)
Less is More: some Computational Principles based on Parcimony, and Limitations of Natural Intelligence
von: Cohen, Laura, et al.
Veröffentlicht: (2025)
von: Cohen, Laura, et al.
Veröffentlicht: (2025)
A Spiking Neural Network based on Neural Manifold for Augmenting Intracortical Brain-Computer Interface Data
von: Zheng, Shengjie, et al.
Veröffentlicht: (2022)
von: Zheng, Shengjie, et al.
Veröffentlicht: (2022)
Naturalistic Computational Cognitive Science: Towards generalizable models and theories that capture the full range of natural behavior
von: Carvalho, Wilka, et al.
Veröffentlicht: (2025)
von: Carvalho, Wilka, et al.
Veröffentlicht: (2025)
Understanding Human Limits in Pattern Recognition: A Computational Model of Sequential Reasoning in Rock, Paper, Scissors
von: Cross, Logan, et al.
Veröffentlicht: (2025)
von: Cross, Logan, et al.
Veröffentlicht: (2025)
A Unified Cortical Circuit Model with Divisive Normalization and Self-Excitation for Robust Representation and Memory Maintenance
von: Su, Jie, et al.
Veröffentlicht: (2025)
von: Su, Jie, et al.
Veröffentlicht: (2025)
Computing with Canonical Microcircuits
von: Douglas, PK
Veröffentlicht: (2025)
von: Douglas, PK
Veröffentlicht: (2025)
Crafting Interpretable Embeddings by Asking LLMs Questions
von: Benara, Vinamra, et al.
Veröffentlicht: (2024)
von: Benara, Vinamra, et al.
Veröffentlicht: (2024)
Interpretable Dual-Filter Fuzzy Neural Networks for Affective Brain-Computer Interfaces
von: Jiang, Xiaowei, et al.
Veröffentlicht: (2025)
von: Jiang, Xiaowei, et al.
Veröffentlicht: (2025)
Continual Developmental Neurosimulation Using Embodied Computational Agents
von: Alicea, Bradly, et al.
Veröffentlicht: (2021)
von: Alicea, Bradly, et al.
Veröffentlicht: (2021)
Language Models Are Capable of Metacognitive Monitoring and Control of Their Internal Activations
von: Ji-An, Li, et al.
Veröffentlicht: (2025)
von: Ji-An, Li, et al.
Veröffentlicht: (2025)
Learning Internal Biological Neuron Parameters and Complexity-Based Encoding for Improved Spiking Neural Networks Performance
von: Rudnicka, Zofia, et al.
Veröffentlicht: (2025)
von: Rudnicka, Zofia, et al.
Veröffentlicht: (2025)
Representational Similarity via Interpretable Visual Concepts
von: Kondapaneni, Neehar, et al.
Veröffentlicht: (2025)
von: Kondapaneni, Neehar, et al.
Veröffentlicht: (2025)
Simplicity in Complexity : Explaining Visual Complexity using Deep Segmentation Models
von: Shen, Tingke, et al.
Veröffentlicht: (2024)
von: Shen, Tingke, et al.
Veröffentlicht: (2024)
ELiSe: Efficient Learning of Sequences in Structured Recurrent Networks
von: Kriener, Laura, et al.
Veröffentlicht: (2024)
von: Kriener, Laura, et al.
Veröffentlicht: (2024)
Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners
von: Csaba, Botos, et al.
Veröffentlicht: (2026)
von: Csaba, Botos, et al.
Veröffentlicht: (2026)
NeuroAI and Beyond: Bridging Between Advances in Neuroscience and ArtificialIntelligence
von: Zador, Anthony, et al.
Veröffentlicht: (2026)
von: Zador, Anthony, et al.
Veröffentlicht: (2026)
OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens
von: Willeke, Konstantin F., et al.
Veröffentlicht: (2026)
von: Willeke, Konstantin F., et al.
Veröffentlicht: (2026)
The age of spiritual machines: Language quietus induces synthetic altered states of consciousness in artificial intelligence
von: Skipper, Jeremy I, et al.
Veröffentlicht: (2024)
von: Skipper, Jeremy I, et al.
Veröffentlicht: (2024)
Neural Erosion: Emulating Controlled Neurodegeneration and Aging in AI Systems
von: Alexos, Antonios, et al.
Veröffentlicht: (2024)
von: Alexos, Antonios, et al.
Veröffentlicht: (2024)
The neural correlates of logical-mathematical symbol systems processing resemble that of spatial cognition more than natural language processing
von: Li, Yuannan, et al.
Veröffentlicht: (2024)
von: Li, Yuannan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Position: An Inner Interpretability Framework for AI Inspired by Lessons from Cognitive Neuroscience
von: Vilas, Martina G., et al.
Veröffentlicht: (2024) -
The Algorithmic Regulator
von: Ruffini, Giulio
Veröffentlicht: (2025) -
From Clever Hans to Scientific Discovery: Interpreting EEG Foundational Transformers with LRP
von: Bexten, Justus Meyer zu, et al.
Veröffentlicht: (2026) -
Net2Brain: A Toolbox to compare artificial vision models with human brain responses
von: Bersch, Domenic, et al.
Veröffentlicht: (2022) -
An Affective-Taxis Hypothesis for Alignment and Interpretability
von: Sennesh, Eli, et al.
Veröffentlicht: (2025)