Fine-Tuning Language Models to Know What They Know
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Sangjun, Meyerson, Elliot, Qiu, Xin, Miikkulainen, Risto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How the Stroop Effect Arises from Optimal Response Times in Laterally Connected Self-Organizing Maps
von: Prabhakaran, Divya, et al.
Veröffentlicht: (2025)
von: Prabhakaran, Divya, et al.
Veröffentlicht: (2025)
Decoding the decoder: Contextual sequence-to-sequence modeling for intracortical speech decoding
von: Olak, Michal, et al.
Veröffentlicht: (2026)
von: Olak, Michal, et al.
Veröffentlicht: (2026)
The brain-AI convergence: Predictive and generative world models for general-purpose computation
von: Ohmae, Shogo, et al.
Veröffentlicht: (2025)
von: Ohmae, Shogo, et al.
Veröffentlicht: (2025)
Can "consciousness" be observed from large language model (LLM) internal states? Dissecting LLM representations obtained from Theory of Mind test with Integrated Information Theory and Span Representation analysis
von: Li, Jingkai
Veröffentlicht: (2025)
von: Li, Jingkai
Veröffentlicht: (2025)
A Concept-Value Network as a Brain Model
von: Greer, Kieran
Veröffentlicht: (2019)
von: Greer, Kieran
Veröffentlicht: (2019)
Decoding Cortical Microcircuits: A Generative Model for Latent Space Exploration and Controlled Synthesis
von: Liu, Xingyu, et al.
Veröffentlicht: (2025)
von: Liu, Xingyu, et al.
Veröffentlicht: (2025)
A Unified Cortical Circuit Model with Divisive Normalization and Self-Excitation for Robust Representation and Memory Maintenance
von: Su, Jie, et al.
Veröffentlicht: (2025)
von: Su, Jie, et al.
Veröffentlicht: (2025)
PaceLLM: Brain-Inspired Large Language Models for Long-Context Understanding
von: Li, Kangcong, et al.
Veröffentlicht: (2025)
von: Li, Kangcong, et al.
Veröffentlicht: (2025)
Modularity is the Bedrock of Natural and Artificial Intelligence
von: Salatiello, Alessandro
Veröffentlicht: (2026)
von: Salatiello, Alessandro
Veröffentlicht: (2026)
Self-organized MT Direction Maps Emerge from Spatiotemporal Contrastive Optimization
von: Gu, Zhaotian, et al.
Veröffentlicht: (2026)
von: Gu, Zhaotian, et al.
Veröffentlicht: (2026)
The role of neuromorphic principles in the future of biomedicine and healthcare
von: Hwang, Grace M., et al.
Veröffentlicht: (2026)
von: Hwang, Grace M., et al.
Veröffentlicht: (2026)
The Fast Lane Hypothesis: Von Economo Neurons Implement a Biological Speed-Accuracy Tradeoff
von: Keskin, Esila
Veröffentlicht: (2026)
von: Keskin, Esila
Veröffentlicht: (2026)
Adaptive Spiking with Plasticity for Energy Aware Neuromorphic Systems
von: Calle-Ortiz, Eduardo, et al.
Veröffentlicht: (2025)
von: Calle-Ortiz, Eduardo, et al.
Veröffentlicht: (2025)
NSPDI-SNN: An efficient lightweight SNN based on nonlinear synaptic pruning and dendritic integration
von: Cai, Wuque, et al.
Veröffentlicht: (2025)
von: Cai, Wuque, et al.
Veröffentlicht: (2025)
Evolving Cognitive Architectures
von: Serov, Alexander
Veröffentlicht: (2025)
von: Serov, Alexander
Veröffentlicht: (2025)
Digital twin brain: a bridge between biological intelligence and artificial intelligence
von: Xiong, Hui, et al.
Veröffentlicht: (2023)
von: Xiong, Hui, et al.
Veröffentlicht: (2023)
Implementing engrams from a machine learning perspective: XOR as a basic motif
von: de Lucas, Jesus Marco, et al.
Veröffentlicht: (2024)
von: de Lucas, Jesus Marco, et al.
Veröffentlicht: (2024)
Associative memory and dead neurons
von: Fanaskov, Vladimir, et al.
Veröffentlicht: (2024)
von: Fanaskov, Vladimir, et al.
Veröffentlicht: (2024)
Neuropsychology and Explainability of AI: A Distributional Approach to the Relationship Between Activation Similarity of Neural Categories in Synthetic Cognition
von: Pichat, Michael, et al.
Veröffentlicht: (2024)
von: Pichat, Michael, et al.
Veröffentlicht: (2024)
Biological Processing Units: Leveraging an Insect Connectome to Pioneer Biofidelic Neural Architectures
von: Yu, Siyu, et al.
Veröffentlicht: (2025)
von: Yu, Siyu, et al.
Veröffentlicht: (2025)
Computing with Canonical Microcircuits
von: Douglas, PK
Veröffentlicht: (2025)
von: Douglas, PK
Veröffentlicht: (2025)
NeuroTree: Hierarchical Functional Brain Pathway Decoding for Mental Health Disorders
von: Ding, Jun-En, et al.
Veröffentlicht: (2025)
von: Ding, Jun-En, et al.
Veröffentlicht: (2025)
Phase codes emerge in recurrent neural networks optimized for modular arithmetic
von: Murray, Keith T.
Veröffentlicht: (2023)
von: Murray, Keith T.
Veröffentlicht: (2023)
Neural Manifolds and Cognitive Consistency: A New Approach to Memory Consolidation in Artificial Systems
von: Nguyen, Phuong-Nam
Veröffentlicht: (2025)
von: Nguyen, Phuong-Nam
Veröffentlicht: (2025)
Multiscale fusion enhanced spiking neural network for invasive BCI neural signal decoding
von: Song, Yu, et al.
Veröffentlicht: (2024)
von: Song, Yu, et al.
Veröffentlicht: (2024)
ELiSe: Efficient Learning of Sequences in Structured Recurrent Networks
von: Kriener, Laura, et al.
Veröffentlicht: (2024)
von: Kriener, Laura, et al.
Veröffentlicht: (2024)
A Grid Cell-Inspired Structured Vector Algebra for Cognitive Maps
von: Krausse, Sven, et al.
Veröffentlicht: (2025)
von: Krausse, Sven, et al.
Veröffentlicht: (2025)
Continual Developmental Neurosimulation Using Embodied Computational Agents
von: Alicea, Bradly, et al.
Veröffentlicht: (2021)
von: Alicea, Bradly, et al.
Veröffentlicht: (2021)
Unambiguous Representations in Neural Networks: An Information-Theoretic Approach to Intentionality
von: Lässig, Francesco
Veröffentlicht: (2025)
von: Lässig, Francesco
Veröffentlicht: (2025)
Learning Internal Biological Neuron Parameters and Complexity-Based Encoding for Improved Spiking Neural Networks Performance
von: Rudnicka, Zofia, et al.
Veröffentlicht: (2025)
von: Rudnicka, Zofia, et al.
Veröffentlicht: (2025)
A Bio-Inspired Research Paradigm of Collision Perception Neurons Enabling Neuro-Robotic Integration: The LGMD Case
von: Qin, Ziyan, et al.
Veröffentlicht: (2025)
von: Qin, Ziyan, et al.
Veröffentlicht: (2025)
Exploring Biologically Inspired Mechanisms of Adversarial Robustness
von: Holzhausen, Konstantin, et al.
Veröffentlicht: (2024)
von: Holzhausen, Konstantin, et al.
Veröffentlicht: (2024)
Can Biologically Plausible Temporal Credit Assignment Rules Match BPTT for Neural Similarity? E-prop as an Example
von: Liu, Yuhan Helena, et al.
Veröffentlicht: (2025)
von: Liu, Yuhan Helena, et al.
Veröffentlicht: (2025)
Neuron Platonic Intrinsic Representation From Dynamics Using Contrastive Learning
von: Wu, Wei, et al.
Veröffentlicht: (2025)
von: Wu, Wei, et al.
Veröffentlicht: (2025)
How connectivity structure shapes rich and lazy learning in neural circuits
von: Liu, Yuhan Helena, et al.
Veröffentlicht: (2023)
von: Liu, Yuhan Helena, et al.
Veröffentlicht: (2023)
A differentiable brain simulator bridging brain simulation and brain-inspired computing
von: Wang, Chaoming, et al.
Veröffentlicht: (2023)
von: Wang, Chaoming, et al.
Veröffentlicht: (2023)
Competitive plasticity to reduce the energetic costs of learning
von: van Rossum, Mark CW
Veröffentlicht: (2023)
von: van Rossum, Mark CW
Veröffentlicht: (2023)
How Do Artificial Intelligences Think? The Three Mathematico-Cognitive Factors of Categorical Segmentation Operated by Synthetic Neurons
von: Pichat, Michael, et al.
Veröffentlicht: (2024)
von: Pichat, Michael, et al.
Veröffentlicht: (2024)
A Practical Guide to Tuning Spiking Neuronal Dynamics
von: Gebhardt, William, et al.
Veröffentlicht: (2025)
von: Gebhardt, William, et al.
Veröffentlicht: (2025)
Hierarchical temporal receptive windows and zero-shot timescale generalization in biologically constrained scale-invariant deep networks
von: Sarkar, Aakash, et al.
Veröffentlicht: (2026)
von: Sarkar, Aakash, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
How the Stroop Effect Arises from Optimal Response Times in Laterally Connected Self-Organizing Maps
von: Prabhakaran, Divya, et al.
Veröffentlicht: (2025) -
Decoding the decoder: Contextual sequence-to-sequence modeling for intracortical speech decoding
von: Olak, Michal, et al.
Veröffentlicht: (2026) -
The brain-AI convergence: Predictive and generative world models for general-purpose computation
von: Ohmae, Shogo, et al.
Veröffentlicht: (2025) -
Can "consciousness" be observed from large language model (LLM) internal states? Dissecting LLM representations obtained from Theory of Mind test with Integrated Information Theory and Span Representation analysis
von: Li, Jingkai
Veröffentlicht: (2025) -
A Concept-Value Network as a Brain Model
von: Greer, Kieran
Veröffentlicht: (2019)