RECALL: Library-Like Behavior In Language Models is Enhanced by Self-Referencing Causal Cycles
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nwadike, Munachiso, Iklassov, Zangir, Aremu, Toluwani, Hiraoka, Tatsuya, Bojkovic, Velibor, Heinzerling, Benjamin, Alqaubeh, Hilal, Takáč, Martin, Inui, Kentaro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The AI Data Scientist
von: Akimov, Farkhad, et al.
Veröffentlicht: (2025)
von: Akimov, Farkhad, et al.
Veröffentlicht: (2025)
Measuring AI Reasoning: A Guide for Researchers
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026)
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026)
Mechanistic Insights into Grokking from the Embedding Layer
von: AlquBoj, H. V., et al.
Veröffentlicht: (2025)
von: AlquBoj, H. V., et al.
Veröffentlicht: (2025)
Number Representations in LLMs: A Computational Parallel to Human Perception
von: AlquBoj, H. V., et al.
Veröffentlicht: (2025)
von: AlquBoj, H. V., et al.
Veröffentlicht: (2025)
Sycophancy Hides Linearly in the Attention Heads
von: Genadi, Rifo, et al.
Veröffentlicht: (2026)
von: Genadi, Rifo, et al.
Veröffentlicht: (2026)
Uncovering the Spectral Bias in Diagonal State Space Models
von: Solozabal, Ruben, et al.
Veröffentlicht: (2025)
von: Solozabal, Ruben, et al.
Veröffentlicht: (2025)
WaveSSM: Multiscale State-Space Models for Non-stationary Signal Attention
von: Solozabal, Ruben, et al.
Veröffentlicht: (2026)
von: Solozabal, Ruben, et al.
Veröffentlicht: (2026)
Emergence of Primacy and Recency Effect in Mamba: A Mechanistic Point of View
von: Airlangga, Muhammad Cendekia, et al.
Veröffentlicht: (2025)
von: Airlangga, Muhammad Cendekia, et al.
Veröffentlicht: (2025)
Self-Guiding Exploration for Combinatorial Problems
von: Iklassov, Zangir, et al.
Veröffentlicht: (2024)
von: Iklassov, Zangir, et al.
Veröffentlicht: (2024)
The Geometry of Numerical Reasoning: Language Models Compare Numeric Properties in Linear Subspaces
von: El-Shangiti, Ahmed Oumar, et al.
Veröffentlicht: (2024)
von: El-Shangiti, Ahmed Oumar, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Solving Stochastic Vehicle Routing Problem with Time Windows
von: Iklassov, Zangir, et al.
Veröffentlicht: (2024)
von: Iklassov, Zangir, et al.
Veröffentlicht: (2024)
Spelling-out is not Straightforward: LLMs' Capability of Tokenization from Token to Characters
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2025)
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2025)
Repetition Neurons: How Do Language Models Produce Repetitions?
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2024)
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2024)
Monotonic Representation of Numeric Properties in Language Models
von: Heinzerling, Benjamin, et al.
Veröffentlicht: (2024)
von: Heinzerling, Benjamin, et al.
Veröffentlicht: (2024)
Mitigating Watermark Forgery in Generative Models via Randomized Key Selection
von: Aremu, Toluwani, et al.
Veröffentlicht: (2025)
von: Aremu, Toluwani, et al.
Veröffentlicht: (2025)
SVRPBench: A Realistic Benchmark for Stochastic Vehicle Routing Problem
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
Cell-Based Representation of Relational Binding in Language Models
von: Dai, Qin, et al.
Veröffentlicht: (2026)
von: Dai, Qin, et al.
Veröffentlicht: (2026)
Representational Analysis of Binding in Language Models
von: Dai, Qin, et al.
Veröffentlicht: (2024)
von: Dai, Qin, et al.
Veröffentlicht: (2024)
Understanding and Controlling Repetition Neurons and Induction Heads in In-Context Learning
von: Doan, Nhi Hoai, et al.
Veröffentlicht: (2025)
von: Doan, Nhi Hoai, et al.
Veröffentlicht: (2025)
Improving Personalisation in Valence and Arousal Prediction using Data Augmentation
von: Nwadike, Munachiso, et al.
Veröffentlicht: (2024)
von: Nwadike, Munachiso, et al.
Veröffentlicht: (2024)
TopK Language Models
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2025)
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2025)
Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
LLM-BABYBENCH: Understanding and Evaluating Grounded Planning and Reasoning in LLMs
von: Choukrani, Omar, et al.
Veröffentlicht: (2025)
von: Choukrani, Omar, et al.
Veröffentlicht: (2025)
Temporal Misalignment in ANN-SNN Conversion and Its Mitigation via Probabilistic Spiking Neurons
von: Bojković, Velibor, et al.
Veröffentlicht: (2025)
von: Bojković, Velibor, et al.
Veröffentlicht: (2025)
A note on rational surgeries on a Hopf link
von: Bojković, Velibor, et al.
Veröffentlicht: (2024)
von: Bojković, Velibor, et al.
Veröffentlicht: (2024)
A Decade of Deep Learning: A Survey on The Magnificent Seven
von: Azizov, Dilshod, et al.
Veröffentlicht: (2024)
von: Azizov, Dilshod, et al.
Veröffentlicht: (2024)
Optimizing Adaptive Attacks against Watermarks for Language Models
von: Diaa, Abdulrahman, et al.
Veröffentlicht: (2024)
von: Diaa, Abdulrahman, et al.
Veröffentlicht: (2024)
Watermarking Should Be Treated as a Monitoring Primitive
von: Aremu, Toluwani, et al.
Veröffentlicht: (2026)
von: Aremu, Toluwani, et al.
Veröffentlicht: (2026)
Linear Representations of Hierarchical Concepts in Language Models
von: Sakata, Masaki, et al.
Veröffentlicht: (2026)
von: Sakata, Masaki, et al.
Veröffentlicht: (2026)
On Entity Identification in Language Models
von: Sakata, Masaki, et al.
Veröffentlicht: (2025)
von: Sakata, Masaki, et al.
Veröffentlicht: (2025)
The Curse of Popularity: Popular Entities have Catastrophic Side Effects when Deleting Knowledge from Language Models
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2024)
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2024)
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
SPACT18: Spiking Human Action Recognition Benchmark Dataset with Complementary RGB and Thermal Modalities
von: Ashraf, Yasser, et al.
Veröffentlicht: (2025)
von: Ashraf, Yasser, et al.
Veröffentlicht: (2025)
Do LLMs Need to Think in One Language? Correlation between Latent Language and Task Performance
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
Hidden Failures in Robustness: Why Supervised Uncertainty Quantification Needs Better Evaluation
von: Stacey, Joe, et al.
Veröffentlicht: (2026)
von: Stacey, Joe, et al.
Veröffentlicht: (2026)
Robust Safety Monitoring of Language Models via Activation Watermarking
von: Aremu, Toluwani, et al.
Veröffentlicht: (2026)
von: Aremu, Toluwani, et al.
Veröffentlicht: (2026)
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference
von: Kamoda, Go, et al.
Veröffentlicht: (2025)
von: Kamoda, Go, et al.
Veröffentlicht: (2025)
MirrorCheck: Efficient Adversarial Defense for Vision-Language Models
von: Fares, Samar, et al.
Veröffentlicht: (2024)
von: Fares, Samar, et al.
Veröffentlicht: (2024)
FTBC: Forward Temporal Bias Correction for Optimizing ANN-SNN Conversion
von: Wu, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Wu, Xiaofeng, et al.
Veröffentlicht: (2024)
UNICON: UNIfied CONtinual Learning for Medical Foundational Models
von: Qazi, Mohammad Areeb, et al.
Veröffentlicht: (2025)
von: Qazi, Mohammad Areeb, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The AI Data Scientist
von: Akimov, Farkhad, et al.
Veröffentlicht: (2025) -
Measuring AI Reasoning: A Guide for Researchers
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026) -
Mechanistic Insights into Grokking from the Embedding Layer
von: AlquBoj, H. V., et al.
Veröffentlicht: (2025) -
Number Representations in LLMs: A Computational Parallel to Human Perception
von: AlquBoj, H. V., et al.
Veröffentlicht: (2025) -
Sycophancy Hides Linearly in the Attention Heads
von: Genadi, Rifo, et al.
Veröffentlicht: (2026)