Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Someya, Taiga, Yoshida, Ryo, Yanaka, Hitomi, Oseki, Yohei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
by: Yoshida, Ryo, et al.
Published: (2024)
by: Yoshida, Ryo, et al.
Published: (2024)
Language Acquisition Device in Large Language Models
by: Mita, Masato, et al.
Published: (2026)
by: Mita, Masato, et al.
Published: (2026)
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
by: Yoshida, Ryo, et al.
Published: (2026)
by: Yoshida, Ryo, et al.
Published: (2026)
If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
by: Yoshida, Ryo, et al.
Published: (2025)
by: Yoshida, Ryo, et al.
Published: (2025)
Fine-Grained Analysis of Shared Syntactic Mechanisms in Language Models
by: Kumon, Ryoma, et al.
Published: (2026)
by: Kumon, Ryoma, et al.
Published: (2026)
Composition, Attention, or Both?
by: Yoshida, Ryo, et al.
Published: (2022)
by: Yoshida, Ryo, et al.
Published: (2022)
Modeling Human Sentence Processing with Left-Corner Recurrent Neural Network Grammars
by: Yoshida, Ryo, et al.
Published: (2021)
by: Yoshida, Ryo, et al.
Published: (2021)
Developmentally-plausible Working Memory Shapes a Critical Period for Language Acquisition
by: Mita, Masato, et al.
Published: (2025)
by: Mita, Masato, et al.
Published: (2025)
Emergent Word Order Universals from Cognitively-Motivated Language Models
by: Kuribayashi, Tatsuki, et al.
Published: (2024)
by: Kuribayashi, Tatsuki, et al.
Published: (2024)
Rethinking the Relationship between the Power Law and Hierarchical Structures
by: Nakaishi, Kai, et al.
Published: (2025)
by: Nakaishi, Kai, et al.
Published: (2025)
Evaluating Structural Generalization in Neural Machine Translation
by: Kumon, Ryoma, et al.
Published: (2024)
by: Kumon, Ryoma, et al.
Published: (2024)
Do Large Vision-Language Models Distinguish between the Actual and Apparent Features of Illusions?
by: Shinozaki, Taiga, et al.
Published: (2025)
by: Shinozaki, Taiga, et al.
Published: (2025)
What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?
by: Ryu, Koki, et al.
Published: (2026)
by: Ryu, Koki, et al.
Published: (2026)
Comprehensive Evaluation of Large Language Models for Topic Modeling
by: Doi, Tomoki, et al.
Published: (2024)
by: Doi, Tomoki, et al.
Published: (2024)
Can Large Language Models Robustly Perform Natural Language Inference for Japanese Comparatives?
by: Mikami, Yosuke, et al.
Published: (2025)
by: Mikami, Yosuke, et al.
Published: (2025)
Investigating Training and Generalization in Faithful Self-Explanations of Large Language Models
by: Doi, Tomoki, et al.
Published: (2025)
by: Doi, Tomoki, et al.
Published: (2025)
Dual Alignment Between Language Model Layers and Human Sentence Processing
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
Analyzing the Inner Workings of Transformers in Compositional Generalization
by: Kumon, Ryoma, et al.
Published: (2025)
by: Kumon, Ryoma, et al.
Published: (2025)
Enhancing Rating Prediction with Off-the-Shelf LLMs Using In-Context User Reviews
by: Ryu, Koki, et al.
Published: (2025)
by: Ryu, Koki, et al.
Published: (2025)
NeuronMoE: Neuron-Guided Mixture-of-Experts for Efficient Multilingual LLM Extension
by: Li, Rongzhi, et al.
Published: (2026)
by: Li, Rongzhi, et al.
Published: (2026)
Neuron-Level Analysis of Cultural Understanding in Large Language Models
by: Yamamoto, Taisei, et al.
Published: (2025)
by: Yamamoto, Taisei, et al.
Published: (2025)
Bridging Perception and Language: A Systematic Benchmark for LVLMs' Understanding of Amodal Completion Reports
by: Watahiki, Amane, et al.
Published: (2025)
by: Watahiki, Amane, et al.
Published: (2025)
Psychometric Predictive Power of Large Language Models
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
Is Structure Dependence Shaped for Efficient Communication?: A Case Study on Coordination
by: Kajikawa, Kohei, et al.
Published: (2024)
by: Kajikawa, Kohei, et al.
Published: (2024)
LLMs Struggle with NLI for Perfect Aspect: A Cross-Linguistic Study in Chinese and Japanese
by: Lu, Jie, et al.
Published: (2025)
by: Lu, Jie, et al.
Published: (2025)
Implementing a Logical Inference System for Japanese Comparatives
by: Mikami, Yosuke, et al.
Published: (2025)
by: Mikami, Yosuke, et al.
Published: (2025)
Developing a Guideline for the Labovian-Structural Analysis of Oral Narratives in Japanese
by: Watahiki, Amane, et al.
Published: (2026)
by: Watahiki, Amane, et al.
Published: (2026)
On the Multilingual Ability of Decoder-based Pre-trained Language Models: Finding and Controlling Language-Specific Neurons
by: Kojima, Takeshi, et al.
Published: (2024)
by: Kojima, Takeshi, et al.
Published: (2024)
Syntactic Learnability of Echo State Neural Language Models at Scale
by: Ueda, Ryo, et al.
Published: (2025)
by: Ueda, Ryo, et al.
Published: (2025)
Exploring Intra and Inter-language Consistency in Embeddings with ICA
by: Li, Rongzhi, et al.
Published: (2024)
by: Li, Rongzhi, et al.
Published: (2024)
Information Locality as an Inductive Bias for Neural Language Models
by: Someya, Taiga, et al.
Published: (2025)
by: Someya, Taiga, et al.
Published: (2025)
JBBQ: Japanese Bias Benchmark for Analyzing Social Biases in Large Language Models
by: Yanaka, Hitomi, et al.
Published: (2024)
by: Yanaka, Hitomi, et al.
Published: (2024)
Can Language Models Learn Typologically Implausible Languages?
by: Xu, Tianyang, et al.
Published: (2025)
by: Xu, Tianyang, et al.
Published: (2025)
Large Language Models Are Human-Like Internally
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency
by: Haga, Akari, et al.
Published: (2024)
by: Haga, Akari, et al.
Published: (2024)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
by: Yamamoto, Taisei, et al.
Published: (2025)
by: Yamamoto, Taisei, et al.
Published: (2025)
Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality
by: Harada, Yuto, et al.
Published: (2025)
by: Harada, Yuto, et al.
Published: (2025)
Layer-wise Regularized Dropout for Neural Language Models
by: Ni, Shiwen, et al.
Published: (2024)
by: Ni, Shiwen, et al.
Published: (2024)
J-CHAT: Japanese Large-scale Spoken Dialogue Corpus for Spoken Dialogue Language Modeling
by: Nakata, Wataru, et al.
Published: (2024)
by: Nakata, Wataru, et al.
Published: (2024)
Can Language Models Induce Grammatical Knowledge from Indirect Evidence?
by: Oba, Miyu, et al.
Published: (2024)
by: Oba, Miyu, et al.
Published: (2024)
Similar Items
-
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
by: Yoshida, Ryo, et al.
Published: (2024) -
Language Acquisition Device in Large Language Models
by: Mita, Masato, et al.
Published: (2026) -
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
by: Yoshida, Ryo, et al.
Published: (2026) -
If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
by: Yoshida, Ryo, et al.
Published: (2025) -
Fine-Grained Analysis of Shared Syntactic Mechanisms in Language Models
by: Kumon, Ryoma, et al.
Published: (2026)