Emergent Word Order Universals from Cognitively-Motivated Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kuribayashi, Tatsuki, Ueda, Ryo, Yoshida, Ryo, Oseki, Yohei, Briscoe, Ted, Baldwin, Timothy |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Psychometric Predictive Power of Large Language Models
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
by: Yoshida, Ryo, et al.
Published: (2026)
by: Yoshida, Ryo, et al.
Published: (2026)
Which Word Orders Facilitate Length Generalization in LMs? An Investigation with GCG-Based Artificial Languages
by: El-Naggar, Nadine, et al.
Published: (2025)
by: El-Naggar, Nadine, et al.
Published: (2025)
Large Language Models Are Human-Like Internally
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
Composition, Attention, or Both?
by: Yoshida, Ryo, et al.
Published: (2022)
by: Yoshida, Ryo, et al.
Published: (2022)
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
by: Yoshida, Ryo, et al.
Published: (2024)
by: Yoshida, Ryo, et al.
Published: (2024)
What Kind of Language is Easy to Language-Model Under Curriculum Learning?
by: El-Naggar, Nadine, et al.
Published: (2026)
by: El-Naggar, Nadine, et al.
Published: (2026)
Developmentally-plausible Working Memory Shapes a Critical Period for Language Acquisition
by: Mita, Masato, et al.
Published: (2025)
by: Mita, Masato, et al.
Published: (2025)
Modeling Human Sentence Processing with Left-Corner Recurrent Neural Network Grammars
by: Yoshida, Ryo, et al.
Published: (2021)
by: Yoshida, Ryo, et al.
Published: (2021)
Language Acquisition Device in Large Language Models
by: Mita, Masato, et al.
Published: (2026)
by: Mita, Masato, et al.
Published: (2026)
Syntactic Learnability of Echo State Neural Language Models at Scale
by: Ueda, Ryo, et al.
Published: (2025)
by: Ueda, Ryo, et al.
Published: (2025)
Does Vision Accelerate Hierarchical Generalization in Neural Language Learners?
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
Dual Alignment Between Language Model Layers and Human Sentence Processing
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
by: Someya, Taiga, et al.
Published: (2025)
by: Someya, Taiga, et al.
Published: (2025)
Can Language Models Learn Typologically Implausible Languages?
by: Xu, Tianyang, et al.
Published: (2025)
by: Xu, Tianyang, et al.
Published: (2025)
If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
by: Yoshida, Ryo, et al.
Published: (2025)
by: Yoshida, Ryo, et al.
Published: (2025)
Rethinking the Relationship between the Power Law and Hierarchical Structures
by: Nakaishi, Kai, et al.
Published: (2025)
by: Nakaishi, Kai, et al.
Published: (2025)
Lewis's Signaling Game as beta-VAE For Natural Word Lengths and Segments
by: Ueda, Ryo, et al.
Published: (2023)
by: Ueda, Ryo, et al.
Published: (2023)
On Representational Dissociation of Language and Arithmetic in Large Language Models
by: Kisako, Riku, et al.
Published: (2025)
by: Kisako, Riku, et al.
Published: (2025)
Generative Emergent Communication: Large Language Model is a Collective World Model
by: Taniguchi, Tadahiro, et al.
Published: (2024)
by: Taniguchi, Tadahiro, et al.
Published: (2024)
Is Structure Dependence Shaped for Efficient Communication?: A Case Study on Coordination
by: Kajikawa, Kohei, et al.
Published: (2024)
by: Kajikawa, Kohei, et al.
Published: (2024)
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency
by: Haga, Akari, et al.
Published: (2024)
by: Haga, Akari, et al.
Published: (2024)
Why Are Parsing Actions for Understanding Message Hierarchies Not Random?
by: Kato, Daichi, et al.
Published: (2025)
by: Kato, Daichi, et al.
Published: (2025)
Can Language Models Induce Grammatical Knowledge from Indirect Evidence?
by: Oba, Miyu, et al.
Published: (2024)
by: Oba, Miyu, et al.
Published: (2024)
Can Input Attributions Explain Inductive Reasoning in In-Context Learning?
by: Ye, Mengyu, et al.
Published: (2024)
by: Ye, Mengyu, et al.
Published: (2024)
Analyzing Feed-Forward Blocks in Transformers through the Lens of Attention Maps
by: Kobayashi, Goro, et al.
Published: (2023)
by: Kobayashi, Goro, et al.
Published: (2023)
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
by: Aoki, Yoichi, et al.
Published: (2024)
by: Aoki, Yoichi, et al.
Published: (2024)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
by: Baumgärtner, Tim, et al.
Published: (2025)
by: Baumgärtner, Tim, et al.
Published: (2025)
A Little Leak Will Sink a Great Ship: Survey of Transparency for Large Language Models from Start to Finish
by: Kaneko, Masahiro, et al.
Published: (2024)
by: Kaneko, Masahiro, et al.
Published: (2024)
Exclusive Unlearning
by: Sasaki, Mutsumi, et al.
Published: (2026)
by: Sasaki, Mutsumi, et al.
Published: (2026)
To Drop or Not to Drop? Predicting Argument Ellipsis Judgments: A Case Study in Japanese
by: Ishizuki, Yukiko, et al.
Published: (2024)
by: Ishizuki, Yukiko, et al.
Published: (2024)
Metropolis-Hastings Captioning Game: Knowledge Fusion of Vision Language Models via Decentralized Bayesian Inference
by: Matsui, Yuta, et al.
Published: (2025)
by: Matsui, Yuta, et al.
Published: (2025)
The Good, the Bad and the Constructive: Automatically Measuring Peer Review's Utility for Authors
by: Sadallah, Abdelrahman, et al.
Published: (2025)
by: Sadallah, Abdelrahman, et al.
Published: (2025)
Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality
by: Harada, Yuto, et al.
Published: (2025)
by: Harada, Yuto, et al.
Published: (2025)
Enhancing Arabic Automated Essay Scoring with Synthetic Data and Error Injection
by: Qwaider, Chatrine, et al.
Published: (2025)
by: Qwaider, Chatrine, et al.
Published: (2025)
ARWI: Arabic Write and Improve
by: Chirkunov, Kirill, et al.
Published: (2025)
by: Chirkunov, Kirill, et al.
Published: (2025)
Multilingual Gradient Word-Order Typology from Universal Dependencies
by: Baylor, Emi, et al.
Published: (2024)
by: Baylor, Emi, et al.
Published: (2024)
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
by: Kaneko, Masahiro, et al.
Published: (2025)
by: Kaneko, Masahiro, et al.
Published: (2025)
Likelihood Variance as Text Importance for Resampling Texts to Map Language Models
by: Oyama, Momose, et al.
Published: (2025)
by: Oyama, Momose, et al.
Published: (2025)
Don't Ignore the Tail: Decoupling top-K Probabilities for Efficient Language Model Distillation
by: Dasgupta, Sayantan, et al.
Published: (2026)
by: Dasgupta, Sayantan, et al.
Published: (2026)
Similar Items
-
Psychometric Predictive Power of Large Language Models
by: Kuribayashi, Tatsuki, et al.
Published: (2023) -
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
by: Yoshida, Ryo, et al.
Published: (2026) -
Which Word Orders Facilitate Length Generalization in LMs? An Investigation with GCG-Based Artificial Languages
by: El-Naggar, Nadine, et al.
Published: (2025) -
Large Language Models Are Human-Like Internally
by: Kuribayashi, Tatsuki, et al.
Published: (2025) -
Composition, Attention, or Both?
by: Yoshida, Ryo, et al.
Published: (2022)