BabyLM's First Constructions: Causal probing provides a signal of learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rozner, Joshua, Weissweiler, Leonie, Shain, Cory |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Constructions are Revealed in Word Distributions
von: Rozner, Joshua, et al.
Veröffentlicht: (2025)
von: Rozner, Joshua, et al.
Veröffentlicht: (2025)
Perturbation: A simple and efficient adversarial tracer for representation learning in language models
von: Rozner, Joshua, et al.
Veröffentlicht: (2026)
von: Rozner, Joshua, et al.
Veröffentlicht: (2026)
BabyLM Turns 3: Call for papers for the 2025 BabyLM workshop
von: Charpentier, Lucas, et al.
Veröffentlicht: (2025)
von: Charpentier, Lucas, et al.
Veröffentlicht: (2025)
BabyLM Turns 4 and Goes Multilingual: Call for Papers for the 2026 BabyLM Workshop
von: Choshen, Leshem, et al.
Veröffentlicht: (2026)
von: Choshen, Leshem, et al.
Veröffentlicht: (2026)
BabyLM's First Words: Word Segmentation as a Phonological Probing Task
von: Goriely, Zébulon, et al.
Veröffentlicht: (2025)
von: Goriely, Zébulon, et al.
Veröffentlicht: (2025)
BAMBINO-LM: (Bilingual-)Human-Inspired Continual Pretraining of BabyLM
von: Shen, Zhewen, et al.
Veröffentlicht: (2024)
von: Shen, Zhewen, et al.
Veröffentlicht: (2024)
Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
von: Warstadt, Alex, et al.
Veröffentlicht: (2025)
von: Warstadt, Alex, et al.
Veröffentlicht: (2025)
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency
von: Haga, Akari, et al.
Veröffentlicht: (2024)
von: Haga, Akari, et al.
Veröffentlicht: (2024)
Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
von: Hu, Michael Y., et al.
Veröffentlicht: (2024)
von: Hu, Michael Y., et al.
Veröffentlicht: (2024)
Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions
von: Scivetti, Wesley, et al.
Veröffentlicht: (2026)
von: Scivetti, Wesley, et al.
Veröffentlicht: (2026)
Dialogue Is Not Enough to Make a Communicative BabyLM (But Neither Is Developmentally Inspired Reinforcement Learning)
von: Padovani, Francesca, et al.
Veröffentlicht: (2025)
von: Padovani, Francesca, et al.
Veröffentlicht: (2025)
Teacher Demonstrations in a BabyLM's Zone of Proximal Development for Contingent Multi-Turn Interaction
von: Salhan, Suchir, et al.
Veröffentlicht: (2025)
von: Salhan, Suchir, et al.
Veröffentlicht: (2025)
Are BabyLMs Second Language Learners?
von: Edman, Lukas, et al.
Veröffentlicht: (2024)
von: Edman, Lukas, et al.
Veröffentlicht: (2024)
[Call for Papers] The 2nd BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena
von: Weissweiler, Leonie, et al.
Veröffentlicht: (2024)
von: Weissweiler, Leonie, et al.
Veröffentlicht: (2024)
Do Construction Distributions Shape Formal Language Learning In German BabyLMs?
von: Bunzeck, Bastian, et al.
Veröffentlicht: (2025)
von: Bunzeck, Bastian, et al.
Veröffentlicht: (2025)
Bringing Up a Bilingual BabyLM: Investigating Multilingual Language Acquisition Using Small-Scale Models
von: Zeng, Linda, et al.
Veröffentlicht: (2026)
von: Zeng, Linda, et al.
Veröffentlicht: (2026)
Child-directed speech facilitates production, not comprehension, in BabyLMs
von: Bunzeck, Bastian, et al.
Veröffentlicht: (2026)
von: Bunzeck, Bastian, et al.
Veröffentlicht: (2026)
Linguistic Generalizations are not Rules: Impacts on Evaluation of LMs
von: Weissweiler, Leonie, et al.
Veröffentlicht: (2025)
von: Weissweiler, Leonie, et al.
Veröffentlicht: (2025)
CLASS-IT: Conversational and Lecture-Aligned Small-Scale Instruction Tuning for BabyLMs
von: Capone, Luca, et al.
Veröffentlicht: (2025)
von: Capone, Luca, et al.
Veröffentlicht: (2025)
Mask and You Shall Receive: Optimizing Masked Language Modeling For Pretraining BabyLMs
von: Edman, Lukas, et al.
Veröffentlicht: (2025)
von: Edman, Lukas, et al.
Veröffentlicht: (2025)
BabyLMs for isiXhosa: Data-Efficient Language Modelling in a Low-Resource Context
von: Matzopoulos, Alexis, et al.
Veröffentlicht: (2025)
von: Matzopoulos, Alexis, et al.
Veröffentlicht: (2025)
Graded strength of comparative illusions is explained by Bayesian inference
von: Zhang, Yuhan, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhan, et al.
Veröffentlicht: (2025)
Are BabyLMs Deaf to Gricean Maxims? A Pragmatic Evaluation of Sample-efficient Language Models
von: Askari, Raha, et al.
Veröffentlicht: (2025)
von: Askari, Raha, et al.
Veröffentlicht: (2025)
Both Direct and Indirect Evidence Contribute to Dative Alternation Preferences in Language Models
von: Yao, Qing, et al.
Veröffentlicht: (2025)
von: Yao, Qing, et al.
Veröffentlicht: (2025)
MultiBLiMP 1.0: A Massively Multilingual Benchmark of Linguistic Minimal Pairs
von: Jumelet, Jaap, et al.
Veröffentlicht: (2025)
von: Jumelet, Jaap, et al.
Veröffentlicht: (2025)
Bias Dynamics in BabyLMs: Towards a Compute-Efficient Sandbox for Democratising Pre-Training Debiasing
von: Trhlik, Filip, et al.
Veröffentlicht: (2026)
von: Trhlik, Filip, et al.
Veröffentlicht: (2026)
Constructions Are So Difficult That Even Large Language Models Get Them Right for the Wrong Reasons
von: Zhou, Shijia, et al.
Veröffentlicht: (2024)
von: Zhou, Shijia, et al.
Veröffentlicht: (2024)
Models Can and Should Embrace the Communicative Nature of Human-Generated Math
von: Boguraev, Sasha, et al.
Veröffentlicht: (2024)
von: Boguraev, Sasha, et al.
Veröffentlicht: (2024)
Artificial Aphasias in Lesioned Language Models
von: Roll, Nathan, et al.
Veröffentlicht: (2026)
von: Roll, Nathan, et al.
Veröffentlicht: (2026)
CausalLM is not optimal for in-context learning
von: Ding, Nan, et al.
Veröffentlicht: (2023)
von: Ding, Nan, et al.
Veröffentlicht: (2023)
Verbing Weirds Language (Models): Evaluation of English Zero-Derivation in Five LLMs
von: Mortensen, David R., et al.
Veröffentlicht: (2024)
von: Mortensen, David R., et al.
Veröffentlicht: (2024)
SYNTHEVAL: Hybrid Behavioral Testing of NLP Models with Synthetic CheckLists
von: Zhao, Raoyuan, et al.
Veröffentlicht: (2024)
von: Zhao, Raoyuan, et al.
Veröffentlicht: (2024)
BabyBabelLM: A Multilingual Benchmark of Developmentally Plausible Training Data
von: Jumelet, Jaap, et al.
Veröffentlicht: (2025)
von: Jumelet, Jaap, et al.
Veröffentlicht: (2025)
Derivational Morphology Reveals Analogical Generalization in Large Language Models
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024)
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024)
Independent-Component-Based Encoding Models of Brain Activity During Story Comprehension
von: Hari, Kamya, et al.
Veröffentlicht: (2026)
von: Hari, Kamya, et al.
Veröffentlicht: (2026)
Small Language Models Also Work With Small Vocabularies: Probing the Linguistic Abilities of Grapheme- and Phoneme-Based Baby Llamas
von: Bunzeck, Bastian, et al.
Veröffentlicht: (2024)
von: Bunzeck, Bastian, et al.
Veröffentlicht: (2024)
UCxn: Typologically Informed Annotation of Constructions Atop Universal Dependencies
von: Weissweiler, Leonie, et al.
Veröffentlicht: (2024)
von: Weissweiler, Leonie, et al.
Veröffentlicht: (2024)
AntLM: Bridging Causal and Masked Language Models
von: Yu, Xinru, et al.
Veröffentlicht: (2024)
von: Yu, Xinru, et al.
Veröffentlicht: (2024)
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP
von: Rozner, Josh, et al.
Veröffentlicht: (2021)
von: Rozner, Josh, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Constructions are Revealed in Word Distributions
von: Rozner, Joshua, et al.
Veröffentlicht: (2025) -
Perturbation: A simple and efficient adversarial tracer for representation learning in language models
von: Rozner, Joshua, et al.
Veröffentlicht: (2026) -
BabyLM Turns 3: Call for papers for the 2025 BabyLM workshop
von: Charpentier, Lucas, et al.
Veröffentlicht: (2025) -
BabyLM Turns 4 and Goes Multilingual: Call for Papers for the 2026 BabyLM Workshop
von: Choshen, Leshem, et al.
Veröffentlicht: (2026) -
BabyLM's First Words: Word Segmentation as a Phonological Probing Task
von: Goriely, Zébulon, et al.
Veröffentlicht: (2025)