Bringing Up a Bilingual BabyLM: Investigating Multilingual Language Acquisition Using Small-Scale Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Zeng, Linda, Feng, Steven Y., Frank, Michael C. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Baby Scale: Investigating Models Trained on Individual Children's Language Input
di: Feng, Steven Y., et al.
Pubblicazione: (2026)
di: Feng, Steven Y., et al.
Pubblicazione: (2026)
BabyLM Turns 4 and Goes Multilingual: Call for Papers for the 2026 BabyLM Workshop
di: Choshen, Leshem, et al.
Pubblicazione: (2026)
di: Choshen, Leshem, et al.
Pubblicazione: (2026)
BAMBINO-LM: (Bilingual-)Human-Inspired Continual Pretraining of BabyLM
di: Shen, Zhewen, et al.
Pubblicazione: (2024)
di: Shen, Zhewen, et al.
Pubblicazione: (2024)
BabyLM Turns 3: Call for papers for the 2025 BabyLM workshop
di: Charpentier, Lucas, et al.
Pubblicazione: (2025)
di: Charpentier, Lucas, et al.
Pubblicazione: (2025)
Bias Dynamics in BabyLMs: Towards a Compute-Efficient Sandbox for Democratising Pre-Training Debiasing
di: Trhlik, Filip, et al.
Pubblicazione: (2026)
di: Trhlik, Filip, et al.
Pubblicazione: (2026)
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency
di: Haga, Akari, et al.
Pubblicazione: (2024)
di: Haga, Akari, et al.
Pubblicazione: (2024)
Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
di: Hu, Michael Y., et al.
Pubblicazione: (2024)
di: Hu, Michael Y., et al.
Pubblicazione: (2024)
Are BabyLMs Second Language Learners?
di: Edman, Lukas, et al.
Pubblicazione: (2024)
di: Edman, Lukas, et al.
Pubblicazione: (2024)
How does a Multilingual LM Handle Multiple Languages?
di: Kakarla, Santhosh, et al.
Pubblicazione: (2025)
di: Kakarla, Santhosh, et al.
Pubblicazione: (2025)
BabyLM's First Constructions: Causal probing provides a signal of learning
di: Rozner, Joshua, et al.
Pubblicazione: (2025)
di: Rozner, Joshua, et al.
Pubblicazione: (2025)
Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
di: Warstadt, Alex, et al.
Pubblicazione: (2025)
di: Warstadt, Alex, et al.
Pubblicazione: (2025)
BabyLM's First Words: Word Segmentation as a Phonological Probing Task
di: Goriely, Zébulon, et al.
Pubblicazione: (2025)
di: Goriely, Zébulon, et al.
Pubblicazione: (2025)
CLASS-IT: Conversational and Lecture-Aligned Small-Scale Instruction Tuning for BabyLMs
di: Capone, Luca, et al.
Pubblicazione: (2025)
di: Capone, Luca, et al.
Pubblicazione: (2025)
Is Child-Directed Speech Effective Training Data for Language Models?
di: Feng, Steven Y., et al.
Pubblicazione: (2024)
di: Feng, Steven Y., et al.
Pubblicazione: (2024)
[Call for Papers] The 2nd BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus
di: Choshen, Leshem, et al.
Pubblicazione: (2024)
di: Choshen, Leshem, et al.
Pubblicazione: (2024)
SUTRA: Scalable Multilingual Language Model Architecture
di: Bendale, Abhijit, et al.
Pubblicazione: (2024)
di: Bendale, Abhijit, et al.
Pubblicazione: (2024)
Can Small Language Models Help Large Language Models Reason Better?: LM-Guided Chain-of-Thought
di: Lee, Jooyoung, et al.
Pubblicazione: (2024)
di: Lee, Jooyoung, et al.
Pubblicazione: (2024)
PonderLM: Pretraining Language Models to Ponder in Continuous Space
di: Zeng, Boyi, et al.
Pubblicazione: (2025)
di: Zeng, Boyi, et al.
Pubblicazione: (2025)
Dialogue Is Not Enough to Make a Communicative BabyLM (But Neither Is Developmentally Inspired Reinforcement Learning)
di: Padovani, Francesca, et al.
Pubblicazione: (2025)
di: Padovani, Francesca, et al.
Pubblicazione: (2025)
Teacher Demonstrations in a BabyLM's Zone of Proximal Development for Contingent Multi-Turn Interaction
di: Salhan, Suchir, et al.
Pubblicazione: (2025)
di: Salhan, Suchir, et al.
Pubblicazione: (2025)
Beyond Bilingual Transfer: Multilingual Code-Switching in Instruction Tuning
di: Asano, Shunta, et al.
Pubblicazione: (2026)
di: Asano, Shunta, et al.
Pubblicazione: (2026)
CogLM: Tracking Cognitive Development of Large Language Models
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
CaLM: Contrasting Large and Small Language Models to Verify Grounded Generation
di: Hsu, I-Hung, et al.
Pubblicazione: (2024)
di: Hsu, I-Hung, et al.
Pubblicazione: (2024)
Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions
di: Scivetti, Wesley, et al.
Pubblicazione: (2026)
di: Scivetti, Wesley, et al.
Pubblicazione: (2026)
Personal Intelligence System UniLM: Hybrid On-Device Small Language Model and Server-Based Large Language Model for Malay Nusantara
di: Nazri, Azree, et al.
Pubblicazione: (2024)
di: Nazri, Azree, et al.
Pubblicazione: (2024)
Group then Scale: Dynamic Mixture-of-Experts Multilingual Language Model
di: Li, Chong, et al.
Pubblicazione: (2025)
di: Li, Chong, et al.
Pubblicazione: (2025)
PhoneLM:an Efficient and Capable Small Language Model Family through Principled Pre-training
di: Yi, Rongjie, et al.
Pubblicazione: (2024)
di: Yi, Rongjie, et al.
Pubblicazione: (2024)
Brainstorming Brings Power to Large Language Models of Knowledge Reasoning
di: Qin, Zining, et al.
Pubblicazione: (2024)
di: Qin, Zining, et al.
Pubblicazione: (2024)
EuroBERT: Scaling Multilingual Encoders for European Languages
di: Boizard, Nicolas, et al.
Pubblicazione: (2025)
di: Boizard, Nicolas, et al.
Pubblicazione: (2025)
Evaluating Neural Language Models as Cognitive Models of Language Acquisition
di: Martínez, Héctor Javier Vázquez, et al.
Pubblicazione: (2023)
di: Martínez, Héctor Javier Vázquez, et al.
Pubblicazione: (2023)
A Language-agnostic Model of Child Language Acquisition
di: Mahon, Louis, et al.
Pubblicazione: (2024)
di: Mahon, Louis, et al.
Pubblicazione: (2024)
Is Multilingual LLM Watermarking Truly Multilingual? Scaling Robustness to 100+ Languages via Back-Translation
di: Mohamed, Asim, et al.
Pubblicazione: (2025)
di: Mohamed, Asim, et al.
Pubblicazione: (2025)
All Languages Matter: On the Multilingual Safety of Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
CCI4.0: A Bilingual Pretraining Dataset for Enhancing Reasoning in Large Language Models
di: Liu, Guang, et al.
Pubblicazione: (2025)
di: Liu, Guang, et al.
Pubblicazione: (2025)
TinyLlama: An Open-Source Small Language Model
di: Zhang, Peiyuan, et al.
Pubblicazione: (2024)
di: Zhang, Peiyuan, et al.
Pubblicazione: (2024)
TernaryLM: Memory-Efficient Language Modeling via Native 1.5-Bit Quantization with Adaptive Layer-wise Scaling
di: Nargund, Nisharg, et al.
Pubblicazione: (2026)
di: Nargund, Nisharg, et al.
Pubblicazione: (2026)
UrduLM: A Resource-Efficient Monolingual Urdu Language Model
di: Ali, Syed Muhammad, et al.
Pubblicazione: (2026)
di: Ali, Syed Muhammad, et al.
Pubblicazione: (2026)
JudgeLM: Fine-tuned Large Language Models are Scalable Judges
di: Zhu, Lianghui, et al.
Pubblicazione: (2023)
di: Zhu, Lianghui, et al.
Pubblicazione: (2023)
REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning
di: Qureshi, Rameez, et al.
Pubblicazione: (2024)
di: Qureshi, Rameez, et al.
Pubblicazione: (2024)
A Comparative Analysis of Bilingual and Trilingual Wav2Vec Models for Automatic Speech Recognition in Multilingual Oral History Archives
di: Lehečka, Jan, et al.
Pubblicazione: (2024)
di: Lehečka, Jan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Baby Scale: Investigating Models Trained on Individual Children's Language Input
di: Feng, Steven Y., et al.
Pubblicazione: (2026) -
BabyLM Turns 4 and Goes Multilingual: Call for Papers for the 2026 BabyLM Workshop
di: Choshen, Leshem, et al.
Pubblicazione: (2026) -
BAMBINO-LM: (Bilingual-)Human-Inspired Continual Pretraining of BabyLM
di: Shen, Zhewen, et al.
Pubblicazione: (2024) -
BabyLM Turns 3: Call for papers for the 2025 BabyLM workshop
di: Charpentier, Lucas, et al.
Pubblicazione: (2025) -
Bias Dynamics in BabyLMs: Towards a Compute-Efficient Sandbox for Democratising Pre-Training Debiasing
di: Trhlik, Filip, et al.
Pubblicazione: (2026)