No Such Thing as a General Learner: Language models and their dual optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Chemla, Emmanuel, Nefdt, Ryan M. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Biasless Language Models Learn Unnaturally: How LLMs Fail to Distinguish the Possible from the Impossible
by: Ziv, Imry, et al.
Published: (2025)
by: Ziv, Imry, et al.
Published: (2025)
Bridging the Empirical-Theoretical Gap in Neural Network Formal Language Learning Using Minimum Description Length
by: Lan, Nur, et al.
Published: (2024)
by: Lan, Nur, et al.
Published: (2024)
Improving Spoken Language Modeling with Phoneme Classification: A Simple Fine-tuning Approach
by: Poli, Maxime, et al.
Published: (2024)
by: Poli, Maxime, et al.
Published: (2024)
Large Language Models as Proxies for Theories of Human Linguistic Cognition
by: Ziv, Imry, et al.
Published: (2025)
by: Ziv, Imry, et al.
Published: (2025)
fastabx: A library for efficient computation of ABX discriminability
by: Poli, Maxime, et al.
Published: (2025)
by: Poli, Maxime, et al.
Published: (2025)
What Makes Two Language Models Think Alike?
by: Salle, Jeanne, et al.
Published: (2024)
by: Salle, Jeanne, et al.
Published: (2024)
The Impact of Syntactic and Semantic Proximity on Machine Translation with Back-Translation
by: Guerin, Nicolas, et al.
Published: (2024)
by: Guerin, Nicolas, et al.
Published: (2024)
Probing Syntax in Large Language Models: Successes and Remaining Challenges
by: Diego-Simón, Pablo J., et al.
Published: (2025)
by: Diego-Simón, Pablo J., et al.
Published: (2025)
Metric Learning Encoding Models: A Multivariate Framework for Interpreting Neural Representations
by: Jalouzot, Louis, et al.
Published: (2024)
by: Jalouzot, Louis, et al.
Published: (2024)
A polar coordinate system represents syntax in large language models
by: Diego-Simón, Pablo, et al.
Published: (2024)
by: Diego-Simón, Pablo, et al.
Published: (2024)
A Neural Model for Word Repetition
by: Dager, Daniel, et al.
Published: (2025)
by: Dager, Daniel, et al.
Published: (2025)
A Minimum Description Length Approach to Regularization in Neural Networks
by: Abudy, Matan, et al.
Published: (2025)
by: Abudy, Matan, et al.
Published: (2025)
Polar probe linearly decodes semantic structures from LLMs
by: Diego-Simón, Pablo J., et al.
Published: (2026)
by: Diego-Simón, Pablo J., et al.
Published: (2026)
DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units
by: Poli, Maxime, et al.
Published: (2026)
by: Poli, Maxime, et al.
Published: (2026)
Does Vision Accelerate Hierarchical Generalization in Neural Language Learners?
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
Are BabyLMs Second Language Learners?
by: Edman, Lukas, et al.
Published: (2024)
by: Edman, Lukas, et al.
Published: (2024)
Annotating Errors in English Learners' Written Language Production: Advancing Automated Written Feedback Systems
by: Coyne, Steven, et al.
Published: (2025)
by: Coyne, Steven, et al.
Published: (2025)
Quantifier Scope Interpretation in Language Learners and LLMs
by: Fang, Shaohua, et al.
Published: (2025)
by: Fang, Shaohua, et al.
Published: (2025)
BERTs are Generative In-Context Learners
by: Samuel, David
Published: (2024)
by: Samuel, David
Published: (2024)
LLMs are Frequency Pattern Learners in Natural Language Inference
by: Cheng, Liang, et al.
Published: (2025)
by: Cheng, Liang, et al.
Published: (2025)
Language Models as Artificial Learners: Investigating Crosslinguistic Influence
by: Issam, Abderrahmane, et al.
Published: (2026)
by: Issam, Abderrahmane, et al.
Published: (2026)
LMCD: Language Models are Zeroshot Cognitive Diagnosis Learners
by: He, Yu, et al.
Published: (2025)
by: He, Yu, et al.
Published: (2025)
Large Language Models Often Say One Thing and Do Another
by: Xu, Ruoxi, et al.
Published: (2025)
by: Xu, Ruoxi, et al.
Published: (2025)
Language Models are Symbolic Learners in Arithmetic
by: Deng, Chunyuan, et al.
Published: (2024)
by: Deng, Chunyuan, et al.
Published: (2024)
BAGEL: Benchmarking Animal Knowledge Expertise in Language Models
by: Shen, Jiacheng, et al.
Published: (2026)
by: Shen, Jiacheng, et al.
Published: (2026)
TAIA: Large Language Models are Out-of-Distribution Data Learners
by: Jiang, Shuyang, et al.
Published: (2024)
by: Jiang, Shuyang, et al.
Published: (2024)
Seal: Advancing Speech Language Models to be Few-Shot Learners
by: Lei, Shuyu, et al.
Published: (2024)
by: Lei, Shuyu, et al.
Published: (2024)
Instruction Pre-Training: Language Models are Supervised Multitask Learners
by: Cheng, Daixuan, et al.
Published: (2024)
by: Cheng, Daixuan, et al.
Published: (2024)
What Makes Diffusion Language Models Super Data Learners?
by: Gao, Zitian, et al.
Published: (2025)
by: Gao, Zitian, et al.
Published: (2025)
Evaluating Adaptive Personalization of Educational Readings with Simulated Learners
by: Woo, Ryan T., et al.
Published: (2026)
by: Woo, Ryan T., et al.
Published: (2026)
Doing Things with Words: Rethinking Theory of Mind Simulation in Large Language Models
by: Lombardi, Agnese, et al.
Published: (2025)
by: Lombardi, Agnese, et al.
Published: (2025)
Say What You Mean: Natural Language Access Control with Large Language Models for Internet of Things
by: Cheng, Ye, et al.
Published: (2025)
by: Cheng, Ye, et al.
Published: (2025)
Large Language Models are Biased Reinforcement Learners
by: Hayes, William M., et al.
Published: (2024)
by: Hayes, William M., et al.
Published: (2024)
Multilingual Embedding Probes Fail to Generalize Across Learner Corpora
by: Lyngbaek, Laurits, et al.
Published: (2026)
by: Lyngbaek, Laurits, et al.
Published: (2026)
Towards Automated Lexicography: Generating and Evaluating Definitions for Learner's Dictionaries
by: Ide, Yusuke, et al.
Published: (2026)
by: Ide, Yusuke, et al.
Published: (2026)
Efficient Prompting for LLM-based Generative Internet of Things
by: Xiao, Bin, et al.
Published: (2024)
by: Xiao, Bin, et al.
Published: (2024)
Large Language Models are In-Context Molecule Learners
by: Li, Jiatong, et al.
Published: (2024)
by: Li, Jiatong, et al.
Published: (2024)
Large Language Models Could Be Rote Learners
by: Xu, Yuyang, et al.
Published: (2025)
by: Xu, Yuyang, et al.
Published: (2025)
LLM-Generated Feedback Supports Learning If Learners Choose to Use It
by: Thomas, Danielle R., et al.
Published: (2025)
by: Thomas, Danielle R., et al.
Published: (2025)
Artificial Intelligence Bias on English Language Learners in Automatic Scoring
by: Guo, Shuchen, et al.
Published: (2025)
by: Guo, Shuchen, et al.
Published: (2025)
Similar Items
-
Biasless Language Models Learn Unnaturally: How LLMs Fail to Distinguish the Possible from the Impossible
by: Ziv, Imry, et al.
Published: (2025) -
Bridging the Empirical-Theoretical Gap in Neural Network Formal Language Learning Using Minimum Description Length
by: Lan, Nur, et al.
Published: (2024) -
Improving Spoken Language Modeling with Phoneme Classification: A Simple Fine-tuning Approach
by: Poli, Maxime, et al.
Published: (2024) -
Large Language Models as Proxies for Theories of Human Linguistic Cognition
by: Ziv, Imry, et al.
Published: (2025) -
fastabx: A library for efficient computation of ABX discriminability
by: Poli, Maxime, et al.
Published: (2025)