Guardado en:
| Autores principales: | Michaelov, James A., Arnett, Catherine, Bergen, Benjamin K. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2404.19178 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Acquisition of Shared Grammatical Representations in Bilingual Language Models
por: Arnett, Catherine, et al.
Publicado: (2025)
por: Arnett, Catherine, et al.
Publicado: (2025)
Disaggregation Reveals Hidden Training Dynamics: The Case of Agreement Attraction
por: Michaelov, James A., et al.
Publicado: (2025)
por: Michaelov, James A., et al.
Publicado: (2025)
Emergent inabilities? Inverse scaling over the course of pretraining
por: Michaelov, James A., et al.
Publicado: (2023)
por: Michaelov, James A., et al.
Publicado: (2023)
Language Model Behavioral Phases are Consistent Across Architecture, Training Data, and Scale
por: Michaelov, James A., et al.
Publicado: (2025)
por: Michaelov, James A., et al.
Publicado: (2025)
Why do language models perform worse for morphologically complex languages?
por: Arnett, Catherine, et al.
Publicado: (2024)
por: Arnett, Catherine, et al.
Publicado: (2024)
Goldfish: Monolingual Language Models for 350 Languages
por: Chang, Tyler A., et al.
Publicado: (2024)
por: Chang, Tyler A., et al.
Publicado: (2024)
A Bit of a Problem: Measurement Disparities in Dataset Sizes Across Languages
por: Arnett, Catherine, et al.
Publicado: (2024)
por: Arnett, Catherine, et al.
Publicado: (2024)
Not quite Sherlock Holmes: Language model predictions do not reliably differentiate impossible from improbable events
por: Michaelov, James A., et al.
Publicado: (2025)
por: Michaelov, James A., et al.
Publicado: (2025)
N-gram-like Language Models Predict Reading Time Best
por: Michaelov, James A., et al.
Publicado: (2026)
por: Michaelov, James A., et al.
Publicado: (2026)
How Open Must Language Models be to Enable Reliable Scientific Inference?
por: Michaelov, James A., et al.
Publicado: (2026)
por: Michaelov, James A., et al.
Publicado: (2026)
Explaining and Mitigating Crosslingual Tokenizer Inequities
por: Arnett, Catherine, et al.
Publicado: (2025)
por: Arnett, Catherine, et al.
Publicado: (2025)
Bigram Subnetworks: Mapping to Next Tokens in Transformer Language Models
por: Chang, Tyler A., et al.
Publicado: (2025)
por: Chang, Tyler A., et al.
Publicado: (2025)
Large Language Models Pass the Turing Test
por: Jones, Cameron R., et al.
Publicado: (2025)
por: Jones, Cameron R., et al.
Publicado: (2025)
Do Large Language Models Exhibit Spontaneous Rational Deception?
por: Taylor, Samuel M., et al.
Publicado: (2025)
por: Taylor, Samuel M., et al.
Publicado: (2025)
Characterizing Learning Curves During Language Model Pre-Training: Learning, Forgetting, and Stability
por: Chang, Tyler A., et al.
Publicado: (2023)
por: Chang, Tyler A., et al.
Publicado: (2023)
BPE Stays on SCRIPT: Structured Encoding for Robust Multilingual Pretokenization
por: Land, Sander, et al.
Publicado: (2025)
por: Land, Sander, et al.
Publicado: (2025)
Evaluating Morphological Alignment of Tokenizers in 70 Languages
por: Arnett, Catherine, et al.
Publicado: (2025)
por: Arnett, Catherine, et al.
Publicado: (2025)
Language Statistics and False Belief Reasoning: Evidence from 41 Open-Weight LMs
por: Trott, Sean, et al.
Publicado: (2026)
por: Trott, Sean, et al.
Publicado: (2026)
Lies, Damned Lies, and Distributional Language Statistics: Persuasion and Deception with Large Language Models
por: Jones, Cameron R., et al.
Publicado: (2024)
por: Jones, Cameron R., et al.
Publicado: (2024)
Computational Sentence-level Metrics Predicting Human Sentence Comprehension
por: Sun, Kun, et al.
Publicado: (2024)
por: Sun, Kun, et al.
Publicado: (2024)
EVOKE: Emotion Vocabulary Of Korean and English
por: Jung, Yoonwon, et al.
Publicado: (2026)
por: Jung, Yoonwon, et al.
Publicado: (2026)
Does GPT-4 pass the Turing test?
por: Jones, Cameron R., et al.
Publicado: (2023)
por: Jones, Cameron R., et al.
Publicado: (2023)
Weight Tying Biases Token Embeddings Towards the Output Space
por: Lopardo, Antonio, et al.
Publicado: (2026)
por: Lopardo, Antonio, et al.
Publicado: (2026)
Different Tokenization Schemes Lead to Comparable Performance in Spanish Number Agreement
por: Arnett, Catherine, et al.
Publicado: (2024)
por: Arnett, Catherine, et al.
Publicado: (2024)
Intra-Layer Recurrence in Transformers for Language Modeling
por: Nguyen, Anthony, et al.
Publicado: (2025)
por: Nguyen, Anthony, et al.
Publicado: (2025)
BPE Gets Picky: Efficient Vocabulary Refinement During Tokenizer Training
por: Chizhov, Pavel, et al.
Publicado: (2024)
por: Chizhov, Pavel, et al.
Publicado: (2024)
Autoregressive + Chain of Thought = Recurrent: Recurrence's Role in Language Models' Computability and a Revisit of Recurrent Transformer
por: Zhang, Xiang, et al.
Publicado: (2024)
por: Zhang, Xiang, et al.
Publicado: (2024)
Toxicity of the Commons: Curating Open-Source Pre-Training Data
por: Arnett, Catherine, et al.
Publicado: (2024)
por: Arnett, Catherine, et al.
Publicado: (2024)
Dissecting the Ullman Variations with a SCALPEL: Why do LLMs fail at Trivial Alterations to the False Belief Task?
por: Pi, Zhiqiang, et al.
Publicado: (2024)
por: Pi, Zhiqiang, et al.
Publicado: (2024)
Will Large Language Models Transform Clinical Prediction?
por: Yildiz, Yusuf, et al.
Publicado: (2025)
por: Yildiz, Yusuf, et al.
Publicado: (2025)
GPT-4 is judged more human than humans in displaced and inverted Turing tests
por: Rathi, Ishika, et al.
Publicado: (2024)
por: Rathi, Ishika, et al.
Publicado: (2024)
Diverging Transformer Predictions for Human Sentence Processing: A Comprehensive Analysis of Agreement Attraction Effects
por: von der Malsburg, Titus, et al.
Publicado: (2026)
por: von der Malsburg, Titus, et al.
Publicado: (2026)
RecurrentGemma: Moving Past Transformers for Efficient Open Language Models
por: Botev, Aleksandar, et al.
Publicado: (2024)
por: Botev, Aleksandar, et al.
Publicado: (2024)
Re-defining Humor Data Objects for AI Humor Research
por: Arnett, Anna, et al.
Publicado: (2026)
por: Arnett, Anna, et al.
Publicado: (2026)
A Comprehensive Evaluation of Semantic Relation Knowledge of Pretrained Language Models and Humans
por: Cao, Zhihan, et al.
Publicado: (2024)
por: Cao, Zhihan, et al.
Publicado: (2024)
Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation
por: Dejl, Adam, et al.
Publicado: (2025)
por: Dejl, Adam, et al.
Publicado: (2025)
An Algebraic View of the Expressivity of Recurrent Language Models
por: Nowak, Franz, et al.
Publicado: (2026)
por: Nowak, Franz, et al.
Publicado: (2026)
Can Large Language Models Match the Conclusions of Systematic Reviews?
por: Polzak, Christopher, et al.
Publicado: (2025)
por: Polzak, Christopher, et al.
Publicado: (2025)
Unraveling the Dominance of Large Language Models Over Transformer Models for Bangla Natural Language Inference: A Comprehensive Study
por: Faria, Fatema Tuj Johora, et al.
Publicado: (2024)
por: Faria, Fatema Tuj Johora, et al.
Publicado: (2024)
LLMs and people both learn to form conventions -- just not with each other
por: Jones, Cameron R., et al.
Publicado: (2026)
por: Jones, Cameron R., et al.
Publicado: (2026)
Ejemplares similares
-
On the Acquisition of Shared Grammatical Representations in Bilingual Language Models
por: Arnett, Catherine, et al.
Publicado: (2025) -
Disaggregation Reveals Hidden Training Dynamics: The Case of Agreement Attraction
por: Michaelov, James A., et al.
Publicado: (2025) -
Emergent inabilities? Inverse scaling over the course of pretraining
por: Michaelov, James A., et al.
Publicado: (2023) -
Language Model Behavioral Phases are Consistent Across Architecture, Training Data, and Scale
por: Michaelov, James A., et al.
Publicado: (2025) -
Why do language models perform worse for morphologically complex languages?
por: Arnett, Catherine, et al.
Publicado: (2024)