Saved in:
| Main Author: | Hill, Felix |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2408.03855 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is your multimodal large language model a good science tutor?
by: Liu, Ming, et al.
Published: (2025)
by: Liu, Ming, et al.
Published: (2025)
Why do language models perform worse for morphologically complex languages?
by: Arnett, Catherine, et al.
Published: (2024)
by: Arnett, Catherine, et al.
Published: (2024)
Large language models are good medical coders, if provided with tools
by: Kwan, Keith
Published: (2024)
by: Kwan, Keith
Published: (2024)
Efficient transformer with reinforced position embedding for language models
by: Hsiao, Yen-Che, et al.
Published: (2024)
by: Hsiao, Yen-Che, et al.
Published: (2024)
Geopolitical biases in LLMs: what are the "good" and the "bad" countries according to contemporary language models
by: Salnikov, Mikhail, et al.
Published: (2025)
by: Salnikov, Mikhail, et al.
Published: (2025)
NLD-LLM: A systematic framework for evaluating small language transformer models on natural language description
by: Jelodar, Hamed, et al.
Published: (2025)
by: Jelodar, Hamed, et al.
Published: (2025)
Lost without translation -- Can transformer (language models) understand mood states?
by: Shivaprakash, Prakrithi, et al.
Published: (2025)
by: Shivaprakash, Prakrithi, et al.
Published: (2025)
Mixture-of-Depths: Dynamically allocating compute in transformer-based language models
by: Raposo, David, et al.
Published: (2024)
by: Raposo, David, et al.
Published: (2024)
Physical models realizing the transformer architecture of large language models
by: Chen, Zeqian
Published: (2025)
by: Chen, Zeqian
Published: (2025)
Why are language models less surprised than humans? Testing the Parse Multiplicity Mismatch Hypothesis
by: Timkey, William, et al.
Published: (2026)
by: Timkey, William, et al.
Published: (2026)
Comparison of different Unique hard attention transformer models by the formal languages they can recognize
by: Ryvkin, Leonid
Published: (2025)
by: Ryvkin, Leonid
Published: (2025)
Humans and transformer LMs: Abstraction drives language learning
by: Jian, Jasper, et al.
Published: (2026)
by: Jian, Jasper, et al.
Published: (2026)
Why do small language models underperform? Studying Language Model Saturation via the Softmax Bottleneck
by: Godey, Nathan, et al.
Published: (2024)
by: Godey, Nathan, et al.
Published: (2024)
Human-like fleeting memory improves language learning but impairs reading time prediction in transformer language models
by: Thamma, Abishek, et al.
Published: (2025)
by: Thamma, Abishek, et al.
Published: (2025)
Re-evaluating Theory of Mind evaluation in large language models
by: Hu, Jennifer, et al.
Published: (2025)
by: Hu, Jennifer, et al.
Published: (2025)
Spatio-temporal transformer to support automatic sign language translation
by: Ruiz, Christian, et al.
Published: (2025)
by: Ruiz, Christian, et al.
Published: (2025)
Why is constrained neural language generation particularly challenging?
by: Garbacea, Cristina, et al.
Published: (2022)
by: Garbacea, Cristina, et al.
Published: (2022)
Attention-based transformer models for image captioning across languages: An in-depth survey and evaluation
by: Albadarneh, Israa A., et al.
Published: (2025)
by: Albadarneh, Israa A., et al.
Published: (2025)
A path to natural language through tokenisation and transformers
by: Berman, David S., et al.
Published: (2026)
by: Berman, David S., et al.
Published: (2026)
Detecting out-of-distribution text using topological features of transformer-based language models
by: Pollano, Andres, et al.
Published: (2023)
by: Pollano, Andres, et al.
Published: (2023)
Enhanced Arabic-language cyberbullying detection: deep embedding and transformer (BERT) approaches
by: Aljohani, Ebtesam Jaber, et al.
Published: (2025)
by: Aljohani, Ebtesam Jaber, et al.
Published: (2025)
A multitask transformer to sign language translation using motion gesture primitives
by: López, Fredy Alejandro Mendoza, et al.
Published: (2025)
by: López, Fredy Alejandro Mendoza, et al.
Published: (2025)
Large language models have learned to use language
by: Lupyan, Gary
Published: (2025)
by: Lupyan, Gary
Published: (2025)
Tiny language models
by: Gross, Ronit D., et al.
Published: (2025)
by: Gross, Ronit D., et al.
Published: (2025)
Child vs. machine language learning: Can the logical structure of human language unleash LLMs?
by: Sauerland, Uli, et al.
Published: (2025)
by: Sauerland, Uli, et al.
Published: (2025)
Can we teach language models to gloss endangered languages?
by: Ginn, Michael, et al.
Published: (2024)
by: Ginn, Michael, et al.
Published: (2024)
Retrieval augmentation of large language models for lay language generation
by: Guo, Yue, et al.
Published: (2022)
by: Guo, Yue, et al.
Published: (2022)
Do large language models resemble humans in language use?
by: Cai, Zhenguang G., et al.
Published: (2023)
by: Cai, Zhenguang G., et al.
Published: (2023)
Studies with impossible languages falsify LMs as models of human language
by: Bowers, Jeffrey S., et al.
Published: (2025)
by: Bowers, Jeffrey S., et al.
Published: (2025)
Large language models are not about natural language
by: Bolhuis, Johan J., et al.
Published: (2025)
by: Bolhuis, Johan J., et al.
Published: (2025)
Dissociating language and thought in large language models
by: Mahowald, Kyle, et al.
Published: (2023)
by: Mahowald, Kyle, et al.
Published: (2023)
Simulated patient systems powered by large language model-based AI agents offer potential for transforming medical education
by: Yu, Huizi, et al.
Published: (2024)
by: Yu, Huizi, et al.
Published: (2024)
Prompting language influences diagnostic reasoning and accuracy of large language models
by: Bazoge, Adrien, et al.
Published: (2026)
by: Bazoge, Adrien, et al.
Published: (2026)
Zero-shot data citation function classification using transformer-based large language models (LLMs)
by: Byers, Neil, et al.
Published: (2025)
by: Byers, Neil, et al.
Published: (2025)
Are LLMs reliable? An exploration of the reliability of large language models in clinical note generation
by: Carandang, Kristine Ann M., et al.
Published: (2025)
by: Carandang, Kristine Ann M., et al.
Published: (2025)
Dynamic data sampler for cross-language transfer learning in large language models
by: Li, Yudong, et al.
Published: (2024)
by: Li, Yudong, et al.
Published: (2024)
InkubaLM: A small language model for low-resource African languages
by: Tonja, Atnafu Lambebo, et al.
Published: (2024)
by: Tonja, Atnafu Lambebo, et al.
Published: (2024)
Multilingual large language models leak human stereotypes across language boundaries
by: Cao, Yang Trista, et al.
Published: (2023)
by: Cao, Yang Trista, et al.
Published: (2023)
Effective vocabulary expanding of multilingual language models for extremely low-resource languages
by: Zheng, Jianyu
Published: (2026)
by: Zheng, Jianyu
Published: (2026)
Anthropocentric bias in language model evaluation
by: Millière, Raphaël, et al.
Published: (2024)
by: Millière, Raphaël, et al.
Published: (2024)
Similar Items
-
Is your multimodal large language model a good science tutor?
by: Liu, Ming, et al.
Published: (2025) -
Why do language models perform worse for morphologically complex languages?
by: Arnett, Catherine, et al.
Published: (2024) -
Large language models are good medical coders, if provided with tools
by: Kwan, Keith
Published: (2024) -
Efficient transformer with reinforced position embedding for language models
by: Hsiao, Yen-Che, et al.
Published: (2024) -
Geopolitical biases in LLMs: what are the "good" and the "bad" countries according to contemporary language models
by: Salnikov, Mikhail, et al.
Published: (2025)