Development of Cognitive Intelligence in Pre-trained Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Shah, Raj Sanjay, Bhardwaj, Khushi, Varma, Sashank |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pre-training LLMs using human-like development data corpus
by: Bhardwaj, Khushi, et al.
Published: (2023)
by: Bhardwaj, Khushi, et al.
Published: (2023)
The potential -- and the pitfalls -- of using pre-trained language models as cognitive science theories
by: Shah, Raj Sanjay, et al.
Published: (2025)
by: Shah, Raj Sanjay, et al.
Published: (2025)
How Well Do Deep Learning Models Capture Human Concepts? The Case of the Typicality Effect
by: Vemuri, Siddhartha K., et al.
Published: (2024)
by: Vemuri, Siddhartha K., et al.
Published: (2024)
Human Behavioral Benchmarking: Numeric Magnitude Comparison Effects in Large Language Models
by: Shah, Raj Sanjay, et al.
Published: (2023)
by: Shah, Raj Sanjay, et al.
Published: (2023)
Incremental Comprehension of Garden-Path Sentences by Large Language Models: Semantic Interpretation, Syntactic Re-Analysis, and Attention
by: Li, Andrew, et al.
Published: (2024)
by: Li, Andrew, et al.
Published: (2024)
The World According to LLMs: How Geographic Origin Influences LLMs' Entity Deduction Capabilities
by: Lalai, Harsh Nishant, et al.
Published: (2025)
by: Lalai, Harsh Nishant, et al.
Published: (2025)
Are More Tokens Rational? Inference-Time Scaling in Language Models as Adaptive Resource Rationality
by: Hu, Zhimin, et al.
Published: (2026)
by: Hu, Zhimin, et al.
Published: (2026)
Modeling Understanding of Story-Based Analogies Using Large Language Models
by: Inani, Kalit, et al.
Published: (2025)
by: Inani, Kalit, et al.
Published: (2025)
When Visuals Aren't the Problem: Evaluating Vision-Language Models on Misleading Data Visualizations
by: Lalai, Harsh Nishant, et al.
Published: (2026)
by: Lalai, Harsh Nishant, et al.
Published: (2026)
Understanding Graphical Perception in Data Visualization through Zero-shot Prompting of Vision-Language Models
by: Guo, Grace, et al.
Published: (2024)
by: Guo, Grace, et al.
Published: (2024)
Towards a Path Dependent Account of Category Fluency
by: Heineman, David, et al.
Published: (2024)
by: Heineman, David, et al.
Published: (2024)
The Representational Geometry of Number
by: Hu, Zhimin, et al.
Published: (2026)
by: Hu, Zhimin, et al.
Published: (2026)
Computer Vision Modeling of the Development of Geometric and Numerical Concepts in Humans
by: Wang, Zekun, et al.
Published: (2025)
by: Wang, Zekun, et al.
Published: (2025)
PrahokBART: A Pre-trained Sequence-to-Sequence Model for Khmer Natural Language Generation
by: Kaing, Hour, et al.
Published: (2025)
by: Kaing, Hour, et al.
Published: (2025)
Genetic Auto-prompt Learning for Pre-trained Code Intelligence Language Models
by: Feng, Chengzhe, et al.
Published: (2024)
by: Feng, Chengzhe, et al.
Published: (2024)
MS-HuBERT: Mitigating Pre-training and Inference Mismatch in Masked Language Modelling methods for learning Speech Representations
by: Yadav, Hemant, et al.
Published: (2024)
by: Yadav, Hemant, et al.
Published: (2024)
From Intentions to Techniques: A Comprehensive Taxonomy and Challenges in Text Watermarking for Large Language Models
by: Lalai, Harsh Nishant, et al.
Published: (2024)
by: Lalai, Harsh Nishant, et al.
Published: (2024)
Natural Mitigation of Catastrophic Interference: Continual Learning in Power-Law Learning Environments
by: Gandhi, Atith, et al.
Published: (2024)
by: Gandhi, Atith, et al.
Published: (2024)
A Neural Network Model of Complementary Learning Systems: Pattern Separation and Completion for Continual Learning
by: Jun, James P, et al.
Published: (2025)
by: Jun, James P, et al.
Published: (2025)
Metadata Conditioning Accelerates Language Model Pre-training
by: Gao, Tianyu, et al.
Published: (2025)
by: Gao, Tianyu, et al.
Published: (2025)
Evaluating Discourse Cohesion in Pre-trained Language Models
by: He, Jie, et al.
Published: (2025)
by: He, Jie, et al.
Published: (2025)
Fine-tuning Pre-trained Language Models for Few-shot Intent Detection: Supervised Pre-training and Isotropization
by: Zhang, Haode, et al.
Published: (2022)
by: Zhang, Haode, et al.
Published: (2022)
Stable Language Model Pre-training by Reducing Embedding Variability
by: Chung, Woojin, et al.
Published: (2024)
by: Chung, Woojin, et al.
Published: (2024)
Tracr-Injection: Distilling Algorithms into Pre-trained Language Models
by: Vergara-Browne, Tomás, et al.
Published: (2025)
by: Vergara-Browne, Tomás, et al.
Published: (2025)
Rethinking the Outlier Distribution in Large Language Models: An In-depth Study
by: Raman, Rahul, et al.
Published: (2025)
by: Raman, Rahul, et al.
Published: (2025)
Model Merging in Pre-training of Large Language Models
by: Li, Yunshui, et al.
Published: (2025)
by: Li, Yunshui, et al.
Published: (2025)
Calibrating Pre-trained Language Classifiers on LLM-generated Noisy Labels via Iterative Refinement
by: Ye, Liqin, et al.
Published: (2025)
by: Ye, Liqin, et al.
Published: (2025)
Multi-Level Feedback Generation with Large Language Models for Empowering Novice Peer Counselors
by: Chaszczewicz, Alicja, et al.
Published: (2024)
by: Chaszczewicz, Alicja, et al.
Published: (2024)
Probing Language Models for Pre-training Data Detection
by: Liu, Zhenhua, et al.
Published: (2024)
by: Liu, Zhenhua, et al.
Published: (2024)
DEPT: Decoupled Embeddings for Pre-training Language Models
by: Iacob, Alex, et al.
Published: (2024)
by: Iacob, Alex, et al.
Published: (2024)
Cross-layer Attention Sharing for Pre-trained Large Language Models
by: Mu, Yongyu, et al.
Published: (2024)
by: Mu, Yongyu, et al.
Published: (2024)
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models
by: Jukić, Josip, et al.
Published: (2024)
by: Jukić, Josip, et al.
Published: (2024)
SparseLLM: Towards Global Pruning for Pre-trained Language Models
by: Bai, Guangji, et al.
Published: (2024)
by: Bai, Guangji, et al.
Published: (2024)
Projective Methods for Mitigating Gender Bias in Pre-trained Language Models
by: Dawkins, Hillary, et al.
Published: (2024)
by: Dawkins, Hillary, et al.
Published: (2024)
Examining Forgetting in Continual Pre-training of Aligned Large Language Models
by: Li, Chen-An, et al.
Published: (2024)
by: Li, Chen-An, et al.
Published: (2024)
A Survey of Pre-trained Language Models for Processing Scientific Text
by: Ho, Xanh, et al.
Published: (2024)
by: Ho, Xanh, et al.
Published: (2024)
STEP: Staged Parameter-Efficient Pre-training for Large Language Models
by: Yano, Kazuki, et al.
Published: (2025)
by: Yano, Kazuki, et al.
Published: (2025)
KoCo: Conditioning Language Model Pre-training on Knowledge Coordinates
by: Li, Yudong, et al.
Published: (2026)
by: Li, Yudong, et al.
Published: (2026)
Knowledge-augmented Pre-trained Language Models for Biomedical Relation Extraction
by: Sänger, Mario, et al.
Published: (2025)
by: Sänger, Mario, et al.
Published: (2025)
Spoken Language Identification with Pre-trained Models and Margin Loss
by: Fang, Zhihua, et al.
Published: (2026)
by: Fang, Zhihua, et al.
Published: (2026)
Similar Items
-
Pre-training LLMs using human-like development data corpus
by: Bhardwaj, Khushi, et al.
Published: (2023) -
The potential -- and the pitfalls -- of using pre-trained language models as cognitive science theories
by: Shah, Raj Sanjay, et al.
Published: (2025) -
How Well Do Deep Learning Models Capture Human Concepts? The Case of the Typicality Effect
by: Vemuri, Siddhartha K., et al.
Published: (2024) -
Human Behavioral Benchmarking: Numeric Magnitude Comparison Effects in Large Language Models
by: Shah, Raj Sanjay, et al.
Published: (2023) -
Incremental Comprehension of Garden-Path Sentences by Large Language Models: Semantic Interpretation, Syntactic Re-Analysis, and Attention
by: Li, Andrew, et al.
Published: (2024)