Non-literal Understanding of Number Words by Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Tsvilodub, Polina, Gandhi, Kanishk, Zhao, Haoran, Fränken, Jan-Philipp, Franke, Michael, Goodman, Noah D. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Procedural Dilemma Generation for Evaluating Moral Reasoning in Humans and Language Models
by: Fränken, Jan-Philipp, et al.
Published: (2024)
by: Fränken, Jan-Philipp, et al.
Published: (2024)
Self-Supervised Alignment with Mutual Information: Learning to Follow Principles without Preference Labels
by: Fränken, Jan-Philipp, et al.
Published: (2024)
by: Fränken, Jan-Philipp, et al.
Published: (2024)
Bayesian Statistical Modeling with Predictors from LLMs
by: Franke, Michael, et al.
Published: (2024)
by: Franke, Michael, et al.
Published: (2024)
Cognitive Modeling with Scaffolded LLMs: A Case Study of Referential Expression Generation
by: Tsvilodub, Polina, et al.
Published: (2024)
by: Tsvilodub, Polina, et al.
Published: (2024)
Integrating Neural and Symbolic Components in a Model of Pragmatic Question-Answering
by: Tsvilodub, Polina, et al.
Published: (2025)
by: Tsvilodub, Polina, et al.
Published: (2025)
Learning to Simulate Human Dialogue
by: Gandhi, Kanishk, et al.
Published: (2026)
by: Gandhi, Kanishk, et al.
Published: (2026)
On Emergent Social World Models -- Evidence for Functional Integration of Theory of Mind and Pragmatic Reasoning in Language Models
by: Tsvilodub, Polina, et al.
Published: (2026)
by: Tsvilodub, Polina, et al.
Published: (2026)
Human-like Affective Cognition in Foundation Models
by: Gandhi, Kanishk, et al.
Published: (2024)
by: Gandhi, Kanishk, et al.
Published: (2024)
Predictions from language models for multiple-choice tasks are not robust under variation of scoring methods
by: Tsvilodub, Polina, et al.
Published: (2024)
by: Tsvilodub, Polina, et al.
Published: (2024)
STaR-GATE: Teaching Language Models to Ask Clarifying Questions
by: Andukuri, Chinmaya, et al.
Published: (2024)
by: Andukuri, Chinmaya, et al.
Published: (2024)
Act or Clarify? Modeling Sensitivity to Uncertainty and Cost in Communication
by: Tsvilodub, Polina, et al.
Published: (2026)
by: Tsvilodub, Polina, et al.
Published: (2026)
Endless Terminals: Scaling RL Environments for Terminal Agents
by: Gandhi, Kanishk, et al.
Published: (2026)
by: Gandhi, Kanishk, et al.
Published: (2026)
Experimental Pragmatics with Machines: Testing LLM Predictions for the Inferences of Plain and Embedded Disjunctions
by: Tsvilodub, Polina, et al.
Published: (2024)
by: Tsvilodub, Polina, et al.
Published: (2024)
Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
by: Gandhi, Kanishk, et al.
Published: (2025)
by: Gandhi, Kanishk, et al.
Published: (2025)
Psychometric Alignment: Capturing Human Knowledge Distributions via Language Models
by: He-Yueya, Joy, et al.
Published: (2024)
by: He-Yueya, Joy, et al.
Published: (2024)
Stream of Search (SoS): Learning to Search in Language
by: Gandhi, Kanishk, et al.
Published: (2024)
by: Gandhi, Kanishk, et al.
Published: (2024)
Scaling up the think-aloud method
by: Wurgaft, Daniel, et al.
Published: (2025)
by: Wurgaft, Daniel, et al.
Published: (2025)
Automated Statistical Model Discovery with Language Models
by: Li, Michael Y., et al.
Published: (2024)
by: Li, Michael Y., et al.
Published: (2024)
From Next-Token to Mathematics: The Learning Dynamics of Mathematical Reasoning in Language Models
by: Mishra, Shubhra, et al.
Published: (2024)
by: Mishra, Shubhra, et al.
Published: (2024)
Evaluating and Optimizing Educational Content with Large Language Model Judgments
by: He-Yueya, Joy, et al.
Published: (2024)
by: He-Yueya, Joy, et al.
Published: (2024)
Is Child-Directed Speech Effective Training Data for Language Models?
by: Feng, Steven Y., et al.
Published: (2024)
by: Feng, Steven Y., et al.
Published: (2024)
Large Language Model Reasoning Failures
by: Song, Peiyang, et al.
Published: (2026)
by: Song, Peiyang, et al.
Published: (2026)
CriticAL: Critic Automation with Language Models
by: Li, Michael Y., et al.
Published: (2024)
by: Li, Michael Y., et al.
Published: (2024)
Integrating Symbolic Natural Language Understanding and Language Models for Word Sense Disambiguation
by: Zhao, Kexin, et al.
Published: (2025)
by: Zhao, Kexin, et al.
Published: (2025)
Bifocal Attention: Harmonizing Geometric and Spectral Positional Embeddings for Algorithmic Generalization
by: Awadhiya, Kanishk
Published: (2026)
by: Awadhiya, Kanishk
Published: (2026)
Words at Play: Benchmarking Audio Pun Understanding in Large Audio-Language Models
by: Su, Yuchen, et al.
Published: (2026)
by: Su, Yuchen, et al.
Published: (2026)
Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought
by: Xiang, Violet, et al.
Published: (2025)
by: Xiang, Violet, et al.
Published: (2025)
Large Language Models Lack Understanding of Character Composition of Words
by: Shin, Andrew, et al.
Published: (2024)
by: Shin, Andrew, et al.
Published: (2024)
Benchmarking Vision Language Models for Cultural Understanding
by: Nayak, Shravan, et al.
Published: (2024)
by: Nayak, Shravan, et al.
Published: (2024)
Can Language Model Understand Word Semantics as A Chatbot? An Empirical Study of Language Model Internal External Mismatch
by: Zhao, Jinman, et al.
Published: (2024)
by: Zhao, Jinman, et al.
Published: (2024)
Do Large Language Models Understand Word Senses?
by: Meconi, Domenico, et al.
Published: (2025)
by: Meconi, Domenico, et al.
Published: (2025)
Typoglycemia under the Hood: Investigating Language Models' Understanding of Scrambled Words
by: Sperduti, Gianluca, et al.
Published: (2025)
by: Sperduti, Gianluca, et al.
Published: (2025)
Number Cookbook: Number Understanding of Language Models and How to Improve It
by: Yang, Haotong, et al.
Published: (2024)
by: Yang, Haotong, et al.
Published: (2024)
PERSONA: A Reproducible Testbed for Pluralistic Alignment
by: Castricato, Louis, et al.
Published: (2024)
by: Castricato, Louis, et al.
Published: (2024)
Bayesian Preference Elicitation with Language Models
by: Handa, Kunal, et al.
Published: (2024)
by: Handa, Kunal, et al.
Published: (2024)
pyvene: A Library for Understanding and Improving PyTorch Models via Interventions
by: Wu, Zhengxuan, et al.
Published: (2024)
by: Wu, Zhengxuan, et al.
Published: (2024)
Hypothesis Search: Inductive Reasoning with Language Models
by: Wang, Ruocheng, et al.
Published: (2023)
by: Wang, Ruocheng, et al.
Published: (2023)
Learning to Compress Prompts with Gist Tokens
by: Mu, Jesse, et al.
Published: (2023)
by: Mu, Jesse, et al.
Published: (2023)
Tug-of-war between idioms' figurative and literal interpretations in LLMs
by: Oh, Soyoung, et al.
Published: (2025)
by: Oh, Soyoung, et al.
Published: (2025)
Probing Language Models' Gesture Understanding for Enhanced Human-AI Interaction
by: Wicke, Philipp
Published: (2024)
by: Wicke, Philipp
Published: (2024)
Similar Items
-
Procedural Dilemma Generation for Evaluating Moral Reasoning in Humans and Language Models
by: Fränken, Jan-Philipp, et al.
Published: (2024) -
Self-Supervised Alignment with Mutual Information: Learning to Follow Principles without Preference Labels
by: Fränken, Jan-Philipp, et al.
Published: (2024) -
Bayesian Statistical Modeling with Predictors from LLMs
by: Franke, Michael, et al.
Published: (2024) -
Cognitive Modeling with Scaffolded LLMs: A Case Study of Referential Expression Generation
by: Tsvilodub, Polina, et al.
Published: (2024) -
Integrating Neural and Symbolic Components in a Model of Pragmatic Question-Answering
by: Tsvilodub, Polina, et al.
Published: (2025)