Leveraging Human Production-Interpretation Asymmetries to Test LLM Cognitive Plausibility
Fuente:
arXiv
Saved in:
| Main Authors: | Lam, Suet-Ying, Zeng, Qingcheng, Wu, Jingyi, Voigt, Rob |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Pragmatic Mind of Machines: Tracing the Emergence of Pragmatic Competence in Large Language Models
by: Yu, Kefan, et al.
Published: (2025)
by: Yu, Kefan, et al.
Published: (2025)
Thinking Out Loud: Do Reasoning Models Know When They're Right?
by: Zeng, Qingcheng, et al.
Published: (2025)
by: Zeng, Qingcheng, et al.
Published: (2025)
Causal Micro-Narratives
by: Heddaya, Mourad, et al.
Published: (2024)
by: Heddaya, Mourad, et al.
Published: (2024)
Good Intentions Beyond ACL: Who Does NLP for Social Good, and Where?
by: LeFevre, Grace, et al.
Published: (2025)
by: LeFevre, Grace, et al.
Published: (2025)
Sympathy over Polarization: A Computational Discourse Analysis of Social Media Posts about the July 2024 Trump Assassination Attempt
by: Zeng, Qingcheng, et al.
Published: (2025)
by: Zeng, Qingcheng, et al.
Published: (2025)
Everything is Plausible: Investigating the Impact of LLM Rationales on Human Notions of Plausibility
by: Palta, Shramay, et al.
Published: (2025)
by: Palta, Shramay, et al.
Published: (2025)
Plausibility Vaccine: Injecting LLM Knowledge for Event Plausibility
by: Chmura, Jacob, et al.
Published: (2025)
by: Chmura, Jacob, et al.
Published: (2025)
If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
by: Yoshida, Ryo, et al.
Published: (2025)
by: Yoshida, Ryo, et al.
Published: (2025)
Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility
by: Wang, Sheng-Fu, et al.
Published: (2025)
by: Wang, Sheng-Fu, et al.
Published: (2025)
Lost in Interpretation: The Plausibility-Faithfulness Trade-off in Cross-Lingual Explanations
by: Banerjee, Somnath, et al.
Published: (2026)
by: Banerjee, Somnath, et al.
Published: (2026)
B4: Towards Optimal Assessment of Plausible Code Solutions with Plausible Tests
by: Chen, Mouxiang, et al.
Published: (2024)
by: Chen, Mouxiang, et al.
Published: (2024)
A Dataset for Physical and Abstract Plausibility and Sources of Human Disagreement
by: Eichel, Annerose, et al.
Published: (2024)
by: Eichel, Annerose, et al.
Published: (2024)
Verified Critical Step Optimization for LLM Agents
by: Li, Mukai, et al.
Published: (2026)
by: Li, Mukai, et al.
Published: (2026)
DeepSieve: Information Sieving via LLM-as-a-Knowledge-Router
by: Guo, Minghao, et al.
Published: (2025)
by: Guo, Minghao, et al.
Published: (2025)
Toward Equitable Access: Leveraging Crowdsourced Reviews to Investigate Public Perceptions of Health Resource Accessibility
by: Xue, Zhaoqian, et al.
Published: (2025)
by: Xue, Zhaoqian, et al.
Published: (2025)
Objectifying the Subjective: Cognitive Biases in Topic Interpretations
by: Hingmire, Swapnil, et al.
Published: (2025)
by: Hingmire, Swapnil, et al.
Published: (2025)
A Closed-Loop Personalized Learning Agent Integrating Neural Cognitive Diagnosis, Bounded-Ability Adaptive Testing, and LLM-Driven Feedback
by: Wang, Zhifeng, et al.
Published: (2025)
by: Wang, Zhifeng, et al.
Published: (2025)
ThinkBench: Dynamic Out-of-Distribution Evaluation for Robust LLM Reasoning
by: Huang, Shulin, et al.
Published: (2025)
by: Huang, Shulin, et al.
Published: (2025)
Support-Contra Asymmetry in LLM Explanations
by: Patil, Avinash
Published: (2025)
by: Patil, Avinash
Published: (2025)
Decoding Emotions in Abstract Art: Cognitive Plausibility of CLIP in Recognizing Color-Emotion Associations
by: Widhoelzl, Hanna-Sophia, et al.
Published: (2024)
by: Widhoelzl, Hanna-Sophia, et al.
Published: (2024)
Triangulating LLM Progress through Benchmarks, Games, and Cognitive Tests
by: Momentè, Filippo, et al.
Published: (2025)
by: Momentè, Filippo, et al.
Published: (2025)
Plausible-Parrots @ MSP2023: Enhancing Semantic Plausibility Modeling using Entity and Event Knowledge
by: Shen, Chong, et al.
Published: (2024)
by: Shen, Chong, et al.
Published: (2024)
Understanding the Cognitive Complexity in Language Elicited by Product Images
by: Chen, Yan-Ying, et al.
Published: (2024)
by: Chen, Yan-Ying, et al.
Published: (2024)
Do Cognitively Interpretable Reasoning Traces Improve LLM Performance?
by: Bhambri, Siddhant, et al.
Published: (2025)
by: Bhambri, Siddhant, et al.
Published: (2025)
Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not
by: Karakaş, Sercan
Published: (2026)
by: Karakaş, Sercan
Published: (2026)
COGNAC at SemEval-2026 Task 5: LLM Ensembles for Human-Level Word Sense Plausibility Rating in Challenging Narratives
by: Islam, Azwad Anjum, et al.
Published: (2026)
by: Islam, Azwad Anjum, et al.
Published: (2026)
Exploring Multilingual Probing in Large Language Models: A Cross-Language Analysis
by: Li, Daoyang, et al.
Published: (2024)
by: Li, Daoyang, et al.
Published: (2024)
Leveraging a Cognitive Model to Measure Subjective Similarity of Human and GPT-4 Written Content
by: Malloy, Tailia, et al.
Published: (2024)
by: Malloy, Tailia, et al.
Published: (2024)
HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns
by: Wang, Xintao, et al.
Published: (2026)
by: Wang, Xintao, et al.
Published: (2026)
Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility
by: Lepori, Michael A., et al.
Published: (2025)
by: Lepori, Michael A., et al.
Published: (2025)
The Confidence Dichotomy: Analyzing and Mitigating Miscalibration in Tool-Use Agents
by: Xuan, Weihao, et al.
Published: (2026)
by: Xuan, Weihao, et al.
Published: (2026)
LLM4CD: Leveraging Large Language Models for Open-World Knowledge Augmented Cognitive Diagnosis
by: Zhang, Weiming, et al.
Published: (2025)
by: Zhang, Weiming, et al.
Published: (2025)
Bias Beware: The Impact of Cognitive Biases on LLM-Driven Product Recommendations
by: Filandrianos, Giorgos, et al.
Published: (2025)
by: Filandrianos, Giorgos, et al.
Published: (2025)
Exploring and Testing Skill-Based Behavioral Profile Annotation: Human Operability and LLM Feasibility under Schema-Guided Execution
by: Wu, Yufeng
Published: (2026)
by: Wu, Yufeng
Published: (2026)
Directional Optimization Asymmetry in Transformers: A Synthetic Stress Test
by: Sahasrabudhe, Mihir
Published: (2025)
by: Sahasrabudhe, Mihir
Published: (2025)
Modelling Adjectival Modification Effects on Semantic Plausibility
by: Golub, Anna, et al.
Published: (2025)
by: Golub, Anna, et al.
Published: (2025)
Large Language Models for Psycholinguistic Plausibility Pretesting
by: Amouyal, Samuel Joseph, et al.
Published: (2024)
by: Amouyal, Samuel Joseph, et al.
Published: (2024)
Enhancing LLM's Cognition via Structurization
by: Liu, Kai, et al.
Published: (2024)
by: Liu, Kai, et al.
Published: (2024)
Language of Bargaining
by: Heddaya, Mourad, et al.
Published: (2023)
by: Heddaya, Mourad, et al.
Published: (2023)
Interpretable Stylistic Variation in Human and LLM Writing Across Genres, Models, and Decoding Strategies
by: Rallapalli, Swati, et al.
Published: (2026)
by: Rallapalli, Swati, et al.
Published: (2026)
Similar Items
-
The Pragmatic Mind of Machines: Tracing the Emergence of Pragmatic Competence in Large Language Models
by: Yu, Kefan, et al.
Published: (2025) -
Thinking Out Loud: Do Reasoning Models Know When They're Right?
by: Zeng, Qingcheng, et al.
Published: (2025) -
Causal Micro-Narratives
by: Heddaya, Mourad, et al.
Published: (2024) -
Good Intentions Beyond ACL: Who Does NLP for Social Good, and Where?
by: LeFevre, Grace, et al.
Published: (2025) -
Sympathy over Polarization: A Computational Discourse Analysis of Social Media Posts about the July 2024 Trump Assassination Attempt
by: Zeng, Qingcheng, et al.
Published: (2025)