Can LLMs Simulate Human Behavioral Variability? A Case Study in the Phonemic Fluency Task
Fuente:
arXiv
Saved in:
| Main Authors: | Qiu, Mengyang, Brisebois, Zoe, Sun, Siena |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Formal Framework for Fluency-based Multi-Reference Evaluation in Grammatical Error Correction
by: Klinger, Eitan, et al.
Published: (2025)
by: Klinger, Eitan, et al.
Published: (2025)
LLM-Powered Grapheme-to-Phoneme Conversion: Benchmark and Case Study
by: Qharabagh, Mahta Fetrat, et al.
Published: (2024)
by: Qharabagh, Mahta Fetrat, et al.
Published: (2024)
Can LLMs Simulate Social Media Engagement? A Study on Action-Guided Response Generation
by: Qiu, Zhongyi, et al.
Published: (2025)
by: Qiu, Zhongyi, et al.
Published: (2025)
Fluency and Faithfulness in Human and Machine Literary Translation
by: Griebel, Sarah, et al.
Published: (2026)
by: Griebel, Sarah, et al.
Published: (2026)
Synthetic Fluency: Hallucinations, Confabulations, and the Creation of Irish Words in LLM-Generated Translations
by: Castilho, Sheila, et al.
Published: (2025)
by: Castilho, Sheila, et al.
Published: (2025)
Shared Lexical Task Representations Explain Behavioral Variability In LLMs
by: Yang, Zhuonan, et al.
Published: (2026)
by: Yang, Zhuonan, et al.
Published: (2026)
Frustratingly Simple Prompting-based Text Denoising
by: Park, Jungyeul, et al.
Published: (2024)
by: Park, Jungyeul, et al.
Published: (2024)
Can Large Language Models Simulate Human Cognition Beyond Behavioral Imitation?
by: Gu, Yuxuan, et al.
Published: (2026)
by: Gu, Yuxuan, et al.
Published: (2026)
Can A Society of Generative Agents Simulate Human Behavior and Inform Public Health Policy? A Case Study on Vaccine Hesitancy
by: Hou, Abe Bohan, et al.
Published: (2025)
by: Hou, Abe Bohan, et al.
Published: (2025)
Teaching Values to Machines: Simulating Human-Like Behavior in LLMs
by: Yehudai, Asaf, et al.
Published: (2026)
by: Yehudai, Asaf, et al.
Published: (2026)
System Report for CCL24-Eval Task 7: Multi-Error Modeling and Fluency-Targeted Pre-training for Chinese Essay Evaluation
by: Zhang, Jingshen, et al.
Published: (2024)
by: Zhang, Jingshen, et al.
Published: (2024)
Can Large Language Model Agents Simulate Human Trust Behavior?
by: Xie, Chengxing, et al.
Published: (2024)
by: Xie, Chengxing, et al.
Published: (2024)
Can LLM Agents Simulate Multi-Turn Human Behavior? Evidence from Real Online Customer Behavior Data
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
Can LLMs Truly Embody Human Personality? Analyzing AI and Human Behavior Alignment in Dispute Resolution
by: Kwon, Deuksin, et al.
Published: (2026)
by: Kwon, Deuksin, et al.
Published: (2026)
Can LLMs Interpret and Leverage Structured Linguistic Representations? A Case Study with AMRs
by: Raut, Ankush, et al.
Published: (2025)
by: Raut, Ankush, et al.
Published: (2025)
How Far Are LLMs from Believable AI? A Benchmark for Evaluating the Believability of Human Behavior Simulation
by: Xiao, Yang, et al.
Published: (2023)
by: Xiao, Yang, et al.
Published: (2023)
Can LLM Reasoning Be Trusted? A Comparative Study: Using Human Benchmarking on Statistical Tasks
by: Nagarkar, Crish, et al.
Published: (2026)
by: Nagarkar, Crish, et al.
Published: (2026)
Knowledge Collapse in LLMs: When Fluency Survives but Facts Fail under Recursive Synthetic Training
by: Keisha, Figarri, et al.
Published: (2025)
by: Keisha, Figarri, et al.
Published: (2025)
Explore the Potential of LLMs in Misinformation Detection: An Empirical Study
by: Chen, Mengyang, et al.
Published: (2023)
by: Chen, Mengyang, et al.
Published: (2023)
From Babbling to Fluency: Evaluating the Evolution of Language Models in Terms of Human Language Acquisition
by: Yang, Qiyuan, et al.
Published: (2024)
by: Yang, Qiyuan, et al.
Published: (2024)
OLaPh: Optimal Language Phonemizer
by: Wirth, Johannes
Published: (2025)
by: Wirth, Johannes
Published: (2025)
OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
Shop-R1: Rewarding LLMs to Simulate Human Behavior in Online Shopping via Reinforcement Learning
by: Zhang, Yimeng, et al.
Published: (2025)
by: Zhang, Yimeng, et al.
Published: (2025)
Prompting with Phonemes: Enhancing LLMs' Multilinguality for Non-Latin Script Languages
by: Nguyen, Hoang H, et al.
Published: (2024)
by: Nguyen, Hoang H, et al.
Published: (2024)
Can LLMs Capture Human Preferences?
by: Goli, Ali, et al.
Published: (2023)
by: Goli, Ali, et al.
Published: (2023)
Performance of Children in Phonemic and Semantic Verbal Fluency Tasks
by: Gilmara de Lucena Leite
Published: (2016)
by: Gilmara de Lucena Leite
Published: (2016)
Modelling the Diachronic Emergence of Phoneme Frequency Distributions
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)
Can Large Language Models Simulate Human Responses? A Case Study of Stated Preference Experiments in the Context of Heating-related Choices
by: Wang, Han, et al.
Published: (2025)
by: Wang, Han, et al.
Published: (2025)
Is my model perplexed for the right reason? Contrasting LLMs' Benchmark Behavior with Token-Level Perplexity
by: Prins, Zoë, et al.
Published: (2026)
by: Prins, Zoë, et al.
Published: (2026)
Simpson's Paradox and the Accuracy-Fluency Tradeoff in Translation
by: Lim, Zheng Wei, et al.
Published: (2024)
by: Lim, Zheng Wei, et al.
Published: (2024)
Reinforcing Human Behavior Simulation via Verbal Feedback
by: Sun, Weiwei, et al.
Published: (2026)
by: Sun, Weiwei, et al.
Published: (2026)
Graph-Based Alternatives to LLMs for Human Simulation
by: Suh, Joseph, et al.
Published: (2025)
by: Suh, Joseph, et al.
Published: (2025)
Investigating Agency of LLMs in Human-AI Collaboration Tasks
by: Sharma, Ashish, et al.
Published: (2023)
by: Sharma, Ashish, et al.
Published: (2023)
Can LLMs Faithfully Explain Themselves in Low-Resource Languages? A Case Study on Emotion Detection in Persian
by: Mehrazar, Mobina, et al.
Published: (2025)
by: Mehrazar, Mobina, et al.
Published: (2025)
Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
by: Zengaffinen, Yanick, et al.
Published: (2026)
by: Zengaffinen, Yanick, et al.
Published: (2026)
Can LLMs Reliably Simulate Human Learner Actions? A Simulation Authoring Framework for Open-Ended Learning Environments
by: Mannekote, Amogh, et al.
Published: (2024)
by: Mannekote, Amogh, et al.
Published: (2024)
AraS2P: Arabic Speech-to-Phonemes System
by: Matar, Bassam, et al.
Published: (2025)
by: Matar, Bassam, et al.
Published: (2025)
LLMs instead of Human Judges? A Large Scale Empirical Study across 20 NLP Evaluation Tasks
by: Bavaresco, Anna, et al.
Published: (2024)
by: Bavaresco, Anna, et al.
Published: (2024)
CoMMET: To What Extent Can LLMs Perform Theory of Mind Tasks?
by: Chen, Ruirui, et al.
Published: (2026)
by: Chen, Ruirui, et al.
Published: (2026)
LLMs as Span Annotators: A Comparative Study of LLMs and Humans
by: Kasner, Zdeněk, et al.
Published: (2025)
by: Kasner, Zdeněk, et al.
Published: (2025)
Similar Items
-
A Formal Framework for Fluency-based Multi-Reference Evaluation in Grammatical Error Correction
by: Klinger, Eitan, et al.
Published: (2025) -
LLM-Powered Grapheme-to-Phoneme Conversion: Benchmark and Case Study
by: Qharabagh, Mahta Fetrat, et al.
Published: (2024) -
Can LLMs Simulate Social Media Engagement? A Study on Action-Guided Response Generation
by: Qiu, Zhongyi, et al.
Published: (2025) -
Fluency and Faithfulness in Human and Machine Literary Translation
by: Griebel, Sarah, et al.
Published: (2026) -
Synthetic Fluency: Hallucinations, Confabulations, and the Creation of Irish Words in LLM-Generated Translations
by: Castilho, Sheila, et al.
Published: (2025)