Real or Robotic? Assessing Whether LLMs Accurately Simulate Qualities of Human Responses in Dialogue
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ivey, Jonathan, Kumar, Shivani, Liu, Jiayu, Shen, Hua, Rakshit, Sushrita, Raju, Rohan, Zhang, Haotian, Ananthasubramaniam, Aparna, Kim, Junghwan, Yi, Bowen, Wright, Dustin, Israeli, Abraham, Møller, Anders Giovanni, Zhang, Lechen, Jurgens, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Role of Network and Identity in the Diffusion of Hashtags
von: Ananthasubramaniam, Aparna, et al.
Veröffentlicht: (2024)
von: Ananthasubramaniam, Aparna, et al.
Veröffentlicht: (2024)
Leveraging Multilingual Training for Authorship Representation: Enhancing Generalization across Languages and Domains
von: Kim, Junghwan, et al.
Veröffentlicht: (2025)
von: Kim, Junghwan, et al.
Veröffentlicht: (2025)
A Test of Time: Predicting the Sustainable Success of Online Collaboration in Wikipedia
von: Israeli, Abraham, et al.
Veröffentlicht: (2024)
von: Israeli, Abraham, et al.
Veröffentlicht: (2024)
Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions
von: Rakshit, Sushrita, et al.
Veröffentlicht: (2026)
von: Rakshit, Sushrita, et al.
Veröffentlicht: (2026)
KODIS: A Multicultural Dispute Resolution Dialogue Corpus
von: Hale, James, et al.
Veröffentlicht: (2025)
von: Hale, James, et al.
Veröffentlicht: (2025)
Cross-Lingual Prompt Steerability: Towards Accurate and Robust LLM Behavior across Languages
von: Zhang, Lechen, et al.
Veröffentlicht: (2025)
von: Zhang, Lechen, et al.
Veröffentlicht: (2025)
Modeling Public Perceptions of Science in Media
von: Pei, Jiaxin, et al.
Veröffentlicht: (2025)
von: Pei, Jiaxin, et al.
Veröffentlicht: (2025)
NUTMEG: Separating Signal From Noise in Annotator Disagreement
von: Ivey, Jonathan, et al.
Veröffentlicht: (2025)
von: Ivey, Jonathan, et al.
Veröffentlicht: (2025)
SPRIG: Improving Large Language Model Performance by System Prompt Optimization
von: Zhang, Lechen, et al.
Veröffentlicht: (2024)
von: Zhang, Lechen, et al.
Veröffentlicht: (2024)
Are Rules Meant to be Broken? Understanding Multilingual Moral Reasoning as a Computational Pipeline with UniMoral
von: Kumar, Shivani, et al.
Veröffentlicht: (2025)
von: Kumar, Shivani, et al.
Veröffentlicht: (2025)
Cooperative Profiles Predict Multi-Agent LLM Team Performance in AI for Science Workflows
von: Kumar, Shivani, et al.
Veröffentlicht: (2026)
von: Kumar, Shivani, et al.
Veröffentlicht: (2026)
CUDA-Accelerated Soft Robot Neural Evolution with Large Language Model Supervision
von: Zhang, Lechen
Veröffentlicht: (2024)
von: Zhang, Lechen
Veröffentlicht: (2024)
Emotionally-Aware Agents for Dispute Resolution
von: Rakshit, Sushrita, et al.
Veröffentlicht: (2025)
von: Rakshit, Sushrita, et al.
Veröffentlicht: (2025)
Social Hatred: Efficient Multimodal Detection of Hatemongers
von: Marzea, Tom, et al.
Veröffentlicht: (2025)
von: Marzea, Tom, et al.
Veröffentlicht: (2025)
AggregHate: An Efficient Aggregative Approach for the Detection of Hatemongers on Social Platforms
von: Marzea, Tom, et al.
Veröffentlicht: (2024)
von: Marzea, Tom, et al.
Veröffentlicht: (2024)
Unstructured Evidence Attribution for Long Context Query Focused Summarization
von: Wright, Dustin, et al.
Veröffentlicht: (2025)
von: Wright, Dustin, et al.
Veröffentlicht: (2025)
The Impact of Generative AI on Social Media: An Experimental Study
von: Møller, Anders Giovanni, et al.
Veröffentlicht: (2025)
von: Møller, Anders Giovanni, et al.
Veröffentlicht: (2025)
Machine Understanding of Scientific Language
von: Wright, Dustin
Veröffentlicht: (2025)
von: Wright, Dustin
Veröffentlicht: (2025)
Characterizing Online Activities Contributing to Suicide Mortality among Youth
von: Ananthasubramaniam, Aparna, et al.
Veröffentlicht: (2025)
von: Ananthasubramaniam, Aparna, et al.
Veröffentlicht: (2025)
You don't need a personality test to know these models are unreliable: Assessing the Reliability of Large Language Models on Psychometric Instruments
von: Shu, Bangzhao, et al.
Veröffentlicht: (2023)
von: Shu, Bangzhao, et al.
Veröffentlicht: (2023)
Harmonizing Code-mixed Conversations: Personality-assisted Code-mixed Response Generation in Dialogues
von: Kumar, Shivani, et al.
Veröffentlicht: (2024)
von: Kumar, Shivani, et al.
Veröffentlicht: (2024)
Whether to trust: the ML leap of faith
von: Frame, Tory, et al.
Veröffentlicht: (2024)
von: Frame, Tory, et al.
Veröffentlicht: (2024)
Aggregating Soft Labels from Crowd Annotations Improves Uncertainty Estimation Under Distribution Shift
von: Wright, Dustin, et al.
Veröffentlicht: (2022)
von: Wright, Dustin, et al.
Veröffentlicht: (2022)
Skill-Aware Data Selection and Fine-Tuning for Data-Efficient Reasoning Distillation
von: Zhang, Lechen, et al.
Veröffentlicht: (2026)
von: Zhang, Lechen, et al.
Veröffentlicht: (2026)
The Real Her? Exploring Whether Young Adults Accept Human-AI Love
von: Zhang, Shuning, et al.
Veröffentlicht: (2025)
von: Zhang, Shuning, et al.
Veröffentlicht: (2025)
One Model, Many Morals: Uncovering Cross-Linguistic Misalignments in Computational Moral Reasoning
von: Farid, Sualeha, et al.
Veröffentlicht: (2025)
von: Farid, Sualeha, et al.
Veröffentlicht: (2025)
The Language of Bargaining: Linguistic Effects in LLM Negotiations
von: Sinha, Stuti, et al.
Veröffentlicht: (2026)
von: Sinha, Stuti, et al.
Veröffentlicht: (2026)
Herd Mentality in Augmentation -- Not a Good Idea! A Robust Multi-stage Approach towards Deepfake Detection
von: Monu, et al.
Veröffentlicht: (2024)
von: Monu, et al.
Veröffentlicht: (2024)
FactBench: A Dynamic Benchmark for In-the-Wild Language Model Factuality Evaluation
von: Bayat, Farima Fatahi, et al.
Veröffentlicht: (2024)
von: Bayat, Farima Fatahi, et al.
Veröffentlicht: (2024)
Prompt Refinement or Fine-tuning? Best Practices for using LLMs in Computational Social Science Tasks
von: Møller, Anders Giovanni, et al.
Veröffentlicht: (2024)
von: Møller, Anders Giovanni, et al.
Veröffentlicht: (2024)
Inductive-Deductive Strategy Reuse for Multi-Turn Instructional Dialogues
von: Ou, Jiao, et al.
Veröffentlicht: (2024)
von: Ou, Jiao, et al.
Veröffentlicht: (2024)
Multi-Intent Recognition in Dialogue Understanding: A Comparison Between Smaller Open-Source LLMs
von: Ahmad, Adnan, et al.
Veröffentlicht: (2025)
von: Ahmad, Adnan, et al.
Veröffentlicht: (2025)
CoDial: Interpretable Task-Oriented Dialogue Systems Through Dialogue Flow Alignment
von: Shayanfar, Radin, et al.
Veröffentlicht: (2025)
von: Shayanfar, Radin, et al.
Veröffentlicht: (2025)
VeriFact: Enhancing Long-Form Factuality Evaluation with Refined Fact Extraction and Reference Facts
von: Liu, Xin, et al.
Veröffentlicht: (2025)
von: Liu, Xin, et al.
Veröffentlicht: (2025)
Enhancing the Preference Extractor in Multi-turn Dialogues: From Annotating Disasters to Accurate Preference Extraction
von: Wang, Cheng, et al.
Veröffentlicht: (2025)
von: Wang, Cheng, et al.
Veröffentlicht: (2025)
Mapping the Podcast Ecosystem with the Structured Podcast Research Corpus
von: Litterer, Benjamin, et al.
Veröffentlicht: (2024)
von: Litterer, Benjamin, et al.
Veröffentlicht: (2024)
Deep Value Benchmark: Measuring Whether Models Generalize Deep Values or Shallow Preferences
von: Ashkinaze, Joshua, et al.
Veröffentlicht: (2025)
von: Ashkinaze, Joshua, et al.
Veröffentlicht: (2025)
Simulating User Diversity in Task-Oriented Dialogue Systems using Large Language Models
von: Ahmad, Adnan, et al.
Veröffentlicht: (2025)
von: Ahmad, Adnan, et al.
Veröffentlicht: (2025)
LW2G: Learning Whether to Grow for Prompt-based Continual Learning
von: Feng, Qian, et al.
Veröffentlicht: (2024)
von: Feng, Qian, et al.
Veröffentlicht: (2024)
Digital Diasporas: How Origin Characteristics and Host-Native Distance Shape Immigrants' Online Cultural Retention
von: Khatua, Aparup, et al.
Veröffentlicht: (2025)
von: Khatua, Aparup, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Role of Network and Identity in the Diffusion of Hashtags
von: Ananthasubramaniam, Aparna, et al.
Veröffentlicht: (2024) -
Leveraging Multilingual Training for Authorship Representation: Enhancing Generalization across Languages and Domains
von: Kim, Junghwan, et al.
Veröffentlicht: (2025) -
A Test of Time: Predicting the Sustainable Success of Online Collaboration in Wikipedia
von: Israeli, Abraham, et al.
Veröffentlicht: (2024) -
Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions
von: Rakshit, Sushrita, et al.
Veröffentlicht: (2026) -
KODIS: A Multicultural Dispute Resolution Dialogue Corpus
von: Hale, James, et al.
Veröffentlicht: (2025)