Towards Simulating Social Media Users with LLMs: Evaluating the Operational Validity of Conditioned Comment Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schwager, Nils, Münker, Simon, Plum, Alistair, Rettinger, Achim |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Next Reply Prediction X Dataset: Linguistic Discrepancies in Naively Generated Content
von: Münker, Simon, et al.
Veröffentlicht: (2026)
von: Münker, Simon, et al.
Veröffentlicht: (2026)
Don't Trust Generative Agents to Mimic Communication on Social Networks Unless You Benchmarked their Empirical Realism
von: Münker, Simon, et al.
Veröffentlicht: (2025)
von: Münker, Simon, et al.
Veröffentlicht: (2025)
Zero-shot prompt-based classification: topic labeling in times of foundation models in German Tweets
von: Münker, Simon, et al.
Veröffentlicht: (2024)
von: Münker, Simon, et al.
Veröffentlicht: (2024)
Political Bias in LLMs: Unaligned Moral Values in Agent-centric Simulations
von: Münker, Simon
Veröffentlicht: (2024)
von: Münker, Simon
Veröffentlicht: (2024)
Cultural Bias in Large Language Models: Evaluating AI Agents through Moral Questionnaires
von: Münker, Simon
Veröffentlicht: (2025)
von: Münker, Simon
Veröffentlicht: (2025)
LLM Agents Predict Social Media Reactions but Do Not Outperform Text Classifiers: Benchmarking Simulation Accuracy Using 120K+ Personas of 1511 Humans
von: Bojic, Ljubisa, et al.
Veröffentlicht: (2026)
von: Bojic, Ljubisa, et al.
Veröffentlicht: (2026)
Do LLMs Judge Distantly Supervised Named Entity Labels Well? Constructing the JudgeWEL Dataset
von: Plum, Alistair, et al.
Veröffentlicht: (2026)
von: Plum, Alistair, et al.
Veröffentlicht: (2026)
InvBERT: Reconstructing Text from Contextualized Word Embeddings by inverting the BERT pipeline
von: Kugler, Kai, et al.
Veröffentlicht: (2021)
von: Kugler, Kai, et al.
Veröffentlicht: (2021)
Identity-Aware Large Language Models require Cultural Reasoning
von: Plum, Alistair, et al.
Veröffentlicht: (2025)
von: Plum, Alistair, et al.
Veröffentlicht: (2025)
Agent-Based Simulations of Online Political Discussions: A Case Study on Elections in Germany
von: Sittar, Abdul, et al.
Veröffentlicht: (2025)
von: Sittar, Abdul, et al.
Veröffentlicht: (2025)
A Subword Embedding Approach for Variation Detection in Luxembourgish User Comments
von: Lutgen, Anne-Marie, et al.
Veröffentlicht: (2026)
von: Lutgen, Anne-Marie, et al.
Veröffentlicht: (2026)
$\texttt{BluePrint}$: A Social Media User Dataset for LLM Persona Evaluation and Training
von: Bück-Kaeffer, Aurélien, et al.
Veröffentlicht: (2025)
von: Bück-Kaeffer, Aurélien, et al.
Veröffentlicht: (2025)
Evaluation of Multilingual LLMs Personalized Text Generation Capabilities Targeting Groups and Social-Media Platforms
von: Macko, Dominik
Veröffentlicht: (2026)
von: Macko, Dominik
Veröffentlicht: (2026)
User Behavior Prediction as a Generic, Robust, Scalable, and Low-Cost Evaluation Strategy for Estimating Generalization in LLMs
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
K-SENSE: A Knowledge-Guided Self-Augmented Encoder for Neuro-Semantic Evaluation of Mental Health Conditions on Social Media
von: Yadav, Vijay
Veröffentlicht: (2026)
von: Yadav, Vijay
Veröffentlicht: (2026)
PlatoLM: Teaching LLMs in Multi-Round Dialogue via a User Simulator
von: Kong, Chuyi, et al.
Veröffentlicht: (2023)
von: Kong, Chuyi, et al.
Veröffentlicht: (2023)
Development and Validation of Engagement and Rapport Scales for Evaluating User Experience in Multimodal Dialogue Systems
von: Kurata, Fuma, et al.
Veröffentlicht: (2025)
von: Kurata, Fuma, et al.
Veröffentlicht: (2025)
The Need for a Socially-Grounded Persona Framework for User Simulation
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2026)
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2026)
Leveraging Implicit Sentiments: Enhancing Reliability and Validity in Psychological Trait Evaluation of LLMs
von: Ma, Huanhuan, et al.
Veröffentlicht: (2025)
von: Ma, Huanhuan, et al.
Veröffentlicht: (2025)
Prosa: Rubric-Based Evaluation of LLMs on Real User Chats in Brazilian Portuguese
von: Junior, Roseval Malaquias, et al.
Veröffentlicht: (2026)
von: Junior, Roseval Malaquias, et al.
Veröffentlicht: (2026)
Moral Lenses, Political Coordinates: Towards Ideological Positioning of Morally Conditioned LLMs
von: Yuan, Chenchen, et al.
Veröffentlicht: (2026)
von: Yuan, Chenchen, et al.
Veröffentlicht: (2026)
Data Analysis and Performance Evaluation of Simulation Deduction Based on LLMs
von: Zhang, Shansi, et al.
Veröffentlicht: (2025)
von: Zhang, Shansi, et al.
Veröffentlicht: (2025)
Towards Reliable Evaluation of Behavior Steering Interventions in LLMs
von: Pres, Itamar, et al.
Veröffentlicht: (2024)
von: Pres, Itamar, et al.
Veröffentlicht: (2024)
Contradiction Detection in RAG Systems: Evaluating LLMs as Context Validators for Improved Information Consistency
von: Gokul, Vignesh, et al.
Veröffentlicht: (2025)
von: Gokul, Vignesh, et al.
Veröffentlicht: (2025)
A Comprehensive Evaluation of Cognitive Biases in LLMs
von: Malberg, Simon, et al.
Veröffentlicht: (2024)
von: Malberg, Simon, et al.
Veröffentlicht: (2024)
Beyond Cooperative Simulators: Generating Realistic User Personas for Robust Evaluation of LLM Agents
von: Chopra, Harshita, et al.
Veröffentlicht: (2026)
von: Chopra, Harshita, et al.
Veröffentlicht: (2026)
Is this the real life? Is this just fantasy? The Misleading Success of Simulating Social Interactions With LLMs
von: Zhou, Xuhui, et al.
Veröffentlicht: (2024)
von: Zhou, Xuhui, et al.
Veröffentlicht: (2024)
Fine-Grained and Thematic Evaluation of LLMs in Social Deduction Game
von: Kim, Byungjun, et al.
Veröffentlicht: (2024)
von: Kim, Byungjun, et al.
Veröffentlicht: (2024)
WHBench: Evaluating Frontier LLMs with Expert-in-the-Loop Validation on Women's Health Topics
von: Maurya, Sneha, et al.
Veröffentlicht: (2026)
von: Maurya, Sneha, et al.
Veröffentlicht: (2026)
Unveiling Social Media Comments with a Novel Named Entity Recognition System for Identity Groups
von: Carvallo, Andrés, et al.
Veröffentlicht: (2024)
von: Carvallo, Andrés, et al.
Veröffentlicht: (2024)
User-Aware Multilingual Abusive Content Detection in Social Media
von: Rehman, Mohammad Zia Ur, et al.
Veröffentlicht: (2024)
von: Rehman, Mohammad Zia Ur, et al.
Veröffentlicht: (2024)
Fine-Grained Behavior Simulation with Role-Playing Large Language Model on Social Media
von: Li, Kun, et al.
Veröffentlicht: (2024)
von: Li, Kun, et al.
Veröffentlicht: (2024)
RLPF: Reinforcement Learning from Prediction Feedback for User Summarization with LLMs
von: Wu, Jiaxing, et al.
Veröffentlicht: (2024)
von: Wu, Jiaxing, et al.
Veröffentlicht: (2024)
Using LLM-as-a-Judge/Jury to Advance Scalable, Clinically-Validated Safety Evaluations of Model Responses to Users Demonstrating Psychosis
von: Reese, May Lynn, et al.
Veröffentlicht: (2026)
von: Reese, May Lynn, et al.
Veröffentlicht: (2026)
YouTube Comments Decoded: Leveraging LLMs for Low Resource Language Classification
von: Deroy, Aniket, et al.
Veröffentlicht: (2024)
von: Deroy, Aniket, et al.
Veröffentlicht: (2024)
PSI-Bench: Towards Clinically Grounded and Interpretable Evaluation of Depression Patient Simulators
von: Hoang, Nguyen Khoi, et al.
Veröffentlicht: (2026)
von: Hoang, Nguyen Khoi, et al.
Veröffentlicht: (2026)
LiveCLKTBench: Towards Reliable Evaluation of Cross-Lingual Knowledge Transfer in Multilingual LLMs
von: Guo, Pei-Fu, et al.
Veröffentlicht: (2025)
von: Guo, Pei-Fu, et al.
Veröffentlicht: (2025)
ConsistencyAI: A Benchmark to Assess LLMs' Factual Consistency When Responding to Different Demographic Groups
von: Banyas, Peter, et al.
Veröffentlicht: (2025)
von: Banyas, Peter, et al.
Veröffentlicht: (2025)
Validating Political Position Predictions of Arguments
von: Robinson, Jordan, et al.
Veröffentlicht: (2026)
von: Robinson, Jordan, et al.
Veröffentlicht: (2026)
Understanding Mental Health Content on Social Media and Its Effect Towards Suicidal Ideation
von: Bhuiyan, Mohaiminul Islam, et al.
Veröffentlicht: (2025)
von: Bhuiyan, Mohaiminul Islam, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Next Reply Prediction X Dataset: Linguistic Discrepancies in Naively Generated Content
von: Münker, Simon, et al.
Veröffentlicht: (2026) -
Don't Trust Generative Agents to Mimic Communication on Social Networks Unless You Benchmarked their Empirical Realism
von: Münker, Simon, et al.
Veröffentlicht: (2025) -
Zero-shot prompt-based classification: topic labeling in times of foundation models in German Tweets
von: Münker, Simon, et al.
Veröffentlicht: (2024) -
Political Bias in LLMs: Unaligned Moral Values in Agent-centric Simulations
von: Münker, Simon
Veröffentlicht: (2024) -
Cultural Bias in Large Language Models: Evaluating AI Agents through Moral Questionnaires
von: Münker, Simon
Veröffentlicht: (2025)