DEBATE: A Large-Scale Benchmark for Evaluating Opinion Dynamics in Role-Playing LLM Agents
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Chuang, Yun-Shiuan, Tu, Ruixuan, Dai, Chengtao, Li, You, Vasani, Smit, Yao, Binwei, Tessler, Michael Henry, Yang, Sijia, Shah, Dhavan, Hawkins, Robert, Hu, Junjie, Rogers, Timothy T. |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Simulating Opinion Dynamics with Networks of LLM-based Agents
par: Chuang, Yun-Shiuan, et autres
Publié: (2023)
par: Chuang, Yun-Shiuan, et autres
Publié: (2023)
The Wisdom of Partisan Crowds: Comparing Collective Intelligence in Humans and LLM-based Agents
par: Chuang, Yun-Shiuan, et autres
Publié: (2023)
par: Chuang, Yun-Shiuan, et autres
Publié: (2023)
Beyond Demographics: Aligning Role-playing LLM-based Agents Using Human Belief Networks
par: Chuang, Yun-Shiuan, et autres
Publié: (2024)
par: Chuang, Yun-Shiuan, et autres
Publié: (2024)
The Delusional Hedge Algorithm as a Model of Human Learning from Diverse Opinions
par: Chuang, Yun-Shiuan, et autres
Publié: (2024)
par: Chuang, Yun-Shiuan, et autres
Publié: (2024)
Optimizing Social Media Annotation of HPV Vaccine Skepticism and Misinformation Using Large Language Models: An Experimental Evaluation of In-Context Learning and Fine-Tuning Stance Detection Across Multiple Models
par: Sun, Luhang, et autres
Publié: (2024)
par: Sun, Luhang, et autres
Publié: (2024)
No Preference Left Behind: Group Distributional Preference Optimization
par: Yao, Binwei, et autres
Publié: (2024)
par: Yao, Binwei, et autres
Publié: (2024)
Adoption and implication of the Biased-Annotator Competence Estimation (BACE) model into COVID-19 vaccine Twitter data: Human annotation for latent message features
par: Sun, Luhang, et autres
Publié: (2023)
par: Sun, Luhang, et autres
Publié: (2023)
Purrfessor: A Fine-tuned Multimodal LLaVA Diet Health Chatbot
par: Lu, Linqi, et autres
Publié: (2024)
par: Lu, Linqi, et autres
Publié: (2024)
Probing LLM World Models: Enhancing Guesstimation with Wisdom of Crowds Decoding
par: Chuang, Yun-Shiuan, et autres
Publié: (2025)
par: Chuang, Yun-Shiuan, et autres
Publié: (2025)
Benchmarking Machine Translation with Cultural Awareness
par: Yao, Binwei, et autres
Publié: (2023)
par: Yao, Binwei, et autres
Publié: (2023)
CharacterEval: A Chinese Benchmark for Role-Playing Conversational Agent Evaluation
par: Tu, Quan, et autres
Publié: (2024)
par: Tu, Quan, et autres
Publié: (2024)
MINDECHO: Role-Playing Language Agents for Key Opinion Leaders
par: Xu, Rui, et autres
Publié: (2024)
par: Xu, Rui, et autres
Publié: (2024)
Rethinking Role-Playing Evaluation: Anonymous Benchmarking and a Systematic Study of Personality Effects
par: Peng, Ji-Lun, et autres
Publié: (2026)
par: Peng, Ji-Lun, et autres
Publié: (2026)
Geometric phases of reduced states in the transverse-field Ising chain
par: Vasani, Chiragkumar R., et autres
Publié: (2026)
par: Vasani, Chiragkumar R., et autres
Publié: (2026)
SpeechRole: A Large-Scale Dataset and Benchmark for Evaluating Speech Role-Playing Agents
par: Jiang, Changhao, et autres
Publié: (2025)
par: Jiang, Changhao, et autres
Publié: (2025)
Social Science Research in the Arab World and Beyond
par: Tessler, Mark
Publié: (2022)
par: Tessler, Mark
Publié: (2022)
A history of the israeli-palestinian conflict / Mark Tessler
par: Tessler, Mark
Publié: (1994)
par: Tessler, Mark
Publié: (1994)
Firm Pigmented Nodule on the Posterior Thigh in an Adolescent
par: Resham Vasani, et autres
Publié: (2025)
par: Resham Vasani, et autres
Publié: (2025)
AI as a deliberative partner fosters intercultural empathy for Americans but fails for Latin American participants
par: Villanueva, Isabel, et autres
Publié: (2025)
par: Villanueva, Isabel, et autres
Publié: (2025)
Fabrication and characterization of high filled tea oil camellia shells/polypropylene composites: Effect of high filler and maleic anhydride grafted polypropylene
par: Peijun Chen, et autres
Publié: (2025)
par: Peijun Chen, et autres
Publié: (2025)
Rehearse With User: Personalized Opinion Summarization via Role-Playing based on Large Language Models
par: Zhang, Yanyue, et autres
Publié: (2025)
par: Zhang, Yanyue, et autres
Publié: (2025)
The Impact of Parent–Child Relationships on the Mental Health of Returnee Children: The Role of Social Support and Psychological Resilience
par: Binrong Dai, et autres
Publié: (2025)
par: Binrong Dai, et autres
Publié: (2025)
Sparse domination for rough multilinear singular integrals
par: Dan, Binwei, et autres
Publié: (2025)
par: Dan, Binwei, et autres
Publié: (2025)
RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents
par: Lai, Huayi, et autres
Publié: (2026)
par: Lai, Huayi, et autres
Publié: (2026)
What Role Do Social Organizations Play in Stimulating Green and Low‐Carbon Consumption in China?
par: Chuang Li, et autres
Publié: (2025)
par: Chuang Li, et autres
Publié: (2025)
Notes on the one-loop amplituhedron and its BCFW tiling
par: Tessler, Ran J.
Publié: (2025)
par: Tessler, Ran J.
Publié: (2025)
How I Passed the CCNA After a Frustrating Start, Better Practice, and a More Honest Study Routine
par: Rogers, Henry
Publié: (2026)
par: Rogers, Henry
Publié: (2026)
Climate, Energy, and Health: A Systematic Review of Global Risks and Resilience Pathways
par: Amirdha Vasani Sankarkumar, et autres
Publié: (2025)
par: Amirdha Vasani Sankarkumar, et autres
Publié: (2025)
The More Similar, the Better? Associations between Latent Semantic Similarity and Emotional Experiences Differ across Conversation Contexts
par: Yu, Chen-Wei, et autres
Publié: (2023)
par: Yu, Chen-Wei, et autres
Publié: (2023)
Is Semantic Chunking Worth the Computational Cost?
par: Qu, Renyi, et autres
Publié: (2024)
par: Qu, Renyi, et autres
Publié: (2024)
Bleach suit for atopic dermatitis
par: Manish K. Shah, et autres
Publié: (2024)
par: Manish K. Shah, et autres
Publié: (2024)
RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models
par: Wang, Zekun Moore, et autres
Publié: (2023)
par: Wang, Zekun Moore, et autres
Publié: (2023)
VoxRole: A Comprehensive Benchmark for Evaluating Speech-Based Role-Playing Agents
par: Wu, Weihao, et autres
Publié: (2025)
par: Wu, Weihao, et autres
Publié: (2025)
RoleMRC: A Fine-Grained Composite Benchmark for Role-Playing and Instruction-Following
par: Lu, Junru, et autres
Publié: (2025)
par: Lu, Junru, et autres
Publié: (2025)
Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models
par: Shah, Arya, et autres
Publié: (2026)
par: Shah, Arya, et autres
Publié: (2026)
Memory-Driven Role-Playing: Evaluation and Enhancement of Persona Knowledge Utilization in LLMs
par: Wang, Kai, et autres
Publié: (2026)
par: Wang, Kai, et autres
Publié: (2026)
Attention Augmented GNN RNN-Attention Models for Advanced Cybersecurity Intrusion Detection
par: Biradar, Jayant, et autres
Publié: (2025)
par: Biradar, Jayant, et autres
Publié: (2025)
Academic Institutions Must Play a Greater Role in Combating the Predatory Publishing and Conference Industry
par: Henry H. Woo, et autres
Publié: (2025)
par: Henry H. Woo, et autres
Publié: (2025)
Topological expansion for posets and the homological $k$-connectivity of random $q$-complexes
par: Tessler, Ran, et autres
Publié: (2023)
par: Tessler, Ran, et autres
Publié: (2023)
A jumping terrestrial leech from Madagascar
par: Mai Fahmy, et autres
Publié: (2024)
par: Mai Fahmy, et autres
Publié: (2024)
Documents similaires
-
Simulating Opinion Dynamics with Networks of LLM-based Agents
par: Chuang, Yun-Shiuan, et autres
Publié: (2023) -
The Wisdom of Partisan Crowds: Comparing Collective Intelligence in Humans and LLM-based Agents
par: Chuang, Yun-Shiuan, et autres
Publié: (2023) -
Beyond Demographics: Aligning Role-playing LLM-based Agents Using Human Belief Networks
par: Chuang, Yun-Shiuan, et autres
Publié: (2024) -
The Delusional Hedge Algorithm as a Model of Human Learning from Diverse Opinions
par: Chuang, Yun-Shiuan, et autres
Publié: (2024) -
Optimizing Social Media Annotation of HPV Vaccine Skepticism and Misinformation Using Large Language Models: An Experimental Evaluation of In-Context Learning and Fine-Tuning Stance Detection Across Multiple Models
par: Sun, Luhang, et autres
Publié: (2024)