Diverse, but Divisive: LLMs Can Exaggerate Gender Differences in Opinion Related to Harms of Misinformation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Neumann, Terrence, Lee, Sooyong, De-Arteaga, Maria, Fazelpour, Sina, Lease, Matthew |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Should you use LLMs to simulate opinions? Quality checks for early-stage deliberation
par: Neumann, Terrence, et autres
Publié: (2025)
par: Neumann, Terrence, et autres
Publié: (2025)
Disciplining Deliberation: A Sociotechnical Perspective on Machine Learning Trade-offs
par: Fazelpour, Sina
Publié: (2024)
par: Fazelpour, Sina
Publié: (2024)
The Value of Disagreement in AI Design, Evaluation, and Alignment
par: Fazelpour, Sina, et autres
Publié: (2025)
par: Fazelpour, Sina, et autres
Publié: (2025)
Aspirational Affordances of AI
par: Fazelpour, Sina, et autres
Publié: (2025)
par: Fazelpour, Sina, et autres
Publié: (2025)
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
par: Dorn, Rebecca, et autres
Publié: (2024)
par: Dorn, Rebecca, et autres
Publié: (2024)
Ambiguity Collapse by LLMs: A Taxonomy of Epistemic Risks
par: Gur-Arieh, Shira, et autres
Publié: (2026)
par: Gur-Arieh, Shira, et autres
Publié: (2026)
Take Caution in Using LLMs as Human Surrogates: Scylla Ex Machina
par: Gao, Yuan, et autres
Publié: (2024)
par: Gao, Yuan, et autres
Publié: (2024)
Evaluating the Simulation of Human Personality-Driven Susceptibility to Misinformation with LLMs
par: Pratelli, Manuel, et autres
Publié: (2025)
par: Pratelli, Manuel, et autres
Publié: (2025)
On the Sensitivity of Instruction-tuned LLMs to Harmful Sentences in Long Inputs
par: Ghorbanpour, Faeze, et autres
Publié: (2025)
par: Ghorbanpour, Faeze, et autres
Publié: (2025)
Authenticity and exclusion: social media algorithms and the dynamics of belonging in epistemic communities
par: Akpinar, Nil-Jana, et autres
Publié: (2024)
par: Akpinar, Nil-Jana, et autres
Publié: (2024)
Explore the Potential of LLMs in Misinformation Detection: An Empirical Study
par: Chen, Mengyang, et autres
Publié: (2023)
par: Chen, Mengyang, et autres
Publié: (2023)
Between Help and Harm: An Evaluation of Mental Health Crisis Handling by LLMs
par: Arnaiz-Rodriguez, Adrian, et autres
Publié: (2025)
par: Arnaiz-Rodriguez, Adrian, et autres
Publié: (2025)
Do Prevalent Bias Metrics Capture Allocational Harms from LLMs?
par: Cyberey, Hannah, et autres
Publié: (2024)
par: Cyberey, Hannah, et autres
Publié: (2024)
Are Rationales Necessary and Sufficient? Tuning LLMs for Explainable Misinformation Detection
par: Wang, Bing, et autres
Publié: (2026)
par: Wang, Bing, et autres
Publié: (2026)
When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms
par: Chun, Chaewan, et autres
Publié: (2026)
par: Chun, Chaewan, et autres
Publié: (2026)
Human-centered NLP Fact-checking: Co-Designing with Fact-checkers using Matchmaking for AI
par: Liu, Houjiang, et autres
Publié: (2023)
par: Liu, Houjiang, et autres
Publié: (2023)
Generative Debunking of Climate Misinformation
par: Zanartu, Francisco, et autres
Publié: (2024)
par: Zanartu, Francisco, et autres
Publié: (2024)
HRIPBench: Benchmarking LLMs in Harm Reduction Information Provision to Support People Who Use Drugs
par: Wang, Kaixuan, et autres
Publié: (2025)
par: Wang, Kaixuan, et autres
Publié: (2025)
Unintended Harms of Value-Aligned LLMs: Psychological and Empirical Insights
par: Choi, Sooyung, et autres
Publié: (2025)
par: Choi, Sooyung, et autres
Publié: (2025)
Taxonomizing Representational Harms using Speech Act Theory
par: Corvi, Emily, et autres
Publié: (2025)
par: Corvi, Emily, et autres
Publié: (2025)
Finding Pareto Trade-offs in Fair and Accurate Detection of Toxic Speech
par: Gupta, Soumyajit, et autres
Publié: (2022)
par: Gupta, Soumyajit, et autres
Publié: (2022)
Robust Misinformation Detection by Visiting Potential Commonsense Conflict
par: Wang, Bing, et autres
Publié: (2025)
par: Wang, Bing, et autres
Publié: (2025)
Unveiling Scoring Processes: Dissecting the Differences between LLMs and Human Graders in Automatic Scoring
par: Wu, Xuansheng, et autres
Publié: (2024)
par: Wu, Xuansheng, et autres
Publié: (2024)
Survival at Any Cost? LLMs and the Choice Between Self-Preservation and Human Harm
par: Mohamadi, Alireza, et autres
Publié: (2025)
par: Mohamadi, Alireza, et autres
Publié: (2025)
The Psychology of Falsehood: A Human-Centric Survey of Misinformation Detection
par: Nandi, Arghodeep, et autres
Publié: (2025)
par: Nandi, Arghodeep, et autres
Publié: (2025)
Careless Whisper: Speech-to-Text Hallucination Harms
par: Koenecke, Allison, et autres
Publié: (2024)
par: Koenecke, Allison, et autres
Publié: (2024)
Cultural Value Differences of LLMs: Prompt, Language, and Model Size
par: Zhong, Qishuai, et autres
Publié: (2024)
par: Zhong, Qishuai, et autres
Publié: (2024)
Accommodation and Epistemic Vigilance: A Pragmatic Account of Why LLMs Fail to Challenge Harmful Beliefs
par: Cheng, Myra, et autres
Publié: (2026)
par: Cheng, Myra, et autres
Publié: (2026)
Expected Harm: Rethinking Safety Evaluation of (Mis)Aligned LLMs
par: Chen, Yen-Shan, et autres
Publié: (2026)
par: Chen, Yen-Shan, et autres
Publié: (2026)
LLM-based Semantic Augmentation for Harmful Content Detection
par: Meguellati, Elyas, et autres
Publié: (2025)
par: Meguellati, Elyas, et autres
Publié: (2025)
Why They Disagree: Decoding Differences in Opinions about AI Risk on the Lex Fridman Podcast
par: Truong, Nghi, et autres
Publié: (2025)
par: Truong, Nghi, et autres
Publié: (2025)
PluriHarms: Benchmarking the Full Spectrum of Human Judgments on AI Harm
par: Li, Jing-Jing, et autres
Publié: (2026)
par: Li, Jing-Jing, et autres
Publié: (2026)
Fairness through Difference Awareness: Measuring Desired Group Discrimination in LLMs
par: Wang, Angelina, et autres
Publié: (2025)
par: Wang, Angelina, et autres
Publié: (2025)
A Capabilities Approach to Studying Bias and Harm in Language Technologies
par: Nigatu, Hellina Hailu, et autres
Publié: (2024)
par: Nigatu, Hellina Hailu, et autres
Publié: (2024)
The Earth is Flat because...: Investigating LLMs' Belief towards Misinformation via Persuasive Conversation
par: Xu, Rongwu, et autres
Publié: (2023)
par: Xu, Rongwu, et autres
Publié: (2023)
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
par: Zhou, Yuhang, et autres
Publié: (2025)
par: Zhou, Yuhang, et autres
Publié: (2025)
SoK: Machine Learning for Misinformation Detection
par: Xiao, Madelyne, et autres
Publié: (2023)
par: Xiao, Madelyne, et autres
Publié: (2023)
Speak Easy: Eliciting Harmful Jailbreaks from LLMs with Simple Interactions
par: Chan, Yik Siu, et autres
Publié: (2025)
par: Chan, Yik Siu, et autres
Publié: (2025)
Rethinking Hate Speech Detection on Social Media: Can LLMs Replace Traditional Models?
par: Singh, Daman Deep, et autres
Publié: (2025)
par: Singh, Daman Deep, et autres
Publié: (2025)
LLMs Can Infer Political Alignment from Online Conversations
par: Lee, Byunghwee, et autres
Publié: (2026)
par: Lee, Byunghwee, et autres
Publié: (2026)
Documents similaires
-
Should you use LLMs to simulate opinions? Quality checks for early-stage deliberation
par: Neumann, Terrence, et autres
Publié: (2025) -
Disciplining Deliberation: A Sociotechnical Perspective on Machine Learning Trade-offs
par: Fazelpour, Sina
Publié: (2024) -
The Value of Disagreement in AI Design, Evaluation, and Alignment
par: Fazelpour, Sina, et autres
Publié: (2025) -
Aspirational Affordances of AI
par: Fazelpour, Sina, et autres
Publié: (2025) -
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
par: Dorn, Rebecca, et autres
Publié: (2024)