Representational Harms in LLM-Generated Narratives Against Global Majority Nationalities
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Ilana, Suresh, Harini, Monroe-White, Thema, Shieh, Evan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Psychosocial Impacts of Generative AI Harms
by: Vassel, Faye-Marie, et al.
Published: (2024)
by: Vassel, Faye-Marie, et al.
Published: (2024)
Laissez-Faire Harms: Algorithmic Biases in Generative Language Models
by: Shieh, Evan, et al.
Published: (2024)
by: Shieh, Evan, et al.
Published: (2024)
Which Institutional Frameworks Do Chatbots Assume? Auditing Jurisdictional Defaults in Multilingual LLMs
by: Wang, Zhizhi, et al.
Published: (2026)
by: Wang, Zhizhi, et al.
Published: (2026)
Representation Noising: A Defence Mechanism Against Harmful Finetuning
by: Rosati, Domenic, et al.
Published: (2024)
by: Rosati, Domenic, et al.
Published: (2024)
The Howard-Harvard effect: Institutional reproduction of intersectional inequalities
by: Kozlowski, Diego, et al.
Published: (2024)
by: Kozlowski, Diego, et al.
Published: (2024)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
by: Brown, Hannah, et al.
Published: (2024)
by: Brown, Hannah, et al.
Published: (2024)
Understanding and Meeting Practitioner Needs When Measuring Representational Harms Caused by LLM-Based Systems
by: Harvey, Emma, et al.
Published: (2025)
by: Harvey, Emma, et al.
Published: (2025)
Single Character Perturbations Break LLM Alignment
by: Lin, Leon, et al.
Published: (2024)
by: Lin, Leon, et al.
Published: (2024)
Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM
by: Zhang, Chi, et al.
Published: (2025)
by: Zhang, Chi, et al.
Published: (2025)
HarmMetric Eval: Benchmarking Metrics and Judges for LLM Harmfulness Assessment
by: Yang, Langqi, et al.
Published: (2025)
by: Yang, Langqi, et al.
Published: (2025)
Self-HarmLLM: Can Large Language Model Harm Itself?
by: Kim, Heehwan, et al.
Published: (2025)
by: Kim, Heehwan, et al.
Published: (2025)
Taxonomizing Representational Harms using Speech Act Theory
by: Corvi, Emily, et al.
Published: (2025)
by: Corvi, Emily, et al.
Published: (2025)
More of the Same: Persistent Representational Harms Under Increased Representation
by: Mickel, Jennifer, et al.
Published: (2025)
by: Mickel, Jennifer, et al.
Published: (2025)
Robust LLM Unlearning Against Relearning Attacks: The Minor Components in Representations Matter
by: Xiao, Zeguan, et al.
Published: (2026)
by: Xiao, Zeguan, et al.
Published: (2026)
Generation-Step-Aware Framework for Cross-Modal Representation and Control in Multilingual Speech-Text Models
by: Nakai, Toshiki, et al.
Published: (2026)
by: Nakai, Toshiki, et al.
Published: (2026)
LLM-based Semantic Augmentation for Harmful Content Detection
by: Meguellati, Elyas, et al.
Published: (2025)
by: Meguellati, Elyas, et al.
Published: (2025)
AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents
by: Andriushchenko, Maksym, et al.
Published: (2024)
by: Andriushchenko, Maksym, et al.
Published: (2024)
"Ownership, Not Just Happy Talk": Co-Designing a Participatory Large Language Model for Journalism
by: Tseng, Emily, et al.
Published: (2025)
by: Tseng, Emily, et al.
Published: (2025)
Latent Fusion Jailbreak: Blending Harmful and Harmless Representations to Elicit Unsafe LLM Outputs
by: Xing, Wenpeng, et al.
Published: (2025)
by: Xing, Wenpeng, et al.
Published: (2025)
From Representational Harms to Quality-of-Service Harms: A Case Study on Llama 2 Safety Safeguards
by: Chehbouni, Khaoula, et al.
Published: (2024)
by: Chehbouni, Khaoula, et al.
Published: (2024)
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning
by: Yi, Xin, et al.
Published: (2024)
by: Yi, Xin, et al.
Published: (2024)
Should LLM Safety Be More Than Refusing Harmful Instructions?
by: Maskey, Utsav, et al.
Published: (2025)
by: Maskey, Utsav, et al.
Published: (2025)
Unintended Impacts of LLM Alignment on Global Representation
by: Ryan, Michael J., et al.
Published: (2024)
by: Ryan, Michael J., et al.
Published: (2024)
SocialHarmBench: Revealing LLM Vulnerabilities to Socially Harmful Requests
by: Pandey, Punya Syon, et al.
Published: (2025)
by: Pandey, Punya Syon, et al.
Published: (2025)
Investigating and Alleviating Harm Amplification in LLM Interactions
by: Guo, Ruohao, et al.
Published: (2026)
by: Guo, Ruohao, et al.
Published: (2026)
Coreference Resolution for Vietnamese Narrative Texts
by: Tran, Hieu-Dai, et al.
Published: (2025)
by: Tran, Hieu-Dai, et al.
Published: (2025)
Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation Learning
by: An, Li, et al.
Published: (2025)
by: An, Li, et al.
Published: (2025)
SemEval-2026 Task 4: Narrative Story Similarity and Narrative Representation Learning
by: Hatzel, Hans Ole, et al.
Published: (2026)
by: Hatzel, Hans Ole, et al.
Published: (2026)
WhoSaidIt: Human-LLM Collaborative Annotation for Text-Based Multilingual Speaker-Attribute Classification
by: Gao, Lingyu, et al.
Published: (2026)
by: Gao, Lingyu, et al.
Published: (2026)
A LLM-Based Ranking Method for the Evaluation of Automatic Counter-Narrative Generation
by: Zubiaga, Irune, et al.
Published: (2024)
by: Zubiaga, Irune, et al.
Published: (2024)
LLM Judges Inconsistently Disagree Across Safety Criteria and Harm Categories
by: Vishnubhotla, Krishnapriya, et al.
Published: (2026)
by: Vishnubhotla, Krishnapriya, et al.
Published: (2026)
"They are uncultured": Unveiling Covert Harms and Social Threats in LLM Generated Conversations
by: Dammu, Preetam Prabhu Srikar, et al.
Published: (2024)
by: Dammu, Preetam Prabhu Srikar, et al.
Published: (2024)
People Make Better Edits: Measuring the Efficacy of LLM-Generated Counterfactually Augmented Data for Harmful Language Detection
by: Sen, Indira, et al.
Published: (2023)
by: Sen, Indira, et al.
Published: (2023)
Revisiting Generalization Across Difficulty Levels: It's Not So Easy
by: Kordi, Yeganeh, et al.
Published: (2025)
by: Kordi, Yeganeh, et al.
Published: (2025)
CHIRON: Rich Character Representations in Long-Form Narratives
by: Gurung, Alexander, et al.
Published: (2024)
by: Gurung, Alexander, et al.
Published: (2024)
LLM-based Detection of Manipulative Political Narratives
by: Schneider, Sinclair, et al.
Published: (2026)
by: Schneider, Sinclair, et al.
Published: (2026)
LongRLVR: Long-Context Reinforcement Learning Requires Verifiable Context Rewards
by: Chen, Guanzheng, et al.
Published: (2026)
by: Chen, Guanzheng, et al.
Published: (2026)
How to Stop Playing Whack-a-Mole: Mapping the Ecosystem of Technologies Facilitating AI-Generated Non-Consensual Intimate Images
by: Ding, Michelle L., et al.
Published: (2026)
by: Ding, Michelle L., et al.
Published: (2026)
LLM for Comparative Narrative Analysis
by: Kampen, Leo, et al.
Published: (2025)
by: Kampen, Leo, et al.
Published: (2025)
The Howard‐Harvard effect: Institutional reproduction of intersectional inequalities
by: Diego Kozlowski, et al.
Published: (2024)
by: Diego Kozlowski, et al.
Published: (2024)
Similar Items
-
The Psychosocial Impacts of Generative AI Harms
by: Vassel, Faye-Marie, et al.
Published: (2024) -
Laissez-Faire Harms: Algorithmic Biases in Generative Language Models
by: Shieh, Evan, et al.
Published: (2024) -
Which Institutional Frameworks Do Chatbots Assume? Auditing Jurisdictional Defaults in Multilingual LLMs
by: Wang, Zhizhi, et al.
Published: (2026) -
Representation Noising: A Defence Mechanism Against Harmful Finetuning
by: Rosati, Domenic, et al.
Published: (2024) -
The Howard-Harvard effect: Institutional reproduction of intersectional inequalities
by: Kozlowski, Diego, et al.
Published: (2024)