The Impact of Persona-based Political Perspectives on Hateful Content Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Civelli, Stefano, Bernardelle, Pietro, Demartini, Gianluca |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Mapping and Influencing the Political Ideology of Large Language Models using Synthetic Personas
por: Bernardelle, Pietro, et al.
Publicado: (2024)
por: Bernardelle, Pietro, et al.
Publicado: (2024)
A Shared Geometry of Difficulty in Multilingual Language Models
por: Civelli, Stefano, et al.
Publicado: (2026)
por: Civelli, Stefano, et al.
Publicado: (2026)
SubData: Bridging Heterogeneous Datasets to Enable Theory-Driven Evaluation of Political and Demographic Perspectives in LLMs
por: Bernardelle, Pietro, et al.
Publicado: (2024)
por: Bernardelle, Pietro, et al.
Publicado: (2024)
Towards Detecting Persuasion on Social Media: From Model Development to Insights on Persuasion Strategies
por: Meguellati, Elyas, et al.
Publicado: (2025)
por: Meguellati, Elyas, et al.
Publicado: (2025)
Ideology-Based LLMs for Content Moderation
por: Civelli, Stefano, et al.
Publicado: (2025)
por: Civelli, Stefano, et al.
Publicado: (2025)
Context Shapes LLMs Retrieval-Augmented Fact-Checking Effectiveness
por: Bernardelle, Pietro, et al.
Publicado: (2026)
por: Bernardelle, Pietro, et al.
Publicado: (2026)
Political Advertising on Facebook During the 2022 Australian Federal Election: A Social Identity Perspective
por: Civelli, Stefano, et al.
Publicado: (2025)
por: Civelli, Stefano, et al.
Publicado: (2025)
Optimizing LLMs with Direct Preferences: A Data Efficiency Perspective
por: Bernardelle, Pietro, et al.
Publicado: (2024)
por: Bernardelle, Pietro, et al.
Publicado: (2024)
Political Ideology Shifts in Large Language Models
por: Bernardelle, Pietro, et al.
Publicado: (2025)
por: Bernardelle, Pietro, et al.
Publicado: (2025)
Query-Document Dense Vectors for LLM Relevance Judgment Bias Analysis
por: Mohtadi, Samaneh, et al.
Publicado: (2026)
por: Mohtadi, Samaneh, et al.
Publicado: (2026)
The Effect of Document Summarization on LLM-Based Relevance Judgments
por: Mohtadi, Samaneh, et al.
Publicado: (2025)
por: Mohtadi, Samaneh, et al.
Publicado: (2025)
Identification of Regulatory Requirements Relevant to Business Processes: A Comparative Study on Generative AI, Embedding-based Ranking, Crowd and Expert-driven Methods
por: Sai, Catherine, et al.
Publicado: (2024)
por: Sai, Catherine, et al.
Publicado: (2024)
HatePrototypes: Interpretable and Transferable Representations for Implicit and Explicit Hate Speech Detection
por: Proskurina, Irina, et al.
Publicado: (2025)
por: Proskurina, Irina, et al.
Publicado: (2025)
HateSieve: A Contrastive Learning Framework for Detecting and Segmenting Hateful Content in Multimodal Memes
por: Su, Xuanyu, et al.
Publicado: (2024)
por: Su, Xuanyu, et al.
Publicado: (2024)
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models
por: Van, Minh-Hao, et al.
Publicado: (2025)
por: Van, Minh-Hao, et al.
Publicado: (2025)
LLM-based Semantic Augmentation for Harmful Content Detection
por: Meguellati, Elyas, et al.
Publicado: (2025)
por: Meguellati, Elyas, et al.
Publicado: (2025)
xList-Hate: A Checklist-Based Framework for Interpretable and Generalizable Hate Speech Detection
por: Girón, Adrián, et al.
Publicado: (2026)
por: Girón, Adrián, et al.
Publicado: (2026)
LLM-Generated Ads: From Personalization Parity to Persuasion Superiority
por: Meguellati, Elyas, et al.
Publicado: (2025)
por: Meguellati, Elyas, et al.
Publicado: (2025)
Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation
por: Lee, Huije, et al.
Publicado: (2026)
por: Lee, Huije, et al.
Publicado: (2026)
Natural Language Processing for the Legal Domain: A Survey of Tasks, Datasets, Models, and Challenges
por: Ariai, Farid, et al.
Publicado: (2024)
por: Ariai, Farid, et al.
Publicado: (2024)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
por: Kumarage, Tharindu, et al.
Publicado: (2024)
por: Kumarage, Tharindu, et al.
Publicado: (2024)
Selective Demonstration Retrieval for Improved Implicit Hate Speech Detection
por: Kim, Yumin, et al.
Publicado: (2025)
por: Kim, Yumin, et al.
Publicado: (2025)
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
por: Fröhling, Leon, et al.
Publicado: (2024)
por: Fröhling, Leon, et al.
Publicado: (2024)
Towards Fairness Assessment of Dutch Hate Speech Detection
por: Bauer, Julie, et al.
Publicado: (2025)
por: Bauer, Julie, et al.
Publicado: (2025)
Embracing Dialectic Intersubjectivity: Coordination of Different Perspectives in Content Analysis with LLM Persona Simulation
por: Kang, Taewoo, et al.
Publicado: (2025)
por: Kang, Taewoo, et al.
Publicado: (2025)
Efficient Models for the Detection of Hate, Abuse and Profanity
por: Tillmann, Christoph, et al.
Publicado: (2024)
por: Tillmann, Christoph, et al.
Publicado: (2024)
HateXScore: A Metric Suite for Evaluating Reasoning Quality in Hate Speech Explanations
por: Hu, Yujia, et al.
Publicado: (2026)
por: Hu, Yujia, et al.
Publicado: (2026)
Conditioning Large Language Models on Legal Systems? Detecting Punishable Hate Speech
por: Ludwig, Florian, et al.
Publicado: (2025)
por: Ludwig, Florian, et al.
Publicado: (2025)
Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages
por: Prome, Ruhina Tabasshum, et al.
Publicado: (2025)
por: Prome, Ruhina Tabasshum, et al.
Publicado: (2025)
A Federated Approach to Few-Shot Hate Speech Detection for Marginalized Communities
por: Ye, Haotian, et al.
Publicado: (2024)
por: Ye, Haotian, et al.
Publicado: (2024)
Hate Speech Detection with Generalizable Target-aware Fairness
por: Chen, Tong, et al.
Publicado: (2024)
por: Chen, Tong, et al.
Publicado: (2024)
A Systematic Analysis of the Impact of Persona Steering on LLM Capabilities
por: Chen, Jiaqi, et al.
Publicado: (2026)
por: Chen, Jiaqi, et al.
Publicado: (2026)
MoCoRP: Modeling Consistent Relations between Persona and Response for Persona-based Dialogue
por: Lee, Kyungro, et al.
Publicado: (2025)
por: Lee, Kyungro, et al.
Publicado: (2025)
Hate Speech Detection using Large Language Models with Data Augmentation and Feature Enhancement
por: Nge, Brian Jing Hong, et al.
Publicado: (2026)
por: Nge, Brian Jing Hong, et al.
Publicado: (2026)
The Impact of Steering Large Language Models with Persona Vectors in Educational Applications
por: Wu, Yongchao, et al.
Publicado: (2026)
por: Wu, Yongchao, et al.
Publicado: (2026)
OSPC: Artificial VLM Features for Hateful Meme Detection
por: Grönquist, Peter
Publicado: (2024)
por: Grönquist, Peter
Publicado: (2024)
Disentangling Hate Across Target Identities
por: Jin, Yiping, et al.
Publicado: (2024)
por: Jin, Yiping, et al.
Publicado: (2024)
Can LLMs Evaluate What They Cannot Annotate? Revisiting LLM Reliability in Hate Speech Detection
por: Piot, Paloma, et al.
Publicado: (2025)
por: Piot, Paloma, et al.
Publicado: (2025)
Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?
por: Wang, Yifan, et al.
Publicado: (2025)
por: Wang, Yifan, et al.
Publicado: (2025)
Hierarchical Sentiment Analysis Framework for Hate Speech Detection: Implementing Binary and Multiclass Classification Strategy
por: Naznin, Faria, et al.
Publicado: (2024)
por: Naznin, Faria, et al.
Publicado: (2024)
Ejemplares similares
-
Mapping and Influencing the Political Ideology of Large Language Models using Synthetic Personas
por: Bernardelle, Pietro, et al.
Publicado: (2024) -
A Shared Geometry of Difficulty in Multilingual Language Models
por: Civelli, Stefano, et al.
Publicado: (2026) -
SubData: Bridging Heterogeneous Datasets to Enable Theory-Driven Evaluation of Political and Demographic Perspectives in LLMs
por: Bernardelle, Pietro, et al.
Publicado: (2024) -
Towards Detecting Persuasion on Social Media: From Model Development to Insights on Persuasion Strategies
por: Meguellati, Elyas, et al.
Publicado: (2025) -
Ideology-Based LLMs for Content Moderation
por: Civelli, Stefano, et al.
Publicado: (2025)