PartisanLens: A Multilingual Dataset of Hyperpartisan and Conspiratorial Immigration Narratives in European Media

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Maggini, Michele Joshua, Piot, Paloma, Pérez, Anxo, Marino, Erik Bran, Montesinos, Lúa Santamaría, Lisboa, Ana, Abuín, Marta Vázquez, Parapar, Javier, Gamallo, Pablo
Formato: Preprint
Publicado: 2026
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866914238274994176
author Maggini, Michele Joshua
Piot, Paloma
Pérez, Anxo
Marino, Erik Bran
Montesinos, Lúa Santamaría
Lisboa, Ana
Abuín, Marta Vázquez
Parapar, Javier
Gamallo, Pablo
author_facet Maggini, Michele Joshua
Piot, Paloma
Pérez, Anxo
Marino, Erik Bran
Montesinos, Lúa Santamaría
Lisboa, Ana
Abuín, Marta Vázquez
Parapar, Javier
Gamallo, Pablo
contents Detecting hyperpartisan narratives and Population Replacement Conspiracy Theories (PRCT) is essential to addressing the spread of misinformation. These complex narratives pose a significant threat, as hyperpartisanship drives political polarisation and institutional distrust, while PRCTs directly motivate real-world extremist violence, making their identification critical for social cohesion and public safety. However, existing resources are scarce, predominantly English-centric, and often analyse hyperpartisanship, stance, and rhetorical bias in isolation rather than as interrelated aspects of political discourse. To bridge this gap, we introduce \textsc{PartisanLens}, the first multilingual dataset of \num{1617} hyperpartisan news headlines in Spanish, Italian, and Portuguese, annotated in multiple political discourse aspects. We first evaluate the classification performance of widely used Large Language Models (LLMs) on this dataset, establishing robust baselines for the classification of hyperpartisan and PRCT narratives. In addition, we assess the viability of using LLMs as automatic annotators for this task, analysing their ability to approximate human annotation. Results highlight both their potential and current limitations. Next, moving beyond standard judgments, we explore whether LLMs can emulate human annotation patterns by conditioning them on socio-economic and ideological profiles that simulate annotator perspectives. At last, we provide our resources and evaluation, \textsc{PartisanLens} supports future research on detecting partisan and conspiratorial narratives in European contexts.
format Preprint
id arxiv_https___arxiv_org_abs_2601_03860
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle PartisanLens: A Multilingual Dataset of Hyperpartisan and Conspiratorial Immigration Narratives in European Media
Maggini, Michele Joshua
Piot, Paloma
Pérez, Anxo
Marino, Erik Bran
Montesinos, Lúa Santamaría
Lisboa, Ana
Abuín, Marta Vázquez
Parapar, Javier
Gamallo, Pablo
Computation and Language
Detecting hyperpartisan narratives and Population Replacement Conspiracy Theories (PRCT) is essential to addressing the spread of misinformation. These complex narratives pose a significant threat, as hyperpartisanship drives political polarisation and institutional distrust, while PRCTs directly motivate real-world extremist violence, making their identification critical for social cohesion and public safety. However, existing resources are scarce, predominantly English-centric, and often analyse hyperpartisanship, stance, and rhetorical bias in isolation rather than as interrelated aspects of political discourse. To bridge this gap, we introduce \textsc{PartisanLens}, the first multilingual dataset of \num{1617} hyperpartisan news headlines in Spanish, Italian, and Portuguese, annotated in multiple political discourse aspects. We first evaluate the classification performance of widely used Large Language Models (LLMs) on this dataset, establishing robust baselines for the classification of hyperpartisan and PRCT narratives. In addition, we assess the viability of using LLMs as automatic annotators for this task, analysing their ability to approximate human annotation. Results highlight both their potential and current limitations. Next, moving beyond standard judgments, we explore whether LLMs can emulate human annotation patterns by conditioning them on socio-economic and ideological profiles that simulate annotator perspectives. At last, we provide our resources and evaluation, \textsc{PartisanLens} supports future research on detecting partisan and conspiratorial narratives in European contexts.
title PartisanLens: A Multilingual Dataset of Hyperpartisan and Conspiratorial Immigration Narratives in European Media
topic Computation and Language
url https://arxiv.org/abs/2601.03860