SPOT: An Annotated French Corpus and Benchmark for Detecting Critical Interventions in Online Conversations
Fuente:
arXiv
Saved in:
| Main Authors: | Berriche, Manon, Nouri, Célia, Clavel, Chloée, Cointet, Jean-Philippe |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Graphically Speaking: Unmasking Abuse in Social Media with Conversation Insights
by: Nouri, Célia, et al.
Published: (2025)
by: Nouri, Célia, et al.
Published: (2025)
Algorithmic Approaches to Opinion Selection for Online Deliberation: A Comparative Study
by: Hafid, Salim, et al.
Published: (2026)
by: Hafid, Salim, et al.
Published: (2026)
Tu crois que c'est vrai ? Diversite des regimes d'enonciation face aux fake news et mecanismes d'autoregulation conversationnelle
by: Berriche, Manon
Published: (2025)
by: Berriche, Manon
Published: (2025)
What is Stigma Attributed to? A Theory-Grounded, Expert-Annotated Interview Corpus for Demystifying Mental-Health Stigma
by: Meng, Han, et al.
Published: (2025)
by: Meng, Han, et al.
Published: (2025)
ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection
by: Delaval, Axel, et al.
Published: (2025)
by: Delaval, Axel, et al.
Published: (2025)
CAPC-CG: A Large-Scale, Expert-Directed LLM-Annotated Corpus of Adaptive Policy Communication in China
by: Sun, Bolun, et al.
Published: (2025)
by: Sun, Bolun, et al.
Published: (2025)
WhatsApp Vaccine Discourse (WhaVax): An Expert-Annotated Dataset and Benchmark for Health Misinformation Detection
by: Santos, Jônatas H. dos, et al.
Published: (2026)
by: Santos, Jônatas H. dos, et al.
Published: (2026)
Mapping the Podcast Ecosystem with the Structured Podcast Research Corpus
by: Litterer, Benjamin, et al.
Published: (2024)
by: Litterer, Benjamin, et al.
Published: (2024)
Who and What? Using Linguistic Features and Annotator Characteristics to Analyze Annotation Variation
by: Maurer, Maximilian, et al.
Published: (2026)
by: Maurer, Maximilian, et al.
Published: (2026)
Towards Better Inclusivity: A Diverse Tweet Corpus of English Varieties
by: Pham, Nhi, et al.
Published: (2024)
by: Pham, Nhi, et al.
Published: (2024)
ConVerse: Benchmarking Contextual Safety in Agent-to-Agent Conversations
by: Gomaa, Amr, et al.
Published: (2025)
by: Gomaa, Amr, et al.
Published: (2025)
The Moral Foundations Reddit Corpus
by: Trager, Jackson, et al.
Published: (2022)
by: Trager, Jackson, et al.
Published: (2022)
LLMs Can Infer Political Alignment from Online Conversations
by: Lee, Byunghwee, et al.
Published: (2026)
by: Lee, Byunghwee, et al.
Published: (2026)
GPT is Not an Annotator: The Necessity of Human Annotation in Fairness Benchmark Construction
by: Felkner, Virginia K., et al.
Published: (2024)
by: Felkner, Virginia K., et al.
Published: (2024)
Politicians vs ChatGPT. A study of presuppositions in French and Italian political communication
by: Garassino, Davide, et al.
Published: (2024)
by: Garassino, Davide, et al.
Published: (2024)
An Annotated Reading of 'The Singer of Tales' in the LLM Era
by: Varshney, Kush R.
Published: (2025)
by: Varshney, Kush R.
Published: (2025)
The Cambridge Law Corpus: A Dataset for Legal AI Research
by: Östling, Andreas, et al.
Published: (2023)
by: Östling, Andreas, et al.
Published: (2023)
ChatGPT as speechwriter for the French presidents
by: Labbé, Dominique, et al.
Published: (2024)
by: Labbé, Dominique, et al.
Published: (2024)
Benchmark on Peer Review Toxic Detection: A Challenging Task with a New Dataset
by: Luo, Man, et al.
Published: (2025)
by: Luo, Man, et al.
Published: (2025)
Corpus-Based Approaches to Igbo Diacritic Restoration
by: Ezeani, Ignatius
Published: (2026)
by: Ezeani, Ignatius
Published: (2026)
The Algorithmic Unconscious: Structural Mechanisms and Implicit Biases in Large Language Models
by: Boisnard, Philippe
Published: (2026)
by: Boisnard, Philippe
Published: (2026)
Extracting O*NET Features from the NLx Corpus to Build Public Use Aggregate Labor Market Data
by: Meisenbacher, Stephen, et al.
Published: (2025)
by: Meisenbacher, Stephen, et al.
Published: (2025)
Measuring Large Language Models Capacity to Annotate Journalistic Sourcing
by: Vincent, Subramaniam, et al.
Published: (2024)
by: Vincent, Subramaniam, et al.
Published: (2024)
ARCADE: A City-Scale Corpus for Fine-Grained Arabic Dialect Tagging
by: Nacar, Omer, et al.
Published: (2026)
by: Nacar, Omer, et al.
Published: (2026)
Conversational Alignment with Artificial Intelligence in Context
by: Sterken, Rachel Katharine, et al.
Published: (2025)
by: Sterken, Rachel Katharine, et al.
Published: (2025)
Fine-tuning with Hierarchical Prompting for Robust Propaganda Classification Across Annotation Schemas
by: Stähelin, Lukas, et al.
Published: (2026)
by: Stähelin, Lukas, et al.
Published: (2026)
Examining Identity Drift in Conversations of LLM Agents
by: Choi, Junhyuk, et al.
Published: (2024)
by: Choi, Junhyuk, et al.
Published: (2024)
Generative AI Advertising as a Problem of Trustworthy Commercial Intervention
by: Qiu, Jingyi, et al.
Published: (2026)
by: Qiu, Jingyi, et al.
Published: (2026)
IDEAlign: Comparing Large Language Models to Human Experts in Open-ended Interpretive Annotations
by: Nam, Hyunji, et al.
Published: (2025)
by: Nam, Hyunji, et al.
Published: (2025)
Navigating the Risks of Using Large Language Models for Text Annotation in Social Science Research
by: Lin, Hao, et al.
Published: (2025)
by: Lin, Hao, et al.
Published: (2025)
SynBullying: A Multi LLM Synthetic Conversational Dataset for Cyberbullying Detection
by: Kazemi, Arefeh, et al.
Published: (2025)
by: Kazemi, Arefeh, et al.
Published: (2025)
Toward Automated Detection of Biased Social Signals from the Content of Clinical Conversations
by: Chen, Feng, et al.
Published: (2024)
by: Chen, Feng, et al.
Published: (2024)
Computational Analysis of Conversation Dynamics through Participant Responsivity
by: Hughes, Margaret, et al.
Published: (2025)
by: Hughes, Margaret, et al.
Published: (2025)
JiraiBench: A Bilingual Benchmark for Evaluating Large Language Models' Detection of Human Self-Destructive Behavior Content in Jirai Community
by: Xiao, Yunze, et al.
Published: (2025)
by: Xiao, Yunze, et al.
Published: (2025)
Catalysts of Conversation: Examining Interaction Dynamics Between Topic Initiators and Commentors in Alzheimer's Disease Online Communities
by: Ni, Congning, et al.
Published: (2024)
by: Ni, Congning, et al.
Published: (2024)
Don't Blame the Data, Blame the Model: Understanding Noise and Bias When Learning from Subjective Annotations
by: Anand, Abhishek, et al.
Published: (2024)
by: Anand, Abhishek, et al.
Published: (2024)
Ethical Concern Identification in NLP: A Corpus of ACL Anthology Ethics Statements
by: Karamolegkou, Antonia, et al.
Published: (2024)
by: Karamolegkou, Antonia, et al.
Published: (2024)
When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms
by: Chun, Chaewan, et al.
Published: (2026)
by: Chun, Chaewan, et al.
Published: (2026)
Cultural Compass: A Framework for Organizing Societal Norms to Detect Violations in Human-AI Conversations
by: Cheng, Myra, et al.
Published: (2026)
by: Cheng, Myra, et al.
Published: (2026)
Gendered Divides in Online Discussions about Reproductive Rights
by: Rao, Ashwin, et al.
Published: (2025)
by: Rao, Ashwin, et al.
Published: (2025)
Similar Items
-
Graphically Speaking: Unmasking Abuse in Social Media with Conversation Insights
by: Nouri, Célia, et al.
Published: (2025) -
Algorithmic Approaches to Opinion Selection for Online Deliberation: A Comparative Study
by: Hafid, Salim, et al.
Published: (2026) -
Tu crois que c'est vrai ? Diversite des regimes d'enonciation face aux fake news et mecanismes d'autoregulation conversationnelle
by: Berriche, Manon
Published: (2025) -
What is Stigma Attributed to? A Theory-Grounded, Expert-Annotated Interview Corpus for Demystifying Mental-Health Stigma
by: Meng, Han, et al.
Published: (2025) -
ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection
by: Delaval, Axel, et al.
Published: (2025)