MFTCXplain: A Multilingual Benchmark Dataset for Evaluating the Moral Reasoning of LLMs through Multi-hop Hate Speech Explanation
Fuente:
arXiv
Saved in:
| Main Authors: | Trager, Jackson, Vargas, Francielle, Alves, Diego, Guida, Matteo, Ngueajio, Mikel K., Agrawal, Ameeta, Daryani, Yalda, Karimi-Malekabadi, Farzan, Plaza-del-Arco, Flor Miriam |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self-Explaining Hate Speech Detection with Moral Rationales
by: Vargas, Francielle, et al.
Published: (2026)
by: Vargas, Francielle, et al.
Published: (2026)
Think Like a Person Before Responding: A Multi-Faceted Evaluation of Persona-Guided LLMs for Countering Hate
by: Ngueajio, Mikel K., et al.
Published: (2025)
by: Ngueajio, Mikel K., et al.
Published: (2025)
Theory Trace Card: Theory-Driven Socio-Cognitive Evaluation of LLMs
by: Karimi-Malekabadi, Farzan, et al.
Published: (2026)
by: Karimi-Malekabadi, Farzan, et al.
Published: (2026)
Scaling Item-to-Standard Alignment with Large Language Models: Accuracy, Limits, and Solutions
by: Karimi-Malekabadi, Farzan, et al.
Published: (2025)
by: Karimi-Malekabadi, Farzan, et al.
Published: (2025)
COPOS: Corpus Of Patient Opinions in Spanish. Application of Sentiment Analysis Techniques
by: Flor Miriam Plaza-del-Arco
Published: (2016)
by: Flor Miriam Plaza-del-Arco
Published: (2016)
MTQ-Eval: Multilingual Text Quality Evaluation for Language Models
by: Pokharel, Rhitabrat, et al.
Published: (2025)
by: Pokharel, Rhitabrat, et al.
Published: (2025)
Exploring the Maze of Multilingual Modeling
by: Nezhad, Sina Bagheri, et al.
Published: (2023)
by: Nezhad, Sina Bagheri, et al.
Published: (2023)
What Drives Performance in Multilingual Language Models?
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
Tracing Moral Foundations in Large Language Models
by: Yu, Chenxiao, et al.
Published: (2026)
by: Yu, Chenxiao, et al.
Published: (2026)
Enhancing Large Language Models with Neurosymbolic Reasoning for Multilingual Tasks
by: Nezhad, Sina Bagheri, et al.
Published: (2025)
by: Nezhad, Sina Bagheri, et al.
Published: (2025)
Bangla Hate Speech Classification with Fine-tuned Transformer Models
by: Jafari, Yalda Keivan, et al.
Published: (2025)
by: Jafari, Yalda Keivan, et al.
Published: (2025)
Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization
by: Giabbanelli, Philippe J., et al.
Published: (2025)
by: Giabbanelli, Philippe J., et al.
Published: (2025)
Realistic threat perception drives intergroup conflict: A causal, dynamic analysis using generative-agent simulations
by: Abdurahman, Suhaib, et al.
Published: (2025)
by: Abdurahman, Suhaib, et al.
Published: (2025)
The Shrinking Landscape of Linguistic Diversity in the Age of Large Language Models
by: Sourati, Zhivar, et al.
Published: (2025)
by: Sourati, Zhivar, et al.
Published: (2025)
Cross-Lingual Activation Steering for Multilingual Language Models
by: Pokharel, Rhitabrat, et al.
Published: (2026)
by: Pokharel, Rhitabrat, et al.
Published: (2026)
CAPO: Confidence Aware Preference Optimization Learning for Multilingual Preferences
by: Pokharel, Rhitabrat, et al.
Published: (2025)
by: Pokharel, Rhitabrat, et al.
Published: (2025)
Wisdom of Instruction-Tuned Language Model Crowds. Exploring Model Label Variation
by: Plaza-del-Arco, Flor Miriam, et al.
Published: (2023)
by: Plaza-del-Arco, Flor Miriam, et al.
Published: (2023)
The Moral Foundations Reddit Corpus
by: Trager, Jackson, et al.
Published: (2022)
by: Trager, Jackson, et al.
Published: (2022)
Beyond Data Quantity: Key Factors Driving Performance in Multilingual Language Models
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
Reasoning on a Spectrum: Aligning LLMs to System 1 and System 2 Thinking
by: Ziabari, Alireza S., et al.
Published: (2025)
by: Ziabari, Alireza S., et al.
Published: (2025)
Exploring Subjective Tasks in Farsi: A Survey Analysis and Evaluation of Language Models
by: Rooein, Donya, et al.
Published: (2025)
by: Rooein, Donya, et al.
Published: (2025)
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
by: Eilertsen, Brage, et al.
Published: (2025)
by: Eilertsen, Brage, et al.
Published: (2025)
Anti-Toxoplasma activities of methanolic extract of Sambucus nigra (Caprifoliaceae) fruits and leaves
by: Ahmad Daryani
Published: (2015)
by: Ahmad Daryani
Published: (2015)
Multilingual Relative Clause Attachment Ambiguity Resolution in Large Language Models
by: Lee, So Young, et al.
Published: (2025)
by: Lee, So Young, et al.
Published: (2025)
No-Worse Context-Aware Decoding: Preventing Neutral Regression in Context-Conditioned Generation
by: Tao, Yufei, et al.
Published: (2026)
by: Tao, Yufei, et al.
Published: (2026)
Understanding Position Bias Effects on Fairness in Social Multi-Document Summarization
by: Olabisi, Olubusayo, et al.
Published: (2024)
by: Olabisi, Olubusayo, et al.
Published: (2024)
franciellevargas/HateBR: v7.0.0
by: Francielle Vargas, et al.
Published: (2025)
by: Francielle Vargas, et al.
Published: (2025)
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection
by: Chan, Fai Leui, et al.
Published: (2024)
by: Chan, Fai Leui, et al.
Published: (2024)
Language Model Council: Democratically Benchmarking Foundation Models on Highly Subjective Tasks
by: Zhao, Justin, et al.
Published: (2024)
by: Zhao, Justin, et al.
Published: (2024)
SINAI at eRisk@CLEF 2023: Approaching Early Detection of Gambling with Natural Language Processing
by: Marmol-Romero, Alba Maria, et al.
Published: (2025)
by: Marmol-Romero, Alba Maria, et al.
Published: (2025)
Emotion Analysis in NLP: Trends, Gaps and Roadmap for Future Directions
by: Plaza-del-Arco, Flor Miriam, et al.
Published: (2024)
by: Plaza-del-Arco, Flor Miriam, et al.
Published: (2024)
Mitigation of Gender and Ethnicity Bias in AI-Generated Stories through Model Explanations
by: Dimgba, Martha O., et al.
Published: (2025)
by: Dimgba, Martha O., et al.
Published: (2025)
Evaluating Multilingual Long-Context Models for Retrieval and Reasoning
by: Agrawal, Ameeta, et al.
Published: (2024)
by: Agrawal, Ameeta, et al.
Published: (2024)
Incorporating Human Explanations for Robust Hate Speech Detection
by: Chen, Jennifer L., et al.
Published: (2024)
by: Chen, Jennifer L., et al.
Published: (2024)
HateXScore: A Metric Suite for Evaluating Reasoning Quality in Hate Speech Explanations
by: Hu, Yujia, et al.
Published: (2026)
by: Hu, Yujia, et al.
Published: (2026)
From Chatbots to Confidants: A Cross-Cultural Study of LLM Adoption for Emotional Support
by: Amat-Lefort, Natalia, et al.
Published: (2026)
by: Amat-Lefort, Natalia, et al.
Published: (2026)
AfriHate: A Multilingual Collection of Hate Speech and Abusive Language Datasets for African Languages
by: Muhammad, Shamsuddeen Hassan, et al.
Published: (2025)
by: Muhammad, Shamsuddeen Hassan, et al.
Published: (2025)
Multi3Hate: Multimodal, Multilingual, and Multicultural Hate Speech Detection with Vision-Language Models
by: Bui, Minh Duc, et al.
Published: (2024)
by: Bui, Minh Duc, et al.
Published: (2024)
FLANS at SemEval-2026 Task 7: RAG with Open-Sourced Smaller LLMs for Everyday Knowledge Across Diverse Languages and Cultures
by: Bogdanova, Liliia, et al.
Published: (2026)
by: Bogdanova, Liliia, et al.
Published: (2026)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
by: Plaza-del-Arco, Flor Miriam, et al.
Published: (2024)
by: Plaza-del-Arco, Flor Miriam, et al.
Published: (2024)
Similar Items
-
Self-Explaining Hate Speech Detection with Moral Rationales
by: Vargas, Francielle, et al.
Published: (2026) -
Think Like a Person Before Responding: A Multi-Faceted Evaluation of Persona-Guided LLMs for Countering Hate
by: Ngueajio, Mikel K., et al.
Published: (2025) -
Theory Trace Card: Theory-Driven Socio-Cognitive Evaluation of LLMs
by: Karimi-Malekabadi, Farzan, et al.
Published: (2026) -
Scaling Item-to-Standard Alignment with Large Language Models: Accuracy, Limits, and Solutions
by: Karimi-Malekabadi, Farzan, et al.
Published: (2025) -
COPOS: Corpus Of Patient Opinions in Spanish. Application of Sentiment Analysis Techniques
by: Flor Miriam Plaza-del-Arco
Published: (2016)