Explaining News Bias Detection: A Comparative SHAP Analysis of Transformer Model Decision Mechanisms
Fuente:
arXiv
Guardado en:
| Autor principal: | Ghosh, Himel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
To Bias or Not to Bias: Detecting bias in News with bias-detector
por: Ghosh, Himel, et al.
Publicado: (2025)
por: Ghosh, Himel, et al.
Publicado: (2025)
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
por: Ghosh, Himel, et al.
Publicado: (2026)
por: Ghosh, Himel, et al.
Publicado: (2026)
Less Back-and-Forth: A Comparative Study of Structured Prompting
por: Ghosh, Saurav, et al.
Publicado: (2026)
por: Ghosh, Saurav, et al.
Publicado: (2026)
Are Today's LLMs Ready to Explain Well-Being Concepts?
por: Jiang, Bohan, et al.
Publicado: (2025)
por: Jiang, Bohan, et al.
Publicado: (2025)
When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration
por: Shi, Quan, et al.
Publicado: (2025)
por: Shi, Quan, et al.
Publicado: (2025)
From Scratch to Fine-Tuned: A Comparative Study of Transformer Training Strategies for Legal Machine Translation
por: Barman, Amit, et al.
Publicado: (2025)
por: Barman, Amit, et al.
Publicado: (2025)
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments
por: Daynauth, Roland, et al.
Publicado: (2024)
por: Daynauth, Roland, et al.
Publicado: (2024)
User-Assistant Bias in LLMs
por: Pan, Xu, et al.
Publicado: (2025)
por: Pan, Xu, et al.
Publicado: (2025)
Humans and Large Language Models in Clinical Decision Support: A Study with Medical Calculators
por: Wan, Nicholas, et al.
Publicado: (2024)
por: Wan, Nicholas, et al.
Publicado: (2024)
Generating Pedagogically Meaningful Visuals for Math Word Problems: A New Benchmark and Analysis of Text-to-Image Models
por: Wang, Junling, et al.
Publicado: (2025)
por: Wang, Junling, et al.
Publicado: (2025)
Cognitive Bias Detection Using Advanced Prompt Engineering
por: Lemieux, Frederic, et al.
Publicado: (2025)
por: Lemieux, Frederic, et al.
Publicado: (2025)
Visualization Literacy of Multimodal Large Language Models: A Comparative Study
por: Li, Zhimin, et al.
Publicado: (2024)
por: Li, Zhimin, et al.
Publicado: (2024)
Actions Speak Louder than Words: Agent Decisions Reveal Implicit Biases in Language Models
por: Li, Yuxuan, et al.
Publicado: (2025)
por: Li, Yuxuan, et al.
Publicado: (2025)
AI as Teammate or Tool? A Review of Human-AI Interaction in Decision Support
por: Samu, Most. Sharmin Sultana, et al.
Publicado: (2026)
por: Samu, Most. Sharmin Sultana, et al.
Publicado: (2026)
More is More: Addition Bias in Large Language Models
por: Santagata, Luca, et al.
Publicado: (2024)
por: Santagata, Luca, et al.
Publicado: (2024)
Efficient Models for the Detection of Hate, Abuse and Profanity
por: Tillmann, Christoph, et al.
Publicado: (2024)
por: Tillmann, Christoph, et al.
Publicado: (2024)
'Since Lawyers are Males..': Examining Implicit Gender Bias in Hindi Language Generation by LLMs
por: Joshi, Ishika, et al.
Publicado: (2024)
por: Joshi, Ishika, et al.
Publicado: (2024)
Thematic Analysis with Open-Source Generative AI and Machine Learning: A New Method for Inductive Qualitative Codebook Development
por: Katz, Andrew, et al.
Publicado: (2024)
por: Katz, Andrew, et al.
Publicado: (2024)
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
por: Chaleshtori, Fateme Hashemi, et al.
Publicado: (2024)
por: Chaleshtori, Fateme Hashemi, et al.
Publicado: (2024)
Hybrid Decision Making via Conformal VLM-generated Guidance
por: Banerjee, Debodeep, et al.
Publicado: (2026)
por: Banerjee, Debodeep, et al.
Publicado: (2026)
How Performance Pressure Influences AI-Assisted Decision Making
por: Haduong, Nikita, et al.
Publicado: (2024)
por: Haduong, Nikita, et al.
Publicado: (2024)
Comparing the Efficacy of GPT-4 and Chat-GPT in Mental Health Care: A Blind Assessment of Large Language Models for Psychological Support
por: Moell, Birger
Publicado: (2024)
por: Moell, Birger
Publicado: (2024)
Human Bias in the Face of AI: Examining Human Judgment Against Text Labeled as AI Generated
por: Zhu, Tiffany, et al.
Publicado: (2024)
por: Zhu, Tiffany, et al.
Publicado: (2024)
WatChat: Explaining perplexing programs by debugging mental models
por: Chandra, Kartik, et al.
Publicado: (2024)
por: Chandra, Kartik, et al.
Publicado: (2024)
Transforming Tuberculosis Care: Optimizing Large Language Models For Enhanced Clinician-Patient Communication
por: Filienko, Daniil, et al.
Publicado: (2025)
por: Filienko, Daniil, et al.
Publicado: (2025)
Assessing the Creativity of Large Language Models: Testing, Limits, and New Frontiers
por: Schapiro, Samuel, et al.
Publicado: (2026)
por: Schapiro, Samuel, et al.
Publicado: (2026)
Large Language Models for Automatic Milestone Detection in Group Discussions
por: Duan, Zhuoxu, et al.
Publicado: (2024)
por: Duan, Zhuoxu, et al.
Publicado: (2024)
Evaluating the Application of ChatGPT in Outpatient Triage Guidance: A Comparative Study
por: Liu, Dou, et al.
Publicado: (2024)
por: Liu, Dou, et al.
Publicado: (2024)
Can Large Language Models Detect Verbal Indicators of Romantic Attraction?
por: Matz, Sandra C., et al.
Publicado: (2024)
por: Matz, Sandra C., et al.
Publicado: (2024)
Automating Customer Needs Analysis: A Comparative Study of Large Language Models in the Travel Industry
por: Barandoni, Simone, et al.
Publicado: (2024)
por: Barandoni, Simone, et al.
Publicado: (2024)
LLMs for XAI: Future Directions for Explaining Explanations
por: Zytek, Alexandra, et al.
Publicado: (2024)
por: Zytek, Alexandra, et al.
Publicado: (2024)
Multimodal Transformer Models for Turn-taking Prediction: Effects on Conversational Dynamics of Human-Agent Interaction during Cooperative Gameplay
por: Bae, Young-Ho, et al.
Publicado: (2025)
por: Bae, Young-Ho, et al.
Publicado: (2025)
Comparative Analysis of Large Language Models for the Machine-Assisted Resolution of User Intentions
por: Flerlage, Justus, et al.
Publicado: (2025)
por: Flerlage, Justus, et al.
Publicado: (2025)
Benchmarking Gender and Political Bias in Large Language Models
por: Yang, Jinrui, et al.
Publicado: (2025)
por: Yang, Jinrui, et al.
Publicado: (2025)
Supporting Student Decisions on Learning Recommendations: An LLM-Based Chatbot with Knowledge Graph Contextualization for Conversational Explainability and Mentoring
por: Abu-Rasheed, Hasan, et al.
Publicado: (2024)
por: Abu-Rasheed, Hasan, et al.
Publicado: (2024)
Prompt Engineering Techniques for Mitigating Cultural Bias Against Arabs and Muslims in Large Language Models: A Systematic Review
por: Asseri, Bushra, et al.
Publicado: (2025)
por: Asseri, Bushra, et al.
Publicado: (2025)
ValueCompass: A Framework for Measuring Contextual Value Alignment Between Human and LLMs
por: Shen, Hua, et al.
Publicado: (2024)
por: Shen, Hua, et al.
Publicado: (2024)
Learning to Generate and Evaluate Fact-checking Explanations with Transformers
por: Feher, Darius, et al.
Publicado: (2024)
por: Feher, Darius, et al.
Publicado: (2024)
A Comprehensive Survey of Bias in LLMs: Current Landscape and Future Directions
por: Ranjan, Rajesh, et al.
Publicado: (2024)
por: Ranjan, Rajesh, et al.
Publicado: (2024)
Prompts Matter: Comparing ML/GAI Approaches for Generating Inductive Qualitative Coding Results
por: Chen, John, et al.
Publicado: (2024)
por: Chen, John, et al.
Publicado: (2024)
Ejemplares similares
-
To Bias or Not to Bias: Detecting bias in News with bias-detector
por: Ghosh, Himel, et al.
Publicado: (2025) -
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
por: Ghosh, Himel, et al.
Publicado: (2026) -
Less Back-and-Forth: A Comparative Study of Structured Prompting
por: Ghosh, Saurav, et al.
Publicado: (2026) -
Are Today's LLMs Ready to Explain Well-Being Concepts?
por: Jiang, Bohan, et al.
Publicado: (2025) -
When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration
por: Shi, Quan, et al.
Publicado: (2025)