BEADs: Bias Evaluation Across Domains
Fuente:
arXiv
Guardado en:
| Autores principales: | Raza, Shaina, Rahman, Mizanur, Zhang, Michael R. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
por: Raza, Shaina, et al.
Publicado: (2023)
por: Raza, Shaina, et al.
Publicado: (2023)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
por: Chatrath, Veronica, et al.
Publicado: (2024)
por: Chatrath, Veronica, et al.
Publicado: (2024)
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
por: Raza, Shaina, et al.
Publicado: (2023)
por: Raza, Shaina, et al.
Publicado: (2023)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
por: Sapkota, Ranjan, et al.
Publicado: (2025)
por: Sapkota, Ranjan, et al.
Publicado: (2025)
LinguaMark: Do Multimodal Models Speak Fairly? A Benchmark-Based Evaluation
por: Raval, Ananya, et al.
Publicado: (2025)
por: Raval, Ananya, et al.
Publicado: (2025)
Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
por: Powers, Maximus, et al.
Publicado: (2024)
por: Powers, Maximus, et al.
Publicado: (2024)
Reliability Analysis of Psychological Concept Extraction and Classification in User-penned Text
por: Garg, Muskan, et al.
Publicado: (2024)
por: Garg, Muskan, et al.
Publicado: (2024)
The Rise of Small Language Models in Healthcare: A Comprehensive Survey
por: Garg, Muskan, et al.
Publicado: (2025)
por: Garg, Muskan, et al.
Publicado: (2025)
Academic case reports lack diversity: Assessing the presence and diversity of sociodemographic and behavioral factors related to Post COVID-19 Condition
por: Florez, Juan Andres Medina, et al.
Publicado: (2025)
por: Florez, Juan Andres Medina, et al.
Publicado: (2025)
Benchmarking GPT-4 against Human Translators: A Comprehensive Evaluation Across Languages, Domains, and Expertise Levels
por: Yan, Jianhao, et al.
Publicado: (2024)
por: Yan, Jianhao, et al.
Publicado: (2024)
LLM-Based Data Science Agents: A Survey of Capabilities, Challenges, and Future Directions
por: Rahman, Mizanur, et al.
Publicado: (2025)
por: Rahman, Mizanur, et al.
Publicado: (2025)
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
por: Raza, Shaina, et al.
Publicado: (2025)
por: Raza, Shaina, et al.
Publicado: (2025)
Framing Political Bias in Multilingual LLMs Across Pakistani Languages
por: Nadeem, Afrozah, et al.
Publicado: (2025)
por: Nadeem, Afrozah, et al.
Publicado: (2025)
FakeWatch: A Framework for Detecting Fake News to Ensure Credible Elections
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
por: Ramprasad, Sanjana, et al.
Publicado: (2024)
por: Ramprasad, Sanjana, et al.
Publicado: (2024)
Retrieval Augmented Generation-based Large Language Models for Bridging Transportation Cybersecurity Legal Knowledge Gaps
por: Akbar, Khandakar Ashrafi, et al.
Publicado: (2025)
por: Akbar, Khandakar Ashrafi, et al.
Publicado: (2025)
Are Bias Evaluation Methods Biased ?
por: Berrayana, Lina, et al.
Publicado: (2025)
por: Berrayana, Lina, et al.
Publicado: (2025)
Acceptance Dynamics Across Cognitive Domains in Speculative Decoding
por: Mahmoud, Saif
Publicado: (2026)
por: Mahmoud, Saif
Publicado: (2026)
Bias in Large Language Models Across Clinical Applications: A Systematic Review
por: Suenghataiphorn, Thanathip, et al.
Publicado: (2025)
por: Suenghataiphorn, Thanathip, et al.
Publicado: (2025)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
por: Fernandes, Gustavo Lúcius, et al.
Publicado: (2026)
por: Fernandes, Gustavo Lúcius, et al.
Publicado: (2026)
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
The Deepfakes We Missed: We Built Detectors for a Threat That Didn't Arrive
por: Raza, Shaina
Publicado: (2026)
por: Raza, Shaina
Publicado: (2026)
Comparative Evaluation of ChatGPT and DeepSeek Across Key NLP Tasks: Strengths, Weaknesses, and Domain-Specific Performance
por: Etaiwi, Wael, et al.
Publicado: (2025)
por: Etaiwi, Wael, et al.
Publicado: (2025)
Evaluation Ethics of LLMs in Legal Domain
por: Zhang, Ruizhe, et al.
Publicado: (2024)
por: Zhang, Ruizhe, et al.
Publicado: (2024)
DIVERS-Bench: Evaluating Language Identification Across Domain Shifts and Code-Switching
por: Ojo, Jessica, et al.
Publicado: (2025)
por: Ojo, Jessica, et al.
Publicado: (2025)
PakBBQ: A Culturally Adapted Bias Benchmark for QA
por: Hashmat, Abdullah, et al.
Publicado: (2025)
por: Hashmat, Abdullah, et al.
Publicado: (2025)
Explainability in Practice: A Survey of Explainable NLP Across Various Domains
por: Mohammadi, Hadi, et al.
Publicado: (2025)
por: Mohammadi, Hadi, et al.
Publicado: (2025)
No LLM is Free From Bias: A Comprehensive Study of Bias Evaluation in Large Language Models
por: Kumar, Charaka Vinayak, et al.
Publicado: (2025)
por: Kumar, Charaka Vinayak, et al.
Publicado: (2025)
Does Reasoning Introduce Bias? A Study of Social Bias Evaluation and Mitigation in LLM Reasoning
por: Wu, Xuyang, et al.
Publicado: (2025)
por: Wu, Xuyang, et al.
Publicado: (2025)
SQL-Exchange: Transforming SQL Queries Across Domains
por: Daviran, Mohammadreza, et al.
Publicado: (2025)
por: Daviran, Mohammadreza, et al.
Publicado: (2025)
A Comprehensive Review of Recommender Systems: Transitioning from Theory to Practice
por: Raza, Shaina, et al.
Publicado: (2024)
por: Raza, Shaina, et al.
Publicado: (2024)
Extending Minimal Pairs with Ordinal Surprisal Curves and Entropy Across Applied Domains
por: Katz, Andrew
Publicado: (2026)
por: Katz, Andrew
Publicado: (2026)
Feature-Augmented Transformers for Robust AI-Text Detection Across Domains and Generators
por: Mady, Mohamed, et al.
Publicado: (2026)
por: Mady, Mohamed, et al.
Publicado: (2026)
Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents
por: Kim, Kangsan, et al.
Publicado: (2026)
por: Kim, Kangsan, et al.
Publicado: (2026)
Conformal Prediction for Risk-Controlled Medical Entity Extraction Across Clinical Domains
por: Shrestha, Manil, et al.
Publicado: (2026)
por: Shrestha, Manil, et al.
Publicado: (2026)
Beyond Correctness: Evaluating Subjective Writing Preferences Across Cultures
por: Ying, Shuangshuang, et al.
Publicado: (2025)
por: Ying, Shuangshuang, et al.
Publicado: (2025)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026)
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026)
Enabling Scalable Evaluation of Bias Patterns in Medical LLMs
por: Fayyaz, Hamed, et al.
Publicado: (2024)
por: Fayyaz, Hamed, et al.
Publicado: (2024)
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
por: Oi, Masanari, et al.
Publicado: (2024)
por: Oi, Masanari, et al.
Publicado: (2024)
Ejemplares similares
-
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
por: Raza, Shaina, et al.
Publicado: (2023) -
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
por: Raza, Shaina, et al.
Publicado: (2024) -
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
por: Chatrath, Veronica, et al.
Publicado: (2024) -
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
por: Raza, Shaina, et al.
Publicado: (2023) -
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
por: Sapkota, Ranjan, et al.
Publicado: (2025)