Responsible AI in NLP: GUS-Net Span-Level Bias Detection Dataset and Benchmark for Generalizations, Unfairness, and Stereotypes
Fuente:
arXiv
Saved in:
| Main Authors: | Powers, Maximus, Raza, Shaina, Chang, Alex, Riaz, Rehana, Mavani, Umang, Jonala, Harshitha Reddy, Tiwari, Ansh, Wei, Hua |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
Bias in the Picture: Benchmarking VLMs with Social-Cue News Images and LLM-as-Judge Assessment
by: Narayanan, Aravind, et al.
Published: (2025)
by: Narayanan, Aravind, et al.
Published: (2025)
The Deepfakes We Missed: We Built Detectors for a Threat That Didn't Arrive
by: Raza, Shaina
Published: (2026)
by: Raza, Shaina
Published: (2026)
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
BEADs: Bias Evaluation Across Domains
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
ViLBias: Detecting and Reasoning about Bias in Multimodal Content
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Investigating and Mitigating Stereotype-aware Unfairness in LLM-based Recommendations
by: Zhao, Zihuai, et al.
Published: (2025)
by: Zhao, Zihuai, et al.
Published: (2025)
The Unfairness of Multifactorial Bias in Recommendation
by: Mansoury, Masoud, et al.
Published: (2026)
by: Mansoury, Masoud, et al.
Published: (2026)
A Narrative Review of Identity, Data, and Location Privacy Techniques in Edge Computing and Mobile Crowdsourcing
by: Bashir, Syed Raza, et al.
Published: (2024)
by: Bashir, Syed Raza, et al.
Published: (2024)
Shielded RecRL: Explanation Generation for Recommender Systems without Ranking Degradation
by: Tiwari, Ansh, et al.
Published: (2025)
by: Tiwari, Ansh, et al.
Published: (2025)
Local Timescale Gates for Timescale-Robust Continual Spiking Neural Networks
by: Tiwari, Ansh, et al.
Published: (2025)
by: Tiwari, Ansh, et al.
Published: (2025)
Exploring Bias and Prediction Metrics to Characterise the Fairness of Machine Learning for Equity-Centered Public Health Decision-Making: A Narrative Review
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
LinguaMark: Do Multimodal Models Speak Fairly? A Benchmark-Based Evaluation
by: Raval, Ananya, et al.
Published: (2025)
by: Raval, Ananya, et al.
Published: (2025)
Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models?
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
by: Chatrath, Veronica, et al.
Published: (2024)
by: Chatrath, Veronica, et al.
Published: (2024)
Detecting Deception, Not Deepfakes: Why Media Forensics Needs Social Theories
by: Ho, Jessee, et al.
Published: (2026)
by: Ho, Jessee, et al.
Published: (2026)
Analyzing the Impact of Fake News on the Anticipated Outcome of the 2024 Election Ahead of Time
by: Raza, Shaina, et al.
Published: (2023)
by: Raza, Shaina, et al.
Published: (2023)
BBQ-V: Benchmarking Visual Stereotype Bias in Large Multimodal Models
by: Narnaware, Vishal, et al.
Published: (2025)
by: Narnaware, Vishal, et al.
Published: (2025)
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
by: Raza, Shaina, et al.
Published: (2023)
by: Raza, Shaina, et al.
Published: (2023)
ConvFill: Model Collaboration for Responsive Conversational Voice Agents
by: Srinivas, Vidya, et al.
Published: (2025)
by: Srinivas, Vidya, et al.
Published: (2025)
FAIR Enough: How Can We Develop and Assess a FAIR-Compliant Dataset for Large Language Models' Training?
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
CRAiG: Contextual Retrieval Augmented Generation
by: Adeola, Maximus
Published: (2026)
by: Adeola, Maximus
Published: (2026)
TuneGenie: Reasoning-based LLM agents for preferential music generation
by: Pandey, Amitesh, et al.
Published: (2025)
by: Pandey, Amitesh, et al.
Published: (2025)
Equity in Healthcare: Analyzing Disparities in Machine Learning Predictions of Diabetic Patient Readmissions
by: Al-Zanbouri, Zainab, et al.
Published: (2024)
by: Al-Zanbouri, Zainab, et al.
Published: (2024)
Fake News Detection: Comparative Evaluation of BERT-like Models and Large Language Models with Generative AI-Annotated Data
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
RESTLESS DE GUS VAN SANT
by:
Published: (2011)
by:
Published: (2011)
SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding
by: Radwan, Ahmed Y., et al.
Published: (2026)
by: Radwan, Ahmed Y., et al.
Published: (2026)
Blind Men and the Elephant: Diverse Perspectives on Gender Stereotypes in Benchmark Datasets
by: Zakizadeh, Mahdi, et al.
Published: (2025)
by: Zakizadeh, Mahdi, et al.
Published: (2025)
Bias and Unfairness in Information Retrieval Systems: New Challenges in the LLM Era
by: Dai, Sunhao, et al.
Published: (2024)
by: Dai, Sunhao, et al.
Published: (2024)
AfriStereo: A Culturally Grounded Dataset for Evaluating Stereotypical Bias in Large Language Models
by: Beux, Yann Le, et al.
Published: (2025)
by: Beux, Yann Le, et al.
Published: (2025)
Reconsidering Sentence-Level Sign Language Translation
by: Tanzer, Garrett, et al.
Published: (2024)
by: Tanzer, Garrett, et al.
Published: (2024)
PARANOID PARK DE GUS VAN SANT
by: Jesús Miguel Sáez González
Published: (2009)
by: Jesús Miguel Sáez González
Published: (2009)
LAST DAYS DE GUS VAN SANT
by: Jesús Miguel Sáez González
Published: (2007)
by: Jesús Miguel Sáez González
Published: (2007)
Assessing Educational Unfair Inequalities at a Regional Level in Colombia
by: Luis Gamboa
Published: (2015)
by: Luis Gamboa
Published: (2015)
The Big Bad Wolf and Stereotype and Bias in the Media.
by: Robinson, Julia
Published: (1998)
by: Robinson, Julia
Published: (1998)
PCS: Perceived Confidence Scoring of Black Box LLMs with Metamorphic Relations
by: Salimian, Sina, et al.
Published: (2025)
by: Salimian, Sina, et al.
Published: (2025)
TRiSM for Agentic AI: A Review of Trust, Risk, and Security Management in LLM-based Agentic Multi-Agent Systems
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
EQUATOR: A Deterministic Framework for Evaluating LLM Reasoning with Open-Ended Questions. # v1.0.0-beta
by: Bernard, Raymond, et al.
Published: (2024)
by: Bernard, Raymond, et al.
Published: (2024)
A new record of Excirolana Orientalis (Dana, 1853), a cirolanid genus and species (Isopoda, Flabellifera) from the Pakistan coast
by: Yasmeen, Rehana
Published: (2002)
by: Yasmeen, Rehana
Published: (2002)
Similar Items
-
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
by: Raza, Shaina, et al.
Published: (2025) -
Bias in the Picture: Benchmarking VLMs with Social-Cue News Images and LLM-as-Judge Assessment
by: Narayanan, Aravind, et al.
Published: (2025) -
The Deepfakes We Missed: We Built Detectors for a Threat That Didn't Arrive
by: Raza, Shaina
Published: (2026) -
MBIAS: Mitigating Bias in Large Language Models While Retaining Context
by: Raza, Shaina, et al.
Published: (2024) -
BEADs: Bias Evaluation Across Domains
by: Raza, Shaina, et al.
Published: (2024)