BharatBBQ: A Multilingual Bias Benchmark for Question Answering in the Indian Context
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tomar, Aditya, Sahoo, Nihar Ranjan, Bhattacharyya, Pushpak |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mathematics Isn't Culture-Free: Probing Cultural Gaps via Entity and Scenario Perturbations
von: Tomar, Aditya, et al.
Veröffentlicht: (2025)
von: Tomar, Aditya, et al.
Veröffentlicht: (2025)
IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context
von: Sahoo, Nihar Ranjan, et al.
Veröffentlicht: (2024)
von: Sahoo, Nihar Ranjan, et al.
Veröffentlicht: (2024)
EsBBQ and CaBBQ: The Spanish and Catalan Bias Benchmarks for Question Answering
von: Ruiz-Fernández, Valle, et al.
Veröffentlicht: (2025)
von: Ruiz-Fernández, Valle, et al.
Veröffentlicht: (2025)
Stereotype Detection as a Catalyst for Enhanced Bias Detection: A Multi-Task Learning Approach
von: Tomar, Aditya, et al.
Veröffentlicht: (2025)
von: Tomar, Aditya, et al.
Veröffentlicht: (2025)
KoBBQ: Korean Bias Benchmark for Question Answering
von: Jin, Jiho, et al.
Veröffentlicht: (2023)
von: Jin, Jiho, et al.
Veröffentlicht: (2023)
GG-BBQ: German Gender Bias Benchmark for Question Answering
von: Satheesh, Shalaka, et al.
Veröffentlicht: (2025)
von: Satheesh, Shalaka, et al.
Veröffentlicht: (2025)
Robust Bias Evaluation with FilBBQ: A Filipino Bias Benchmark for Question-Answering Language Models
von: Gamboa, Lance Calvin Lim, et al.
Veröffentlicht: (2026)
von: Gamboa, Lance Calvin Lim, et al.
Veröffentlicht: (2026)
Enhancing Food-Domain Question Answering with a Multimodal Knowledge Graph: Hybrid QA Generation and Diversity Analysis
von: B, Srihari K, et al.
Veröffentlicht: (2025)
von: B, Srihari K, et al.
Veröffentlicht: (2025)
Evaluating Dialect Robustness of Language Models via Conversation Understanding
von: Srirag, Dipankar, et al.
Veröffentlicht: (2024)
von: Srirag, Dipankar, et al.
Veröffentlicht: (2024)
Precision Empowers, Excess Distracts: Visual Question Answering With Dynamically Infused Knowledge In Language Models
von: Jhalani, Manas, et al.
Veröffentlicht: (2024)
von: Jhalani, Manas, et al.
Veröffentlicht: (2024)
Open-DeBias: Toward Mitigating Open-Set Bias in Language Models
von: Rani, Arti, et al.
Veröffentlicht: (2025)
von: Rani, Arti, et al.
Veröffentlicht: (2025)
PakBBQ: A Culturally Adapted Bias Benchmark for QA
von: Hashmat, Abdullah, et al.
Veröffentlicht: (2025)
von: Hashmat, Abdullah, et al.
Veröffentlicht: (2025)
Together We Can: Multilingual Automatic Post-Editing for Low-Resource Languages
von: Deoghare, Sourabh, et al.
Veröffentlicht: (2024)
von: Deoghare, Sourabh, et al.
Veröffentlicht: (2024)
MULTITAT: Benchmarking Multilingual Table-and-Text Question Answering
von: Zhang, Xuanliang, et al.
Veröffentlicht: (2025)
von: Zhang, Xuanliang, et al.
Veröffentlicht: (2025)
Social Bias in Popular Question-Answering Benchmarks
von: Kraft, Angelie, et al.
Veröffentlicht: (2025)
von: Kraft, Angelie, et al.
Veröffentlicht: (2025)
Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding
von: De, Anik, et al.
Veröffentlicht: (2025)
von: De, Anik, et al.
Veröffentlicht: (2025)
Unlocking Markets: A Multilingual Benchmark to Cross-Market Question Answering
von: Yuan, Yifei, et al.
Veröffentlicht: (2024)
von: Yuan, Yifei, et al.
Veröffentlicht: (2024)
Towards Emotion Consistency Analysis of Large Language Models in Emotional Conversational Contexts
von: Oram, Sneha, et al.
Veröffentlicht: (2026)
von: Oram, Sneha, et al.
Veröffentlicht: (2026)
ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues
von: Vedula, Bhaskara Hanuma, et al.
Veröffentlicht: (2026)
von: Vedula, Bhaskara Hanuma, et al.
Veröffentlicht: (2026)
XLQA: A Benchmark for Locale-Aware Multilingual Open-Domain Question Answering
von: Roh, Keon-Woo, et al.
Veröffentlicht: (2025)
von: Roh, Keon-Woo, et al.
Veröffentlicht: (2025)
How effective is Multi-source pivoting for Translation of Low Resource Indian Languages?
von: Gaikwad, Pranav, et al.
Veröffentlicht: (2024)
von: Gaikwad, Pranav, et al.
Veröffentlicht: (2024)
Reconsidering SMT Over NMT for Closely Related Languages: A Case Study of Persian-Hindi Pair
von: Yousofi, Waisullah, et al.
Veröffentlicht: (2024)
von: Yousofi, Waisullah, et al.
Veröffentlicht: (2024)
Expect the unexpected: Harnessing Sentence Completion for Sarcasm Detection
von: Joshi, Aditya, et al.
Veröffentlicht: (2017)
von: Joshi, Aditya, et al.
Veröffentlicht: (2017)
FAMMA: A Benchmark for Financial Domain Multilingual Multimodal Question Answering
von: Xue, Siqiao, et al.
Veröffentlicht: (2024)
von: Xue, Siqiao, et al.
Veröffentlicht: (2024)
P-ReMIS: Pragmatic Reasoning in Mental Health and a Social Implication
von: Oram, Sneha, et al.
Veröffentlicht: (2025)
von: Oram, Sneha, et al.
Veröffentlicht: (2025)
Main Predicate and Their Arguments as Explanation Signals For Intent Classification
von: Pimparkhede, Sameer, et al.
Veröffentlicht: (2025)
von: Pimparkhede, Sameer, et al.
Veröffentlicht: (2025)
Facts-and-Feelings: Capturing both Objectivity and Subjectivity in Table-to-Text Generation
von: Dey, Tathagata, et al.
Veröffentlicht: (2024)
von: Dey, Tathagata, et al.
Veröffentlicht: (2024)
We Care: Multimodal Depression Detection and Knowledge Infused Mental Health Therapeutic Response Generation
von: Moon, Palash, et al.
Veröffentlicht: (2024)
von: Moon, Palash, et al.
Veröffentlicht: (2024)
IndicSQuAD: A Comprehensive Multilingual Question Answering Dataset for Indic Languages
von: Endait, Sharvi, et al.
Veröffentlicht: (2025)
von: Endait, Sharvi, et al.
Veröffentlicht: (2025)
CVQA: Culturally-diverse Multilingual Visual Question Answering Benchmark
von: Romero, David, et al.
Veröffentlicht: (2024)
von: Romero, David, et al.
Veröffentlicht: (2024)
RoundTripOCR: A Data Generation Technique for Enhancing Post-OCR Error Correction in Low-Resource Devanagari Languages
von: Kashid, Harshvivek, et al.
Veröffentlicht: (2024)
von: Kashid, Harshvivek, et al.
Veröffentlicht: (2024)
VoiceBBQ: Investigating Effect of Content and Acoustics in Social Bias of Spoken Language Model
von: Choi, Junhyuk, et al.
Veröffentlicht: (2025)
von: Choi, Junhyuk, et al.
Veröffentlicht: (2025)
MedExpQA: Multilingual Benchmarking of Large Language Models for Medical Question Answering
von: Alonso, Iñigo, et al.
Veröffentlicht: (2024)
von: Alonso, Iñigo, et al.
Veröffentlicht: (2024)
EVJVQA Challenge: Multilingual Visual Question Answering
von: Nguyen, Ngan Luu-Thuy, et al.
Veröffentlicht: (2023)
von: Nguyen, Ngan Luu-Thuy, et al.
Veröffentlicht: (2023)
Multilingual Question Answering in Low-Resource Settings: A Dzongkha-English Benchmark for Foundation Models
von: Hosain, Md. Tanzib, et al.
Veröffentlicht: (2025)
von: Hosain, Md. Tanzib, et al.
Veröffentlicht: (2025)
On the Calibration of Multilingual Question Answering LLMs
von: Yang, Yahan, et al.
Veröffentlicht: (2023)
von: Yang, Yahan, et al.
Veröffentlicht: (2023)
TableEval: A Real-World Benchmark for Complex, Multilingual, and Multi-Structured Table Question Answering
von: Zhu, Junnan, et al.
Veröffentlicht: (2025)
von: Zhu, Junnan, et al.
Veröffentlicht: (2025)
A Case Study on Context-Aware Neural Machine Translation with Multi-Task Learning
von: Appicharla, Ramakrishna, et al.
Veröffentlicht: (2024)
von: Appicharla, Ramakrishna, et al.
Veröffentlicht: (2024)
PARAM-1 BharatGen 2.9B Model
von: Pundalik, Kundeshwar, et al.
Veröffentlicht: (2025)
von: Pundalik, Kundeshwar, et al.
Veröffentlicht: (2025)
Striking a Balance between Classical and Deep Learning Approaches in Natural Language Processing Pedagogy
von: Joshi, Aditya, et al.
Veröffentlicht: (2024)
von: Joshi, Aditya, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Mathematics Isn't Culture-Free: Probing Cultural Gaps via Entity and Scenario Perturbations
von: Tomar, Aditya, et al.
Veröffentlicht: (2025) -
IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context
von: Sahoo, Nihar Ranjan, et al.
Veröffentlicht: (2024) -
EsBBQ and CaBBQ: The Spanish and Catalan Bias Benchmarks for Question Answering
von: Ruiz-Fernández, Valle, et al.
Veröffentlicht: (2025) -
Stereotype Detection as a Catalyst for Enhanced Bias Detection: A Multi-Task Learning Approach
von: Tomar, Aditya, et al.
Veröffentlicht: (2025) -
KoBBQ: Korean Bias Benchmark for Question Answering
von: Jin, Jiho, et al.
Veröffentlicht: (2023)