Towards Region-aware Bias Evaluation Metrics
Fuente:
arXiv
Guardado en:
| Autores principales: | Borah, Angana, Garimella, Aparna, Mihalcea, Rada |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
por: Borah, Angana, et al.
Publicado: (2024)
por: Borah, Angana, et al.
Publicado: (2024)
The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models
por: Arif, Samee, et al.
Publicado: (2026)
por: Arif, Samee, et al.
Publicado: (2026)
The Curious Case of Curiosity across Human Cultures and LLMs
por: Borah, Angana, et al.
Publicado: (2025)
por: Borah, Angana, et al.
Publicado: (2025)
Mind the (Belief) Gap: Group Identity in the World of LLMs
por: Borah, Angana, et al.
Publicado: (2025)
por: Borah, Angana, et al.
Publicado: (2025)
Persuasion at Play: Understanding Misinformation Dynamics in Demographic-Aware Human-LLM Interactions
por: Borah, Angana, et al.
Publicado: (2025)
por: Borah, Angana, et al.
Publicado: (2025)
Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility
por: Borah, Angana, et al.
Publicado: (2026)
por: Borah, Angana, et al.
Publicado: (2026)
The Power of Many: Multi-Agent Multimodal Models for Cultural Image Captioning
por: Bai, Longju, et al.
Publicado: (2024)
por: Bai, Longju, et al.
Publicado: (2024)
Are Human Interactions Replicable by Generative Agents? A Case Study on Pronoun Usage in Hierarchical Interactions
por: Deng, Naihao, et al.
Publicado: (2025)
por: Deng, Naihao, et al.
Publicado: (2025)
Modeling Contextual Passage Utility for Multihop Question Answering
por: Jain, Akriti, et al.
Publicado: (2025)
por: Jain, Akriti, et al.
Publicado: (2025)
Knowing What's Missing: Assessing Information Sufficiency in Question Answering
por: Jain, Akriti, et al.
Publicado: (2025)
por: Jain, Akriti, et al.
Publicado: (2025)
Is This a Bad Table? A Closer Look at the Evaluation of Table Generation from Text
por: Ramu, Pritika, et al.
Publicado: (2024)
por: Ramu, Pritika, et al.
Publicado: (2024)
Towards Dog Bark Decoding: Leveraging Human Speech Processing for Automated Bark Classification
por: Abzaliev, Artem, et al.
Publicado: (2024)
por: Abzaliev, Artem, et al.
Publicado: (2024)
Rethinking Table Instruction Tuning
por: Deng, Naihao, et al.
Publicado: (2025)
por: Deng, Naihao, et al.
Publicado: (2025)
Causally Testing Gender Bias in LLMs: A Case Study on Occupational Bias
por: Chen, Yuen, et al.
Publicado: (2022)
por: Chen, Yuen, et al.
Publicado: (2022)
Application Specific Compression of Deep Learning Models
por: Rai, Rohit Raj, et al.
Publicado: (2024)
por: Rai, Rohit Raj, et al.
Publicado: (2024)
TabReX : Tabular Referenceless eXplainable Evaluation
por: Anvekar, Tejas, et al.
Publicado: (2025)
por: Anvekar, Tejas, et al.
Publicado: (2025)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
por: Stewart, Ian, et al.
Publicado: (2024)
por: Stewart, Ian, et al.
Publicado: (2024)
Towards Algorithmic Fidelity: Mental Health Representation across Demographics in Synthetic vs. Human-generated Data
por: Mori, Shinka, et al.
Publicado: (2024)
por: Mori, Shinka, et al.
Publicado: (2024)
Are Word Embedding Methods Stable and Should We Care About It?
por: Borah, Angana, et al.
Publicado: (2021)
por: Borah, Angana, et al.
Publicado: (2021)
MAiDE-up: Multilingual Deception Detection of GPT-generated Hotel Reviews
por: Ignat, Oana, et al.
Publicado: (2024)
por: Ignat, Oana, et al.
Publicado: (2024)
Patient-Centered RAG for Oncology Visit Aid Following the Ottawa Decision Guide
por: Liu, Siyang, et al.
Publicado: (2025)
por: Liu, Siyang, et al.
Publicado: (2025)
The Generation Gap: Exploring Age Bias in the Value Systems of Large Language Models
por: Liu, Siyang, et al.
Publicado: (2024)
por: Liu, Siyang, et al.
Publicado: (2024)
Cross-cultural Inspiration Detection and Analysis in Real and LLM-generated Social Media Data
por: Ignat, Oana, et al.
Publicado: (2024)
por: Ignat, Oana, et al.
Publicado: (2024)
Measuring Spurious Correlation in Classification: 'Clever Hans' in Translationese
por: Borah, Angana, et al.
Publicado: (2023)
por: Borah, Angana, et al.
Publicado: (2023)
Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions
por: Hong, Pengfei, et al.
Publicado: (2024)
por: Hong, Pengfei, et al.
Publicado: (2024)
Doc2Chart: Intent-Driven Zero-Shot Chart Generation from Documents
por: Jain, Akriti, et al.
Publicado: (2025)
por: Jain, Akriti, et al.
Publicado: (2025)
An Answer is just the Start: Related Insight Generation for Open-Ended Document-Grounded QA
por: Sharma, Saransh, et al.
Publicado: (2026)
por: Sharma, Saransh, et al.
Publicado: (2026)
One Word at a Time: Incremental Completion Decomposition Breaks LLM Safety
por: Arif, Samee, et al.
Publicado: (2026)
por: Arif, Samee, et al.
Publicado: (2026)
IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context
por: Sahoo, Nihar Ranjan, et al.
Publicado: (2024)
por: Sahoo, Nihar Ranjan, et al.
Publicado: (2024)
CliniDial: A Naturally Occurring Multimodal Dialogue Dataset for Team Reflection in Action During Clinical Operation
por: Deng, Naihao, et al.
Publicado: (2025)
por: Deng, Naihao, et al.
Publicado: (2025)
Annotations on a Budget: Leveraging Geo-Data Similarity to Balance Model Performance and Annotation Cost
por: Ignat, Oana, et al.
Publicado: (2024)
por: Ignat, Oana, et al.
Publicado: (2024)
Presentations are not always linear! GNN meets LLM for Document-to-Presentation Transformation with Attribution
por: Maheshwari, Himanshu, et al.
Publicado: (2024)
por: Maheshwari, Himanshu, et al.
Publicado: (2024)
Uplifting Lower-Income Data: Strategies for Socioeconomic Perspective Shifts in Large Multi-modal Models
por: Nwatu, Joan, et al.
Publicado: (2024)
por: Nwatu, Joan, et al.
Publicado: (2024)
SciDoc2Diagrammer-MAF: Towards Generation of Scientific Diagrams from Documents guided by Multi-Aspect Feedback Refinement
por: Mondal, Ishani, et al.
Publicado: (2024)
por: Mondal, Ishani, et al.
Publicado: (2024)
Human Action Co-occurrence in Lifestyle Vlogs using Graph Link Prediction
por: Ignat, Oana, et al.
Publicado: (2023)
por: Ignat, Oana, et al.
Publicado: (2023)
Why AI Is WEIRD and Should Not Be This Way: Towards AI For Everyone, With Everyone, By Everyone
por: Mihalcea, Rada, et al.
Publicado: (2024)
por: Mihalcea, Rada, et al.
Publicado: (2024)
Moneyball with LLMs: Analyzing Tabular Summarization in Sports Narratives
por: Upadhyay, Ritam, et al.
Publicado: (2025)
por: Upadhyay, Ritam, et al.
Publicado: (2025)
Infogen: Generating Complex Statistical Infographics from Documents
por: Ghosh, Akash, et al.
Publicado: (2025)
por: Ghosh, Akash, et al.
Publicado: (2025)
Evaluating Metrics for Bias in Word Embeddings
por: Schröder, Sarah, et al.
Publicado: (2021)
por: Schröder, Sarah, et al.
Publicado: (2021)
Democratic or Authoritarian? Probing a New Dimension of Political Biases in Large Language Models
por: Piedrahita, David Guzman, et al.
Publicado: (2025)
por: Piedrahita, David Guzman, et al.
Publicado: (2025)
Ejemplares similares
-
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
por: Borah, Angana, et al.
Publicado: (2024) -
The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models
por: Arif, Samee, et al.
Publicado: (2026) -
The Curious Case of Curiosity across Human Cultures and LLMs
por: Borah, Angana, et al.
Publicado: (2025) -
Mind the (Belief) Gap: Group Identity in the World of LLMs
por: Borah, Angana, et al.
Publicado: (2025) -
Persuasion at Play: Understanding Misinformation Dynamics in Demographic-Aware Human-LLM Interactions
por: Borah, Angana, et al.
Publicado: (2025)