Evaluating Bias in LLMs for Job-Resume Matching: Gender, Race, and Education
Fuente:
arXiv
Guardado en:
| Autores principales: | Iso, Hayate, Pezeshkpour, Pouya, Bhutani, Nikita, Hruschka, Estevam |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Less is More for Long Document Summary Evaluation by LLMs
por: Wu, Yunshu, et al.
Publicado: (2023)
por: Wu, Yunshu, et al.
Publicado: (2023)
From Single to Multi: How LLMs Hallucinate in Multi-Document Summarization
por: Belem, Catarina G., et al.
Publicado: (2024)
por: Belem, Catarina G., et al.
Publicado: (2024)
Insight-RAG: Enhancing LLMs with Insight-Driven Augmentation
por: Pezeshkpour, Pouya, et al.
Publicado: (2025)
por: Pezeshkpour, Pouya, et al.
Publicado: (2025)
Learning Beyond the Surface: How Far Can Continual Pre-Training with LoRA Enhance LLMs' Domain-Specific Insight Learning?
por: Pezeshkpour, Pouya, et al.
Publicado: (2025)
por: Pezeshkpour, Pouya, et al.
Publicado: (2025)
AutoPyVerifier: Learning Compact Executable Verifiers for Large Language Model Outputs
por: Pezeshkpour, Pouya, et al.
Publicado: (2026)
por: Pezeshkpour, Pouya, et al.
Publicado: (2026)
From Task Solving to Robust Real-World Adaptation in LLM Agents
por: Pezeshkpour, Pouya, et al.
Publicado: (2026)
por: Pezeshkpour, Pouya, et al.
Publicado: (2026)
Multi-Conditional Ranking with Large Language Models
por: Pezeshkpour, Pouya, et al.
Publicado: (2024)
por: Pezeshkpour, Pouya, et al.
Publicado: (2024)
The Rarity Blind Spot: A Framework for Evaluating Statistical Reasoning in LLMs
por: Maekawa, Seiji, et al.
Publicado: (2025)
por: Maekawa, Seiji, et al.
Publicado: (2025)
Reasoning Capacity in Multi-Agent Systems: Limitations, Challenges and Human-Centered Solutions
por: Pezeshkpour, Pouya, et al.
Publicado: (2024)
por: Pezeshkpour, Pouya, et al.
Publicado: (2024)
Holistic Reasoning with Long-Context LMs: A Benchmark for Database Operations on Massive Textual Data
por: Maekawa, Seiji, et al.
Publicado: (2024)
por: Maekawa, Seiji, et al.
Publicado: (2024)
From Proof to Program: Characterizing Tool-Induced Reasoning Hallucinations in Large Language Models
por: Bayat, Farima Fatahi, et al.
Publicado: (2025)
por: Bayat, Farima Fatahi, et al.
Publicado: (2025)
Natural Language Processing for Human Resources: A Survey
por: Otani, Naoki, et al.
Publicado: (2024)
por: Otani, Naoki, et al.
Publicado: (2024)
Mixed Signals: Decoding VLMs' Reasoning and Underlying Bias in Vision-Language Conflict
por: Pezeshkpour, Pouya, et al.
Publicado: (2025)
por: Pezeshkpour, Pouya, et al.
Publicado: (2025)
Towards Reliable Benchmarking: A Contamination Free, Controllable Evaluation Framework for Multi-step LLM Function Calling
por: Maekawa, Seiji, et al.
Publicado: (2025)
por: Maekawa, Seiji, et al.
Publicado: (2025)
Align then Train: Efficient Retrieval Adapter Learning
por: Maekawa, Seiji, et al.
Publicado: (2026)
por: Maekawa, Seiji, et al.
Publicado: (2026)
XATU: A Fine-grained Instruction-based Benchmark for Explainable Text Updates
por: Zhang, Haopeng, et al.
Publicado: (2023)
por: Zhang, Haopeng, et al.
Publicado: (2023)
Retrieval Helps or Hurts? A Deeper Dive into the Efficacy of Retrieval Augmentation to Language Models
por: Maekawa, Seiji, et al.
Publicado: (2024)
por: Maekawa, Seiji, et al.
Publicado: (2024)
Do Agents Need to Plan Step-by-Step? Rethinking Planning Horizon in Data-Centric Tool Calling
por: Otani, Naoki, et al.
Publicado: (2026)
por: Otani, Naoki, et al.
Publicado: (2026)
AutoTemplate: A Simple Recipe for Lexically Constrained Text Generation
por: Iso, Hayate
Publicado: (2022)
por: Iso, Hayate
Publicado: (2022)
AmbigNLG: Addressing Task Ambiguity in Instruction for NLG
por: Niwa, Ayana, et al.
Publicado: (2024)
por: Niwa, Ayana, et al.
Publicado: (2024)
Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval
por: Wilson, Kyra, et al.
Publicado: (2024)
por: Wilson, Kyra, et al.
Publicado: (2024)
LLMs Are Not Intelligent Thinkers: Introducing Mathematical Topic Tree Benchmark for Comprehensive Evaluation of LLMs
por: Davoodi, Arash Gholami, et al.
Publicado: (2024)
por: Davoodi, Arash Gholami, et al.
Publicado: (2024)
Noisy Pairing and Partial Supervision for Stylized Opinion Summarization
por: Iso, Hayate, et al.
Publicado: (2022)
por: Iso, Hayate, et al.
Publicado: (2022)
FAIRE: Assessing Racial and Gender Bias in AI-Driven Resume Evaluations
por: Wen, Athena, et al.
Publicado: (2025)
por: Wen, Athena, et al.
Publicado: (2025)
Measuring Gender Bias in Job Title Matching for Grammatical Gender Languages
por: García-Sardiña, Laura, et al.
Publicado: (2025)
por: García-Sardiña, Laura, et al.
Publicado: (2025)
Leveraging Large Language Models for Career Mobility Analysis: A Study of Gender, Race, and Job Change Using U.S. Online Resume Profiles
por: Achananuparp, Palakorn, et al.
Publicado: (2025)
por: Achananuparp, Palakorn, et al.
Publicado: (2025)
A Blueprint Architecture of Compound AI Systems for Enterprise
por: Kandogan, Eser, et al.
Publicado: (2024)
por: Kandogan, Eser, et al.
Publicado: (2024)
OmniTQA: A Cost-Aware System for Hybrid Query Processing over Semi-Structured Data
por: Shahbazi, Nima, et al.
Publicado: (2026)
por: Shahbazi, Nima, et al.
Publicado: (2026)
ConFit v2: Improving Resume-Job Matching using Hypothetical Resume Embedding and Runner-Up Hard-Negative Mining
por: Yu, Xiao, et al.
Publicado: (2025)
por: Yu, Xiao, et al.
Publicado: (2025)
A Dynamic Self-Evolving Extraction System
por: Amin-Naseri, Moin, et al.
Publicado: (2026)
por: Amin-Naseri, Moin, et al.
Publicado: (2026)
ConFit: Improving Resume-Job Matching using Data Augmentation and Contrastive Learning
por: Yu, Xiao, et al.
Publicado: (2024)
por: Yu, Xiao, et al.
Publicado: (2024)
CareerBERT: Matching Resumes to ESCO Jobs in a Shared Embedding Space for Generic Job Recommendations
por: Rosenberger, Julian, et al.
Publicado: (2025)
por: Rosenberger, Julian, et al.
Publicado: (2025)
ConFit v3: Improving Resume-Job Matching with LLM-based Re-Ranking
por: Yu, Xiao, et al.
Publicado: (2026)
por: Yu, Xiao, et al.
Publicado: (2026)
Towards Probabilistic Question Answering Over Tabular Data
por: Shen, Chen, et al.
Publicado: (2025)
por: Shen, Chen, et al.
Publicado: (2025)
RECAP: REwriting Conversations for Intent Understanding in Agentic Planning
por: Mitra, Kushan, et al.
Publicado: (2025)
por: Mitra, Kushan, et al.
Publicado: (2025)
FactLens: Benchmarking Fine-Grained Fact Verification
por: Mitra, Kushan, et al.
Publicado: (2024)
por: Mitra, Kushan, et al.
Publicado: (2024)
Evaluating Gender Bias of LLMs in Making Morality Judgements
por: Bajaj, Divij, et al.
Publicado: (2024)
por: Bajaj, Divij, et al.
Publicado: (2024)
Same Content, Different Representations: A Controlled Study for Table QA
por: Zhang, Yue, et al.
Publicado: (2025)
por: Zhang, Yue, et al.
Publicado: (2025)
Verification-Aware Planning for Multi-Agent Systems
por: Xu, Tianyang, et al.
Publicado: (2025)
por: Xu, Tianyang, et al.
Publicado: (2025)
Orchestrating Agents and Data for Enterprise: A Blueprint Architecture for Compound AI
por: Kandogan, Eser, et al.
Publicado: (2025)
por: Kandogan, Eser, et al.
Publicado: (2025)
Ejemplares similares
-
Less is More for Long Document Summary Evaluation by LLMs
por: Wu, Yunshu, et al.
Publicado: (2023) -
From Single to Multi: How LLMs Hallucinate in Multi-Document Summarization
por: Belem, Catarina G., et al.
Publicado: (2024) -
Insight-RAG: Enhancing LLMs with Insight-Driven Augmentation
por: Pezeshkpour, Pouya, et al.
Publicado: (2025) -
Learning Beyond the Surface: How Far Can Continual Pre-Training with LoRA Enhance LLMs' Domain-Specific Insight Learning?
por: Pezeshkpour, Pouya, et al.
Publicado: (2025) -
AutoPyVerifier: Learning Compact Executable Verifiers for Large Language Model Outputs
por: Pezeshkpour, Pouya, et al.
Publicado: (2026)