Small Changes, Large Consequences: Analyzing the Allocational Fairness of LLMs in Hiring Contexts
Fuente:
arXiv
Saved in:
| Main Authors: | Seshadri, Preethi, Chen, Hongyu, Singh, Sameer, Goldfarb-Tarrant, Seraphina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Safer or Luckier? LLMs as Safety Evaluators Are Not Robust to Artifacts
by: Chen, Hongyu, et al.
Published: (2025)
by: Chen, Hongyu, et al.
Published: (2025)
Lost in Simulation: LLM-Simulated Users are Unreliable Proxies for Human Users in Agentic Evaluations
by: Seshadri, Preethi, et al.
Published: (2026)
by: Seshadri, Preethi, et al.
Published: (2026)
Scalpel vs. Hammer: GRPO Amplifies Existing Capabilities, SFT Replaces Them
by: Rajani, Neel, et al.
Published: (2025)
by: Rajani, Neel, et al.
Published: (2025)
MultiContrievers: Analysis of Dense Retrieval Representations
by: Goldfarb-Tarrant, Seraphina, et al.
Published: (2024)
by: Goldfarb-Tarrant, Seraphina, et al.
Published: (2024)
A SMART Mnemonic Sounds like "Glue Tonic": Mixing LLMs with Student Feedback to Make Mnemonic Learning Stick
by: Balepur, Nishant, et al.
Published: (2024)
by: Balepur, Nishant, et al.
Published: (2024)
Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning
by: Aakanksha, et al.
Published: (2024)
by: Aakanksha, et al.
Published: (2024)
Are Models Biased on Text without Gender-related Language?
by: Belém, Catarina G, et al.
Published: (2024)
by: Belém, Catarina G, et al.
Published: (2024)
Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism
by: Orgad, Hadas, et al.
Published: (2026)
by: Orgad, Hadas, et al.
Published: (2026)
The Multilingual Alignment Prism: Aligning Global and Local Preferences to Reduce Harm
by: Aakanksha, et al.
Published: (2024)
by: Aakanksha, et al.
Published: (2024)
A Good Plan is Hard to Find: Aligning Models with Preferences is Misaligned with What Helps Users
by: Balepur, Nishant, et al.
Published: (2025)
by: Balepur, Nishant, et al.
Published: (2025)
MAPS: A Multilingual Benchmark for Agent Performance and Security
by: Hofman, Omer, et al.
Published: (2025)
by: Hofman, Omer, et al.
Published: (2025)
AI Hiring with LLMs: A Context-Aware and Explainable Multi-Agent Framework for Resume Screening
by: Lo, Frank P. -W., et al.
Published: (2025)
by: Lo, Frank P. -W., et al.
Published: (2025)
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models
by: Wang, Ze, et al.
Published: (2024)
by: Wang, Ze, et al.
Published: (2024)
Learning is Forgetting: LLM Training As Lossy Compression
by: Conklin, Henry C., et al.
Published: (2026)
by: Conklin, Henry C., et al.
Published: (2026)
Evaluating the Promise and Pitfalls of LLMs in Hiring Decisions
by: Anzenberg, Eitan, et al.
Published: (2025)
by: Anzenberg, Eitan, et al.
Published: (2025)
Hire a Linguist!: Learning Endangered Languages with In-Context Linguistic Descriptions
by: Zhang, Kexun, et al.
Published: (2024)
by: Zhang, Kexun, et al.
Published: (2024)
DAST: Context-Aware Compression in LLMs via Dynamic Allocation of Soft Tokens
by: Chen, Shaoshen, et al.
Published: (2025)
by: Chen, Shaoshen, et al.
Published: (2025)
Med-HEAL: Analyzing and Mitigating Hallucinations in Medical LLMs with Hallucination-Aware In-Context Learning
by: Liao, Yiming, et al.
Published: (2026)
by: Liao, Yiming, et al.
Published: (2026)
Speech LLMs in Low-Resource Scenarios: Data Volume Requirements and the Impact of Pretraining on High-Resource Languages
by: Fong, Seraphina, et al.
Published: (2025)
by: Fong, Seraphina, et al.
Published: (2025)
Agree to Disagree? A Meta-Evaluation of LLM Misgendering
by: Subramonian, Arjun, et al.
Published: (2025)
by: Subramonian, Arjun, et al.
Published: (2025)
The Multilingual Divide and Its Impact on Global AI Safety
by: Peppin, Aidan, et al.
Published: (2025)
by: Peppin, Aidan, et al.
Published: (2025)
Nudging: Inference-time Alignment of LLMs via Guided Decoding
by: Fei, Yu, et al.
Published: (2024)
by: Fei, Yu, et al.
Published: (2024)
LLMs on Trial: Evaluating Judicial Fairness for Large Language Models
by: Hu, Yiran, et al.
Published: (2025)
by: Hu, Yiran, et al.
Published: (2025)
Breaking Quadratic Barriers: A Non-Attention LLM for Ultra-Long Context Horizons
by: Kiruluta, Andrew, et al.
Published: (2025)
by: Kiruluta, Andrew, et al.
Published: (2025)
Leveraging In-Context Learning for Language Model Agents
by: Gupta, Shivanshu, et al.
Published: (2025)
by: Gupta, Shivanshu, et al.
Published: (2025)
Analyzing the Role of Context in Forecasting with Large Language Models
by: Mutschlechner, Gerrit, et al.
Published: (2025)
by: Mutschlechner, Gerrit, et al.
Published: (2025)
"You Are Rejected!": An Empirical Study of Large Language Models Taking Hiring Evaluations
by: Fu, Dingjie, et al.
Published: (2025)
by: Fu, Dingjie, et al.
Published: (2025)
FairMT-Bench: Benchmarking Fairness for Multi-turn Dialogue in Conversational LLMs
by: Fan, Zhiting, et al.
Published: (2024)
by: Fan, Zhiting, et al.
Published: (2024)
Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation
by: Pan, Ruotong, et al.
Published: (2024)
by: Pan, Ruotong, et al.
Published: (2024)
Do Large Language Models Discriminate in Hiring Decisions on the Basis of Race, Ethnicity, and Gender?
by: An, Haozhe, et al.
Published: (2024)
by: An, Haozhe, et al.
Published: (2024)
SMITE: Enhancing Fairness in LLMs through Optimal In-Context Example Selection via Dynamic Validation
by: Chhikara, Garima, et al.
Published: (2025)
by: Chhikara, Garima, et al.
Published: (2025)
Evaluating LLM Behavior in Hiring: Implicit Weights, Fairness Across Groups, and Alignment with Human Preferences
by: Hoffmann, Morgane, et al.
Published: (2026)
by: Hoffmann, Morgane, et al.
Published: (2026)
Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models
by: Ismailov, Altynbek, et al.
Published: (2025)
by: Ismailov, Altynbek, et al.
Published: (2025)
Benchmark Illusion: Disagreement among LLMs and Its Scientific Consequences
by: Yang, Eddie, et al.
Published: (2026)
by: Yang, Eddie, et al.
Published: (2026)
Do Large Language Models Rank Fairly? An Empirical Study on the Fairness of LLMs as Rankers
by: Wang, Yuan, et al.
Published: (2024)
by: Wang, Yuan, et al.
Published: (2024)
Entropy-Aligned Decoding of LMs for Better Writing and Reasoning
by: Ahmed, Kareem, et al.
Published: (2026)
by: Ahmed, Kareem, et al.
Published: (2026)
XITE: Cross-lingual Interpolation for Transfer using Embeddings
by: Fazili, Barah, et al.
Published: (2026)
by: Fazili, Barah, et al.
Published: (2026)
Invisible Filters: Cultural Bias in Hiring Evaluations Using Large Language Models
by: Rao, Pooja S. B., et al.
Published: (2025)
by: Rao, Pooja S. B., et al.
Published: (2025)
S2LPP: Small-to-Large Prompt Prediction across LLMs
by: Cheng, Liang, et al.
Published: (2025)
by: Cheng, Liang, et al.
Published: (2025)
Monte Carlo Sampling for Analyzing In-Context Examples
by: Schoch, Stephanie, et al.
Published: (2025)
by: Schoch, Stephanie, et al.
Published: (2025)
Similar Items
-
Safer or Luckier? LLMs as Safety Evaluators Are Not Robust to Artifacts
by: Chen, Hongyu, et al.
Published: (2025) -
Lost in Simulation: LLM-Simulated Users are Unreliable Proxies for Human Users in Agentic Evaluations
by: Seshadri, Preethi, et al.
Published: (2026) -
Scalpel vs. Hammer: GRPO Amplifies Existing Capabilities, SFT Replaces Them
by: Rajani, Neel, et al.
Published: (2025) -
MultiContrievers: Analysis of Dense Retrieval Representations
by: Goldfarb-Tarrant, Seraphina, et al.
Published: (2024) -
A SMART Mnemonic Sounds like "Glue Tonic": Mixing LLMs with Student Feedback to Make Mnemonic Learning Stick
by: Balepur, Nishant, et al.
Published: (2024)