On the Mutual Influence of Gender and Occupation in LLM Representations
Fuente:
arXiv
Guardado en:
| Autores principales: | An, Haozhe, Baumler, Connor, Sancheti, Abhilasha, Rudinger, Rachel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Influence of Gender and Race in Romantic Relationship Prediction from Large Language Models
por: Sancheti, Abhilasha, et al.
Publicado: (2024)
por: Sancheti, Abhilasha, et al.
Publicado: (2024)
How much reliable is ChatGPT's prediction on Information Extraction under Input Perturbations?
por: Mondal, Ishani, et al.
Publicado: (2024)
por: Mondal, Ishani, et al.
Publicado: (2024)
Artifacts or Abduction: How Do LLMs Answer Multiple-Choice Questions Without the Question?
por: Balepur, Nishant, et al.
Publicado: (2024)
por: Balepur, Nishant, et al.
Publicado: (2024)
Susu Box or Piggy Bank: Assessing Cultural Commonsense Knowledge between Ghana and the U.S
por: Acquaye, Christabel, et al.
Publicado: (2024)
por: Acquaye, Christabel, et al.
Publicado: (2024)
Do Large Language Models Discriminate in Hiring Decisions on the Basis of Race, Ethnicity, and Gender?
por: An, Haozhe, et al.
Publicado: (2024)
por: An, Haozhe, et al.
Publicado: (2024)
Post-Hoc Answer Attribution for Grounded and Trustworthy Long Document Comprehension: Task, Insights, and Challenges
por: Sancheti, Abhilasha, et al.
Publicado: (2024)
por: Sancheti, Abhilasha, et al.
Publicado: (2024)
When Stereotypes GTG: The Impact of Predictive Text Suggestions on Gender Bias in Human-AI Co-Writing
por: Baumler, Connor, et al.
Publicado: (2024)
por: Baumler, Connor, et al.
Publicado: (2024)
Reverse Question Answering: Can an LLM Write a Question so Hard (or Bad) that it Can't Answer?
por: Balepur, Nishant, et al.
Publicado: (2024)
por: Balepur, Nishant, et al.
Publicado: (2024)
NLI under the Microscope: What Atomic Hypothesis Decomposition Reveals
por: Srikanth, Neha, et al.
Publicado: (2025)
por: Srikanth, Neha, et al.
Publicado: (2025)
Is Your Large Language Model Knowledgeable or a Choices-Only Cheater?
por: Balepur, Nishant, et al.
Publicado: (2024)
por: Balepur, Nishant, et al.
Publicado: (2024)
Multiple LLM Agents Debate for Equitable Cultural Alignment
por: Ki, Dayeon, et al.
Publicado: (2025)
por: Ki, Dayeon, et al.
Publicado: (2025)
Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations
por: Acquaye, Christabel, et al.
Publicado: (2026)
por: Acquaye, Christabel, et al.
Publicado: (2026)
Everything is Plausible: Investigating the Impact of LLM Rationales on Human Notions of Plausibility
por: Palta, Shramay, et al.
Publicado: (2025)
por: Palta, Shramay, et al.
Publicado: (2025)
Test-Time Reasoners Are Strategic Multiple-Choice Test-Takers
por: Balepur, Nishant, et al.
Publicado: (2025)
por: Balepur, Nishant, et al.
Publicado: (2025)
It's Not Easy Being Wrong: Large Language Models Struggle with Process of Elimination Reasoning
por: Balepur, Nishant, et al.
Publicado: (2023)
por: Balepur, Nishant, et al.
Publicado: (2023)
How often are errors in natural language reasoning due to paraphrastic variability?
por: Srikanth, Neha, et al.
Publicado: (2024)
por: Srikanth, Neha, et al.
Publicado: (2024)
Can You Make It Sound Like You? Post-Editing LLM-Generated Text for Personal Style
por: Baumler, Connor, et al.
Publicado: (2026)
por: Baumler, Connor, et al.
Publicado: (2026)
DiscoTrace: Representing and Comparing Answering Strategies of Humans and LLMs in Information-Seeking Question Answering
por: Srikanth, Neha, et al.
Publicado: (2026)
por: Srikanth, Neha, et al.
Publicado: (2026)
Which of These Best Describes Multiple Choice Evaluation with LLMs? A) Forced B) Flawed C) Fixable D) All of the Above
por: Balepur, Nishant, et al.
Publicado: (2025)
por: Balepur, Nishant, et al.
Publicado: (2025)
Language Models Predict Empathy Gaps Between Social In-groups and Out-groups
por: Hou, Yu, et al.
Publicado: (2025)
por: Hou, Yu, et al.
Publicado: (2025)
Exploring Gender Bias Beyond Occupational Titles
por: Sabir, Ahmed, et al.
Publicado: (2025)
por: Sabir, Ahmed, et al.
Publicado: (2025)
Speaking the Right Language: The Impact of Expertise Alignment in User-AI Interactions
por: Palta, Shramay, et al.
Publicado: (2025)
por: Palta, Shramay, et al.
Publicado: (2025)
Assumed Identities: Quantifying Gender Bias in Machine Translation of Gender-Ambiguous Occupational Terms
por: Mastromichalakis, Orfeas Menis, et al.
Publicado: (2025)
por: Mastromichalakis, Orfeas Menis, et al.
Publicado: (2025)
LABOR-LLM: Language-Based Occupational Representations with Large Language Models
por: Athey, Susan, et al.
Publicado: (2024)
por: Athey, Susan, et al.
Publicado: (2024)
HALoGEN: Fantastic LLM Hallucinations and Where to Find Them
por: Ravichander, Abhilasha, et al.
Publicado: (2025)
por: Ravichander, Abhilasha, et al.
Publicado: (2025)
Understanding Common Ground Misalignment in Goal-Oriented Dialog: A Case-Study with Ubuntu Chat Logs
por: Sarkar, Rupak, et al.
Publicado: (2025)
por: Sarkar, Rupak, et al.
Publicado: (2025)
'Rich Dad, Poor Lad': How do Large Language Models Contextualize Socioeconomic Factors in College Admission ?
por: Nghiem, Huy, et al.
Publicado: (2025)
por: Nghiem, Huy, et al.
Publicado: (2025)
Are Female Carpenters like Blue Bananas? A Corpus Investigation of Occupation Gender Typicality
por: Ju, Da, et al.
Publicado: (2024)
por: Ju, Da, et al.
Publicado: (2024)
Whose Boat Does it Float? Improving Personalization in Preference Tuning via Inferred User Personas
por: Balepur, Nishant, et al.
Publicado: (2025)
por: Balepur, Nishant, et al.
Publicado: (2025)
FRIDA to the Rescue! Analyzing Synthetic Data Effectiveness in Object-Based Common Sense Reasoning for Disaster Response
por: Shichman, Mollie, et al.
Publicado: (2025)
por: Shichman, Mollie, et al.
Publicado: (2025)
Plausibly Problematic Questions in Multiple-Choice Benchmarks for Commonsense Reasoning
por: Palta, Shramay, et al.
Publicado: (2024)
por: Palta, Shramay, et al.
Publicado: (2024)
Colombian Waitresses y Jueces canadienses: Gender and Country Biases in Occupation Recommendations from LLMs
por: Rodríguez, Elisa Forcada, et al.
Publicado: (2025)
por: Rodríguez, Elisa Forcada, et al.
Publicado: (2025)
Learning Mutually Informed Representations for Characters and Subwords
por: Wang, Yilin, et al.
Publicado: (2023)
por: Wang, Yilin, et al.
Publicado: (2023)
Multilingual large language models leak human stereotypes across language boundaries
por: Cao, Yang Trista, et al.
Publicado: (2023)
por: Cao, Yang Trista, et al.
Publicado: (2023)
Natural Language Inference Improves Compositionality in Vision-Language Models
por: Cascante-Bonilla, Paola, et al.
Publicado: (2024)
por: Cascante-Bonilla, Paola, et al.
Publicado: (2024)
Causally Testing Gender Bias in LLMs: A Case Study on Occupational Bias
por: Chen, Yuen, et al.
Publicado: (2022)
por: Chen, Yuen, et al.
Publicado: (2022)
Reheat Nachos for Dinner? Evaluating AI Support for Cross-Cultural Communication of Neologisms
por: Ki, Dayeon, et al.
Publicado: (2026)
por: Ki, Dayeon, et al.
Publicado: (2026)
Poor-Supervised Evaluation for SuperLLM via Mutual Consistency
por: Yuan, Peiwen, et al.
Publicado: (2024)
por: Yuan, Peiwen, et al.
Publicado: (2024)
The Causal Influence of Grammatical Gender on Distributional Semantics
por: Stańczak, Karolina, et al.
Publicado: (2023)
por: Stańczak, Karolina, et al.
Publicado: (2023)
GOSt-MT: A Knowledge Graph for Occupation-related Gender Biases in Machine Translation
por: Mastromichalakis, Orfeas Menis, et al.
Publicado: (2024)
por: Mastromichalakis, Orfeas Menis, et al.
Publicado: (2024)
Ejemplares similares
-
On the Influence of Gender and Race in Romantic Relationship Prediction from Large Language Models
por: Sancheti, Abhilasha, et al.
Publicado: (2024) -
How much reliable is ChatGPT's prediction on Information Extraction under Input Perturbations?
por: Mondal, Ishani, et al.
Publicado: (2024) -
Artifacts or Abduction: How Do LLMs Answer Multiple-Choice Questions Without the Question?
por: Balepur, Nishant, et al.
Publicado: (2024) -
Susu Box or Piggy Bank: Assessing Cultural Commonsense Knowledge between Ghana and the U.S
por: Acquaye, Christabel, et al.
Publicado: (2024) -
Do Large Language Models Discriminate in Hiring Decisions on the Basis of Race, Ethnicity, and Gender?
por: An, Haozhe, et al.
Publicado: (2024)