In-Contextual Gender Bias Suppression for Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Oba, Daisuke, Kaneko, Masahiro, Bollegala, Danushka |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding
di: Oba, Daisuke, et al.
Pubblicazione: (2026)
di: Oba, Daisuke, et al.
Pubblicazione: (2026)
The Gaps between Pre-train and Downstream Settings in Bias Evaluation and Debiasing
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
Eagle: Ethical Dataset Given from Real Interactions
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
Investigating the Contextualised Word Embedding Dimensions Specified for Contextual and Temporal Semantic Changes
di: Aida, Taichi, et al.
Pubblicazione: (2024)
di: Aida, Taichi, et al.
Pubblicazione: (2024)
Improving Diversity of Commonsense Generation by Large Language Models via In-Context Learning
di: Zhang, Tianhui, et al.
Pubblicazione: (2024)
di: Zhang, Tianhui, et al.
Pubblicazione: (2024)
Annotating Training Data for Conditional Semantic Textual Similarity Measurement using Large Language Models
di: Zhang, Gaifan, et al.
Pubblicazione: (2025)
di: Zhang, Gaifan, et al.
Pubblicazione: (2025)
Neuron-Level Analysis of Cultural Understanding in Large Language Models
di: Yamamoto, Taisei, et al.
Pubblicazione: (2025)
di: Yamamoto, Taisei, et al.
Pubblicazione: (2025)
A Semantic Distance Metric Learning approach for Lexical Semantic Change Detection
di: Aida, Taichi, et al.
Pubblicazione: (2024)
di: Aida, Taichi, et al.
Pubblicazione: (2024)
SCDTour: Embedding Axis Ordering and Merging for Interpretable Semantic Change Detection
di: Aida, Taichi, et al.
Pubblicazione: (2025)
di: Aida, Taichi, et al.
Pubblicazione: (2025)
Map of Encoders -- Mapping Sentence Encoders using Quantum Relative Entropy
di: Zhang, Gaifan, et al.
Pubblicazione: (2026)
di: Zhang, Gaifan, et al.
Pubblicazione: (2026)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
di: Yamamoto, Taisei, et al.
Pubblicazione: (2025)
di: Yamamoto, Taisei, et al.
Pubblicazione: (2025)
Evaluating Short-Term Temporal Fluctuations of Social Biases in Social Media Data and Masked Language Models
di: Zhou, Yi, et al.
Pubblicazione: (2024)
di: Zhou, Yi, et al.
Pubblicazione: (2024)
Synthetic Data Generation for Training Diversified Commonsense Reasoning Models
di: Zhang, Tianhui, et al.
Pubblicazione: (2026)
di: Zhang, Tianhui, et al.
Pubblicazione: (2026)
Evaluating Unsupervised Dimensionality Reduction Methods for Pretrained Sentence Embeddings
di: Zhang, Gaifan, et al.
Pubblicazione: (2024)
di: Zhang, Gaifan, et al.
Pubblicazione: (2024)
Evaluating the Effect of Retrieval Augmentation on Social Biases
di: Zhang, Tianhui, et al.
Pubblicazione: (2025)
di: Zhang, Tianhui, et al.
Pubblicazione: (2025)
CASE -- Condition-Aware Sentence Embeddings for Conditional Semantic Textual Similarity Measurement
di: Zhang, Gaifan, et al.
Pubblicazione: (2025)
di: Zhang, Gaifan, et al.
Pubblicazione: (2025)
Evaluating the Evaluation of Diversity in Commonsense Generation
di: Zhang, Tianhui, et al.
Pubblicazione: (2025)
di: Zhang, Tianhui, et al.
Pubblicazione: (2025)
Evaluating Gender Bias of Pre-trained Language Models in Natural Language Inference by Considering All Labels
di: Anantaprayoon, Panatchakorn, et al.
Pubblicazione: (2023)
di: Anantaprayoon, Panatchakorn, et al.
Pubblicazione: (2023)
Social Bias Evaluation for Large Language Models Requires Prompt Variations
di: Hida, Rem, et al.
Pubblicazione: (2024)
di: Hida, Rem, et al.
Pubblicazione: (2024)
Improving Unsupervised Constituency Parsing via Maximizing Semantic Information
di: Chen, Junjie, et al.
Pubblicazione: (2024)
di: Chen, Junjie, et al.
Pubblicazione: (2024)
Unsupervised Parsing by Searching for Frequent Word Sequences among Sentences with Equivalent Predicate-Argument Structures
di: Chen, Junjie, et al.
Pubblicazione: (2024)
di: Chen, Junjie, et al.
Pubblicazione: (2024)
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
di: Oi, Masanari, et al.
Pubblicazione: (2024)
di: Oi, Masanari, et al.
Pubblicazione: (2024)
Improving Pre-trained Language Model Sensitivity via Mask Specific losses: A case study on Biomedical NER
di: Abaho, Micheal, et al.
Pubblicazione: (2024)
di: Abaho, Micheal, et al.
Pubblicazione: (2024)
A Little Leak Will Sink a Great Ship: Survey of Transparency for Large Language Models from Start to Finish
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
On the Alignment of Large Language Models with Global Human Opinion
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
JUBAKU: An Adversarial Benchmark for Exposing Culturally Grounded Stereotypes in Japanese LLMs
di: Shiotani, Taihei, et al.
Pubblicazione: (2026)
di: Shiotani, Taihei, et al.
Pubblicazione: (2026)
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
di: Kaneko, Masahiro, et al.
Pubblicazione: (2025)
di: Kaneko, Masahiro, et al.
Pubblicazione: (2025)
Evaluating Gender Bias in Large Language Models
di: Döll, Michael, et al.
Pubblicazione: (2024)
di: Döll, Michael, et al.
Pubblicazione: (2024)
What Matters in Memorizing and Recalling Facts? Multifaceted Benchmarks for Knowledge Probing in Language Models
di: Zhao, Xin, et al.
Pubblicazione: (2024)
di: Zhao, Xin, et al.
Pubblicazione: (2024)
Intent-Aware Self-Correction for Mitigating Social Biases in Large Language Models
di: Anantaprayoon, Panatchakorn, et al.
Pubblicazione: (2025)
di: Anantaprayoon, Panatchakorn, et al.
Pubblicazione: (2025)
A Japanese Benchmark for Evaluating Social Bias in Reasoning Based on Attribution Theory
di: Shiotani, Taihei, et al.
Pubblicazione: (2026)
di: Shiotani, Taihei, et al.
Pubblicazione: (2026)
Gender Bias in Large Language Models across Multiple Languages
di: Zhao, Jinman, et al.
Pubblicazione: (2024)
di: Zhao, Jinman, et al.
Pubblicazione: (2024)
Drifting Objectives for Refining Discrete Diffusion Language Models
di: Oba, Daisuke, et al.
Pubblicazione: (2026)
di: Oba, Daisuke, et al.
Pubblicazione: (2026)
Gender Bias in Emotion Recognition by Large Language Models
di: Herbert, Maureen, et al.
Pubblicazione: (2025)
di: Herbert, Maureen, et al.
Pubblicazione: (2025)
Diffusion-State Policy Optimization for Masked Diffusion Language Models
di: Oba, Daisuke, et al.
Pubblicazione: (2026)
di: Oba, Daisuke, et al.
Pubblicazione: (2026)
Leveraging Large Language Models to Measure Gender Representation Bias in Gendered Language Corpora
di: Derner, Erik, et al.
Pubblicazione: (2024)
di: Derner, Erik, et al.
Pubblicazione: (2024)
Tracing the Roots of Facts in Multilingual Language Models: Independent, Shared, and Transferred Knowledge
di: Zhao, Xin, et al.
Pubblicazione: (2024)
di: Zhao, Xin, et al.
Pubblicazione: (2024)
Mitigation of Gender and Ethnicity Bias in AI-Generated Stories through Model Explanations
di: Dimgba, Martha O., et al.
Pubblicazione: (2025)
di: Dimgba, Martha O., et al.
Pubblicazione: (2025)
Detection, Classification, and Mitigation of Gender Bias in Large Language Models
di: Cheng, Xiaoqing, et al.
Pubblicazione: (2025)
di: Cheng, Xiaoqing, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024) -
Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding
di: Oba, Daisuke, et al.
Pubblicazione: (2026) -
The Gaps between Pre-train and Downstream Settings in Bias Evaluation and Debiasing
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024) -
Eagle: Ethical Dataset Given from Real Interactions
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024) -
Investigating the Contextualised Word Embedding Dimensions Specified for Contextual and Temporal Semantic Changes
di: Aida, Taichi, et al.
Pubblicazione: (2024)