Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wen, Yuchen, Bi, Keping, Chen, Wei, Guo, Jiafeng, Cheng, Xueqi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Estimating Commonsense Plausibility through Semantic Shifts
von: Cui, Wanqing, et al.
Veröffentlicht: (2025)
von: Cui, Wanqing, et al.
Veröffentlicht: (2025)
An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
von: Zhang, Hengran, et al.
Veröffentlicht: (2024)
von: Zhang, Hengran, et al.
Veröffentlicht: (2024)
LLM-Specific Utility: A New Perspective for Retrieval-Augmented Generation
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
Bagging-Based Model Merging for Robust General Text Embeddings
von: Zhang, Hengran, et al.
Veröffentlicht: (2026)
von: Zhang, Hengran, et al.
Veröffentlicht: (2026)
Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
A Comparative Study of Specialized LLMs as Dense Retrievers
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
CLIPure: Purification in Latent Space via CLIP for Adversarially Robust Zero-Shot Classification
von: Zhang, Mingkun, et al.
Veröffentlicht: (2025)
von: Zhang, Mingkun, et al.
Veröffentlicht: (2025)
Distilling a Small Utility-Based Passage Selector to Enhance Retrieval-Augmented Generation
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
Locating and Mitigating Gender Bias in Large Language Models
von: Cai, Yuchen, et al.
Veröffentlicht: (2024)
von: Cai, Yuchen, et al.
Veröffentlicht: (2024)
Utility-Focused LLM Annotation for Retrieval and Retrieval-Augmented Generation
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
Beyond Relevance: Utility-Centric Retrieval in the LLM Era
von: Zhang, Hengran, et al.
Veröffentlicht: (2026)
von: Zhang, Hengran, et al.
Veröffentlicht: (2026)
MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning
von: Cui, Wanqing, et al.
Veröffentlicht: (2024)
von: Cui, Wanqing, et al.
Veröffentlicht: (2024)
How Knowledge Popularity Influences and Enhances LLM Knowledge Boundary Perception
von: Ni, Shiyu, et al.
Veröffentlicht: (2025)
von: Ni, Shiyu, et al.
Veröffentlicht: (2025)
When Do LLMs Need Retrieval Augmentation? Mitigating LLMs' Overconfidence Helps Retrieval Augmentation
von: Ni, Shiyu, et al.
Veröffentlicht: (2024)
von: Ni, Shiyu, et al.
Veröffentlicht: (2024)
LINKAGE: Listwise Ranking among Varied-Quality References for Non-Factoid QA Evaluation via LLMs
von: Yang, Sihui, et al.
Veröffentlicht: (2024)
von: Yang, Sihui, et al.
Veröffentlicht: (2024)
Towards Robust Universal Information Extraction: Benchmark, Evaluation, and Solution
von: Zhu, Jizhao, et al.
Veröffentlicht: (2025)
von: Zhu, Jizhao, et al.
Veröffentlicht: (2025)
How Do LLM-Generated Texts Impact Term-Based Retrieval Models?
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Promoting Equality in Large Language Models: Identifying and Mitigating the Implicit Bias based on Bayesian Theory
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues
von: Vedula, Bhaskara Hanuma, et al.
Veröffentlicht: (2026)
von: Vedula, Bhaskara Hanuma, et al.
Veröffentlicht: (2026)
Uncovering Implicit Bias in Large Language Models with Concept Learning Dataset
von: Wang, Leroy Z.
Veröffentlicht: (2025)
von: Wang, Leroy Z.
Veröffentlicht: (2025)
From Implicit to Explicit: Enhancing Self-Recognition in Large Language Models
von: Zhou, Yinghan, et al.
Veröffentlicht: (2025)
von: Zhou, Yinghan, et al.
Veröffentlicht: (2025)
No LLM is Free From Bias: A Comprehensive Study of Bias Evaluation in Large Language Models
von: Kumar, Charaka Vinayak, et al.
Veröffentlicht: (2025)
von: Kumar, Charaka Vinayak, et al.
Veröffentlicht: (2025)
Are Large Language Models More Honest in Their Probabilistic or Verbalized Confidence?
von: Ni, Shiyu, et al.
Veröffentlicht: (2024)
von: Ni, Shiyu, et al.
Veröffentlicht: (2024)
StruEdit: Structured Outputs Enable the Fast and Accurate Knowledge Editing for Large Language Models
von: Bi, Baolong, et al.
Veröffentlicht: (2024)
von: Bi, Baolong, et al.
Veröffentlicht: (2024)
Large Language Model Bias Mitigation from the Perspective of Knowledge Editing
von: Chen, Ruizhe, et al.
Veröffentlicht: (2024)
von: Chen, Ruizhe, et al.
Veröffentlicht: (2024)
QUITO-X: A New Perspective on Context Compression from the Information Bottleneck Theory
von: Wang, Yihang, et al.
Veröffentlicht: (2024)
von: Wang, Yihang, et al.
Veröffentlicht: (2024)
ReAD: Reinforcement-Guided Capability Distillation for Large Language Models
von: Cheng, Xueqi, et al.
Veröffentlicht: (2026)
von: Cheng, Xueqi, et al.
Veröffentlicht: (2026)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs
von: Kumar, Divyanshu, et al.
Veröffentlicht: (2024)
von: Kumar, Divyanshu, et al.
Veröffentlicht: (2024)
Robust Neural Information Retrieval: An Adversarial and Out-of-distribution Perspective
von: Liu, Yu-An, et al.
Veröffentlicht: (2024)
von: Liu, Yu-An, et al.
Veröffentlicht: (2024)
MDIT-Bench: Evaluating the Dual-Implicit Toxicity in Large Multimodal Models
von: Jin, Bohan, et al.
Veröffentlicht: (2025)
von: Jin, Bohan, et al.
Veröffentlicht: (2025)
Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary Perception
von: Ni, Shiyu, et al.
Veröffentlicht: (2025)
von: Ni, Shiyu, et al.
Veröffentlicht: (2025)
ImF: Implicit Fingerprint for Large Language Models
von: Wu, Jiaxuan, et al.
Veröffentlicht: (2025)
von: Wu, Jiaxuan, et al.
Veröffentlicht: (2025)
Psychometric Predictive Power of Large Language Models
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
LPNL: Scalable Link Prediction with Large Language Models
von: Bi, Baolong, et al.
Veröffentlicht: (2024)
von: Bi, Baolong, et al.
Veröffentlicht: (2024)
Mitigating Boundary Ambiguity and Inherent Bias for Text Classification in the Era of Large Language Models
von: Lu, Zhenyi, et al.
Veröffentlicht: (2024)
von: Lu, Zhenyi, et al.
Veröffentlicht: (2024)
Intersectional Bias in Japanese Large Language Models from a Contextualized Perspective
von: Yanaka, Hitomi, et al.
Veröffentlicht: (2025)
von: Yanaka, Hitomi, et al.
Veröffentlicht: (2025)
Think Before You Speak: Cultivating Communication Skills of Large Language Models via Inner Monologue
von: Zhou, Junkai, et al.
Veröffentlicht: (2023)
von: Zhou, Junkai, et al.
Veröffentlicht: (2023)
On the Capacity of Citation Generation by Large Language Models
von: Qian, Haosheng, et al.
Veröffentlicht: (2024)
von: Qian, Haosheng, et al.
Veröffentlicht: (2024)
Who is in the Spotlight: The Hidden Bias Undermining Multimodal Retrieval-Augmented Generation
von: Yao, Jiayu, et al.
Veröffentlicht: (2025)
von: Yao, Jiayu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Estimating Commonsense Plausibility through Semantic Shifts
von: Cui, Wanqing, et al.
Veröffentlicht: (2025) -
An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
von: Zhang, Hengran, et al.
Veröffentlicht: (2024) -
LLM-Specific Utility: A New Perspective for Retrieval-Augmented Generation
von: Zhang, Hengran, et al.
Veröffentlicht: (2025) -
Bagging-Based Model Merging for Robust General Text Embeddings
von: Zhang, Hengran, et al.
Veröffentlicht: (2026) -
Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)