Saved in:
| Main Authors: | Li, Yifei, Chen, Guanyi, He, Tingting |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.31056 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BertaQA: How Much Do Language Models Know About Local Culture?
by: Etxaniz, Julen, et al.
Published: (2024)
by: Etxaniz, Julen, et al.
Published: (2024)
Do Large Language Models Know How Much They Know?
by: Prato, Gabriele, et al.
Published: (2025)
by: Prato, Gabriele, et al.
Published: (2025)
How Do People Quantify Naturally: Evidence from Mandarin Picture Description
by: Zhang, Yayun, et al.
Published: (2026)
by: Zhang, Yayun, et al.
Published: (2026)
Do Large Language Models Judge Error Severity Like Humans?
by: Sun, Diege, et al.
Published: (2025)
by: Sun, Diege, et al.
Published: (2025)
Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups
by: Liu, Geng, et al.
Published: (2025)
by: Liu, Geng, et al.
Published: (2025)
LLMs Know More About Numbers than They Can Say
by: Yuchi, Fengting, et al.
Published: (2026)
by: Yuchi, Fengting, et al.
Published: (2026)
Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
by: Chen, Xingyu, et al.
Published: (2024)
by: Chen, Xingyu, et al.
Published: (2024)
Pronoun Logic
by: Bohrer, Rose, et al.
Published: (2024)
by: Bohrer, Rose, et al.
Published: (2024)
Do LLMs Know to Respect Copyright Notice?
by: Xu, Jialiang, et al.
Published: (2024)
by: Xu, Jialiang, et al.
Published: (2024)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
by: Oğuz, Metehan, et al.
Published: (2024)
by: Oğuz, Metehan, et al.
Published: (2024)
On the Robustness of Knowledge Editing for Detoxification
by: Dong, Ming, et al.
Published: (2026)
by: Dong, Ming, et al.
Published: (2026)
Emotional Supporters often Use Multiple Strategies in a Single Turn
by: Bai, Xin, et al.
Published: (2025)
by: Bai, Xin, et al.
Published: (2025)
Computational Modelling of Plurality and Definiteness in Chinese Noun Phrases
by: Liu, Yuqi, et al.
Published: (2024)
by: Liu, Yuqi, et al.
Published: (2024)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
by: Islam, Saad Obaid ul, et al.
Published: (2025)
by: Islam, Saad Obaid ul, et al.
Published: (2025)
How Much Do Large Language Models Know about Human Motion? A Case Study in 3D Avatar Control
by: Li, Kunhang, et al.
Published: (2025)
by: Li, Kunhang, et al.
Published: (2025)
When Seekers Are Hard to Help: Evaluating Emotional Support Dialogue Systems in Worst-Case Interactions
by: Yang, Jiajie, et al.
Published: (2026)
by: Yang, Jiajie, et al.
Published: (2026)
Evaluating the Unseen Capabilities: How Many Theorems Do LLMs Know?
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
by: Fernandes, Gustavo Lúcius, et al.
Published: (2026)
by: Fernandes, Gustavo Lúcius, et al.
Published: (2026)
What Do LLMs Know About Alzheimer's Disease? Multi-loss Fine-Tuning and Probing for AD Detection
by: Jiang, Lei, et al.
Published: (2026)
by: Jiang, Lei, et al.
Published: (2026)
Robust Pronoun Fidelity with English LLMs: Are they Reasoning, Repeating, or Just Biased?
by: Gautam, Vagrant, et al.
Published: (2024)
by: Gautam, Vagrant, et al.
Published: (2024)
LLMs Learn Constructions That Humans Do Not Know
by: Dunn, Jonathan, et al.
Published: (2025)
by: Dunn, Jonathan, et al.
Published: (2025)
The Expressions of Depression and Anxiety in Chinese Psycho-counseling: Usage of First-person Singular Pronoun and Negative Emotional Words
by: Ma, Lizhi, et al.
Published: (2025)
by: Ma, Lizhi, et al.
Published: (2025)
How Much Do Circuits Tell Us? Measuring the Consistency and Specificity of Language Model Circuits
by: Li, Michael, et al.
Published: (2026)
by: Li, Michael, et al.
Published: (2026)
Mention Attention for Pronoun Translation
by: Tang, Gongbo, et al.
Published: (2024)
by: Tang, Gongbo, et al.
Published: (2024)
Turning English-centric LLMs Into Polyglots: How Much Multilinguality Is Needed?
by: Kew, Tannon, et al.
Published: (2023)
by: Kew, Tannon, et al.
Published: (2023)
Do They Understand Them? An Updated Evaluation on Nonbinary Pronoun Handling in Large Language Models
by: Tang, Xushuo, et al.
Published: (2025)
by: Tang, Xushuo, et al.
Published: (2025)
Knowing But Not Doing: Convergent Morality and Divergent Action in LLMs
by: Huang, Jen-tse, et al.
Published: (2026)
by: Huang, Jen-tse, et al.
Published: (2026)
What Do Self-Supervised Speech Models Know About Words?
by: Pasad, Ankita, et al.
Published: (2023)
by: Pasad, Ankita, et al.
Published: (2023)
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
by: Zhao, Raoyuan, et al.
Published: (2025)
by: Zhao, Raoyuan, et al.
Published: (2025)
A Primer in Post-Training Reasoning Data: What We Know About How It Works
by: Li, Yaoming, et al.
Published: (2026)
by: Li, Yaoming, et al.
Published: (2026)
CCNU at SemEval-2025 Task 3: Leveraging Internal and External Knowledge of Large Language Models for Multilingual Hallucination Annotation
by: Liu, Xu, et al.
Published: (2025)
by: Liu, Xu, et al.
Published: (2025)
How Much of Your Data Can Suck? Thresholds for Domain Performance and Emergent Misalignment in LLMs
by: Ouyang, Jian, et al.
Published: (2025)
by: Ouyang, Jian, et al.
Published: (2025)
GRUFF: LLM Pronoun Fidelity, Reasoning, and Biases in German
by: Mewes, Fabian, et al.
Published: (2026)
by: Mewes, Fabian, et al.
Published: (2026)
Do LLMs Know about Hallucination? An Empirical Investigation of LLM's Hidden States
by: Duan, Hanyu, et al.
Published: (2024)
by: Duan, Hanyu, et al.
Published: (2024)
Do You Know About My Nation? Investigating Multilingual Language Models' Cultural Literacy Through Factual Knowledge
by: Tanwar, Eshaan, et al.
Published: (2025)
by: Tanwar, Eshaan, et al.
Published: (2025)
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation
by: Li, Zhuohang, et al.
Published: (2024)
by: Li, Zhuohang, et al.
Published: (2024)
Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks
by: Pan, Wenbo, et al.
Published: (2025)
by: Pan, Wenbo, et al.
Published: (2025)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
by: Cheang, Chi Seng, et al.
Published: (2025)
by: Cheang, Chi Seng, et al.
Published: (2025)
Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models
by: Madhusudhan, Nishanth, et al.
Published: (2024)
by: Madhusudhan, Nishanth, et al.
Published: (2024)
How Much Annotation is Needed to Compare Summarization Models?
by: Shaib, Chantal, et al.
Published: (2024)
by: Shaib, Chantal, et al.
Published: (2024)
Similar Items
-
BertaQA: How Much Do Language Models Know About Local Culture?
by: Etxaniz, Julen, et al.
Published: (2024) -
Do Large Language Models Know How Much They Know?
by: Prato, Gabriele, et al.
Published: (2025) -
How Do People Quantify Naturally: Evidence from Mandarin Picture Description
by: Zhang, Yayun, et al.
Published: (2026) -
Do Large Language Models Judge Error Severity Like Humans?
by: Sun, Diege, et al.
Published: (2025) -
Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups
by: Liu, Geng, et al.
Published: (2025)