Do You Trust Me? Cognitive-Affective Signatures of Trustworthiness in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yeo, Gerard, Churina, Svetlana, Jaidka, Kokil |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Incivility and Rigidity: Evaluating the Risks of Fine-Tuning LLMs for Political Argumentation
von: Churina, Svetlana, et al.
Veröffentlicht: (2024)
von: Churina, Svetlana, et al.
Veröffentlicht: (2024)
On the Limitations of Steering in Language Model Alignment
von: Niranjan, Chebrolu, et al.
Veröffentlicht: (2025)
von: Niranjan, Chebrolu, et al.
Veröffentlicht: (2025)
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2025)
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2025)
From Passive to Persuasive: Localized Activation Injection for Empathy and Negotiation
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025)
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025)
"Reasoning" with Rhetoric: On the Style-Evidence Tradeoff in LLM-Generated Counter-Arguments
von: Verma, Preetika, et al.
Veröffentlicht: (2024)
von: Verma, Preetika, et al.
Veröffentlicht: (2024)
Disentangling Codemixing in Chats: The NUS ABC Codemixed Corpus
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
Beyond Text: Leveraging Multi-Task Learning and Cognitive Appraisal Theory for Post-Purchase Intention Analysis
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2024)
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2024)
Conversations: Love Them, Hate Them, Steer Them
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025)
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025)
Learning Through Dialogue: Engagement and Efficacy Matter More Than Explanations
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2026)
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2026)
Layer of Truth: Probing Belief Shifts under Continual Pre-Training Poisoning
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
Labels or Input? Rethinking Augmentation in Multimodal Hate Detection
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2025)
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2025)
Althea: Human-AI Collaboration for Fact-Checking and Critical Reasoning
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
Turn-Level Empathy Prediction Using Psychological Indicators
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
The MediaSpin Dataset: Post-Publication News Headline Edits Annotated for Media Bias
von: Verma, Preetika, et al.
Veröffentlicht: (2024)
von: Verma, Preetika, et al.
Veröffentlicht: (2024)
TrustMH-Bench: A Comprehensive Benchmark for Evaluating the Trustworthiness of Large Language Models in Mental Health
von: Xiong, Zixin, et al.
Veröffentlicht: (2026)
von: Xiong, Zixin, et al.
Veröffentlicht: (2026)
Do Large Language Models Mirror Cognitive Language Processing?
von: Ren, Yuqi, et al.
Veröffentlicht: (2024)
von: Ren, Yuqi, et al.
Veröffentlicht: (2024)
Reading Between the Lines: How Electronic Nonverbal Cues shape Emotion Decoding
von: Kumar, Taara, et al.
Veröffentlicht: (2026)
von: Kumar, Taara, et al.
Veröffentlicht: (2026)
TrustLDM: Benchmarking Trustworthiness in Language Diffusion Models
von: Mo, Yichuan, et al.
Veröffentlicht: (2026)
von: Mo, Yichuan, et al.
Veröffentlicht: (2026)
AudioTrust: Benchmarking the Multifaceted Trustworthiness of Audio Large Language Models
von: Li, Kai, et al.
Veröffentlicht: (2025)
von: Li, Kai, et al.
Veröffentlicht: (2025)
Now You Hear Me: Audio Narrative Attacks Against Large Audio-Language Models
von: Yu, Ye, et al.
Veröffentlicht: (2026)
von: Yu, Ye, et al.
Veröffentlicht: (2026)
MultiTrust: A Comprehensive Benchmark Towards Trustworthy Multimodal Large Language Models
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
HealMe: Harnessing Cognitive Reframing in Large Language Models for Psychotherapy
von: Xiao, Mengxi, et al.
Veröffentlicht: (2024)
von: Xiao, Mengxi, et al.
Veröffentlicht: (2024)
Eliciting Trustworthiness Priors of Large Language Models via Economic Games
von: Yan, Siyu, et al.
Veröffentlicht: (2026)
von: Yan, Siyu, et al.
Veröffentlicht: (2026)
To Trust or Not to Trust? Enhancing Large Language Models' Situated Faithfulness to External Contexts
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
More RLHF, More Trust? On The Impact of Preference Alignment On Trustworthiness
von: Li, Aaron J., et al.
Veröffentlicht: (2024)
von: Li, Aaron J., et al.
Veröffentlicht: (2024)
A Comprehensive Survey on the Trustworthiness of Large Language Models in Healthcare
von: Aljohani, Manar, et al.
Veröffentlicht: (2025)
von: Aljohani, Manar, et al.
Veröffentlicht: (2025)
Impact of Decoding Methods on Human Alignment of Conversational LLMs
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
Fusion-Augmented Large Language Models: Boosting Diagnostic Trustworthiness via Model Consensus
von: Siam, Md Kamrul, et al.
Veröffentlicht: (2025)
von: Siam, Md Kamrul, et al.
Veröffentlicht: (2025)
Trust-Oriented Adaptive Guardrails for Large Language Models
von: Hu, Jinwei, et al.
Veröffentlicht: (2024)
von: Hu, Jinwei, et al.
Veröffentlicht: (2024)
Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey
von: Ni, Bo, et al.
Veröffentlicht: (2025)
von: Ni, Bo, et al.
Veröffentlicht: (2025)
LLM4Sweat: A Trustworthy Large Language Model for Hyperhidrosis Support
von: Lin, Wenjie, et al.
Veröffentlicht: (2025)
von: Lin, Wenjie, et al.
Veröffentlicht: (2025)
Why Would You Suggest That? Human Trust in Language Model Responses
von: Sharma, Manasi, et al.
Veröffentlicht: (2024)
von: Sharma, Manasi, et al.
Veröffentlicht: (2024)
DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models
von: Wang, Boxin, et al.
Veröffentlicht: (2023)
von: Wang, Boxin, et al.
Veröffentlicht: (2023)
Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression
von: Hong, Junyuan, et al.
Veröffentlicht: (2024)
von: Hong, Junyuan, et al.
Veröffentlicht: (2024)
Cognitive Memory in Large Language Models
von: Shan, Lianlei, et al.
Veröffentlicht: (2025)
von: Shan, Lianlei, et al.
Veröffentlicht: (2025)
Cognitive Effects in Large Language Models
von: Shaki, Jonathan, et al.
Veröffentlicht: (2023)
von: Shaki, Jonathan, et al.
Veröffentlicht: (2023)
Me LLaMA: Foundation Large Language Models for Medical Applications
von: Xie, Qianqian, et al.
Veröffentlicht: (2024)
von: Xie, Qianqian, et al.
Veröffentlicht: (2024)
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models
von: Tan, Fiona Anting, et al.
Veröffentlicht: (2024)
von: Tan, Fiona Anting, et al.
Veröffentlicht: (2024)
FreeEval: A Modular Framework for Trustworthy and Efficient Evaluation of Large Language Models
von: Yu, Zhuohao, et al.
Veröffentlicht: (2024)
von: Yu, Zhuohao, et al.
Veröffentlicht: (2024)
Towards Tracing Trustworthiness Dynamics: Revisiting Pre-training Period of Large Language Models
von: Qian, Chen, et al.
Veröffentlicht: (2024)
von: Qian, Chen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Incivility and Rigidity: Evaluating the Risks of Fine-Tuning LLMs for Political Argumentation
von: Churina, Svetlana, et al.
Veröffentlicht: (2024) -
On the Limitations of Steering in Language Model Alignment
von: Niranjan, Chebrolu, et al.
Veröffentlicht: (2025) -
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2025) -
From Passive to Persuasive: Localized Activation Injection for Empathy and Negotiation
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025) -
"Reasoning" with Rhetoric: On the Style-Evidence Tradeoff in LLM-Generated Counter-Arguments
von: Verma, Preetika, et al.
Veröffentlicht: (2024)