Evaluating LLM Alignment With Human Trust Models
Fuente:
arXiv
Saved in:
| Main Authors: | Debnath, Anushka, Cranefield, Stephen, Savarimuthu, Bastin Tony Roy, Lorini, Emiliano |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can LLMs Reason About Trust?: A Pilot Study
by: Debnath, Anushka, et al.
Published: (2025)
by: Debnath, Anushka, et al.
Published: (2025)
Social Norm Reasoning in Multimodal Language Models: An Evaluation
by: Chowdhury, Oishik, et al.
Published: (2026)
by: Chowdhury, Oishik, et al.
Published: (2026)
Norm Violation Detection in Multi-Agent Systems using Large Language Models: A Pilot Study
by: He, Shawn, et al.
Published: (2024)
by: He, Shawn, et al.
Published: (2024)
Evolution of Cooperation in LLM-Agent Societies: A Preliminary Study Using Different Punishment Strategies
by: Warnakulasuriya, Kavindu, et al.
Published: (2025)
by: Warnakulasuriya, Kavindu, et al.
Published: (2025)
Harnessing the power of LLMs for normative reasoning in MASs
by: Savarimuthu, Bastin Tony Roy, et al.
Published: (2024)
by: Savarimuthu, Bastin Tony Roy, et al.
Published: (2024)
Review Arcade: On the Human Alignment and Gameability of LLM Reviews
by: Hatzel, Hans Ole, et al.
Published: (2026)
by: Hatzel, Hans Ole, et al.
Published: (2026)
Designing for Accountable Agents: a Viewpoint
by: Cranefield, Stephen, et al.
Published: (2026)
by: Cranefield, Stephen, et al.
Published: (2026)
Co-Alignment: Rethinking Alignment as Bidirectional Human-AI Cognitive Adaptation
by: Li, Yubo, et al.
Published: (2025)
by: Li, Yubo, et al.
Published: (2025)
To Mask or to Mirror: Human-AI Alignment in Collective Reasoning
by: Qian, Crystal, et al.
Published: (2025)
by: Qian, Crystal, et al.
Published: (2025)
Dynamic Trust-Aware Sparse Communication Topology for LLM-Based Multi-Agent Consensus
by: Gou, Wanshuang, et al.
Published: (2026)
by: Gou, Wanshuang, et al.
Published: (2026)
SORA-ATMAS: Adaptive Trust Management and Multi-LLM Aligned Governance for Future Smart Cities
by: Antuley, Usama, et al.
Published: (2025)
by: Antuley, Usama, et al.
Published: (2025)
Autonomous Agents on Blockchains: Standards, Execution Models, and Trust Boundaries
by: Alqithami, Saad
Published: (2026)
by: Alqithami, Saad
Published: (2026)
The Traitors: Deception and Trust in Multi-Agent Language Model Simulations
by: Curvo, Pedro M. P.
Published: (2025)
by: Curvo, Pedro M. P.
Published: (2025)
Ultra Low-Cost Two-Stage Multimodal System for Non-Normative Behavior Detection
by: Lu, Albert, et al.
Published: (2024)
by: Lu, Albert, et al.
Published: (2024)
The Influence of Human-inspired Agentic Sophistication in LLM-driven Strategic Reasoners
by: Trencsenyi, Vince, et al.
Published: (2025)
by: Trencsenyi, Vince, et al.
Published: (2025)
What Is Your Agent's GPA? A Framework for Evaluating Agent Goal-Plan-Action Alignment
by: Jia, Allison Sihan, et al.
Published: (2025)
by: Jia, Allison Sihan, et al.
Published: (2025)
Communication-Efficient Desire Alignment for Embodied Agent-Human Adaptation
by: Wang, Yuanfei, et al.
Published: (2025)
by: Wang, Yuanfei, et al.
Published: (2025)
Interactional Fairness in LLM Multi-Agent Systems: An Evaluation Framework
by: Binkyte, Ruta
Published: (2025)
by: Binkyte, Ruta
Published: (2025)
The Social Laboratory: A Psychometric Framework for Multi-Agent LLM Evaluation
by: Reza, Zarreen
Published: (2025)
by: Reza, Zarreen
Published: (2025)
Inducing Personality in LLM-Based Honeypot Agents: Measuring the Effect on Human-Like Agenda Generation
by: Newsham, Lewis, et al.
Published: (2025)
by: Newsham, Lewis, et al.
Published: (2025)
Don't Trust Stubborn Neighbors: A Security Framework for Agentic Networks
by: Abedini, Samira, et al.
Published: (2026)
by: Abedini, Samira, et al.
Published: (2026)
Evaluating Theory of Mind and Internal Beliefs in LLM-Based Multi-Agent Systems
by: Kostka, Adam, et al.
Published: (2026)
by: Kostka, Adam, et al.
Published: (2026)
Trust-based Consensus in Multi-Agent Reinforcement Learning Systems
by: Fung, Ho Long, et al.
Published: (2022)
by: Fung, Ho Long, et al.
Published: (2022)
OPTAGENT: Optimizing Multi-Agent LLM Interactions Through Verbal Reinforcement Learning for Enhanced Reasoning
by: Bi, Zhenyu, et al.
Published: (2025)
by: Bi, Zhenyu, et al.
Published: (2025)
Agent-based Modeling and Simulation of Human Muscle For Development of Human Gait Analyzer Application
by: Saadati, Sina, et al.
Published: (2022)
by: Saadati, Sina, et al.
Published: (2022)
Who is Helping Whom? Analyzing Inter-dependencies to Evaluate Cooperation in Human-AI Teaming
by: Biswas, Upasana, et al.
Published: (2025)
by: Biswas, Upasana, et al.
Published: (2025)
Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems
by: Zhang, Yuzhe, et al.
Published: (2026)
by: Zhang, Yuzhe, et al.
Published: (2026)
A Blockchain-Monitored Agentic AI Architecture for Trusted Perception-Reasoning-Action Pipelines
by: Jan, Salman, et al.
Published: (2025)
by: Jan, Salman, et al.
Published: (2025)
Trust model of privacy-concerned, emotionally-aware agents in a cooperative logistics problem
by: Carbo, J., et al.
Published: (2024)
by: Carbo, J., et al.
Published: (2024)
Trust-Based Social Learning for Communication (TSLEC) Protocol Evolution in Multi-Agent Reinforcement Learning
by: Weinberg, Abraham Itzhak
Published: (2025)
by: Weinberg, Abraham Itzhak
Published: (2025)
Personalized Recommendation Systems using Multimodal, Autonomous, Multi Agent Systems
by: Thakkar, Param, et al.
Published: (2024)
by: Thakkar, Param, et al.
Published: (2024)
Epistemic Context Learning: Building Trust the Right Way in LLM-Based Multi-Agent Systems
by: Zhou, Ruiwen, et al.
Published: (2026)
by: Zhou, Ruiwen, et al.
Published: (2026)
Distributed Online Life-Long Learning (DOL3) for Multi-agent Trust and Reputation Assessment in E-commerce
by: Ramamoorthy, Hariprasauth, et al.
Published: (2024)
by: Ramamoorthy, Hariprasauth, et al.
Published: (2024)
Building and Measuring Trust between Large Language Models
by: Buyl, Maarten, et al.
Published: (2025)
by: Buyl, Maarten, et al.
Published: (2025)
Toward LLM-Agent-Based Modeling of Transportation Systems: A Conceptual Framework
by: Liu, Tianming, et al.
Published: (2024)
by: Liu, Tianming, et al.
Published: (2024)
Trust, Lies, and Long Memories: Emergent Social Dynamics and Reputation in Multi-Round Avalon with LLM Agents
by: Ellawela, Suveen
Published: (2026)
by: Ellawela, Suveen
Published: (2026)
AssistantX: An LLM-Powered Proactive Assistant in Collaborative Human-Populated Environment
by: Sun, Nan, et al.
Published: (2024)
by: Sun, Nan, et al.
Published: (2024)
Modeling Latent Partner Strategies for Adaptive Zero-Shot Human-Agent Collaboration
by: Li, Benjamin, et al.
Published: (2025)
by: Li, Benjamin, et al.
Published: (2025)
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
by: Sinha, Aarush, et al.
Published: (2026)
by: Sinha, Aarush, et al.
Published: (2026)
AMUSE: Audio-Visual Benchmark and Alignment Framework for Agentic Multi-Speaker Understanding
by: Chowdhury, Sanjoy, et al.
Published: (2025)
by: Chowdhury, Sanjoy, et al.
Published: (2025)
Similar Items
-
Can LLMs Reason About Trust?: A Pilot Study
by: Debnath, Anushka, et al.
Published: (2025) -
Social Norm Reasoning in Multimodal Language Models: An Evaluation
by: Chowdhury, Oishik, et al.
Published: (2026) -
Norm Violation Detection in Multi-Agent Systems using Large Language Models: A Pilot Study
by: He, Shawn, et al.
Published: (2024) -
Evolution of Cooperation in LLM-Agent Societies: A Preliminary Study Using Different Punishment Strategies
by: Warnakulasuriya, Kavindu, et al.
Published: (2025) -
Harnessing the power of LLMs for normative reasoning in MASs
by: Savarimuthu, Bastin Tony Roy, et al.
Published: (2024)