Illusions of Confidence? Diagnosing LLM Truthfulness via Neighborhood Consistency
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Haoming, Zhao, Ningyuan, Yao, Yunzhi, Xu, Weihong, Wang, Hongru, Deng, Xinle, Deng, Shumin, Pan, Jeff Z., Chen, Huajun, Zhang, Ningyu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Model Kinship for Merging Large Language Models
by: Hu, Yedi, et al.
Published: (2024)
by: Hu, Yedi, et al.
Published: (2024)
LightMem: Lightweight and Efficient Memory-Augmented Generation
by: Fang, Jizhan, et al.
Published: (2025)
by: Fang, Jizhan, et al.
Published: (2025)
KnowAgent: Knowledge-Augmented Planning for LLM-Based Agents
by: Zhu, Yuqi, et al.
Published: (2024)
by: Zhu, Yuqi, et al.
Published: (2024)
EasyEdit2: An Easy-to-use Steering Framework for Editing Large Language Models
by: Xu, Ziwen, et al.
Published: (2025)
by: Xu, Ziwen, et al.
Published: (2025)
Exploring Collaboration Mechanisms for LLM Agents: A Social Psychology View
by: Zhang, Jintian, et al.
Published: (2023)
by: Zhang, Jintian, et al.
Published: (2023)
ReLearn: Unlearning via Learning for Large Language Models
by: Xu, Haoming, et al.
Published: (2025)
by: Xu, Haoming, et al.
Published: (2025)
How Do LLMs Acquire New Knowledge? A Knowledge Circuits Perspective on Continual Pre-Training
by: Ou, Yixin, et al.
Published: (2025)
by: Ou, Yixin, et al.
Published: (2025)
AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning
by: Qiao, Shuofei, et al.
Published: (2024)
by: Qiao, Shuofei, et al.
Published: (2024)
Benchmarking Agentic Workflow Generation
by: Qiao, Shuofei, et al.
Published: (2024)
by: Qiao, Shuofei, et al.
Published: (2024)
Stability of AI Governance Systems: A Coupled Dynamics Model of Public Trust and Social Disruptions
by: Lai, Jiaqi, et al.
Published: (2026)
by: Lai, Jiaqi, et al.
Published: (2026)
Detoxifying Large Language Models via Knowledge Editing
by: Wang, Mengru, et al.
Published: (2024)
by: Wang, Mengru, et al.
Published: (2024)
AutoMind: Adaptive Knowledgeable Agent for Automated Data Science
by: Ou, Yixin, et al.
Published: (2025)
by: Ou, Yixin, et al.
Published: (2025)
How to Steer Your Multi-Agent System: Human-LLM Collaborative Planning
by: He, Zeyu, et al.
Published: (2026)
by: He, Zeyu, et al.
Published: (2026)
How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities
by: Xu, Ziwen, et al.
Published: (2026)
by: Xu, Ziwen, et al.
Published: (2026)
DarwinTOD: LLM-driven Lifelong Self-evolution for Task-oriented Dialog Systems
by: Zhang, Shuyu, et al.
Published: (2026)
by: Zhang, Shuyu, et al.
Published: (2026)
LLM-Powered Virtual Patient Agents for Interactive Clinical Skills Training with Automated Feedback
by: Voigt, Henrik, et al.
Published: (2025)
by: Voigt, Henrik, et al.
Published: (2025)
Decoupled Intelligence: A Multi-Agent LLM Framework for Controllable Traffic Scenario Generation in SUMO
by: Li, Shuyang, et al.
Published: (2026)
by: Li, Shuyang, et al.
Published: (2026)
FACET: Teacher-Centred LLM-Based Multi-Agent Systems-Towards Personalized Educational Worksheets
by: Gonnermann-Müller, Jana, et al.
Published: (2025)
by: Gonnermann-Müller, Jana, et al.
Published: (2025)
A Simulation-Based Method for Testing Collaborative Learning Scaffolds Using LLM-Based Multi-Agent Systems
by: Wua, Han, et al.
Published: (2026)
by: Wua, Han, et al.
Published: (2026)
Multi-Stakeholder Alignment in LLM-Powered Collaborative AI Systems: A Multi-Agent Framework for Intelligent Tutoring
by: Uchoa, Alexandre P, et al.
Published: (2025)
by: Uchoa, Alexandre P, et al.
Published: (2025)
A LLM-Driven Multi-Agent Systems for Professional Development of Mathematics Teachers
by: Yang, Kaiqi, et al.
Published: (2025)
by: Yang, Kaiqi, et al.
Published: (2025)
NLI4VolVis: Natural Language Interaction for Volume Visualization via LLM Multi-Agents and Editable 3D Gaussian Splatting
by: Ai, Kuangshi, et al.
Published: (2025)
by: Ai, Kuangshi, et al.
Published: (2025)
Can A Society of Generative Agents Simulate Human Behavior and Inform Public Health Policy? A Case Study on Vaccine Hesitancy
by: Hou, Abe Bohan, et al.
Published: (2025)
by: Hou, Abe Bohan, et al.
Published: (2025)
StructMem: Structured Memory for Long-Horizon Behavior in LLMs
by: Xu, Buqiang, et al.
Published: (2026)
by: Xu, Buqiang, et al.
Published: (2026)
Elicitron: An LLM Agent-Based Simulation Framework for Design Requirements Elicitation
by: Ataei, Mohammadmehdi, et al.
Published: (2024)
by: Ataei, Mohammadmehdi, et al.
Published: (2024)
DialogGuard: Multi-Agent Psychosocial Safety Evaluation of Sensitive LLM Responses
by: Luo, Han, et al.
Published: (2025)
by: Luo, Han, et al.
Published: (2025)
FlowEval: Reference-based Evaluation of Generated User Interfaces
by: Wu, Jason, et al.
Published: (2026)
by: Wu, Jason, et al.
Published: (2026)
Auto-Slides: An Interactive Multi-Agent System for Creating and Customizing Research Presentations
by: Yang, Yuheng, et al.
Published: (2025)
by: Yang, Yuheng, et al.
Published: (2025)
Conversational Self-Play for Discovering and Understanding Psychotherapy Approaches
by: Kampman, Onno P, et al.
Published: (2025)
by: Kampman, Onno P, et al.
Published: (2025)
Too Many Specialists: Emergent Inefficiencies and Bottlenecks for Multi-agent Ad-hoc Collaboration
by: Panny, Benjamin, et al.
Published: (2026)
by: Panny, Benjamin, et al.
Published: (2026)
Data assimilation approach for addressing imperfections in people flow measurement techniques using particle filter
by: Murata, Ryo, et al.
Published: (2024)
by: Murata, Ryo, et al.
Published: (2024)
EmBARDiment: an Embodied AI Agent for Productivity in XR
by: Bovo, Riccardo, et al.
Published: (2024)
by: Bovo, Riccardo, et al.
Published: (2024)
Inject, Fork, Compare: Defining an Interaction Vocabulary for Multi-Agent Simulation Platforms
by: Lee, HwiJoon, et al.
Published: (2025)
by: Lee, HwiJoon, et al.
Published: (2025)
CandorMD: An AI-Assisted Audio Simulation and Feedback System for Training Clinicians for Medical Error Disclosure
by: Lin, Inna Wanyin, et al.
Published: (2026)
by: Lin, Inna Wanyin, et al.
Published: (2026)
Cloud and IoT based Smart Agent-driven Simulation of Human Gait for Detecting Muscles Disorder
by: Saadati, Sina, et al.
Published: (2024)
by: Saadati, Sina, et al.
Published: (2024)
Context-Mediated Domain Adaptation in Multi-Agent Sensemaking Systems
by: Wolter, Anton, et al.
Published: (2026)
by: Wolter, Anton, et al.
Published: (2026)
MAxPrototyper: A Multi-Agent Generation System for Interactive User Interface Prototyping
by: Yuan, Mingyue, et al.
Published: (2024)
by: Yuan, Mingyue, et al.
Published: (2024)
AIPOM: Agent-aware Interactive Planning for Multi-Agent Systems
by: Kim, Hannah, et al.
Published: (2025)
by: Kim, Hannah, et al.
Published: (2025)
A2H: Agent-to-Human Protocol for AI Agent
by: Liang, Zhiyuan, et al.
Published: (2025)
by: Liang, Zhiyuan, et al.
Published: (2025)
CoCre-Sam (Kokkuri-san): Modeling Ouija Board as Collective Langevin Dynamics Sampling from Fused Language Models
by: Taniguchi, Tadahiro, et al.
Published: (2025)
by: Taniguchi, Tadahiro, et al.
Published: (2025)
Similar Items
-
Exploring Model Kinship for Merging Large Language Models
by: Hu, Yedi, et al.
Published: (2024) -
LightMem: Lightweight and Efficient Memory-Augmented Generation
by: Fang, Jizhan, et al.
Published: (2025) -
KnowAgent: Knowledge-Augmented Planning for LLM-Based Agents
by: Zhu, Yuqi, et al.
Published: (2024) -
EasyEdit2: An Easy-to-use Steering Framework for Editing Large Language Models
by: Xu, Ziwen, et al.
Published: (2025) -
Exploring Collaboration Mechanisms for LLM Agents: A Social Psychology View
by: Zhang, Jintian, et al.
Published: (2023)