Saved in:
| Main Authors: | Bao, Yuntai, Zhang, Xuhong, Du, Tianyu, Zhao, Xinkui, Feng, Zhengwen, Peng, Hao, Yin, Jianwei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.00823 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scalable Multi-Stage Influence Function for Large Language Models via Eigenvalue-Corrected Kronecker-Factored Parameterization
by: Bao, Yuntai, et al.
Published: (2025)
by: Bao, Yuntai, et al.
Published: (2025)
HijackRAG: Hijacking Attacks against Retrieval-Augmented Large Language Models
by: Zhang, Yucheng, et al.
Published: (2024)
by: Zhang, Yucheng, et al.
Published: (2024)
The Geometries of Truth Are Orthogonal Across Tasks
by: Azizian, Waiss, et al.
Published: (2025)
by: Azizian, Waiss, et al.
Published: (2025)
Faithful Bi-Directional Model Steering via Distribution Matching and Distributed Interchange Interventions
by: Bao, Yuntai, et al.
Published: (2026)
by: Bao, Yuntai, et al.
Published: (2026)
Testing the Limits of Truth Directions in LLMs
by: Poulis, Angelos, et al.
Published: (2026)
by: Poulis, Angelos, et al.
Published: (2026)
Pruning Weights but Not Truth: Safeguarding Truthfulness While Pruning LLMs
by: Fu, Yao, et al.
Published: (2025)
by: Fu, Yao, et al.
Published: (2025)
How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs
by: Adarsh, Shivam, et al.
Published: (2026)
by: Adarsh, Shivam, et al.
Published: (2026)
Logical Form and Truth-Conditions
by: Andrea IACONA
Published: (2013)
by: Andrea IACONA
Published: (2013)
Towards Steering without Sacrifice: Principled Training of Steering Vectors for Prompt-only Interventions
by: Bao, Yuntai, et al.
Published: (2026)
by: Bao, Yuntai, et al.
Published: (2026)
Truth Claims Across Media
Published: (2024)
Published: (2024)
Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks
by: Karia, Rushang, et al.
Published: (2024)
by: Karia, Rushang, et al.
Published: (2024)
ERA-CoT: Improving Chain-of-Thought through Entity Relationship Analysis
by: Liu, Yanming, et al.
Published: (2024)
by: Liu, Yanming, et al.
Published: (2024)
TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
by: Wei, Zhepei, et al.
Published: (2025)
by: Wei, Zhepei, et al.
Published: (2025)
CoreGuard: Safeguarding Foundational Capabilities of LLMs Against Model Stealing in Edge Deployment
by: Li, Qinfeng, et al.
Published: (2024)
by: Li, Qinfeng, et al.
Published: (2024)
Safeguarding the Truth of High-Value Price Oracle Task: A Dynamically Adjusted Truth Discovery Method
by: Xian, Youquan, et al.
Published: (2024)
by: Xian, Youquan, et al.
Published: (2024)
The Muses of Truth and Transformation
by: Chinen, Allan B.
Published: (2024)
by: Chinen, Allan B.
Published: (2024)
Tool-Planner: Task Planning with Clusters across Multiple Tools
by: Liu, Yanming, et al.
Published: (2024)
by: Liu, Yanming, et al.
Published: (2024)
TransLinkGuard: Safeguarding Transformer Models Against Model Stealing in Edge Deployment
by: Li, Qinfeng, et al.
Published: (2024)
by: Li, Qinfeng, et al.
Published: (2024)
Walking the Schrödinger Bridge: A Direct Trajectory for Text-to-3D Generation
by: Li, Ziying, et al.
Published: (2025)
by: Li, Ziying, et al.
Published: (2025)
Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages
by: Bayes, Edward, et al.
Published: (2024)
by: Bayes, Edward, et al.
Published: (2024)
Compressing LLMs: The Truth is Rarely Pure and Never Simple
by: Jaiswal, Ajay, et al.
Published: (2023)
by: Jaiswal, Ajay, et al.
Published: (2023)
RA-ISF: Learning to Answer and Understand from Retrieval Augmentation via Iterative Self-Feedback
by: Liu, Yanming, et al.
Published: (2024)
by: Liu, Yanming, et al.
Published: (2024)
CollabEdit: Towards Non-destructive Collaborative Knowledge Editing
by: Zheng, Jiamu, et al.
Published: (2024)
by: Zheng, Jiamu, et al.
Published: (2024)
MIDAS: Modeling Ground-Truth Distributions with Dark Knowledge for Domain Generalized Stereo Matching
by: Xu, Peng, et al.
Published: (2025)
by: Xu, Peng, et al.
Published: (2025)
EdgeJury: Cross-Reviewed Small-Model Ensembles for Truthful Question Answering on Serverless Edge Inference
by: Kumar, Aayush
Published: (2025)
by: Kumar, Aayush
Published: (2025)
TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space
by: Zhang, Shaolei, et al.
Published: (2024)
by: Zhang, Shaolei, et al.
Published: (2024)
On the Universal Truthfulness Hyperplane Inside LLMs
by: Liu, Junteng, et al.
Published: (2024)
by: Liu, Junteng, et al.
Published: (2024)
Love the Truth, the Whole Truth, and the Truth about Everything. An Interview with Josef Seifert
by: Rodrigo Guerra López
Published: (2014)
by: Rodrigo Guerra López
Published: (2014)
TruthFlow: Truthful LLM Generation via Representation Flow Correction
by: Wang, Hanyu, et al.
Published: (2025)
by: Wang, Hanyu, et al.
Published: (2025)
Truth
by: Cubitt, Sean
Published: (2024)
by: Cubitt, Sean
Published: (2024)
Queries With Exact Truth Values in Paraconsistent Description Logics
by: Bienvenu, Meghyn, et al.
Published: (2024)
by: Bienvenu, Meghyn, et al.
Published: (2024)
Neural Quantum States in Variational Monte Carlo Method: A Brief Summary
by: Song, Yuntai
Published: (2024)
by: Song, Yuntai
Published: (2024)
Accurate Table Question Answering with Accessible LLMs
by: Jiang, Yangfan, et al.
Published: (2026)
by: Jiang, Yangfan, et al.
Published: (2026)
The Truth, the Whole Truth, and Nothing but the Truth: Automatic Visualization Evaluation from Reconstruction Quality
by: Bujack, Roxana, et al.
Published: (2026)
by: Bujack, Roxana, et al.
Published: (2026)
Ground Truth Generation for Multilingual Historical NLP using LLMs
by: Gladstone, Clovis, et al.
Published: (2025)
by: Gladstone, Clovis, et al.
Published: (2025)
SecCoder: Towards Generalizable and Robust Secure Code Generation
by: Zhang, Boyu, et al.
Published: (2024)
by: Zhang, Boyu, et al.
Published: (2024)
Truth-Aware Decoding: A Program-Logic Approach to Factual Language Generation
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
Truthful Aggregation of LLMs with an Application to Online Advertising
by: Soumalias, Ermis, et al.
Published: (2024)
by: Soumalias, Ermis, et al.
Published: (2024)
Truth is Universal: Robust Detection of Lies in LLMs
by: Bürger, Lennart, et al.
Published: (2024)
by: Bürger, Lennart, et al.
Published: (2024)
Truth Knows No Language: Evaluating Truthfulness Beyond English
by: Figueras, Blanca Calvo, et al.
Published: (2025)
by: Figueras, Blanca Calvo, et al.
Published: (2025)
Similar Items
-
Scalable Multi-Stage Influence Function for Large Language Models via Eigenvalue-Corrected Kronecker-Factored Parameterization
by: Bao, Yuntai, et al.
Published: (2025) -
HijackRAG: Hijacking Attacks against Retrieval-Augmented Large Language Models
by: Zhang, Yucheng, et al.
Published: (2024) -
The Geometries of Truth Are Orthogonal Across Tasks
by: Azizian, Waiss, et al.
Published: (2025) -
Faithful Bi-Directional Model Steering via Distribution Matching and Distributed Interchange Interventions
by: Bao, Yuntai, et al.
Published: (2026) -
Testing the Limits of Truth Directions in LLMs
by: Poulis, Angelos, et al.
Published: (2026)