TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
Fuente:
arXiv
Saved in:
| Main Authors: | Yaldiz, Duygu Nur, Bakman, Yavuz Faruk, Kang, Sungmin, Öziş, Alperen, Yildiz, Hayrettin Eren, Shah, Mitash Ashish, Huang, Zhiqi, Kumar, Anoop, Samuel, Alfy, Liu, Daben, Karimireddy, Sai Praneeth, Avestimehr, Salman |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
by: Bakman, Yavuz, et al.
Published: (2026)
by: Bakman, Yavuz, et al.
Published: (2026)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
by: Bakman, Yavuz, et al.
Published: (2025)
by: Bakman, Yavuz, et al.
Published: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
by: Bakman, Yavuz, et al.
Published: (2025)
by: Bakman, Yavuz, et al.
Published: (2025)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
by: Ziashahabi, Amir, et al.
Published: (2025)
by: Ziashahabi, Amir, et al.
Published: (2025)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
by: Wang, Nien-Shao, et al.
Published: (2025)
by: Wang, Nien-Shao, et al.
Published: (2025)
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
by: Bakman, Yavuz Faruk, et al.
Published: (2024)
by: Bakman, Yavuz Faruk, et al.
Published: (2024)
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
by: Mushtaq, Erum, et al.
Published: (2024)
by: Mushtaq, Erum, et al.
Published: (2024)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
by: Oguz, Metehan, et al.
Published: (2025)
by: Oguz, Metehan, et al.
Published: (2025)
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
by: Yaldiz, Duygu Nur, et al.
Published: (2024)
by: Yaldiz, Duygu Nur, et al.
Published: (2024)
HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
by: Mushtaq, Erum, et al.
Published: (2025)
by: Mushtaq, Erum, et al.
Published: (2025)
FB-RAG: Improving RAG with Forward and Backward Lookup
by: Chawla, Kushal, et al.
Published: (2025)
by: Chawla, Kushal, et al.
Published: (2025)
A Comparison of Independent and Joint Fine-tuning Strategies for Retrieval-Augmented Generation
by: Lawton, Neal Gregory, et al.
Published: (2025)
by: Lawton, Neal Gregory, et al.
Published: (2025)
Confidence-Based Response Abstinence: Improving LLM Trustworthiness via Activation-Based Uncertainty Estimation
by: Huang, Zhiqi, et al.
Published: (2025)
by: Huang, Zhiqi, et al.
Published: (2025)
Play by the Type Rules: Inferring Constraints for LLM Functions in Declarative Programs
by: Glenn, Parker, et al.
Published: (2025)
by: Glenn, Parker, et al.
Published: (2025)
Readability Reconsidered: A Cross-Dataset Analysis of Reference-Free Metrics
by: Belem, Catarina G, et al.
Published: (2025)
by: Belem, Catarina G, et al.
Published: (2025)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
by: Oğuz, Metehan, et al.
Published: (2024)
by: Oğuz, Metehan, et al.
Published: (2024)
LLM Optimization Unlocks Real-Time Pairwise Reranking
by: Wu, Jingyu, et al.
Published: (2025)
by: Wu, Jingyu, et al.
Published: (2025)
GEM: A Scale-Aware and Distribution-Sensitive Sparse Fine-Tuning Framework for Effective Downstream Adaptation
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
EPSVec: Efficient and Private Synthetic Data Generation via Dataset Vectors
by: Banayeeanzade, Amin, et al.
Published: (2026)
by: Banayeeanzade, Amin, et al.
Published: (2026)
Godel Implication on Finite Chains: Truth Tables and Catalan-Bracketing Enumerations
by: Yildiz, Volkan
Published: (2026)
by: Yildiz, Volkan
Published: (2026)
Improving Consistency in Retrieval-Augmented Systems with Group Similarity Rewards
by: Hamman, Faisal, et al.
Published: (2025)
by: Hamman, Faisal, et al.
Published: (2025)
Alignment-Weighted DPO: A principled reasoning approach to improve safety alignment
by: Hu, Mengxuan, et al.
Published: (2026)
by: Hu, Mengxuan, et al.
Published: (2026)
Truth-Aware Decoding: A Program-Logic Approach to Factual Language Generation
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
Optimization with Access to Auxiliary Information
by: Chayti, El Mahdi, et al.
Published: (2022)
by: Chayti, El Mahdi, et al.
Published: (2022)
Love the Truth, the Whole Truth, and the Truth about Everything. An Interview with Josef Seifert
by: Rodrigo Guerra López
Published: (2014)
by: Rodrigo Guerra López
Published: (2014)
On the Limits of Momentum in Decentralized and Federated Optimization
by: Zaccone, Riccardo, et al.
Published: (2025)
by: Zaccone, Riccardo, et al.
Published: (2025)
Reference-Specific Unlearning Metrics Can Hide the Truth: A Reality Check
by: Cho, Sungjun, et al.
Published: (2025)
by: Cho, Sungjun, et al.
Published: (2025)
Truth
by: Cubitt, Sean
Published: (2024)
by: Cubitt, Sean
Published: (2024)
The Truth, the Whole Truth, and Nothing but the Truth: Automatic Visualization Evaluation from Reconstruction Quality
by: Bujack, Roxana, et al.
Published: (2026)
by: Bujack, Roxana, et al.
Published: (2026)
Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models
by: Jain, Neel, et al.
Published: (2024)
by: Jain, Neel, et al.
Published: (2024)
Truth, Justice, and Secrecy: Cake Cutting Under Privacy Constraints
by: Salman, Yaron, et al.
Published: (2025)
by: Salman, Yaron, et al.
Published: (2025)
Nursing Care Behaviours and Challenges Faced During Truth‐Telling: A Phenomenological Study
by: Ayşegül Yıldız İçigen, et al.
Published: (2026)
by: Ayşegül Yıldız İçigen, et al.
Published: (2026)
Truth Knows No Language: Evaluating Truthfulness Beyond English
by: Figueras, Blanca Calvo, et al.
Published: (2025)
by: Figueras, Blanca Calvo, et al.
Published: (2025)
TruthStance: An Annotated Dataset of Conversations on Truth Social
by: Ameen, Fathima, et al.
Published: (2026)
by: Ameen, Fathima, et al.
Published: (2026)
Sampling More, Getting Less: Calibration is the Diversity Bottleneck in LLMs
by: Banayeeanzade, Amin, et al.
Published: (2026)
by: Banayeeanzade, Amin, et al.
Published: (2026)
Collaborative Heterogeneous Causal Inference Beyond Meta-analysis
by: Guo, Tianyu, et al.
Published: (2024)
by: Guo, Tianyu, et al.
Published: (2024)
Defection-Free Collaboration between Competitors in a Learning System
by: Werner, Mariel, et al.
Published: (2024)
by: Werner, Mariel, et al.
Published: (2024)
Do Data Valuations Make Good Data Prices?
by: Fan, Dongyang, et al.
Published: (2025)
by: Fan, Dongyang, et al.
Published: (2025)
The Shape of Truth
by: FERNANDEZ, DAHLIA D.
Published: (2025)
by: FERNANDEZ, DAHLIA D.
Published: (2025)
Similar Items
-
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
by: Bakman, Yavuz, et al.
Published: (2026) -
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
by: Bakman, Yavuz, et al.
Published: (2025) -
Reconsidering LLM Uncertainty Estimation Methods in the Wild
by: Bakman, Yavuz, et al.
Published: (2025) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025) -
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
by: Ziashahabi, Amir, et al.
Published: (2025)