TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yaldiz, Duygu Nur, Bakman, Yavuz Faruk, Kang, Sungmin, Öziş, Alperen, Yildiz, Hayrettin Eren, Shah, Mitash Ashish, Huang, Zhiqi, Kumar, Anoop, Samuel, Alfy, Liu, Daben, Karimireddy, Sai Praneeth, Avestimehr, Salman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
von: Bakman, Yavuz, et al.
Veröffentlicht: (2026)
von: Bakman, Yavuz, et al.
Veröffentlicht: (2026)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
von: Kang, Sungmin, et al.
Veröffentlicht: (2025)
von: Kang, Sungmin, et al.
Veröffentlicht: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
von: Ziashahabi, Amir, et al.
Veröffentlicht: (2025)
von: Ziashahabi, Amir, et al.
Veröffentlicht: (2025)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
von: Wang, Nien-Shao, et al.
Veröffentlicht: (2025)
von: Wang, Nien-Shao, et al.
Veröffentlicht: (2025)
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
von: Bakman, Yavuz Faruk, et al.
Veröffentlicht: (2024)
von: Bakman, Yavuz Faruk, et al.
Veröffentlicht: (2024)
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
von: Mushtaq, Erum, et al.
Veröffentlicht: (2024)
von: Mushtaq, Erum, et al.
Veröffentlicht: (2024)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
von: Oguz, Metehan, et al.
Veröffentlicht: (2025)
von: Oguz, Metehan, et al.
Veröffentlicht: (2025)
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2024)
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2024)
HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
von: Mushtaq, Erum, et al.
Veröffentlicht: (2025)
von: Mushtaq, Erum, et al.
Veröffentlicht: (2025)
FB-RAG: Improving RAG with Forward and Backward Lookup
von: Chawla, Kushal, et al.
Veröffentlicht: (2025)
von: Chawla, Kushal, et al.
Veröffentlicht: (2025)
A Comparison of Independent and Joint Fine-tuning Strategies for Retrieval-Augmented Generation
von: Lawton, Neal Gregory, et al.
Veröffentlicht: (2025)
von: Lawton, Neal Gregory, et al.
Veröffentlicht: (2025)
Confidence-Based Response Abstinence: Improving LLM Trustworthiness via Activation-Based Uncertainty Estimation
von: Huang, Zhiqi, et al.
Veröffentlicht: (2025)
von: Huang, Zhiqi, et al.
Veröffentlicht: (2025)
Play by the Type Rules: Inferring Constraints for LLM Functions in Declarative Programs
von: Glenn, Parker, et al.
Veröffentlicht: (2025)
von: Glenn, Parker, et al.
Veröffentlicht: (2025)
Readability Reconsidered: A Cross-Dataset Analysis of Reference-Free Metrics
von: Belem, Catarina G, et al.
Veröffentlicht: (2025)
von: Belem, Catarina G, et al.
Veröffentlicht: (2025)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
von: Oğuz, Metehan, et al.
Veröffentlicht: (2024)
von: Oğuz, Metehan, et al.
Veröffentlicht: (2024)
LLM Optimization Unlocks Real-Time Pairwise Reranking
von: Wu, Jingyu, et al.
Veröffentlicht: (2025)
von: Wu, Jingyu, et al.
Veröffentlicht: (2025)
GEM: A Scale-Aware and Distribution-Sensitive Sparse Fine-Tuning Framework for Effective Downstream Adaptation
von: Kang, Sungmin, et al.
Veröffentlicht: (2025)
von: Kang, Sungmin, et al.
Veröffentlicht: (2025)
EPSVec: Efficient and Private Synthetic Data Generation via Dataset Vectors
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
Godel Implication on Finite Chains: Truth Tables and Catalan-Bracketing Enumerations
von: Yildiz, Volkan
Veröffentlicht: (2026)
von: Yildiz, Volkan
Veröffentlicht: (2026)
Improving Consistency in Retrieval-Augmented Systems with Group Similarity Rewards
von: Hamman, Faisal, et al.
Veröffentlicht: (2025)
von: Hamman, Faisal, et al.
Veröffentlicht: (2025)
Alignment-Weighted DPO: A principled reasoning approach to improve safety alignment
von: Hu, Mengxuan, et al.
Veröffentlicht: (2026)
von: Hu, Mengxuan, et al.
Veröffentlicht: (2026)
Truth-Aware Decoding: A Program-Logic Approach to Factual Language Generation
von: Alpay, Faruk, et al.
Veröffentlicht: (2025)
von: Alpay, Faruk, et al.
Veröffentlicht: (2025)
Optimization with Access to Auxiliary Information
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2022)
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2022)
Love the Truth, the Whole Truth, and the Truth about Everything. An Interview with Josef Seifert
von: Rodrigo Guerra López
Veröffentlicht: (2014)
von: Rodrigo Guerra López
Veröffentlicht: (2014)
On the Limits of Momentum in Decentralized and Federated Optimization
von: Zaccone, Riccardo, et al.
Veröffentlicht: (2025)
von: Zaccone, Riccardo, et al.
Veröffentlicht: (2025)
Reference-Specific Unlearning Metrics Can Hide the Truth: A Reality Check
von: Cho, Sungjun, et al.
Veröffentlicht: (2025)
von: Cho, Sungjun, et al.
Veröffentlicht: (2025)
Truth
von: Cubitt, Sean
Veröffentlicht: (2024)
von: Cubitt, Sean
Veröffentlicht: (2024)
The Truth, the Whole Truth, and Nothing but the Truth: Automatic Visualization Evaluation from Reconstruction Quality
von: Bujack, Roxana, et al.
Veröffentlicht: (2026)
von: Bujack, Roxana, et al.
Veröffentlicht: (2026)
Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models
von: Jain, Neel, et al.
Veröffentlicht: (2024)
von: Jain, Neel, et al.
Veröffentlicht: (2024)
Truth, Justice, and Secrecy: Cake Cutting Under Privacy Constraints
von: Salman, Yaron, et al.
Veröffentlicht: (2025)
von: Salman, Yaron, et al.
Veröffentlicht: (2025)
Nursing Care Behaviours and Challenges Faced During Truth‐Telling: A Phenomenological Study
von: Ayşegül Yıldız İçigen, et al.
Veröffentlicht: (2026)
von: Ayşegül Yıldız İçigen, et al.
Veröffentlicht: (2026)
Truth Knows No Language: Evaluating Truthfulness Beyond English
von: Figueras, Blanca Calvo, et al.
Veröffentlicht: (2025)
von: Figueras, Blanca Calvo, et al.
Veröffentlicht: (2025)
TruthStance: An Annotated Dataset of Conversations on Truth Social
von: Ameen, Fathima, et al.
Veröffentlicht: (2026)
von: Ameen, Fathima, et al.
Veröffentlicht: (2026)
Sampling More, Getting Less: Calibration is the Diversity Bottleneck in LLMs
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
Collaborative Heterogeneous Causal Inference Beyond Meta-analysis
von: Guo, Tianyu, et al.
Veröffentlicht: (2024)
von: Guo, Tianyu, et al.
Veröffentlicht: (2024)
Defection-Free Collaboration between Competitors in a Learning System
von: Werner, Mariel, et al.
Veröffentlicht: (2024)
von: Werner, Mariel, et al.
Veröffentlicht: (2024)
Do Data Valuations Make Good Data Prices?
von: Fan, Dongyang, et al.
Veröffentlicht: (2025)
von: Fan, Dongyang, et al.
Veröffentlicht: (2025)
The Shape of Truth
von: FERNANDEZ, DAHLIA D.
Veröffentlicht: (2025)
von: FERNANDEZ, DAHLIA D.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
von: Bakman, Yavuz, et al.
Veröffentlicht: (2026) -
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025) -
Reconsidering LLM Uncertainty Estimation Methods in the Wild
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
von: Kang, Sungmin, et al.
Veröffentlicht: (2025) -
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
von: Ziashahabi, Amir, et al.
Veröffentlicht: (2025)