Prompt-Response Semantic Divergence Metrics for Faithfulness Hallucination and Misalignment Detection in Large Language Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Halperin, Igor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantic Faithfulness and Entropy Production Measures to Tame Your LLM Demons and Manage Hallucinations
von: Halperin, Igor
Veröffentlicht: (2025)
von: Halperin, Igor
Veröffentlicht: (2025)
Topic Identification in LLM Input-Output Pairs through the Lens of Information Bottleneck
von: Halperin, Igor
Veröffentlicht: (2025)
von: Halperin, Igor
Veröffentlicht: (2025)
Exploring the Synergy of Quantitative Factors and Newsflow Representations from Large Language Models for Stock Return Prediction
von: Guo, Tian, et al.
Veröffentlicht: (2025)
von: Guo, Tian, et al.
Veröffentlicht: (2025)
A Comparative Study of DSPy Teleprompter Algorithms for Aligning Large Language Models Evaluation Metrics to Human Evaluation
von: Sarmah, Bhaskarjit, et al.
Veröffentlicht: (2024)
von: Sarmah, Bhaskarjit, et al.
Veröffentlicht: (2024)
Temporal Relational Reasoning of Large Language Models for Detecting Stock Portfolio Crashes
von: Koa, Kelvin J. L., et al.
Veröffentlicht: (2024)
von: Koa, Kelvin J. L., et al.
Veröffentlicht: (2024)
Multi-Dimensional Behavioral Evaluation of Agentic Stock Prediction Systems Using Large Language Model Judges with Closed-Loop Reinforcement Learning Feedback
von: Ridhawi, Mohammad Al, et al.
Veröffentlicht: (2026)
von: Ridhawi, Mohammad Al, et al.
Veröffentlicht: (2026)
Distributional Semantics Tracing: A Framework for Explaining Hallucinations in Large Language Models
von: Bhatia, Gagan, et al.
Veröffentlicht: (2025)
von: Bhatia, Gagan, et al.
Veröffentlicht: (2025)
Tx-LLM: A Large Language Model for Therapeutics
von: Chaves, Juan Manuel Zambrano, et al.
Veröffentlicht: (2024)
von: Chaves, Juan Manuel Zambrano, et al.
Veröffentlicht: (2024)
Assessing Consistency and Reproducibility in the Outputs of Large Language Models: Evidence Across Diverse Finance and Accounting Tasks
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2025)
von: Wang, Julian Junyan, et al.
Veröffentlicht: (2025)
PLLaMa: An Open-source Large Language Model for Plant Science
von: Yang, Xianjun, et al.
Veröffentlicht: (2024)
von: Yang, Xianjun, et al.
Veröffentlicht: (2024)
CPR: Mitigating Large Language Model Hallucinations with Curative Prompt Refinement
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
Multi-stage Prompt Refinement for Mitigating Hallucinations in Large Language Models
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
FaithLM: Towards Faithful Explanations for Large Language Models
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
Diabetica: Adapting Large Language Model to Enhance Multiple Medical Tasks in Diabetes Care and Management
von: Wei, Lai, et al.
Veröffentlicht: (2024)
von: Wei, Lai, et al.
Veröffentlicht: (2024)
CAISSON: Concept-Augmented Inference Suite of Self-Organizing Neural Networks
von: Halperin, Igor
Veröffentlicht: (2024)
von: Halperin, Igor
Veröffentlicht: (2024)
Exploring Sentiment Dynamics and Predictive Behaviors in Cryptocurrency Discussions by Few-Shot Learning with Large Language Models
von: Tash, Moein Shahiki, et al.
Veröffentlicht: (2024)
von: Tash, Moein Shahiki, et al.
Veröffentlicht: (2024)
Protein Large Language Models: A Comprehensive Survey
von: Xiao, Yijia, et al.
Veröffentlicht: (2025)
von: Xiao, Yijia, et al.
Veröffentlicht: (2025)
AQUA: A Large Language Model for Aquaculture & Fisheries
von: Narisetty, Praneeth, et al.
Veröffentlicht: (2025)
von: Narisetty, Praneeth, et al.
Veröffentlicht: (2025)
OceanGPT: A Large Language Model for Ocean Science Tasks
von: Bi, Zhen, et al.
Veröffentlicht: (2023)
von: Bi, Zhen, et al.
Veröffentlicht: (2023)
PMPO: Probabilistic Metric Prompt Optimization for Small and Large Language Models
von: Zhao, Chenzhuo, et al.
Veröffentlicht: (2025)
von: Zhao, Chenzhuo, et al.
Veröffentlicht: (2025)
(Im)possibility of Automated Hallucination Detection in Large Language Models
von: Karbasi, Amin, et al.
Veröffentlicht: (2025)
von: Karbasi, Amin, et al.
Veröffentlicht: (2025)
Scalable Token-Level Hallucination Detection in Large Language Models
von: Min, Rui, et al.
Veröffentlicht: (2026)
von: Min, Rui, et al.
Veröffentlicht: (2026)
Evaluating Evaluation Metrics -- The Mirage of Hallucination Detection
von: Kulkarni, Atharva, et al.
Veröffentlicht: (2025)
von: Kulkarni, Atharva, et al.
Veröffentlicht: (2025)
ORLM: A Customizable Framework in Training Large Models for Automated Optimization Modeling
von: Huang, Chenyu, et al.
Veröffentlicht: (2024)
von: Huang, Chenyu, et al.
Veröffentlicht: (2024)
Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance
von: Benhenda, Mostapha
Veröffentlicht: (2026)
von: Benhenda, Mostapha
Veröffentlicht: (2026)
Emoji Driven Crypto Assets Market Reactions
von: Zuo, Xiaorui, et al.
Veröffentlicht: (2024)
von: Zuo, Xiaorui, et al.
Veröffentlicht: (2024)
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach
von: Ni, Haowei, et al.
Veröffentlicht: (2024)
von: Ni, Haowei, et al.
Veröffentlicht: (2024)
Principled Detection of Hallucinations in Large Language Models via Multiple Testing
von: Li, Jiawei, et al.
Veröffentlicht: (2025)
von: Li, Jiawei, et al.
Veröffentlicht: (2025)
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
Manifold-based Sampling for In-Context Hallucination Detection in Large Language Models
von: Vamshi, Bodla Krishna, et al.
Veröffentlicht: (2026)
von: Vamshi, Bodla Krishna, et al.
Veröffentlicht: (2026)
Modeling and Detecting Company Risks from News: A Case Study in Bloomberg News
von: Pei, Jiaxin, et al.
Veröffentlicht: (2025)
von: Pei, Jiaxin, et al.
Veröffentlicht: (2025)
Transformer Circuit Faithfulness Metrics are not Robust
von: Miller, Joseph, et al.
Veröffentlicht: (2024)
von: Miller, Joseph, et al.
Veröffentlicht: (2024)
FiMI: A Domain-Specific Language Model for Indian Finance Ecosystem
von: Kathar, Aboli, et al.
Veröffentlicht: (2026)
von: Kathar, Aboli, et al.
Veröffentlicht: (2026)
Detecting AI Hallucinations in Finance: An Information-Theoretic Method Cuts Hallucination Rate by 92%
von: Singha, Mainak
Veröffentlicht: (2025)
von: Singha, Mainak
Veröffentlicht: (2025)
Regress, Don't Guess -- A Regression-like Loss on Number Tokens for Language Models
von: Zausinger, Jonas, et al.
Veröffentlicht: (2024)
von: Zausinger, Jonas, et al.
Veröffentlicht: (2024)
FinBloom: Knowledge Grounding Large Language Model with Real-time Financial Data
von: Sinha, Ankur, et al.
Veröffentlicht: (2025)
von: Sinha, Ankur, et al.
Veröffentlicht: (2025)
FLAME: Financial Large-Language Model Assessment and Metrics Evaluation
von: Guo, Jiayu, et al.
Veröffentlicht: (2025)
von: Guo, Jiayu, et al.
Veröffentlicht: (2025)
Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations
von: Matton, Katie, et al.
Veröffentlicht: (2025)
von: Matton, Katie, et al.
Veröffentlicht: (2025)
Mitigating Hallucinated Translations in Large Language Models with Hallucination-focused Preference Optimization
von: Tang, Zilu, et al.
Veröffentlicht: (2025)
von: Tang, Zilu, et al.
Veröffentlicht: (2025)
Incorporating Attribution Importance for Improving Faithfulness Metrics
von: Zhao, Zhixue, et al.
Veröffentlicht: (2023)
von: Zhao, Zhixue, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Semantic Faithfulness and Entropy Production Measures to Tame Your LLM Demons and Manage Hallucinations
von: Halperin, Igor
Veröffentlicht: (2025) -
Topic Identification in LLM Input-Output Pairs through the Lens of Information Bottleneck
von: Halperin, Igor
Veröffentlicht: (2025) -
Exploring the Synergy of Quantitative Factors and Newsflow Representations from Large Language Models for Stock Return Prediction
von: Guo, Tian, et al.
Veröffentlicht: (2025) -
A Comparative Study of DSPy Teleprompter Algorithms for Aligning Large Language Models Evaluation Metrics to Human Evaluation
von: Sarmah, Bhaskarjit, et al.
Veröffentlicht: (2024) -
Temporal Relational Reasoning of Large Language Models for Detecting Stock Portfolio Crashes
von: Koa, Kelvin J. L., et al.
Veröffentlicht: (2024)