Semantic Faithfulness and Entropy Production Measures to Tame Your LLM Demons and Manage Hallucinations
Fuente:
arXiv
Saved in:
| Main Author: | Halperin, Igor |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompt-Response Semantic Divergence Metrics for Faithfulness Hallucination and Misalignment Detection in Large Language Models
by: Halperin, Igor
Published: (2025)
by: Halperin, Igor
Published: (2025)
CAISSON: Concept-Augmented Inference Suite of Self-Organizing Neural Networks
by: Halperin, Igor
Published: (2024)
by: Halperin, Igor
Published: (2024)
Topic Identification in LLM Input-Output Pairs through the Lens of Information Bottleneck
by: Halperin, Igor
Published: (2025)
by: Halperin, Igor
Published: (2025)
Know Your Limits: Entropy Estimation Modeling for Compression and Generalization
by: Badger, Benjamin L., et al.
Published: (2025)
by: Badger, Benjamin L., et al.
Published: (2025)
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach
by: Ni, Haowei, et al.
Published: (2024)
by: Ni, Haowei, et al.
Published: (2024)
The Stepwise Informativeness Assumption: Why are Entropy Dynamics and Reasoning Correlated in LLMs?
by: Català, Mar Gonzàlez I, et al.
Published: (2026)
by: Català, Mar Gonzàlez I, et al.
Published: (2026)
FinBloom: Knowledge Grounding Large Language Model with Real-time Financial Data
by: Sinha, Ankur, et al.
Published: (2025)
by: Sinha, Ankur, et al.
Published: (2025)
Measuring Uncertainty in Transformer Circuits with Effective Information Consistency
by: Krasnovsky, Anatoly A.
Published: (2025)
by: Krasnovsky, Anatoly A.
Published: (2025)
Learning is Forgetting: LLM Training As Lossy Compression
by: Conklin, Henry C., et al.
Published: (2026)
by: Conklin, Henry C., et al.
Published: (2026)
A Training-free Method for LLM Text Attribution
by: Radvand, Tara, et al.
Published: (2025)
by: Radvand, Tara, et al.
Published: (2025)
HeavyWater and SimplexWater: Distortion-Free LLM Watermarks for Low-Entropy Next-Token Predictions
by: Tsur, Dor, et al.
Published: (2025)
by: Tsur, Dor, et al.
Published: (2025)
A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability
by: Omidvar, Hamed, et al.
Published: (2026)
by: Omidvar, Hamed, et al.
Published: (2026)
Microservices-Based Framework for Predictive Analytics and Real-time Performance Enhancement in Travel Reservation Systems
by: Barua, Biman, et al.
Published: (2024)
by: Barua, Biman, et al.
Published: (2024)
Conversational Factor Information Retrieval Model (ConFIRM)
by: Choi, Stephen, et al.
Published: (2023)
by: Choi, Stephen, et al.
Published: (2023)
Identifying Banking Transaction Descriptions via Support Vector Machine Short-Text Classification Based on a Specialized Labelled Corpus
by: García-Méndez, Silvia, et al.
Published: (2024)
by: García-Méndez, Silvia, et al.
Published: (2024)
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs
by: Kossen, Jannik, et al.
Published: (2024)
by: Kossen, Jannik, et al.
Published: (2024)
Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance
by: Benhenda, Mostapha
Published: (2026)
by: Benhenda, Mostapha
Published: (2026)
Emoji Driven Crypto Assets Market Reactions
by: Zuo, Xiaorui, et al.
Published: (2024)
by: Zuo, Xiaorui, et al.
Published: (2024)
Tx-LLM: A Large Language Model for Therapeutics
by: Chaves, Juan Manuel Zambrano, et al.
Published: (2024)
by: Chaves, Juan Manuel Zambrano, et al.
Published: (2024)
Exploring the Synergy of Quantitative Factors and Newsflow Representations from Large Language Models for Stock Return Prediction
by: Guo, Tian, et al.
Published: (2025)
by: Guo, Tian, et al.
Published: (2025)
Multi-Dimensional Behavioral Evaluation of Agentic Stock Prediction Systems Using Large Language Model Judges with Closed-Loop Reinforcement Learning Feedback
by: Ridhawi, Mohammad Al, et al.
Published: (2026)
by: Ridhawi, Mohammad Al, et al.
Published: (2026)
Diabetica: Adapting Large Language Model to Enhance Multiple Medical Tasks in Diabetes Care and Management
by: Wei, Lai, et al.
Published: (2024)
by: Wei, Lai, et al.
Published: (2024)
SLIM: Sparse Latent Steering for Interpretable and Property-Directed LLM-Based Molecular Editing
by: Zhang, Mingxu, et al.
Published: (2026)
by: Zhang, Mingxu, et al.
Published: (2026)
OAEI-LLM-T: A TBox Benchmark Dataset for Understanding Large Language Model Hallucinations in Ontology Matching
by: Qiang, Zhangcheng, et al.
Published: (2025)
by: Qiang, Zhangcheng, et al.
Published: (2025)
Semantic In-Domain Product Identification for Search Queries
by: Sharma, Sanat, et al.
Published: (2024)
by: Sharma, Sanat, et al.
Published: (2024)
Measuring Faithfulness Depends on How You Measure: Classifier Sensitivity in LLM Chain-of-Thought Evaluation
by: Young, Richard J.
Published: (2026)
by: Young, Richard J.
Published: (2026)
MOOSE-Chem2: Exploring LLM Limits in Fine-Grained Scientific Hypothesis Discovery via Hierarchical Search
by: Yang, Zonglin, et al.
Published: (2025)
by: Yang, Zonglin, et al.
Published: (2025)
Hallucination is a Consequence of Space-Optimality: A Rate-Distortion Theorem for Membership Testing
by: Guo, Anxin, et al.
Published: (2026)
by: Guo, Anxin, et al.
Published: (2026)
FaithEval: Can Your Language Model Stay Faithful to Context, Even If "The Moon is Made of Marshmallows"
by: Ming, Yifei, et al.
Published: (2024)
by: Ming, Yifei, et al.
Published: (2024)
Mitigating Hallucination with ZeroG: An Advanced Knowledge Management Engine
by: Sharma, Anantha, et al.
Published: (2024)
by: Sharma, Anantha, et al.
Published: (2024)
Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning
by: Mitra, Purbesh, et al.
Published: (2025)
by: Mitra, Purbesh, et al.
Published: (2025)
Measuring and Reducing LLM Hallucination without Gold-Standard Answers
by: Wei, Jiaheng, et al.
Published: (2024)
by: Wei, Jiaheng, et al.
Published: (2024)
RSAT: Structured Attribution Makes Small Language Models Faithful Table Reasoners
by: Gajjar, Jugal, et al.
Published: (2026)
by: Gajjar, Jugal, et al.
Published: (2026)
Visual Language Model based Cross-modal Semantic Communication Systems
by: Jiang, Feibo, et al.
Published: (2024)
by: Jiang, Feibo, et al.
Published: (2024)
Memorization-Compression Cycles Improve Generalization
by: Yu, Fangyuan
Published: (2025)
by: Yu, Fangyuan
Published: (2025)
SPEX: Scaling Feature Interaction Explanations for LLMs
by: Kang, Justin Singh, et al.
Published: (2025)
by: Kang, Justin Singh, et al.
Published: (2025)
Subjective Depth and Timescale Transformers: Learning Where and When to Compute
by: Wieser, Frederico, et al.
Published: (2025)
by: Wieser, Frederico, et al.
Published: (2025)
SQuat: Subspace-orthogonal KV Cache Quantization
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
An Information Theoretic Perspective on Agentic System Design
by: He, Shizhe, et al.
Published: (2025)
by: He, Shizhe, et al.
Published: (2025)
Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Similar Items
-
Prompt-Response Semantic Divergence Metrics for Faithfulness Hallucination and Misalignment Detection in Large Language Models
by: Halperin, Igor
Published: (2025) -
CAISSON: Concept-Augmented Inference Suite of Self-Organizing Neural Networks
by: Halperin, Igor
Published: (2024) -
Topic Identification in LLM Input-Output Pairs through the Lens of Information Bottleneck
by: Halperin, Igor
Published: (2025) -
Know Your Limits: Entropy Estimation Modeling for Compression and Generalization
by: Badger, Benjamin L., et al.
Published: (2025) -
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach
by: Ni, Haowei, et al.
Published: (2024)