Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Wen, Peng, Guangyue, Li, Wei, Wei, Shaohang, Song, Feifan, Wang, Liang, Yang, Nan, Zhang, Xingxing, Jin, Jing, Wei, Furu, Wang, Houfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Odysseus Navigates the Sirens' Song: Dynamic Focus Decoding for Factual and Diverse Open-Ended Text Generation
von: Luo, Wen, et al.
Veröffentlicht: (2025)
von: Luo, Wen, et al.
Veröffentlicht: (2025)
Only Say What You Know: Calibration-Aware Generation for Long-Form Factuality
von: Luo, Wen, et al.
Veröffentlicht: (2026)
von: Luo, Wen, et al.
Veröffentlicht: (2026)
Explanation based In-Context Demonstrations Retrieval for Multilingual Grammatical Error Correction
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding
von: Song, Feifan, et al.
Veröffentlicht: (2025)
von: Song, Feifan, et al.
Veröffentlicht: (2025)
HalluDial: A Large-Scale Benchmark for Automatic Dialogue-Level Hallucination Evaluation
von: Luo, Wen, et al.
Veröffentlicht: (2024)
von: Luo, Wen, et al.
Veröffentlicht: (2024)
TIME: A Multi-level Benchmark for Temporal Reasoning of LLMs in Real-World Scenarios
von: Wei, Shaohang, et al.
Veröffentlicht: (2025)
von: Wei, Shaohang, et al.
Veröffentlicht: (2025)
Mitigating Overthinking through Reasoning Shaping
von: Song, Feifan, et al.
Veröffentlicht: (2025)
von: Song, Feifan, et al.
Veröffentlicht: (2025)
Bootstrap Your Own Context Length
von: Wang, Liang, et al.
Veröffentlicht: (2024)
von: Wang, Liang, et al.
Veröffentlicht: (2024)
Learning to Retrieve In-Context Examples for Large Language Models
von: Wang, Liang, et al.
Veröffentlicht: (2023)
von: Wang, Liang, et al.
Veröffentlicht: (2023)
Detection-Correction Structure via General Language Model for Grammatical Error Correction
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
CiteCheck: Towards Accurate Citation Faithfulness Detection
von: Xu, Ziyao, et al.
Veröffentlicht: (2025)
von: Xu, Ziyao, et al.
Veröffentlicht: (2025)
Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generation
von: Peng, Guangyue, et al.
Veröffentlicht: (2026)
von: Peng, Guangyue, et al.
Veröffentlicht: (2026)
FEA-Bench: A Benchmark for Evaluating Repository-Level Code Generation for Feature Implementation
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
ICDPO: Effectively Borrowing Alignment Capability of Others via In-context Direct Preference Optimization
von: Song, Feifan, et al.
Veröffentlicht: (2024)
von: Song, Feifan, et al.
Veröffentlicht: (2024)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
von: Tang, Zhengyang, et al.
Veröffentlicht: (2024)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2024)
Examining False Positives under Inference Scaling for Mathematical Reasoning
von: Wang, Yu, et al.
Veröffentlicht: (2025)
von: Wang, Yu, et al.
Veröffentlicht: (2025)
Thinking Augmented Pre-training
von: Wang, Liang, et al.
Veröffentlicht: (2025)
von: Wang, Liang, et al.
Veröffentlicht: (2025)
Characterizing Truthfulness in Large Language Model Generations with Local Intrinsic Dimension
von: Yin, Fan, et al.
Veröffentlicht: (2024)
von: Yin, Fan, et al.
Veröffentlicht: (2024)
Large Search Model: Redefining Search Stack in the Era of LLMs
von: Wang, Liang, et al.
Veröffentlicht: (2023)
von: Wang, Liang, et al.
Veröffentlicht: (2023)
Multilingual E5 Text Embeddings: A Technical Report
von: Wang, Liang, et al.
Veröffentlicht: (2024)
von: Wang, Liang, et al.
Veröffentlicht: (2024)
Improving Text Embeddings with Large Language Models
von: Wang, Liang, et al.
Veröffentlicht: (2023)
von: Wang, Liang, et al.
Veröffentlicht: (2023)
LongEmbed: Extending Embedding Models for Long Context Retrieval
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
PoSE: Efficient Context Window Extension of LLMs via Positional Skip-wise Training
von: Zhu, Dawei, et al.
Veröffentlicht: (2023)
von: Zhu, Dawei, et al.
Veröffentlicht: (2023)
Chain-of-Retrieval Augmented Generation
von: Wang, Liang, et al.
Veröffentlicht: (2025)
von: Wang, Liang, et al.
Veröffentlicht: (2025)
QueST: Incentivizing LLMs to Generate Difficult Problems
von: Hu, Hanxu, et al.
Veröffentlicht: (2025)
von: Hu, Hanxu, et al.
Veröffentlicht: (2025)
WildLong: Synthesizing Realistic Long-Context Instruction Data at Scale
von: Li, Jiaxi, et al.
Veröffentlicht: (2025)
von: Li, Jiaxi, et al.
Veröffentlicht: (2025)
Learning to Draft: Adaptive Speculative Decoding with Reinforcement Learning
von: Zhang, Jiebin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiebin, et al.
Veröffentlicht: (2026)
Preference Ranking Optimization for Human Alignment
von: Song, Feifan, et al.
Veröffentlicht: (2023)
von: Song, Feifan, et al.
Veröffentlicht: (2023)
Scaling Data Diversity for Fine-Tuning Language Models in Human Alignment
von: Song, Feifan, et al.
Veröffentlicht: (2024)
von: Song, Feifan, et al.
Veröffentlicht: (2024)
MeNTi: Bridging Medical Calculator and LLM Agent with Nested Tool Calling
von: Zhu, Yakun, et al.
Veröffentlicht: (2024)
von: Zhu, Yakun, et al.
Veröffentlicht: (2024)
Self-Boosting Large Language Models with Synthetic Preference Data
von: Dong, Qingxiu, et al.
Veröffentlicht: (2024)
von: Dong, Qingxiu, et al.
Veröffentlicht: (2024)
MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
P-Aligner: Enabling Pre-Alignment of Language Models via Principled Instruction Synthesis
von: Song, Feifan, et al.
Veröffentlicht: (2025)
von: Song, Feifan, et al.
Veröffentlicht: (2025)
Investigating the (De)Composition Capabilities of Large Language Models in Natural-to-Formal Language Conversion
von: Xu, Ziyao, et al.
Veröffentlicht: (2025)
von: Xu, Ziyao, et al.
Veröffentlicht: (2025)
SPOR: A Comprehensive and Practical Evaluation Method for Compositional Generalization in Data-to-Text Generation
von: Xu, Ziyao, et al.
Veröffentlicht: (2024)
von: Xu, Ziyao, et al.
Veröffentlicht: (2024)
Towards Thinking-Optimal Scaling of Test-Time Compute for LLM Reasoning
von: Yang, Wenkai, et al.
Veröffentlicht: (2025)
von: Yang, Wenkai, et al.
Veröffentlicht: (2025)
BitNet a4.8: 4-bit Activations for 1-bit LLMs
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
BitNet v2: Native 4-bit Activations with Hadamard Transformation for 1-bit LLMs
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
BitNet b1.58 2B4T Technical Report
von: Ma, Shuming, et al.
Veröffentlicht: (2025)
von: Ma, Shuming, et al.
Veröffentlicht: (2025)
Utilizing Local Hierarchy with Adversarial Training for Hierarchical Text Classification
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Odysseus Navigates the Sirens' Song: Dynamic Focus Decoding for Factual and Diverse Open-Ended Text Generation
von: Luo, Wen, et al.
Veröffentlicht: (2025) -
Only Say What You Know: Calibration-Aware Generation for Long-Form Factuality
von: Luo, Wen, et al.
Veröffentlicht: (2026) -
Explanation based In-Context Demonstrations Retrieval for Multilingual Grammatical Error Correction
von: Li, Wei, et al.
Veröffentlicht: (2025) -
Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding
von: Song, Feifan, et al.
Veröffentlicht: (2025) -
HalluDial: A Large-Scale Benchmark for Automatic Dialogue-Level Hallucination Evaluation
von: Luo, Wen, et al.
Veröffentlicht: (2024)