Only Say What You Know: Calibration-Aware Generation for Long-Form Factuality
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Wen, Peng, Guangyue, Wang, Liang, Yang, Nan, Li, Wei, Song, Yuhan, Wei, Shaohang, Song, Feifan, Wei, Furu, Wang, Houfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Odysseus Navigates the Sirens' Song: Dynamic Focus Decoding for Factual and Diverse Open-Ended Text Generation
von: Luo, Wen, et al.
Veröffentlicht: (2025)
von: Luo, Wen, et al.
Veröffentlicht: (2025)
Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations
von: Luo, Wen, et al.
Veröffentlicht: (2026)
von: Luo, Wen, et al.
Veröffentlicht: (2026)
Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding
von: Song, Feifan, et al.
Veröffentlicht: (2025)
von: Song, Feifan, et al.
Veröffentlicht: (2025)
Explanation based In-Context Demonstrations Retrieval for Multilingual Grammatical Error Correction
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
TIME: A Multi-level Benchmark for Temporal Reasoning of LLMs in Real-World Scenarios
von: Wei, Shaohang, et al.
Veröffentlicht: (2025)
von: Wei, Shaohang, et al.
Veröffentlicht: (2025)
Mitigating Overthinking through Reasoning Shaping
von: Song, Feifan, et al.
Veröffentlicht: (2025)
von: Song, Feifan, et al.
Veröffentlicht: (2025)
HalluDial: A Large-Scale Benchmark for Automatic Dialogue-Level Hallucination Evaluation
von: Luo, Wen, et al.
Veröffentlicht: (2024)
von: Luo, Wen, et al.
Veröffentlicht: (2024)
You Know What I'm Saying: Jailbreak Attack via Implicit Reference
von: Wu, Tianyu, et al.
Veröffentlicht: (2024)
von: Wu, Tianyu, et al.
Veröffentlicht: (2024)
Detection-Correction Structure via General Language Model for Grammatical Error Correction
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generation
von: Peng, Guangyue, et al.
Veröffentlicht: (2026)
von: Peng, Guangyue, et al.
Veröffentlicht: (2026)
FEA-Bench: A Benchmark for Evaluating Repository-Level Code Generation for Feature Implementation
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
ICDPO: Effectively Borrowing Alignment Capability of Others via In-context Direct Preference Optimization
von: Song, Feifan, et al.
Veröffentlicht: (2024)
von: Song, Feifan, et al.
Veröffentlicht: (2024)
Think Through Uncertainty: Improving Long-Form Generation Factuality via Reasoning Calibration
von: Liu, Xin, et al.
Veröffentlicht: (2026)
von: Liu, Xin, et al.
Veröffentlicht: (2026)
LongEmbed: Extending Embedding Models for Long Context Retrieval
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
It's Not Just What You Say, but How You Say It: The Effects of Enterprise Social Media on Service Management, Through the Lens of Signaling Theory
von: Alexandra Budjanovcanin, et al.
Veröffentlicht: (2025)
von: Alexandra Budjanovcanin, et al.
Veröffentlicht: (2025)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
von: Park, Young-Jin, et al.
Veröffentlicht: (2025)
von: Park, Young-Jin, et al.
Veröffentlicht: (2025)
Learning to Retrieve In-Context Examples for Large Language Models
von: Wang, Liang, et al.
Veröffentlicht: (2023)
von: Wang, Liang, et al.
Veröffentlicht: (2023)
Seeing What You Say: Expressive Image Generation from Speech
von: Lee, Jiyoung, et al.
Veröffentlicht: (2025)
von: Lee, Jiyoung, et al.
Veröffentlicht: (2025)
From a Bird's Eye View to See: Joint Camera and Subject Registration without the Camera Calibration
von: Qian, Zekun, et al.
Veröffentlicht: (2022)
von: Qian, Zekun, et al.
Veröffentlicht: (2022)
YOCO: You Only Calibrate Once for Accurate Extrinsic Parameter in LiDAR-Camera Systems
von: Zeng, Tianle, et al.
Veröffentlicht: (2024)
von: Zeng, Tianle, et al.
Veröffentlicht: (2024)
You Never Know a Person, You Only Know Their Defenses: Detecting Levels of Psychological Defense Mechanisms in Supportive Conversations
von: Na, Hongbin, et al.
Veröffentlicht: (2025)
von: Na, Hongbin, et al.
Veröffentlicht: (2025)
Just Say What You Want: Only-prompting Self-rewarding Online Preference Optimization
von: Xu, Ruijie, et al.
Veröffentlicht: (2024)
von: Xu, Ruijie, et al.
Veröffentlicht: (2024)
You Only Cache Once: Decoder-Decoder Architectures for Language Models
von: Sun, Yutao, et al.
Veröffentlicht: (2024)
von: Sun, Yutao, et al.
Veröffentlicht: (2024)
Know What You Know: Metacognitive Entropy Calibration for Verifiable RL Reasoning
von: Zhao, Qiannian, et al.
Veröffentlicht: (2026)
von: Zhao, Qiannian, et al.
Veröffentlicht: (2026)
Beyond Transcription: Unified Audio Schema for Perception-Aware AudioLLMs
von: Zhang, Linhao, et al.
Veröffentlicht: (2026)
von: Zhang, Linhao, et al.
Veröffentlicht: (2026)
KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs
von: Song, Yuhan, et al.
Veröffentlicht: (2025)
von: Song, Yuhan, et al.
Veröffentlicht: (2025)
When You Don't Know the Answer, Say So
Veröffentlicht: (2024)
Veröffentlicht: (2024)
Grasp as You Say: Language-guided Dexterous Grasp Generation
von: Wei, Yi-Lin, et al.
Veröffentlicht: (2024)
von: Wei, Yi-Lin, et al.
Veröffentlicht: (2024)
Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance
von: Wang, Zan, et al.
Veröffentlicht: (2024)
von: Wang, Zan, et al.
Veröffentlicht: (2024)
P-Aligner: Enabling Pre-Alignment of Language Models via Principled Instruction Synthesis
von: Song, Feifan, et al.
Veröffentlicht: (2025)
von: Song, Feifan, et al.
Veröffentlicht: (2025)
Beyond Precision: Importance-Aware Recall for Factuality Evaluation in Long-Form LLM Generation
von: Jafari, Nazanin, et al.
Veröffentlicht: (2026)
von: Jafari, Nazanin, et al.
Veröffentlicht: (2026)
Chain-of-Retrieval Augmented Generation
von: Wang, Liang, et al.
Veröffentlicht: (2025)
von: Wang, Liang, et al.
Veröffentlicht: (2025)
Examining False Positives under Inference Scaling for Mathematical Reasoning
von: Wang, Yu, et al.
Veröffentlicht: (2025)
von: Wang, Yu, et al.
Veröffentlicht: (2025)
Preference Ranking Optimization for Human Alignment
von: Song, Feifan, et al.
Veröffentlicht: (2023)
von: Song, Feifan, et al.
Veröffentlicht: (2023)
Scaling Data Diversity for Fine-Tuning Language Models in Human Alignment
von: Song, Feifan, et al.
Veröffentlicht: (2024)
von: Song, Feifan, et al.
Veröffentlicht: (2024)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
von: Cha, Sungguk, et al.
Veröffentlicht: (2024)
von: Cha, Sungguk, et al.
Veröffentlicht: (2024)
What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2026)
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2026)
Think Only When You Need with Large Hybrid-Reasoning Models
von: Jiang, Lingjie, et al.
Veröffentlicht: (2025)
von: Jiang, Lingjie, et al.
Veröffentlicht: (2025)
SPOR: A Comprehensive and Practical Evaluation Method for Compositional Generalization in Data-to-Text Generation
von: Xu, Ziyao, et al.
Veröffentlicht: (2024)
von: Xu, Ziyao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Odysseus Navigates the Sirens' Song: Dynamic Focus Decoding for Factual and Diverse Open-Ended Text Generation
von: Luo, Wen, et al.
Veröffentlicht: (2025) -
Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations
von: Luo, Wen, et al.
Veröffentlicht: (2026) -
Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding
von: Song, Feifan, et al.
Veröffentlicht: (2025) -
Explanation based In-Context Demonstrations Retrieval for Multilingual Grammatical Error Correction
von: Li, Wei, et al.
Veröffentlicht: (2025) -
TIME: A Multi-level Benchmark for Temporal Reasoning of LLMs in Real-World Scenarios
von: Wei, Shaohang, et al.
Veröffentlicht: (2025)