Hallucinate at the Last in Long Response Generation: A Case Study on Long Document Summarization
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Joonho, Yoon, Seunghyun, Chang, Hwan, Kim, Byeongjeong, Lee, Hwanhee |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FIZZ: Factual Inconsistency Detection by Zoom-in Summary and Zoom-out Document
by: Yang, Joonho, et al.
Published: (2024)
by: Yang, Joonho, et al.
Published: (2024)
Chronological Passage Assembling in RAG framework for Temporal Question Answering
by: Kim, Byeongjeong, et al.
Published: (2025)
by: Kim, Byeongjeong, et al.
Published: (2025)
Probing-RAG: Self-Probing to Guide Language Models in Selective Document Retrieval
by: Baek, Ingeol, et al.
Published: (2024)
by: Baek, Ingeol, et al.
Published: (2024)
Which Retain Set Matters for LLM Unlearning? A Case Study on Entity Unlearning
by: Chang, Hwan, et al.
Published: (2025)
by: Chang, Hwan, et al.
Published: (2025)
Enhancing Multilingual RAG Systems with Debiased Language Preference-Guided Query Fusion
by: Park, Jeonghyun, et al.
Published: (2026)
by: Park, Jeonghyun, et al.
Published: (2026)
Personality Editing for Language Models through Adjusting Self-Referential Queries
by: Hwang, Seojin, et al.
Published: (2025)
by: Hwang, Seojin, et al.
Published: (2025)
SAFE-SQL: Self-Augmented In-Context Learning with Fine-grained Example Selection for Text-to-SQL
by: Lee, Jimin, et al.
Published: (2025)
by: Lee, Jimin, et al.
Published: (2025)
Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models
by: Jang, Haeun, et al.
Published: (2026)
by: Jang, Haeun, et al.
Published: (2026)
Crafting the Path: Robust Query Rewriting for Information Retrieval
by: Baek, Ingeol, et al.
Published: (2024)
by: Baek, Ingeol, et al.
Published: (2024)
ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents
by: Chang, Hwan, et al.
Published: (2025)
by: Chang, Hwan, et al.
Published: (2025)
Keep Security! Benchmarking Security Policy Preservation in Large Language Model Contexts Against Indirect Attacks in Question Answering
by: Chang, Hwan, et al.
Published: (2025)
by: Chang, Hwan, et al.
Published: (2025)
Automatic Summarization of Long Documents
by: Chhibbar, Naman, et al.
Published: (2024)
by: Chhibbar, Naman, et al.
Published: (2024)
Low-Resource Cross-Lingual Summarization through Few-Shot Learning with Large Language Models
by: Park, Gyutae, et al.
Published: (2024)
by: Park, Gyutae, et al.
Published: (2024)
Conversational Query Reformulation with the Guidance of Retrieved Documents
by: Park, Jeonghyun, et al.
Published: (2024)
by: Park, Jeonghyun, et al.
Published: (2024)
Learning to Summarize by Learning to Quiz: Adversarial Agentic Collaboration for Long Document Summarization
by: Wang, Weixuan, et al.
Published: (2025)
by: Wang, Weixuan, et al.
Published: (2025)
Context-Aware Hierarchical Merging for Long Document Summarization
by: Ou, Litu, et al.
Published: (2025)
by: Ou, Litu, et al.
Published: (2025)
Agent-as-Judge for Factual Summarization of Long Narratives
by: Jeong, Yeonseok, et al.
Published: (2025)
by: Jeong, Yeonseok, et al.
Published: (2025)
Exploring Persona Sentiment Sensitivity in Personalized Dialogue Generation
by: Jun, Yonghyun, et al.
Published: (2025)
by: Jun, Yonghyun, et al.
Published: (2025)
ToDi: Token-wise Distillation via Fine-Grained Divergence Control
by: Jung, Seongryong, et al.
Published: (2025)
by: Jung, Seongryong, et al.
Published: (2025)
Dynamic Order Template Prediction for Generative Aspect-Based Sentiment Analysis
by: Jun, Yonghyun, et al.
Published: (2024)
by: Jun, Yonghyun, et al.
Published: (2024)
Discourse-Driven Evaluation: Unveiling Factual Inconsistency in Long Document Summarization
by: Zhong, Yang, et al.
Published: (2025)
by: Zhong, Yang, et al.
Published: (2025)
Selective Demonstration Retrieval for Improved Implicit Hate Speech Detection
by: Kim, Yumin, et al.
Published: (2025)
by: Kim, Yumin, et al.
Published: (2025)
Characterizing Multimodal Long-form Summarization: A Case Study on Financial Reports
by: Cao, Tianyu, et al.
Published: (2024)
by: Cao, Tianyu, et al.
Published: (2024)
End-to-End Long Document Summarization using Gradient Caching
by: Saxena, Rohit, et al.
Published: (2025)
by: Saxena, Rohit, et al.
Published: (2025)
LOCOST: State-Space Models for Long Document Abstractive Summarization
by: Bronnec, Florian Le, et al.
Published: (2024)
by: Bronnec, Florian Le, et al.
Published: (2024)
StrucSum: Graph-Structured Reasoning for Long Document Extractive Summarization with LLMs
by: Yuan, Haohan, et al.
Published: (2025)
by: Yuan, Haohan, et al.
Published: (2025)
MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference
by: Park, Jeonghyun, et al.
Published: (2025)
by: Park, Jeonghyun, et al.
Published: (2025)
Hypothetical Documents or Knowledge Leakage? Rethinking LLM-based Query Expansion
by: Yoon, Yejun, et al.
Published: (2025)
by: Yoon, Yejun, et al.
Published: (2025)
From Single to Multi: How LLMs Hallucinate in Multi-Document Summarization
by: Belem, Catarina G., et al.
Published: (2024)
by: Belem, Catarina G., et al.
Published: (2024)
Investigating Language Preference of Multilingual RAG Systems
by: Park, Jeonghyun, et al.
Published: (2025)
by: Park, Jeonghyun, et al.
Published: (2025)
NexusSum: Hierarchical LLM Agents for Long-Form Narrative Summarization
by: Kim, Hyuntak, et al.
Published: (2025)
by: Kim, Hyuntak, et al.
Published: (2025)
Theme-Explanation Structure for Table Summarization using Large Language Models: A Case Study on Korean Tabular Data
by: Kwack, TaeYoon, et al.
Published: (2025)
by: Kwack, TaeYoon, et al.
Published: (2025)
Long Dialog Summarization: An Analysis
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
ARLED: Leveraging LED-based ARMAN Model for Abstractive Summarization of Persian Long Documents
by: Zangooei, Samira, et al.
Published: (2025)
by: Zangooei, Samira, et al.
Published: (2025)
NoLiMa: Long-Context Evaluation Beyond Literal Matching
by: Modarressi, Ali, et al.
Published: (2025)
by: Modarressi, Ali, et al.
Published: (2025)
Stress Testing Factual Consistency Metrics for Long-Document Summarization
by: Mujahid, Zain Muhammad, et al.
Published: (2025)
by: Mujahid, Zain Muhammad, et al.
Published: (2025)
Incorporating Distributions of Discourse Structure for Long Document Abstractive Summarization
by: Liu, Dongqi, et al.
Published: (2023)
by: Liu, Dongqi, et al.
Published: (2023)
Towards Enhancing Coherence in Extractive Summarization: Dataset and Experiments with LLMs
by: Parmar, Mihir, et al.
Published: (2024)
by: Parmar, Mihir, et al.
Published: (2024)
On Positional Bias of Faithfulness for Long-form Summarization
by: Wan, David, et al.
Published: (2024)
by: Wan, David, et al.
Published: (2024)
CORG: Generating Answers from Complex, Interrelated Contexts
by: Lee, Hyunji, et al.
Published: (2025)
by: Lee, Hyunji, et al.
Published: (2025)
Similar Items
-
FIZZ: Factual Inconsistency Detection by Zoom-in Summary and Zoom-out Document
by: Yang, Joonho, et al.
Published: (2024) -
Chronological Passage Assembling in RAG framework for Temporal Question Answering
by: Kim, Byeongjeong, et al.
Published: (2025) -
Probing-RAG: Self-Probing to Guide Language Models in Selective Document Retrieval
by: Baek, Ingeol, et al.
Published: (2024) -
Which Retain Set Matters for LLM Unlearning? A Case Study on Entity Unlearning
by: Chang, Hwan, et al.
Published: (2025) -
Enhancing Multilingual RAG Systems with Debiased Language Preference-Guided Query Fusion
by: Park, Jeonghyun, et al.
Published: (2026)