STORYSUMM: Evaluating Faithfulness in Story Summarization
Fuente:
arXiv
Saved in:
| Main Authors: | Subbiah, Melanie, Ladhak, Faisal, Mishra, Akankshya, Adams, Griffin, Chilton, Lydia B., McKeown, Kathleen |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reading Subtext: Evaluating Large Language Models on Short Story Summarization with Writers
by: Subbiah, Melanie, et al.
Published: (2024)
by: Subbiah, Melanie, et al.
Published: (2024)
Is the Top Still Spinning? Evaluating Subjectivity in Narrative Understanding
by: Subbiah, Melanie, et al.
Published: (2025)
by: Subbiah, Melanie, et al.
Published: (2025)
Counterfactual Simulatability of LLM Explanations for Generation Tasks
by: Limpijankit, Marvin, et al.
Published: (2025)
by: Limpijankit, Marvin, et al.
Published: (2025)
AdvSumm: Adversarial Training for Bias Mitigation in Text Summarization
by: Gupta, Mukur, et al.
Published: (2025)
by: Gupta, Mukur, et al.
Published: (2025)
Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives
by: Subbiah, Melanie, et al.
Published: (2026)
by: Subbiah, Melanie, et al.
Published: (2026)
Enhancing Multimodal Affective Analysis with Learned Live Comment Features
by: Deng, Zhaoyuan, et al.
Published: (2024)
by: Deng, Zhaoyuan, et al.
Published: (2024)
Summarization of Opinionated Political Documents with Varied Perspectives
by: Deas, Nicholas, et al.
Published: (2024)
by: Deas, Nicholas, et al.
Published: (2024)
Reranking-based Generation for Unbiased Perspective Summarization
by: Ri, Narutatsu, et al.
Published: (2025)
by: Ri, Narutatsu, et al.
Published: (2025)
ParaGuide: Guided Diffusion Paraphrasers for Plug-and-Play Textual Style Transfer
by: Horvitz, Zachary, et al.
Published: (2023)
by: Horvitz, Zachary, et al.
Published: (2023)
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization
by: Tang, Liyan, et al.
Published: (2024)
by: Tang, Liyan, et al.
Published: (2024)
Computational Representations of Character Significance in Novels
by: Mian, Haaris, et al.
Published: (2026)
by: Mian, Haaris, et al.
Published: (2026)
See It from My Perspective: How Language Affects Cultural Bias in Image Understanding
by: Ananthram, Amith, et al.
Published: (2024)
by: Ananthram, Amith, et al.
Published: (2024)
Parallel Structures in Pre-training Data Yield In-Context Learning
by: Chen, Yanda, et al.
Published: (2024)
by: Chen, Yanda, et al.
Published: (2024)
On the Relation between Sensitivity and Accuracy in In-context Learning
by: Chen, Yanda, et al.
Published: (2022)
by: Chen, Yanda, et al.
Published: (2022)
Artificial Impressions: Evaluating Large Language Model Behavior Through the Lens of Trait Impressions
by: Deas, Nicholas, et al.
Published: (2025)
by: Deas, Nicholas, et al.
Published: (2025)
Topic-aware Large Language Models for Summarizing the Lived Healthcare Experiences Described in Health Stories
by: Bilalpur, Maneesh, et al.
Published: (2025)
by: Bilalpur, Maneesh, et al.
Published: (2025)
Faithful Chart Summarization with ChaTS-Pi
by: Krichene, Syrine, et al.
Published: (2024)
by: Krichene, Syrine, et al.
Published: (2024)
Getting Serious about Humor: Crafting Humor Datasets with Unfunny Large Language Models
by: Horvitz, Zachary, et al.
Published: (2024)
by: Horvitz, Zachary, et al.
Published: (2024)
Aligning Large Language Models via Fine-grained Supervision
by: Xu, Dehong, et al.
Published: (2024)
by: Xu, Dehong, et al.
Published: (2024)
TrueBrief: Faithful Summarization through Small Language Models
by: Lakara, Kumud, et al.
Published: (2025)
by: Lakara, Kumud, et al.
Published: (2025)
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference
by: Warner, Benjamin, et al.
Published: (2024)
by: Warner, Benjamin, et al.
Published: (2024)
Not Just Novelty: A Longitudinal Study on Utility and Customization of an AI Workflow
by: Long, Tao, et al.
Published: (2024)
by: Long, Tao, et al.
Published: (2024)
FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs
by: Bao, Forrest Sheng, et al.
Published: (2024)
by: Bao, Forrest Sheng, et al.
Published: (2024)
S^2tory: Story Spine Distillation for Movie Script Summarization
by: Lu, Mingzhe, et al.
Published: (2026)
by: Lu, Mingzhe, et al.
Published: (2026)
AgenticSum: An Agentic Inference-Time Framework for Faithful Clinical Text Summarization
by: Piya, Fahmida Liza, et al.
Published: (2026)
by: Piya, Fahmida Liza, et al.
Published: (2026)
ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
by: Qian, Cheng, et al.
Published: (2025)
by: Qian, Cheng, et al.
Published: (2025)
LiveNewsBench: Evaluating LLM Web Search Capabilities with Freshly Curated News
by: Zhang, Yunfan, et al.
Published: (2026)
by: Zhang, Yunfan, et al.
Published: (2026)
Improving Faithfulness of Large Language Models in Summarization via Sliding Generation and Self-Consistency
by: Li, Taiji, et al.
Published: (2024)
by: Li, Taiji, et al.
Published: (2024)
Beyond Overlap Metrics: Rewarding Reasoning and Preferences for Faithful Multi-Role Dialogue Summarization
by: Mei, Xiaoyong, et al.
Published: (2026)
by: Mei, Xiaoyong, et al.
Published: (2026)
Forecasting Conversation Derailments Through Generation
by: Zhang, Yunfan, et al.
Published: (2025)
by: Zhang, Yunfan, et al.
Published: (2025)
iBERT: Interpretable Embeddings via Sense Decomposition
by: Anand, Vishal, et al.
Published: (2025)
by: Anand, Vishal, et al.
Published: (2025)
Many-Turn Jailbreaking
by: Yang, Xianjun, et al.
Published: (2025)
by: Yang, Xianjun, et al.
Published: (2025)
Improving Faithfulness of Abstractive Summarization by Controlling Confounding Effect of Irrelevant Sentences
by: Ghoshal, Asish, et al.
Published: (2022)
by: Ghoshal, Asish, et al.
Published: (2022)
Northeastern Uni at Multilingual Counterspeech Generation: Enhancing Counter Speech Generation with LLM Alignment through Direct Preference Optimization
by: Wadhwa, Sahil, et al.
Published: (2024)
by: Wadhwa, Sahil, et al.
Published: (2024)
Mechanistic Interpretability of GPT-like Models on Summarization Tasks
by: Mishra, Anurag
Published: (2025)
by: Mishra, Anurag
Published: (2025)
Mining Contextualized Visual Associations from Images for Creativity Understanding
by: Sahu, Ananya, et al.
Published: (2025)
by: Sahu, Ananya, et al.
Published: (2025)
MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models
by: Ha, Hyeonjeong, et al.
Published: (2026)
by: Ha, Hyeonjeong, et al.
Published: (2026)
Exploring Chain-of-Thought Reasoning for Steerable Pluralistic Alignment
by: Zhang, Yunfan, et al.
Published: (2025)
by: Zhang, Yunfan, et al.
Published: (2025)
StoryAlign: Evaluating and Training Reward Models for Story Generation
by: Xia, Haotian, et al.
Published: (2026)
by: Xia, Haotian, et al.
Published: (2026)
Evaluate Summarization in Fine-Granularity: Auto Evaluation with LLM
by: Yuan, Dong, et al.
Published: (2024)
by: Yuan, Dong, et al.
Published: (2024)
Similar Items
-
Reading Subtext: Evaluating Large Language Models on Short Story Summarization with Writers
by: Subbiah, Melanie, et al.
Published: (2024) -
Is the Top Still Spinning? Evaluating Subjectivity in Narrative Understanding
by: Subbiah, Melanie, et al.
Published: (2025) -
Counterfactual Simulatability of LLM Explanations for Generation Tasks
by: Limpijankit, Marvin, et al.
Published: (2025) -
AdvSumm: Adversarial Training for Bias Mitigation in Text Summarization
by: Gupta, Mukur, et al.
Published: (2025) -
Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives
by: Subbiah, Melanie, et al.
Published: (2026)