Can one size fit all?: Measuring Failure in Multi-Document Summarization Domain Transfer
Fuente:
arXiv
Saved in:
| Main Authors: | DeLucia, Alexandra, Dredze, Mark |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Anti-LM Decoding for Zero-shot In-context Machine Translation
by: Sia, Suzanna, et al.
Published: (2023)
by: Sia, Suzanna, et al.
Published: (2023)
MedScore: Generalizable Factuality Evaluation of Free-Form Medical Answers by Domain-adapted Claim Decomposition and Verification
by: Huang, Heyuan, et al.
Published: (2025)
by: Huang, Heyuan, et al.
Published: (2025)
Using Natural Language Inference to Improve Persona Extraction from Dialogue in a New Domain
by: DeLucia, Alexandra, et al.
Published: (2024)
by: DeLucia, Alexandra, et al.
Published: (2024)
Same Verdict, Different Reasons: LLM-as-a-Judge and Clinician Disagreement on Medical Chatbot Completeness
by: DeLucia, Alexandra, et al.
Published: (2026)
by: DeLucia, Alexandra, et al.
Published: (2026)
On the Failure of Latent State Persistence in Large Language Models
by: Huang, Jen-tse, et al.
Published: (2025)
by: Huang, Jen-tse, et al.
Published: (2025)
Amuro and Char: Analyzing the Relationship between Pre-Training and Fine-Tuning of Large Language Models
by: Sun, Kaiser, et al.
Published: (2024)
by: Sun, Kaiser, et al.
Published: (2024)
RAG LLMs are Not Safer: A Safety Analysis of Retrieval-Augmented Generation for Large Language Models
by: An, Bang, et al.
Published: (2025)
by: An, Bang, et al.
Published: (2025)
Task Matters: Knowledge Requirements Shape LLM Responses to Context-Memory Conflict
by: Sun, Kaiser, et al.
Published: (2025)
by: Sun, Kaiser, et al.
Published: (2025)
Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles
by: Jahara, Fatima, et al.
Published: (2025)
by: Jahara, Fatima, et al.
Published: (2025)
BERT-VBD: Vietnamese Multi-Document Summarization Framework
by: Vuong, Tuan-Cuong, et al.
Published: (2024)
by: Vuong, Tuan-Cuong, et al.
Published: (2024)
Advantages of Domain Knowledge Injection for Legal Document Summarization: A Case Study on Summarizing Indian Court Judgments in English and Hindi
by: Datta, Debtanu, et al.
Published: (2026)
by: Datta, Debtanu, et al.
Published: (2026)
Towards Multi-dimensional Evaluation of LLM Summarization across Domains and Languages
by: Min, Hyangsuk, et al.
Published: (2025)
by: Min, Hyangsuk, et al.
Published: (2025)
Automatic Summarization of Long Documents
by: Chhibbar, Naman, et al.
Published: (2024)
by: Chhibbar, Naman, et al.
Published: (2024)
Understanding Position Bias Effects on Fairness in Social Multi-Document Summarization
by: Olabisi, Olubusayo, et al.
Published: (2024)
by: Olabisi, Olubusayo, et al.
Published: (2024)
MultiBanAbs: A Comprehensive Multi-Domain Bangla Abstractive Text Summarization Dataset
by: Ferdous, Md. Tanzim, et al.
Published: (2025)
by: Ferdous, Md. Tanzim, et al.
Published: (2025)
MetaSumPerceiver: Multimodal Multi-Document Evidence Summarization for Fact-Checking
by: Chen, Ting-Chih, et al.
Published: (2024)
by: Chen, Ting-Chih, et al.
Published: (2024)
Summarization for Generative Relation Extraction in the Microbiome Domain
by: Khettari, Oumaima El, et al.
Published: (2025)
by: Khettari, Oumaima El, et al.
Published: (2025)
Leveraging Long-Context Large Language Models for Multi-Document Understanding and Summarization in Enterprise Applications
by: Godbole, Aditi, et al.
Published: (2024)
by: Godbole, Aditi, et al.
Published: (2024)
Fine-Tuned Language Models for Domain-Specific Summarization and Tagging
by: Wang, Jun, et al.
Published: (2025)
by: Wang, Jun, et al.
Published: (2025)
Mitigating Hallucination in Abstractive Summarization with Domain-Conditional Mutual Information
by: Chae, Kyubyung, et al.
Published: (2024)
by: Chae, Kyubyung, et al.
Published: (2024)
EROS: Entity-Driven Controlled Policy Document Summarization
by: Singh, Joykirat, et al.
Published: (2024)
by: Singh, Joykirat, et al.
Published: (2024)
Diversity Measures: Domain-Independent Proxies for Failure in Language Model Queries
by: Ngu, Noel, et al.
Published: (2023)
by: Ngu, Noel, et al.
Published: (2023)
Accelerating Scientific Discovery with Multi-Document Summarization of Impact-Ranked Papers
by: Koloveas, Paris, et al.
Published: (2025)
by: Koloveas, Paris, et al.
Published: (2025)
End-to-End Long Document Summarization using Gradient Caching
by: Saxena, Rohit, et al.
Published: (2025)
by: Saxena, Rohit, et al.
Published: (2025)
Key-Element-Informed sLLM Tuning for Document Summarization
by: Ryu, Sangwon, et al.
Published: (2024)
by: Ryu, Sangwon, et al.
Published: (2024)
Can Large Language Model Summarizers Adapt to Diverse Scientific Communication Goals?
by: Fonseca, Marcio, et al.
Published: (2024)
by: Fonseca, Marcio, et al.
Published: (2024)
Coverage-based Fairness in Multi-document Summarization
by: Li, Haoyuan, et al.
Published: (2024)
by: Li, Haoyuan, et al.
Published: (2024)
Discourse-Driven Evaluation: Unveiling Factual Inconsistency in Long Document Summarization
by: Zhong, Yang, et al.
Published: (2025)
by: Zhong, Yang, et al.
Published: (2025)
Revolutionizing API Documentation through Summarization
by: Naghshzan, AmirHossein, et al.
Published: (2024)
by: Naghshzan, AmirHossein, et al.
Published: (2024)
Evaluation of Large Language Models for Summarization Tasks in the Medical Domain: A Narrative Review
by: Croxford, Emma, et al.
Published: (2024)
by: Croxford, Emma, et al.
Published: (2024)
Can Small Language Models Handle Context-Summarized Multi-Turn Customer-Service QA? A Synthetic Data-Driven Comparative Evaluation
by: Cooray, Lakshan, et al.
Published: (2026)
by: Cooray, Lakshan, et al.
Published: (2026)
Can GPT models Follow Human Summarization Guidelines? A Study for Targeted Communication Goals
by: Zhou, Yongxin, et al.
Published: (2023)
by: Zhou, Yongxin, et al.
Published: (2023)
PreSumm: Predicting Summarization Performance Without Summarizing
by: Koniaev, Steven, et al.
Published: (2025)
by: Koniaev, Steven, et al.
Published: (2025)
Multi-Label Clinical Text Eligibility Classification and Summarization System
by: Yerramsetty, Surya Tejaswi, et al.
Published: (2025)
by: Yerramsetty, Surya Tejaswi, et al.
Published: (2025)
Rethinking Transformer-based Multi-document Summarization: An Empirical Investigation
by: Ma, Congbo, et al.
Published: (2024)
by: Ma, Congbo, et al.
Published: (2024)
Understanding LLM Behavior in Multi-Target Cross-Lingual Summarization
by: Ryu, Sangwon, et al.
Published: (2026)
by: Ryu, Sangwon, et al.
Published: (2026)
Multi-Dimensional Optimization for Text Summarization via Reinforcement Learning
by: Ryu, Sangwon, et al.
Published: (2024)
by: Ryu, Sangwon, et al.
Published: (2024)
GraphLSS: Integrating Lexical, Structural, and Semantic Features for Long Document Extractive Summarization
by: Bugueño, Margarita, et al.
Published: (2024)
by: Bugueño, Margarita, et al.
Published: (2024)
AI and Generative AI for Research Discovery and Summarization
by: Glickman, Mark, et al.
Published: (2024)
by: Glickman, Mark, et al.
Published: (2024)
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
by: Ramprasad, Sanjana, et al.
Published: (2024)
by: Ramprasad, Sanjana, et al.
Published: (2024)
Similar Items
-
Anti-LM Decoding for Zero-shot In-context Machine Translation
by: Sia, Suzanna, et al.
Published: (2023) -
MedScore: Generalizable Factuality Evaluation of Free-Form Medical Answers by Domain-adapted Claim Decomposition and Verification
by: Huang, Heyuan, et al.
Published: (2025) -
Using Natural Language Inference to Improve Persona Extraction from Dialogue in a New Domain
by: DeLucia, Alexandra, et al.
Published: (2024) -
Same Verdict, Different Reasons: LLM-as-a-Judge and Clinician Disagreement on Medical Chatbot Completeness
by: DeLucia, Alexandra, et al.
Published: (2026) -
On the Failure of Latent State Persistence in Large Language Models
by: Huang, Jen-tse, et al.
Published: (2025)