Do Multi-Document Summarization Models Synthesize?
Fuente:
arXiv
Guardado en:
| Autores principales: | DeYoung, Jay, Martinez, Stephanie C., Marshall, Iain J., Wallace, Byron C. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Automatically Extracting Numerical Results from Randomized Controlled Trials with Large Language Models
por: Yun, Hye Sun, et al.
Publicado: (2024)
por: Yun, Hye Sun, et al.
Publicado: (2024)
Caught in the Web of Words: Do LLMs Fall for Spin in Medical Literature?
por: Yun, Hye Sun, et al.
Publicado: (2025)
por: Yun, Hye Sun, et al.
Publicado: (2025)
How Much Annotation is Needed to Compare Summarization Models?
por: Shaib, Chantal, et al.
Publicado: (2024)
por: Shaib, Chantal, et al.
Publicado: (2024)
Do Automatic Factuality Metrics Measure Factuality? A Critical Evaluation
por: Ramprasad, Sanjana, et al.
Publicado: (2024)
por: Ramprasad, Sanjana, et al.
Publicado: (2024)
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
por: Ramprasad, Sanjana, et al.
Publicado: (2024)
por: Ramprasad, Sanjana, et al.
Publicado: (2024)
Revisiting Relation Extraction in the era of Large Language Models
por: Wadhwa, Somin, et al.
Publicado: (2023)
por: Wadhwa, Somin, et al.
Publicado: (2023)
Can SAEs reveal and mitigate racial biases of LLMs in healthcare?
por: Ahsan, Hiba, et al.
Publicado: (2025)
por: Ahsan, Hiba, et al.
Publicado: (2025)
FactPICO: Factuality Evaluation for Plain Language Summarization of Medical Evidence
por: Joseph, Sebastian Antony, et al.
Publicado: (2024)
por: Joseph, Sebastian Antony, et al.
Publicado: (2024)
Decide less, communicate more: On the construct validity of end-to-end fact-checking in medicine
por: Joseph, Sebastian, et al.
Publicado: (2025)
por: Joseph, Sebastian, et al.
Publicado: (2025)
Who Taught You That? Tracing Teachers in Model Distillation
por: Wadhwa, Somin, et al.
Publicado: (2025)
por: Wadhwa, Somin, et al.
Publicado: (2025)
Circuit Distillation
por: Wadhwa, Somin, et al.
Publicado: (2025)
por: Wadhwa, Somin, et al.
Publicado: (2025)
Investigating Mysteries of CoT-Augmented Distillation
por: Wadhwa, Somin, et al.
Publicado: (2024)
por: Wadhwa, Somin, et al.
Publicado: (2024)
Do Activation Verbalization Methods Convey Privileged Information?
por: Li, Millicent, et al.
Publicado: (2025)
por: Li, Millicent, et al.
Publicado: (2025)
Question answering systems for health professionals at the point of care -- a systematic review
por: Kell, Gregory, et al.
Publicado: (2024)
por: Kell, Gregory, et al.
Publicado: (2024)
Leveraging ChatGPT in Pharmacovigilance Event Extraction: An Empirical Study
por: Sun, Zhaoyue, et al.
Publicado: (2024)
por: Sun, Zhaoyue, et al.
Publicado: (2024)
Learning the Wrong Lessons: Syntactic-Domain Spurious Correlations in Language Models
por: Shaib, Chantal, et al.
Publicado: (2025)
por: Shaib, Chantal, et al.
Publicado: (2025)
From Single to Multi: How LLMs Hallucinate in Multi-Document Summarization
por: Belem, Catarina G., et al.
Publicado: (2024)
por: Belem, Catarina G., et al.
Publicado: (2024)
Vector Arithmetic in Concept and Token Subspaces
por: Feucht, Sheridan, et al.
Publicado: (2025)
por: Feucht, Sheridan, et al.
Publicado: (2025)
GLIMPSE: Pragmatically Informative Multi-Document Summarization for Scholarly Reviews
por: Darrin, Maxime, et al.
Publicado: (2024)
por: Darrin, Maxime, et al.
Publicado: (2024)
Measuring AI "Slop" in Text
por: Shaib, Chantal, et al.
Publicado: (2025)
por: Shaib, Chantal, et al.
Publicado: (2025)
Detection and Measurement of Syntactic Templates in Generated Text
por: Shaib, Chantal, et al.
Publicado: (2024)
por: Shaib, Chantal, et al.
Publicado: (2024)
Can one size fit all?: Measuring Failure in Multi-Document Summarization Domain Transfer
por: DeLucia, Alexandra, et al.
Publicado: (2025)
por: DeLucia, Alexandra, et al.
Publicado: (2025)
Compared to What? Baselines and Metrics for Counterfactual Prompting
por: Yang, Zihao, et al.
Publicado: (2026)
por: Yang, Zihao, et al.
Publicado: (2026)
Don't Pay Attention, PLANT It: Pretraining Attention via Learning-to-Rank
por: Roy, Debjyoti Saha, et al.
Publicado: (2024)
por: Roy, Debjyoti Saha, et al.
Publicado: (2024)
Topic-Guided Reinforcement Learning with LLMs for Enhancing Multi-Document Summarization
por: Li, Chuyuan, et al.
Publicado: (2025)
por: Li, Chuyuan, et al.
Publicado: (2025)
Input Order Shapes LLM Semantic Alignment in Multi-Document Summarization
por: Ma, Jing
Publicado: (2025)
por: Ma, Jing
Publicado: (2025)
GLIMMER: Incorporating Graph and Lexical Features in Unsupervised Multi-Document Summarization
por: Liu, Ran, et al.
Publicado: (2024)
por: Liu, Ran, et al.
Publicado: (2024)
Document Summarization with Conformal Importance Guarantees
por: Kuwahara, Bruce, et al.
Publicado: (2025)
por: Kuwahara, Bruce, et al.
Publicado: (2025)
BERT-VBD: Vietnamese Multi-Document Summarization Framework
por: Vuong, Tuan-Cuong, et al.
Publicado: (2024)
por: Vuong, Tuan-Cuong, et al.
Publicado: (2024)
Learning from Natural Language Explanations for Generalizable Entity Matching
por: Wadhwa, Somin, et al.
Publicado: (2024)
por: Wadhwa, Somin, et al.
Publicado: (2024)
Distribution-Aware Reward: Reinforcement Learning over Predictive Distributions for LLM Regression
por: Park, Jungsoo, et al.
Publicado: (2026)
por: Park, Jungsoo, et al.
Publicado: (2026)
Improving Attributed Long-form Question Answering with Intent Awareness
por: Zhao, Xinran, et al.
Publicado: (2026)
por: Zhao, Xinran, et al.
Publicado: (2026)
Large Language Models for Summarizing Czech Historical Documents and Beyond
por: Tran, Václav, et al.
Publicado: (2025)
por: Tran, Václav, et al.
Publicado: (2025)
Retrieving Evidence from EHRs with LLMs: Possibilities and Challenges
por: Ahsan, Hiba, et al.
Publicado: (2023)
por: Ahsan, Hiba, et al.
Publicado: (2023)
LightPAL: Lightweight Passage Retrieval for Open Domain Multi-Document Summarization
por: Enomoto, Masafumi, et al.
Publicado: (2024)
por: Enomoto, Masafumi, et al.
Publicado: (2024)
Learning to Summarize by Learning to Quiz: Adversarial Agentic Collaboration for Long Document Summarization
por: Wang, Weixuan, et al.
Publicado: (2025)
por: Wang, Weixuan, et al.
Publicado: (2025)
Cross-Document Event-Keyed Summarization
por: Walden, William, et al.
Publicado: (2024)
por: Walden, William, et al.
Publicado: (2024)
GenAudit: Fixing Factual Errors in Language Model Outputs with Evidence
por: Krishna, Kundan, et al.
Publicado: (2024)
por: Krishna, Kundan, et al.
Publicado: (2024)
Elucidating Mechanisms of Demographic Bias in LLMs for Healthcare
por: Ahsan, Hiba, et al.
Publicado: (2025)
por: Ahsan, Hiba, et al.
Publicado: (2025)
A Mixed-Language Multi-Document News Summarization Dataset and a Graphs-Based Extract-Generate Model
por: Gao, Shengxiang, et al.
Publicado: (2024)
por: Gao, Shengxiang, et al.
Publicado: (2024)
Ejemplares similares
-
Automatically Extracting Numerical Results from Randomized Controlled Trials with Large Language Models
por: Yun, Hye Sun, et al.
Publicado: (2024) -
Caught in the Web of Words: Do LLMs Fall for Spin in Medical Literature?
por: Yun, Hye Sun, et al.
Publicado: (2025) -
How Much Annotation is Needed to Compare Summarization Models?
por: Shaib, Chantal, et al.
Publicado: (2024) -
Do Automatic Factuality Metrics Measure Factuality? A Critical Evaluation
por: Ramprasad, Sanjana, et al.
Publicado: (2024) -
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
por: Ramprasad, Sanjana, et al.
Publicado: (2024)