The State and Fate of Summarization Datasets: A Survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dahan, Noam, Stanovsky, Gabriel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Leveraging Digitized Newspapers to Collect Summarization Data in Low-Resource Languages
von: Dahan, Noam, et al.
Veröffentlicht: (2025)
von: Dahan, Noam, et al.
Veröffentlicht: (2025)
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation
von: Habba, Eliya, et al.
Veröffentlicht: (2025)
von: Habba, Eliya, et al.
Veröffentlicht: (2025)
Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition
von: Goldstein, Ariel, et al.
Veröffentlicht: (2024)
von: Goldstein, Ariel, et al.
Veröffentlicht: (2024)
Comparing Humans and Models on a Similar Scale: Towards Cognitive Gender Bias Evaluation in Coreference Resolution
von: Lior, Gili, et al.
Veröffentlicht: (2023)
von: Lior, Gili, et al.
Veröffentlicht: (2023)
Surveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy, Trends and Metrics Analysis
von: Berger, Uri, et al.
Veröffentlicht: (2024)
von: Berger, Uri, et al.
Veröffentlicht: (2024)
In-Context Learning on a Budget: A Case Study in Token Classification
von: Berger, Uri, et al.
Veröffentlicht: (2024)
von: Berger, Uri, et al.
Veröffentlicht: (2024)
Leveraging Collection-Wide Similarities for Unsupervised Document Structure Extraction
von: Lior, Gili, et al.
Veröffentlicht: (2024)
von: Lior, Gili, et al.
Veröffentlicht: (2024)
Beyond Memorization: Distinguishing between Reductive and Epistemic Reasoning in LLMs using Classic Logic Puzzles
von: Gabay, Adi, et al.
Veröffentlicht: (2026)
von: Gabay, Adi, et al.
Veröffentlicht: (2026)
State of What Art? A Call for Multi-Prompt LLM Evaluation
von: Mizrahi, Moran, et al.
Veröffentlicht: (2023)
von: Mizrahi, Moran, et al.
Veröffentlicht: (2023)
SurveySum: A Dataset for Summarizing Multiple Scientific Articles into a Survey Section
von: Fernandes, Leandro Carísio, et al.
Veröffentlicht: (2024)
von: Fernandes, Leandro Carísio, et al.
Veröffentlicht: (2024)
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
von: Lior, Gili, et al.
Veröffentlicht: (2025)
von: Lior, Gili, et al.
Veröffentlicht: (2025)
Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs
von: Itzhak, Itay, et al.
Veröffentlicht: (2025)
von: Itzhak, Itay, et al.
Veröffentlicht: (2025)
Applying Intrinsic Debiasing on Downstream Tasks: Challenges and Considerations for Machine Translation
von: Iluz, Bar, et al.
Veröffentlicht: (2024)
von: Iluz, Bar, et al.
Veröffentlicht: (2024)
Time to Talk: LLM Agents for Asynchronous Group Communication in Mafia Games
von: Eckhaus, Niv, et al.
Veröffentlicht: (2025)
von: Eckhaus, Niv, et al.
Veröffentlicht: (2025)
DOVE: A Large-Scale Multi-Dimensional Predictions Dataset Towards Meaningful LLM Evaluation
von: Habba, Eliya, et al.
Veröffentlicht: (2025)
von: Habba, Eliya, et al.
Veröffentlicht: (2025)
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments
von: Lior, Gili, et al.
Veröffentlicht: (2025)
von: Lior, Gili, et al.
Veröffentlicht: (2025)
Improving Image Captioning by Mimicking Human Reformulation Feedback at Inference-time
von: Berger, Uri, et al.
Veröffentlicht: (2025)
von: Berger, Uri, et al.
Veröffentlicht: (2025)
Looking Beyond The Top-1: Transformers Determine Top Tokens In Order
von: Lioubashevski, Daria, et al.
Veröffentlicht: (2024)
von: Lioubashevski, Daria, et al.
Veröffentlicht: (2024)
What Am I Missing? Question-Answering as Hidden State Probing
von: Luo, Chu Fei, et al.
Veröffentlicht: (2026)
von: Luo, Chu Fei, et al.
Veröffentlicht: (2026)
More Documents, Same Length: Isolating the Challenge of Multiple Documents in RAG
von: Levy, Shahar, et al.
Veröffentlicht: (2025)
von: Levy, Shahar, et al.
Veröffentlicht: (2025)
Generating Query-Focused Summarization Datasets from Query-Free Summarization Datasets
von: Chali, Yllias, et al.
Veröffentlicht: (2026)
von: Chali, Yllias, et al.
Veröffentlicht: (2026)
Can LLMs Help Uncover Insights about LLMs? A Large-Scale, Evolving Literature Analysis of Frontier LLMs
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
SEAM: A Stochastic Benchmark for Multi-Document Tasks
von: Lior, Gili, et al.
Veröffentlicht: (2024)
von: Lior, Gili, et al.
Veröffentlicht: (2024)
A Dataset and Benchmark for Consumer Healthcare Question Summarization
von: Basu, Abhishek, et al.
Veröffentlicht: (2025)
von: Basu, Abhishek, et al.
Veröffentlicht: (2025)
Anticipatory Evaluation of Language Models
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
Beyond Benchmarks: On The False Promise of AI Regulation
von: Stanovsky, Gabriel, et al.
Veröffentlicht: (2025)
von: Stanovsky, Gabriel, et al.
Veröffentlicht: (2025)
From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs
von: Itzhak, Itay, et al.
Veröffentlicht: (2026)
von: Itzhak, Itay, et al.
Veröffentlicht: (2026)
ScheMatiQ: From Research Question to Structured Data through Interactive Schema Discovery
von: Levy, Shahar, et al.
Veröffentlicht: (2026)
von: Levy, Shahar, et al.
Veröffentlicht: (2026)
Schema-Driven Information Extraction from Heterogeneous Tables
von: Bai, Fan, et al.
Veröffentlicht: (2023)
von: Bai, Fan, et al.
Veröffentlicht: (2023)
A Comprehensive Survey on Legal Summarization: Challenges and Future Directions
von: Akter, Mousumi, et al.
Veröffentlicht: (2025)
von: Akter, Mousumi, et al.
Veröffentlicht: (2025)
PERCS: Persona-Guided Controllable Biomedical Summarization Dataset
von: Salvi, Rohan Charudatt, et al.
Veröffentlicht: (2025)
von: Salvi, Rohan Charudatt, et al.
Veröffentlicht: (2025)
Survey of Query-based Text Summarization
von: Yu, Hang, et al.
Veröffentlicht: (2022)
von: Yu, Hang, et al.
Veröffentlicht: (2022)
CTISum: A New Benchmark Dataset For Cyber Threat Intelligence Summarization
von: Peng, Wei, et al.
Veröffentlicht: (2024)
von: Peng, Wei, et al.
Veröffentlicht: (2024)
ACLSum: A New Dataset for Aspect-based Summarization of Scientific Publications
von: Takeshita, Sotaro, et al.
Veröffentlicht: (2024)
von: Takeshita, Sotaro, et al.
Veröffentlicht: (2024)
A Nurse is Blue and Elephant is Rugby: Cross Domain Alignment in Large Language Models Reveal Human-like Patterns
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
Summarizing Speech: A Comprehensive Survey
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
SAUCE: Synchronous and Asynchronous User-Customizable Environment for Multi-Agent LLM Interaction
von: Neuberger, Shlomo, et al.
Veröffentlicht: (2024)
von: Neuberger, Shlomo, et al.
Veröffentlicht: (2024)
Russian-Language Multimodal Dataset for Automatic Summarization of Scientific Papers
von: Tsanda, Alena, et al.
Veröffentlicht: (2024)
von: Tsanda, Alena, et al.
Veröffentlicht: (2024)
Cross-lingual Cross-temporal Summarization: Dataset, Models, Evaluation
von: Zhang, Ran, et al.
Veröffentlicht: (2023)
von: Zhang, Ran, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Leveraging Digitized Newspapers to Collect Summarization Data in Low-Resource Languages
von: Dahan, Noam, et al.
Veröffentlicht: (2025) -
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation
von: Habba, Eliya, et al.
Veröffentlicht: (2025) -
Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition
von: Goldstein, Ariel, et al.
Veröffentlicht: (2024) -
Comparing Humans and Models on a Similar Scale: Towards Cognitive Gender Bias Evaluation in Coreference Resolution
von: Lior, Gili, et al.
Veröffentlicht: (2023) -
Surveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy, Trends and Metrics Analysis
von: Berger, Uri, et al.
Veröffentlicht: (2024)