CASPR: Automated Evaluation Metric for Contrastive Summarization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ananthamurugan, Nirupan, Duong, Dat, George, Philip, Gupta, Ankita, Tata, Sandeep, Gunel, Beliz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SUMIE: A Synthetic Benchmark for Incremental Entity Summarization
von: Hwang, Eunjeong, et al.
Veröffentlicht: (2024)
von: Hwang, Eunjeong, et al.
Veröffentlicht: (2024)
STRUM-LLM: Attributed and Structured Contrastive Summarization
von: Gunel, Beliz, et al.
Veröffentlicht: (2024)
von: Gunel, Beliz, et al.
Veröffentlicht: (2024)
Enhancing Incremental Summarization with Structured Representations
von: Hwang, EunJeong, et al.
Veröffentlicht: (2024)
von: Hwang, EunJeong, et al.
Veröffentlicht: (2024)
PRISM: Efficient Long-Range Reasoning With Short-Context LLMs
von: Jayalath, Dulhan, et al.
Veröffentlicht: (2024)
von: Jayalath, Dulhan, et al.
Veröffentlicht: (2024)
An Automated Length-Aware Quality Metric for Summarization
von: Foland, Andrew D.
Veröffentlicht: (2025)
von: Foland, Andrew D.
Veröffentlicht: (2025)
APPLS: Evaluating Evaluation Metrics for Plain Language Summarization
von: Guo, Yue, et al.
Veröffentlicht: (2023)
von: Guo, Yue, et al.
Veröffentlicht: (2023)
Calibrating Model-Based Evaluation Metrics for Summarization
von: Liu, Hongye, et al.
Veröffentlicht: (2026)
von: Liu, Hongye, et al.
Veröffentlicht: (2026)
Predicting Task Performance with Context-aware Scaling Laws
von: Montgomery, Kyle, et al.
Veröffentlicht: (2025)
von: Montgomery, Kyle, et al.
Veröffentlicht: (2025)
Q-STRUM Debate: Query-Driven Contrastive Summarization for Recommendation Comparison
von: Saad, George-Kirollos, et al.
Veröffentlicht: (2025)
von: Saad, George-Kirollos, et al.
Veröffentlicht: (2025)
What's under the hood: Investigating Automatic Metrics on Meeting Summarization
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
Rethinking Scientific Summarization Evaluation: Grounding Explainable Metrics on Facet-aware Benchmark
von: Chen, Xiuying, et al.
Veröffentlicht: (2024)
von: Chen, Xiuying, et al.
Veröffentlicht: (2024)
Beyond N-Grams: Rethinking Evaluation Metrics and Strategies for Multilingual Abstractive Summarization
von: Mondshine, Itai, et al.
Veröffentlicht: (2025)
von: Mondshine, Itai, et al.
Veröffentlicht: (2025)
MLAN: Language-Based Instruction Tuning Preserves and Transfers Knowledge in Multimodal Language Models
von: Tu, Jianhong, et al.
Veröffentlicht: (2024)
von: Tu, Jianhong, et al.
Veröffentlicht: (2024)
Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
DeepScholar-Bench: A Live Benchmark and Automated Evaluation for Generative Research Synthesis
von: Patel, Liana, et al.
Veröffentlicht: (2025)
von: Patel, Liana, et al.
Veröffentlicht: (2025)
ContrastScore: Towards Higher Quality, Less Biased, More Efficient Evaluation Metrics with Contrastive Evaluation
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
PlainQAFact: Retrieval-augmented Factual Consistency Evaluation Metric for Biomedical Plain Language Summarization
von: You, Zhiwen, et al.
Veröffentlicht: (2025)
von: You, Zhiwen, et al.
Veröffentlicht: (2025)
Faithful Model Evaluation for Model-Based Metrics
von: Goyal, Palash, et al.
Veröffentlicht: (2023)
von: Goyal, Palash, et al.
Veröffentlicht: (2023)
Discrete Diffusion Language Model for Efficient Text Summarization
von: Dat, Do Huu, et al.
Veröffentlicht: (2024)
von: Dat, Do Huu, et al.
Veröffentlicht: (2024)
Enhancing Argument Summarization: Prioritizing Exhaustiveness in Key Point Generation and Introducing an Automatic Coverage Evaluation Metric
von: Khosravani, Mohammad, et al.
Veröffentlicht: (2024)
von: Khosravani, Mohammad, et al.
Veröffentlicht: (2024)
Evaluating Metrics for Bias in Word Embeddings
von: Schröder, Sarah, et al.
Veröffentlicht: (2021)
von: Schröder, Sarah, et al.
Veröffentlicht: (2021)
Mitigating the Impact of Reference Quality on Evaluation of Summarization Systems with Reference-Free Metrics
von: Gigant, Théo, et al.
Veröffentlicht: (2024)
von: Gigant, Théo, et al.
Veröffentlicht: (2024)
Medical Question Summarization with Entity-driven Contrastive Learning
von: Lu, Wenpeng, et al.
Veröffentlicht: (2023)
von: Lu, Wenpeng, et al.
Veröffentlicht: (2023)
NovAScore: A New Automated Metric for Evaluating Document Level Novelty
von: Ai, Lin, et al.
Veröffentlicht: (2024)
von: Ai, Lin, et al.
Veröffentlicht: (2024)
Iterative Augmentation with Summarization Refinement (IASR) Evaluation for Unstructured Survey data Modeling and Analysis
von: Bhattad, Payal, et al.
Veröffentlicht: (2025)
von: Bhattad, Payal, et al.
Veröffentlicht: (2025)
An Analysis on Automated Metrics for Evaluating Japanese-English Chat Translation
von: Rusli, Andre, et al.
Veröffentlicht: (2024)
von: Rusli, Andre, et al.
Veröffentlicht: (2024)
SteerEval: Inference-time Interventions Strengthen Multilingual Generalization in Neural Summarization Metrics
von: Casola, Silvia, et al.
Veröffentlicht: (2026)
von: Casola, Silvia, et al.
Veröffentlicht: (2026)
Legal Document Summarization: Enhancing Judicial Efficiency through Automation Detection
von: Li, Yongjie, et al.
Veröffentlicht: (2025)
von: Li, Yongjie, et al.
Veröffentlicht: (2025)
Improving Factual Consistency of News Summarization by Contrastive Preference Optimization
von: Feng, Huawen, et al.
Veröffentlicht: (2023)
von: Feng, Huawen, et al.
Veröffentlicht: (2023)
BioACE: An Automated Framework for Biomedical Answer and Citation Evaluations
von: Gupta, Deepak, et al.
Veröffentlicht: (2026)
von: Gupta, Deepak, et al.
Veröffentlicht: (2026)
A Dataset and Benchmark for Consumer Healthcare Question Summarization
von: Basu, Abhishek, et al.
Veröffentlicht: (2025)
von: Basu, Abhishek, et al.
Veröffentlicht: (2025)
Medalyze: Lightweight Medical Report Summarization Application Using FLAN-T5-Large
von: Nguyen, Van-Tinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Van-Tinh, et al.
Veröffentlicht: (2025)
Moneyball with LLMs: Analyzing Tabular Summarization in Sports Narratives
von: Upadhyay, Ritam, et al.
Veröffentlicht: (2025)
von: Upadhyay, Ritam, et al.
Veröffentlicht: (2025)
What do the metrics mean? A critical analysis of the use of Automated Evaluation Metrics in Interpreting
von: Downie, Jonathan, et al.
Veröffentlicht: (2026)
von: Downie, Jonathan, et al.
Veröffentlicht: (2026)
Stress Testing Factual Consistency Metrics for Long-Document Summarization
von: Mujahid, Zain Muhammad, et al.
Veröffentlicht: (2025)
von: Mujahid, Zain Muhammad, et al.
Veröffentlicht: (2025)
On the Role of Summary Content Units in Text Summarization Evaluation
von: Nawrath, Marcel, et al.
Veröffentlicht: (2024)
von: Nawrath, Marcel, et al.
Veröffentlicht: (2024)
Fine-grained and Explainable Factuality Evaluation for Multimodal Summarization
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
PSentScore: Evaluating Sentiment Polarity in Dialogue Summarization
von: Zhou, Yongxin, et al.
Veröffentlicht: (2023)
von: Zhou, Yongxin, et al.
Veröffentlicht: (2023)
Multi-Dimensional Evaluation of Text Summarization with In-Context Learning
von: Jain, Sameer, et al.
Veröffentlicht: (2023)
von: Jain, Sameer, et al.
Veröffentlicht: (2023)
mFACE: Multilingual Summarization with Factual Consistency Evaluation
von: Aharoni, Roee, et al.
Veröffentlicht: (2022)
von: Aharoni, Roee, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
SUMIE: A Synthetic Benchmark for Incremental Entity Summarization
von: Hwang, Eunjeong, et al.
Veröffentlicht: (2024) -
STRUM-LLM: Attributed and Structured Contrastive Summarization
von: Gunel, Beliz, et al.
Veröffentlicht: (2024) -
Enhancing Incremental Summarization with Structured Representations
von: Hwang, EunJeong, et al.
Veröffentlicht: (2024) -
PRISM: Efficient Long-Range Reasoning With Short-Context LLMs
von: Jayalath, Dulhan, et al.
Veröffentlicht: (2024) -
An Automated Length-Aware Quality Metric for Summarization
von: Foland, Andrew D.
Veröffentlicht: (2025)