Blind to the Human Touch: Overlap Bias in LLM-Based Summary Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Fang, Jiangnan, Liu, Cheng-Tse, Deilamsalehy, Hanieh, Ahmed, Nesreen K., Mathur, Puneet, Lipka, Nedim, Dernoncourt, Franck, Rossi, Ryan A. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-LLM Text Summarization
by: Fang, Jiangnan, et al.
Published: (2024)
by: Fang, Jiangnan, et al.
Published: (2024)
Structured Uncertainty guided Clarification for LLM Agents
by: Suri, Manan, et al.
Published: (2025)
by: Suri, Manan, et al.
Published: (2025)
ChartLens: Fine-grained Visual Attribution in Charts
by: Suri, Manan, et al.
Published: (2025)
by: Suri, Manan, et al.
Published: (2025)
Trust but Verify: Introducing DAVinCI -- A Framework for Dual Attribution and Verification in Claim Inference for Language Models
by: Rawte, Vipula, et al.
Published: (2026)
by: Rawte, Vipula, et al.
Published: (2026)
Follow the Flow: Fine-grained Flowchart Attribution with Neurosymbolic Agents
by: Suri, Manan, et al.
Published: (2025)
by: Suri, Manan, et al.
Published: (2025)
PlotGen: Multi-Agent LLM-based Scientific Data Visualization via Multimodal Feedback
by: Goswami, Kanika, et al.
Published: (2025)
by: Goswami, Kanika, et al.
Published: (2025)
PlotEdit: Natural Language-Driven Accessible Chart Editing in PDFs via Multimodal LLM Agents
by: Goswami, Kanika, et al.
Published: (2025)
by: Goswami, Kanika, et al.
Published: (2025)
A Multi-LLM Debiasing Framework
by: Owens, Deonna M., et al.
Published: (2024)
by: Owens, Deonna M., et al.
Published: (2024)
Document Attribution: Examining Citation Relationships using Large Language Models
by: Rawte, Vipula, et al.
Published: (2025)
by: Rawte, Vipula, et al.
Published: (2025)
Cluster-R1: Large Reasoning Models Are Instruction-following Clustering Agents
by: Qing, Peijun, et al.
Published: (2026)
by: Qing, Peijun, et al.
Published: (2026)
ChartCitor: Multi-Agent Framework for Fine-Grained Chart Visual Attribution
by: Goswami, Kanika, et al.
Published: (2025)
by: Goswami, Kanika, et al.
Published: (2025)
MODS: Moderating a Mixture of Document Speakers to Summarize Debatable Queries in Document Collections
by: Balepur, Nishant, et al.
Published: (2025)
by: Balepur, Nishant, et al.
Published: (2025)
Towards Enhancing Coherence in Extractive Summarization: Dataset and Experiments with LLMs
by: Parmar, Mihir, et al.
Published: (2024)
by: Parmar, Mihir, et al.
Published: (2024)
NoLiMa: Long-Context Evaluation Beyond Literal Matching
by: Modarressi, Ali, et al.
Published: (2025)
by: Modarressi, Ali, et al.
Published: (2025)
Test-Time Strategies for More Efficient and Accurate Agentic RAG
by: Zhang, Brian, et al.
Published: (2026)
by: Zhang, Brian, et al.
Published: (2026)
Personalized Graph-Based Retrieval for Large Language Models
by: Au, Steven, et al.
Published: (2025)
by: Au, Steven, et al.
Published: (2025)
Sparse Personalized Text Generation with Multi-Trajectory Reasoning
by: Ni, Bo, et al.
Published: (2026)
by: Ni, Bo, et al.
Published: (2026)
OATS: Opinion Aspect Target Sentiment Quadruple Extraction Dataset for Aspect-Based Sentiment Analysis
by: Chebolu, Siva Uday Sampreeth, et al.
Published: (2023)
by: Chebolu, Siva Uday Sampreeth, et al.
Published: (2023)
Survey of User Interface Design and Interaction Techniques in Generative AI Applications
by: Luera, Reuben, et al.
Published: (2024)
by: Luera, Reuben, et al.
Published: (2024)
Steering MoE LLMs via Expert (De)Activation
by: Fayyaz, Mohsen, et al.
Published: (2025)
by: Fayyaz, Mohsen, et al.
Published: (2025)
ROAST: Review-level Opinion Aspect Sentiment Target Joint Detection for ABSA
by: Chebolu, Siva Uday Sampreeth, et al.
Published: (2024)
by: Chebolu, Siva Uday Sampreeth, et al.
Published: (2024)
A Framework for Fine-Tuning LLMs using Heterogeneous Feedback
by: Aponte, Ryan, et al.
Published: (2024)
by: Aponte, Ryan, et al.
Published: (2024)
VisDoM: Multi-Document QA with Visually Rich Elements Using Multimodal Retrieval-Augmented Generation
by: Suri, Manan, et al.
Published: (2024)
by: Suri, Manan, et al.
Published: (2024)
Mixture of Structural-and-Textual Retrieval over Text-rich Graph Knowledge Bases
by: Lei, Yongjia, et al.
Published: (2025)
by: Lei, Yongjia, et al.
Published: (2025)
DynaSaur: Large Language Agents Beyond Predefined Actions
by: Nguyen, Dang, et al.
Published: (2024)
by: Nguyen, Dang, et al.
Published: (2024)
Taipan: Efficient and Expressive State Space Language Models with Selective Attention
by: Van Nguyen, Chien, et al.
Published: (2024)
by: Van Nguyen, Chien, et al.
Published: (2024)
LongLaMP: A Benchmark for Personalized Long-form Text Generation
by: Kumar, Ishita, et al.
Published: (2024)
by: Kumar, Ishita, et al.
Published: (2024)
MLLM as a UI Judge: Benchmarking Multimodal LLMs for Predicting Human Perception of User Interfaces
by: Luera, Reuben A., et al.
Published: (2025)
by: Luera, Reuben A., et al.
Published: (2025)
Lizard: An Efficient Linearization Framework for Large Language Models
by: Van Nguyen, Chien, et al.
Published: (2025)
by: Van Nguyen, Chien, et al.
Published: (2025)
FigCaps-HF: A Figure-to-Caption Generative Framework and Benchmark with Human Feedback
by: Singh, Ashish, et al.
Published: (2023)
by: Singh, Ashish, et al.
Published: (2023)
Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models
by: Nguyen, Minh, et al.
Published: (2024)
by: Nguyen, Minh, et al.
Published: (2024)
Knowledge Homophily in Large Language Models
by: Sahu, Utkarsh, et al.
Published: (2025)
by: Sahu, Utkarsh, et al.
Published: (2025)
Charts Are Not Images: On the Challenges of Scientific Chart Editing
by: Li, Shawn, et al.
Published: (2025)
by: Li, Shawn, et al.
Published: (2025)
Iterative Critique-Refine Framework for Enhancing LLM Personalization
by: Maram, Durga Prasad, et al.
Published: (2025)
by: Maram, Durga Prasad, et al.
Published: (2025)
Bias and Fairness in Large Language Models: A Survey
by: Gallegos, Isabel O., et al.
Published: (2023)
by: Gallegos, Isabel O., et al.
Published: (2023)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
by: Gallegos, Isabel O., et al.
Published: (2024)
by: Gallegos, Isabel O., et al.
Published: (2024)
Efficient Tree-Structured Deep Research with Adaptive Resource Allocation
by: Nie, Lunyiu, et al.
Published: (2025)
by: Nie, Lunyiu, et al.
Published: (2025)
Quantitative LLM Judges
by: Sahoo, Aishwarya, et al.
Published: (2025)
by: Sahoo, Aishwarya, et al.
Published: (2025)
VipAct: Visual-Perception Enhancement via Specialized VLM Agent Collaboration and Tool-use
by: Zhang, Zhehao, et al.
Published: (2024)
by: Zhang, Zhehao, et al.
Published: (2024)
Reasoning-Based Personalized Generation for Users with Sparse Data
by: Ni, Bo, et al.
Published: (2026)
by: Ni, Bo, et al.
Published: (2026)
Similar Items
-
Multi-LLM Text Summarization
by: Fang, Jiangnan, et al.
Published: (2024) -
Structured Uncertainty guided Clarification for LLM Agents
by: Suri, Manan, et al.
Published: (2025) -
ChartLens: Fine-grained Visual Attribution in Charts
by: Suri, Manan, et al.
Published: (2025) -
Trust but Verify: Introducing DAVinCI -- A Framework for Dual Attribution and Verification in Claim Inference for Language Models
by: Rawte, Vipula, et al.
Published: (2026) -
Follow the Flow: Fine-grained Flowchart Attribution with Neurosymbolic Agents
by: Suri, Manan, et al.
Published: (2025)