Evaluating LLMs' Assessment of Mixed-Context Hallucination Through the Lens of Summarization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qi, Siya, Cao, Rui, He, Yulan, Yuan, Zheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluation of LLMs for Process Model Analysis and Optimization
von: Kumar, Akhil, et al.
Veröffentlicht: (2025)
von: Kumar, Akhil, et al.
Veröffentlicht: (2025)
TQA-Bench: Evaluating LLMs for Multi-Table Question Answering with Scalable Context and Symbolic Extension
von: Qiu, Zipeng, et al.
Veröffentlicht: (2024)
von: Qiu, Zipeng, et al.
Veröffentlicht: (2024)
A Survey of Automatic Hallucination Evaluation on Natural Language Generation
von: Qi, Siya, et al.
Veröffentlicht: (2024)
von: Qi, Siya, et al.
Veröffentlicht: (2024)
Culinary Crossroads: A RAG Framework for Enhancing Diversity in Cross-Cultural Recipe Adaptation
von: Hu, Tianyi, et al.
Veröffentlicht: (2025)
von: Hu, Tianyi, et al.
Veröffentlicht: (2025)
From Facts to Conclusions : Integrating Deductive Reasoning in Retrieval-Augmented LLMs
von: Mishra, Shubham, et al.
Veröffentlicht: (2025)
von: Mishra, Shubham, et al.
Veröffentlicht: (2025)
Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers
von: Liu, Geng, et al.
Veröffentlicht: (2025)
von: Liu, Geng, et al.
Veröffentlicht: (2025)
Enhancing Long Context Performance in LLMs Through Inner Loop Query Mechanism
von: Tang, Yimin, et al.
Veröffentlicht: (2024)
von: Tang, Yimin, et al.
Veröffentlicht: (2024)
Taxonomy and Analysis of Sensitive User Queries in Generative AI Search
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2024)
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2024)
CausalCite: A Causal Formulation of Paper Citations
von: Kumar, Ishan, et al.
Veröffentlicht: (2023)
von: Kumar, Ishan, et al.
Veröffentlicht: (2023)
Retrieval Improvements Do Not Guarantee Better Answers: A Study of RAG for AI Policy QA
von: Mathur, Saahil, et al.
Veröffentlicht: (2026)
von: Mathur, Saahil, et al.
Veröffentlicht: (2026)
Epistemic Diversity and Knowledge Collapse in Large Language Models
von: Wright, Dustin, et al.
Veröffentlicht: (2025)
von: Wright, Dustin, et al.
Veröffentlicht: (2025)
From Data to Behavior: Predicting Unintended Model Behaviors Before Training
von: Wang, Mengru, et al.
Veröffentlicht: (2026)
von: Wang, Mengru, et al.
Veröffentlicht: (2026)
STRUM-LLM: Attributed and Structured Contrastive Summarization
von: Gunel, Beliz, et al.
Veröffentlicht: (2024)
von: Gunel, Beliz, et al.
Veröffentlicht: (2024)
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
von: Yu, Yue, et al.
Veröffentlicht: (2024)
von: Yu, Yue, et al.
Veröffentlicht: (2024)
Towards a Robust Retrieval-Based Summarization System
von: Liu, Shengjie, et al.
Veröffentlicht: (2024)
von: Liu, Shengjie, et al.
Veröffentlicht: (2024)
FG-RAG: Enhancing Query-Focused Summarization with Context-Aware Fine-Grained Graph RAG
von: Hong, Yubin, et al.
Veröffentlicht: (2025)
von: Hong, Yubin, et al.
Veröffentlicht: (2025)
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
von: Xu, Peng, et al.
Veröffentlicht: (2024)
von: Xu, Peng, et al.
Veröffentlicht: (2024)
Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain
von: Liu, Wei, et al.
Veröffentlicht: (2026)
von: Liu, Wei, et al.
Veröffentlicht: (2026)
Thesis: Document Summarization with applications to Keyword extraction and Image Retrieval
von: Sundararaj, Jayaprakash
Veröffentlicht: (2024)
von: Sundararaj, Jayaprakash
Veröffentlicht: (2024)
Leveraging the Power of LLMs: A Fine-Tuning Approach for High-Quality Aspect-Based Summarization
von: Mullick, Ankan, et al.
Veröffentlicht: (2024)
von: Mullick, Ankan, et al.
Veröffentlicht: (2024)
MixLM: High-Throughput and Effective LLM Ranking via Text-Embedding Mix-Interaction
von: Li, Guoyao, et al.
Veröffentlicht: (2025)
von: Li, Guoyao, et al.
Veröffentlicht: (2025)
DQNC2S: DQN-based Cross-stream Crisis event Summarizer
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
Haystack Engineering: Context Engineering for Heterogeneous and Agentic Long-Context Evaluation
von: Li, Mufei, et al.
Veröffentlicht: (2025)
von: Li, Mufei, et al.
Veröffentlicht: (2025)
Knowledge Conflicts for LLMs: A Survey
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
Fine-Grained Table Retrieval Through the Lens of Complex Queries
von: Kosiuk, Wojciech, et al.
Veröffentlicht: (2026)
von: Kosiuk, Wojciech, et al.
Veröffentlicht: (2026)
Evaluating and Enhancing Large Language Models for Novelty Assessment in Scholarly Publications
von: Lin, Ethan, et al.
Veröffentlicht: (2024)
von: Lin, Ethan, et al.
Veröffentlicht: (2024)
SynDy: Synthetic Dynamic Dataset Generation Framework for Misinformation Tasks
von: Shliselberg, Michael, et al.
Veröffentlicht: (2024)
von: Shliselberg, Michael, et al.
Veröffentlicht: (2024)
PatentEdits: Framing Patent Novelty as Textual Entailment
von: Lee, Ryan, et al.
Veröffentlicht: (2024)
von: Lee, Ryan, et al.
Veröffentlicht: (2024)
Stairway to Fairness: Connecting Group and Individual Fairness
von: Rampisela, Theresia Veronika, et al.
Veröffentlicht: (2025)
von: Rampisela, Theresia Veronika, et al.
Veröffentlicht: (2025)
Stop DDoS Attacking the Research Community with AI-Generated Survey Papers
von: Lin, Jianghao, et al.
Veröffentlicht: (2025)
von: Lin, Jianghao, et al.
Veröffentlicht: (2025)
Large Language Model-Based Knowledge Graph System Construction for Sustainable Development Goals: An AI-Based Speculative Design Perspective
von: Lin, Yi-De, et al.
Veröffentlicht: (2025)
von: Lin, Yi-De, et al.
Veröffentlicht: (2025)
Inducing Sustained Creativity and Diversity in Large Language Models
von: Luo, Queenie, et al.
Veröffentlicht: (2026)
von: Luo, Queenie, et al.
Veröffentlicht: (2026)
Scraping the Shadows: Deep Learning Breakthroughs in Dark Web Intelligence
von: Bakermans, Ingmar, et al.
Veröffentlicht: (2025)
von: Bakermans, Ingmar, et al.
Veröffentlicht: (2025)
Document-Level Event Extraction with Definition-Driven ICL
von: Liu, Zhuoyuan, et al.
Veröffentlicht: (2024)
von: Liu, Zhuoyuan, et al.
Veröffentlicht: (2024)
Suicide Phenotyping from Clinical Notes in Safety-Net Psychiatric Hospital Using Multi-Label Classification with Pre-Trained Language Models
von: Li, Zehan, et al.
Veröffentlicht: (2024)
von: Li, Zehan, et al.
Veröffentlicht: (2024)
SentimentLens: Reconciling Sentiment and Ratings via Dual-Modality in the Hospitality Sector
von: Jayakody, Dineth, et al.
Veröffentlicht: (2026)
von: Jayakody, Dineth, et al.
Veröffentlicht: (2026)
Ground-Truth Subgraphs for Better Training and Evaluation of Knowledge Graph Augmented LLMs
von: Cattaneo, Alberto, et al.
Veröffentlicht: (2025)
von: Cattaneo, Alberto, et al.
Veröffentlicht: (2025)
AGENT-CQ: Automatic Generation and Evaluation of Clarifying Questions for Conversational Search with LLMs
von: Siro, Clemencia, et al.
Veröffentlicht: (2024)
von: Siro, Clemencia, et al.
Veröffentlicht: (2024)
DuoLens: A Framework for Robust Detection of Machine-Generated Multilingual Text and Code
von: Agrawal, Shriyansh, et al.
Veröffentlicht: (2025)
von: Agrawal, Shriyansh, et al.
Veröffentlicht: (2025)
A Systematic Examination of Preference Learning through the Lens of Instruction-Following
von: Kim, Joongwon, et al.
Veröffentlicht: (2024)
von: Kim, Joongwon, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Evaluation of LLMs for Process Model Analysis and Optimization
von: Kumar, Akhil, et al.
Veröffentlicht: (2025) -
TQA-Bench: Evaluating LLMs for Multi-Table Question Answering with Scalable Context and Symbolic Extension
von: Qiu, Zipeng, et al.
Veröffentlicht: (2024) -
A Survey of Automatic Hallucination Evaluation on Natural Language Generation
von: Qi, Siya, et al.
Veröffentlicht: (2024) -
Culinary Crossroads: A RAG Framework for Enhancing Diversity in Cross-Cultural Recipe Adaptation
von: Hu, Tianyi, et al.
Veröffentlicht: (2025) -
From Facts to Conclusions : Integrating Deductive Reasoning in Retrieval-Augmented LLMs
von: Mishra, Shubham, et al.
Veröffentlicht: (2025)