FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Bao, Forrest Sheng, Li, Miaoran, Qu, Renyi, Luo, Ge, Wan, Erana, Tang, Yujia, Fan, Weisi, Tamber, Manveer Singh, Kazi, Suleman, Sourabh, Vivek, Qi, Mike, Tu, Ruixuan, Xu, Chenyu, Gonzales, Matthew, Mendelevitch, Ofer, Ahmad, Amin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking LLM Faithfulness in RAG with Evolving Leaderboards
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
Conventional Contrastive Learning Often Falls Short: Improving Dense Retrieval with Cross-Encoder Listwise Distillation and Synthetic Data
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
Teaching Dense Retrieval Models to Specialize with Listwise Distillation and LLM Data Augmentation
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
Illusions of Relevance: Arbitrary Content Injection Attacks Deceive Retrievers, Rerankers, and LLM Judges
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
Is Semantic Chunking Worth the Computational Cost?
by: Qu, Renyi, et al.
Published: (2024)
by: Qu, Renyi, et al.
Published: (2024)
Unifying Adversarial Robustness and Training Across Text Scoring Models
by: Tamber, Manveer Singh, et al.
Published: (2026)
by: Tamber, Manveer Singh, et al.
Published: (2026)
Can't Hide Behind the API: Stealing Black-Box Commercial Embedding Models
by: Tamber, Manveer Singh, et al.
Published: (2024)
by: Tamber, Manveer Singh, et al.
Published: (2024)
MIRAGE-Bench: Automatic Multilingual Benchmark Arena for Retrieval-Augmented Generation Systems
by: Thakur, Nandan, et al.
Published: (2024)
by: Thakur, Nandan, et al.
Published: (2024)
FaithLens: Detecting and Explaining Faithfulness Hallucination
by: Si, Shuzheng, et al.
Published: (2025)
by: Si, Shuzheng, et al.
Published: (2025)
Características del empleo y la remuneración económica en el ejercicio de la Reumatología en diferentes distritos de Argentina
by: Fernando Eraña
Published: (2024)
by: Fernando Eraña
Published: (2024)
Sobre la viabilidad de una epistemología empírica y normativa
by: Ángeles Eraña
Published: (2007)
by: Ángeles Eraña
Published: (2007)
DANIEL QUESADA , ed. 2009. Cuestiones de Teoría del Conocimiento . Madrid: Tecnos.
by: Ángeles Eraña
Published: (2012)
by: Ángeles Eraña
Published: (2012)
¿Es posible la justicia epistémica sin un lugar común? (Hacia una reconceptualización del espacio público y las relaciones sociales)*
by: Ángeles Eraña
Published: (2022)
by: Ángeles Eraña
Published: (2022)
La escucha: una herramienta para combatir la injusticia hermenéutica
by: Ángeles Eraña
Published: (2025)
by: Ángeles Eraña
Published: (2025)
El conocimiento como una actividad colectiva
by: Ángeles Eraña
Published: (2016)
by: Ángeles Eraña
Published: (2016)
Del Río, F. (2020). Las filósofas tienen la palabra. Siglo XXI. 232 pp.
by: Ángeles Eraña
Published: (2021)
by: Ángeles Eraña
Published: (2021)
STORYSUMM: Evaluating Faithfulness in Story Summarization
by: Subbiah, Melanie, et al.
Published: (2024)
by: Subbiah, Melanie, et al.
Published: (2024)
Faithful Chart Summarization with ChaTS-Pi
by: Krichene, Syrine, et al.
Published: (2024)
by: Krichene, Syrine, et al.
Published: (2024)
On Positional Bias of Faithfulness for Long-form Summarization
by: Wan, David, et al.
Published: (2024)
by: Wan, David, et al.
Published: (2024)
Rembrandt – A képek színjátéka. Hermeneutikai kísérletek (online képmelléklet)
by: Rényi, András
Published: (2025)
by: Rényi, András
Published: (2025)
GUISE: Graph GaUssIan Shading watErmark
by: Yang, Renyi
Published: (2024)
by: Yang, Renyi
Published: (2024)
Rembrandt – A képek színjátéka. Hermeneutikai kísérletek
by: András, Rényi
Published: (2025)
by: András, Rényi
Published: (2025)
Thinking, Faithful and Stable: Mitigating Hallucinations in LLMs
by: Zou, Chelsea, et al.
Published: (2025)
by: Zou, Chelsea, et al.
Published: (2025)
Correction with Backtracking Reduces Hallucination in Summarization
by: Liu, Zhenzhen, et al.
Published: (2023)
by: Liu, Zhenzhen, et al.
Published: (2023)
Hallucinating AI Hijacking Attack: Large Language Models and Malicious Code Recommenders
by: Noever, David, et al.
Published: (2024)
by: Noever, David, et al.
Published: (2024)
Exploración del nivel de Neurofobia en estudiantes de medicina en México*
by: Irma Elisa Eraña Rojas
Published: (2018)
by: Irma Elisa Eraña Rojas
Published: (2018)
Single Point, Full Mask: Velocity-Guided Level Set Evolution for End-to-End Amodal Segmentation
by: Li, Zhixuan, et al.
Published: (2025)
by: Li, Zhixuan, et al.
Published: (2025)
Reducing Hallucinations in Summarization via Reinforcement Learning with Entity Hallucination Index
by: Katwe, Praveenkumar, et al.
Published: (2025)
by: Katwe, Praveenkumar, et al.
Published: (2025)
FaithCoT-Bench: Benchmarking Instance-Level Faithfulness of Chain-of-Thought Reasoning
by: Shen, Xu, et al.
Published: (2025)
by: Shen, Xu, et al.
Published: (2025)
TrueBrief: Faithful Summarization through Small Language Models
by: Lakara, Kumud, et al.
Published: (2025)
by: Lakara, Kumud, et al.
Published: (2025)
Learning from Self Critique and Refinement for Faithful LLM Summarization
by: Hu, Ting-Yao, et al.
Published: (2025)
by: Hu, Ting-Yao, et al.
Published: (2025)
Mitigating Large Language Model Hallucination with Faithful Finetuning
by: Hu, Minda, et al.
Published: (2024)
by: Hu, Minda, et al.
Published: (2024)
FaithSCAN: Model-Driven Single-Pass Hallucination Detection for Faithful Visual Question Answering
by: Tong, Chaodong, et al.
Published: (2026)
by: Tong, Chaodong, et al.
Published: (2026)
The Bayesian Reflex: Online Learning as the Autonomic Nervous System of Modern and Future AI
by: Bhattacharya, Durba, et al.
Published: (2026)
by: Bhattacharya, Durba, et al.
Published: (2026)
Modern Thinking
by: Kennedy, Mike
Published: (2006)
by: Kennedy, Mike
Published: (2006)
Radical Carbocyclization Intercepted Reductive C−C Bond Formation Between 1,n‐Enynes or Dienes and Electron‐Poor Olefins/Alkynes
by: Subhadeep Hazra, et al.
Published: (2024)
by: Subhadeep Hazra, et al.
Published: (2024)
COVID-19 dynamical evolution prediction in Mexico, decision making and social implementation: mid/low income countries study
by: Fernandez-Erana, Saasil, et al.
Published: (2020)
by: Fernandez-Erana, Saasil, et al.
Published: (2020)
Improving Faithfulness of Abstractive Summarization by Controlling Confounding Effect of Irrelevant Sentences
by: Ghoshal, Asish, et al.
Published: (2022)
by: Ghoshal, Asish, et al.
Published: (2022)
Enhancing Faithfulness in Abstractive Summarization via Span-Level Fine-Tuning
by: Huang, Sicong, et al.
Published: (2025)
by: Huang, Sicong, et al.
Published: (2025)
Delamination Detection in Layered Waveguides using Ostrovsky Wave Packets
by: Tamber, J. S., et al.
Published: (2024)
by: Tamber, J. S., et al.
Published: (2024)
Similar Items
-
Benchmarking LLM Faithfulness in RAG with Evolving Leaderboards
by: Tamber, Manveer Singh, et al.
Published: (2025) -
Conventional Contrastive Learning Often Falls Short: Improving Dense Retrieval with Cross-Encoder Listwise Distillation and Synthetic Data
by: Tamber, Manveer Singh, et al.
Published: (2025) -
Teaching Dense Retrieval Models to Specialize with Listwise Distillation and LLM Data Augmentation
by: Tamber, Manveer Singh, et al.
Published: (2025) -
Illusions of Relevance: Arbitrary Content Injection Attacks Deceive Retrievers, Rerankers, and LLM Judges
by: Tamber, Manveer Singh, et al.
Published: (2025) -
Is Semantic Chunking Worth the Computational Cost?
by: Qu, Renyi, et al.
Published: (2024)