FactGuard: Leveraging Multi-Agent Systems to Generate Answerable and Unanswerable Questions for Enhanced Long-Context LLM Extraction
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Qian-Wen, Li, Fang, Wang, Jie, Qiao, Lingfeng, Yu, Yifei, Yin, Di, Sun, Xing |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
FactGuard: Event-Centric and Commonsense-Guided Fake News Detection
par: He, Jing, et autres
Publié: (2025)
par: He, Jing, et autres
Publié: (2025)
FactGuard: Agentic Video Misinformation Detection via Reinforcement Learning
par: Li, Zehao, et autres
Publié: (2026)
par: Li, Zehao, et autres
Publié: (2026)
RetinaQA: A Robust Knowledge Base Question Answering Model for both Answerable and Unanswerable Questions
par: Faldu, Prayushi, et autres
Publié: (2024)
par: Faldu, Prayushi, et autres
Publié: (2024)
Sequential-NIAH: A Needle-In-A-Haystack Benchmark for Extracting Sequential Needles from Long Contexts
par: Yu, Yifei, et autres
Publié: (2025)
par: Yu, Yifei, et autres
Publié: (2025)
Automatic Answerability Evaluation for Question Generation
par: Wang, Zifan, et autres
Publié: (2023)
par: Wang, Zifan, et autres
Publié: (2023)
CJEval: A Benchmark for Assessing Large Language Models Using Chinese Junior High School Exam Data
par: Zhang, Qian-Wen, et autres
Publié: (2024)
par: Zhang, Qian-Wen, et autres
Publié: (2024)
Answerability in Retrieval-Augmented Open-Domain Question Answering
par: Abdumalikov, Rustam, et autres
Publié: (2024)
par: Abdumalikov, Rustam, et autres
Publié: (2024)
YTCommentQA: Video Question Answerability in Instructional Videos
par: Yang, Saelyne, et autres
Publié: (2024)
par: Yang, Saelyne, et autres
Publié: (2024)
SNFinLLM: Systematic and Nuanced Financial Domain Adaptation of Chinese Large Language Models
par: Zhao, Shujuan, et autres
Publié: (2024)
par: Zhao, Shujuan, et autres
Publié: (2024)
Answerability Fields: Answerable Location Estimation via Diffusion Models
par: Azuma, Daichi, et autres
Publié: (2024)
par: Azuma, Daichi, et autres
Publié: (2024)
VeriFact: Enhancing Long-Form Factuality Evaluation with Refined Fact Extraction and Reference Facts
par: Liu, Xin, et autres
Publié: (2025)
par: Liu, Xin, et autres
Publié: (2025)
VISREAS: Complex Visual Reasoning with Unanswerable Questions
par: Akter, Syeda Nahida, et autres
Publié: (2024)
par: Akter, Syeda Nahida, et autres
Publié: (2024)
VisionTrap: Unanswerable Questions On Visual Data
par: Saadat, Asir, et autres
Publié: (2025)
par: Saadat, Asir, et autres
Publié: (2025)
The Effect of Intertwined Epidemiologic Concepts on Answerable Research Questions in Perinatal Epidemiology
par: Penelope P. Howards, et autres
Publié: (2025)
par: Penelope P. Howards, et autres
Publié: (2025)
SCARE: A Benchmark for SQL Correction and Question Answerability Classification for Reliable EHR Question Answering
par: Lee, Gyubok, et autres
Publié: (2025)
par: Lee, Gyubok, et autres
Publié: (2025)
I Could've Asked That: Reformulating Unanswerable Questions
par: Zhao, Wenting, et autres
Publié: (2024)
par: Zhao, Wenting, et autres
Publié: (2024)
AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions
par: Kirichenko, Polina, et autres
Publié: (2025)
par: Kirichenko, Polina, et autres
Publié: (2025)
Towards Unbiased Evaluation of Detecting Unanswerable Questions in EHRSQL
par: Yang, Yongjin, et autres
Publié: (2024)
par: Yang, Yongjin, et autres
Publié: (2024)
Unanswered Questions in Colombia’s Foreign Language Education Policy
par: Camilo Andrés Bonilla Carvajal
Publié: (2016)
par: Camilo Andrés Bonilla Carvajal
Publié: (2016)
Search for Coverage: Learning Coverage-Aware Retrieval with Augmented Sub-Question Answerability
par: Ju, Jia-Huei, et autres
Publié: (2026)
par: Ju, Jia-Huei, et autres
Publié: (2026)
Contextual Candor: Enhancing LLM Trustworthiness Through Hierarchical Unanswerability Detection
par: Robinson, Steven, et autres
Publié: (2025)
par: Robinson, Steven, et autres
Publié: (2025)
MAD-Fact: A Multi-Agent Debate Framework for Long-Form Factuality Evaluation in LLMs
par: Ning, Yucheng, et autres
Publié: (2025)
par: Ning, Yucheng, et autres
Publié: (2025)
UAQFact: Evaluating Factual Knowledge Utilization of LLMs on Unanswerable Questions
par: Tan, Chuanyuan, et autres
Publié: (2025)
par: Tan, Chuanyuan, et autres
Publié: (2025)
Update on Menopause Hormone Therapy; Current Indications and Unanswered Questions
par: Annice Mukherjee, et autres
Publié: (2025)
par: Annice Mukherjee, et autres
Publié: (2025)
TP53 ‐Mutated Acute Myeloid Leukemia: Unanswered Questions
par: Antonella Bruzzese, et autres
Publié: (2025)
par: Antonella Bruzzese, et autres
Publié: (2025)
PolitNuggets: Benchmarking Agentic Discovery of Long-Tail Political Facts
par: Zhu, Yifei
Publié: (2026)
par: Zhu, Yifei
Publié: (2026)
Detecting LLM Hallucination Through Layer-wise Information Deficiency: Analysis of Ambiguous Prompts and Unanswerable Questions
par: Kim, Hazel, et autres
Publié: (2024)
par: Kim, Hazel, et autres
Publié: (2024)
QGuard:Question-based Zero-shot Guard for Multi-modal LLM Safety
par: Lee, Taegyeong, et autres
Publié: (2025)
par: Lee, Taegyeong, et autres
Publié: (2025)
TUBench: Benchmarking Large Vision-Language Models on Trustworthiness with Unanswerable Questions
par: He, Xingwei, et autres
Publié: (2024)
par: He, Xingwei, et autres
Publié: (2024)
Benchmarking Visual LLMs Resilience to Unanswerable Questions on Visually Rich Documents
par: Napolitano, Davide, et autres
Publié: (2025)
par: Napolitano, Davide, et autres
Publié: (2025)
Reproductive Ecology and Evolutionary Anthropology: Foundations, Unanswered Questions, and Future Directions
par: R. G. Bribiescas, et autres
Publié: (2026)
par: R. G. Bribiescas, et autres
Publié: (2026)
Query Generation Pipeline with Enhanced Answerability Assessment for Financial Information Retrieval
par: Kim, Hyunkyu, et autres
Publié: (2025)
par: Kim, Hyunkyu, et autres
Publié: (2025)
MoHoBench: Assessing Honesty of Multimodal Large Language Models via Unanswerable Visual Questions
par: Zhu, Yanxu, et autres
Publié: (2025)
par: Zhu, Yanxu, et autres
Publié: (2025)
Youtu-GraphRAG: Vertically Unified Agents for Graph Retrieval-Augmented Complex Reasoning
par: Dong, Junnan, et autres
Publié: (2025)
par: Dong, Junnan, et autres
Publié: (2025)
GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning
par: Xiang, Zhen, et autres
Publié: (2024)
par: Xiang, Zhen, et autres
Publié: (2024)
WebWeaver: Breaking Topology Confidentiality in LLM Multi-Agent Systems with Stealthy Context-Based Inference
par: Xiong, Zixun, et autres
Publié: (2026)
par: Xiong, Zixun, et autres
Publié: (2026)
Leveraging LLMs as Meta-Judges: A Multi-Agent Framework for Evaluating LLM Judgments
par: Li, Yuran, et autres
Publié: (2025)
par: Li, Yuran, et autres
Publié: (2025)
Exploring React Library Related Questions on Stack Overflow: Answered vs. Unanswered
par: Ardity, Vanesya Aura, et autres
Publié: (2025)
par: Ardity, Vanesya Aura, et autres
Publié: (2025)
CLIP-UP: CLIP-Based Unanswerable Problem Detection for Visual Question Answering
par: Vardi, Ben, et autres
Publié: (2025)
par: Vardi, Ben, et autres
Publié: (2025)
Beyond Synergy: Unanswered Questions and Critical Reflections on AI‐Assisted Colonoscopy Efficacy?
par: Zhuoming Guo, et autres
Publié: (2026)
par: Zhuoming Guo, et autres
Publié: (2026)
Documents similaires
-
FactGuard: Event-Centric and Commonsense-Guided Fake News Detection
par: He, Jing, et autres
Publié: (2025) -
FactGuard: Agentic Video Misinformation Detection via Reinforcement Learning
par: Li, Zehao, et autres
Publié: (2026) -
RetinaQA: A Robust Knowledge Base Question Answering Model for both Answerable and Unanswerable Questions
par: Faldu, Prayushi, et autres
Publié: (2024) -
Sequential-NIAH: A Needle-In-A-Haystack Benchmark for Extracting Sequential Needles from Long Contexts
par: Yu, Yifei, et autres
Publié: (2025) -
Automatic Answerability Evaluation for Question Generation
par: Wang, Zifan, et autres
Publié: (2023)