Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent
Fuente:
arXiv
Saved in:
| Main Authors: | Qin, Ziyuan, Cheng, Dongjie, Wang, Haoyu, Yi, Huahui, Shao, Yuting, Fan, Zhiyuan, Li, Kang, Lao, Qicheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
iDPA: Instance Decoupled Prompt Attention for Incremental Medical Object Detection
by: Yi, Huahui, et al.
Published: (2025)
by: Yi, Huahui, et al.
Published: (2025)
Guiding Medical Vision-Language Models with Explicit Visual Prompts: Framework Design and Comprehensive Exploration of Prompt Variations
by: Zhu, Kangyu, et al.
Published: (2025)
by: Zhu, Kangyu, et al.
Published: (2025)
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
by: Lim, Youngsun, et al.
Published: (2024)
by: Lim, Youngsun, et al.
Published: (2024)
Improving Medical Visual Reinforcement Fine-Tuning via Perception and Reasoning Augmentation
by: Yang, Guangjing, et al.
Published: (2026)
by: Yang, Guangjing, et al.
Published: (2026)
MLAE: Masked LoRA Experts for Visual Parameter-Efficient Fine-Tuning
by: Wang, Junjie, et al.
Published: (2024)
by: Wang, Junjie, et al.
Published: (2024)
Scene-Text Grounding for Text-Based Video Question Answering
by: Zhou, Sheng, et al.
Published: (2024)
by: Zhou, Sheng, et al.
Published: (2024)
TV-SAM: Increasing Zero-Shot Segmentation Performance on Multimodal Medical Images Using GPT-4 Generated Descriptive Prompts Without Human Annotation
by: Jiang, Zekun, et al.
Published: (2024)
by: Jiang, Zekun, et al.
Published: (2024)
Combining LLMs and Knowledge Graphs to Reduce Hallucinations in Question Answering
by: Pusch, Larissa, et al.
Published: (2024)
by: Pusch, Larissa, et al.
Published: (2024)
ClueTracer: Question-to-Vision Clue Tracing for Training-Free Hallucination Suppression in Multimodal Reasoning
by: Xi, Gongli, et al.
Published: (2026)
by: Xi, Gongli, et al.
Published: (2026)
SGEdit: Bridging LLM with Text2Image Generative Model for Scene Graph-based Image Editing
by: Zhang, Zhiyuan, et al.
Published: (2024)
by: Zhang, Zhiyuan, et al.
Published: (2024)
Hybrid Graphs for Table-and-Text based Question Answering using LLMs
by: Agarwal, Ankush, et al.
Published: (2025)
by: Agarwal, Ankush, et al.
Published: (2025)
An Effective Data Augmentation Method by Asking Questions about Scene Text Images
by: Yao, Xu, et al.
Published: (2026)
by: Yao, Xu, et al.
Published: (2026)
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering
by: Zhou, Sheng, et al.
Published: (2025)
by: Zhou, Sheng, et al.
Published: (2025)
One-to-Normal: Anomaly Personalization for Few-shot Anomaly Detection
by: Li, Yiyue, et al.
Published: (2025)
by: Li, Yiyue, et al.
Published: (2025)
KG-Guard: Graph-Based Hallucination Detection for Knowledge Base Question Answering
by: Sawczyn, Albert, et al.
Published: (2026)
by: Sawczyn, Albert, et al.
Published: (2026)
Debate over Mixed-knowledge: A Robust Multi-Agent Reasoning Framework for Incomplete Knowledge Graph Question Answering
by: Liu, Jilong, et al.
Published: (2025)
by: Liu, Jilong, et al.
Published: (2025)
On Early Detection of Hallucinations in Factual Question Answering
by: Snyder, Ben, et al.
Published: (2023)
by: Snyder, Ben, et al.
Published: (2023)
Hallucination Benchmark in Medical Visual Question Answering
by: Wu, Jinge, et al.
Published: (2024)
by: Wu, Jinge, et al.
Published: (2024)
FlowX: Towards Explainable Graph Neural Networks via Message Flows
by: Gui, Shurui, et al.
Published: (2022)
by: Gui, Shurui, et al.
Published: (2022)
Adversarial Training with OCR Modality Perturbation for Scene-Text Visual Question Answering
by: Shen, Zhixuan, et al.
Published: (2024)
by: Shen, Zhixuan, et al.
Published: (2024)
Observation of anomalous Floquet non-Abelian topological insulators
by: Qiu, Huahui, et al.
Published: (2025)
by: Qiu, Huahui, et al.
Published: (2025)
FinTextQA: A Dataset for Long-form Financial Question Answering
by: Chen, Jian, et al.
Published: (2024)
by: Chen, Jian, et al.
Published: (2024)
Warehouse Spatial Question Answering with LLM Agent
by: Huang, Hsiang-Wei, et al.
Published: (2025)
by: Huang, Hsiang-Wei, et al.
Published: (2025)
ATLANTIS at SemEval-2025 Task 3: Detecting Hallucinated Text Spans in Question Answering
by: Kobus, Catherine, et al.
Published: (2025)
by: Kobus, Catherine, et al.
Published: (2025)
LP-LM: No Hallucinations in Question Answering with Logic Programming
by: Wu, Katherine, et al.
Published: (2025)
by: Wu, Katherine, et al.
Published: (2025)
Generate-on-Graph: Treat LLM as both Agent and KG in Incomplete Knowledge Graph Question Answering
by: Xu, Yao, et al.
Published: (2024)
by: Xu, Yao, et al.
Published: (2024)
VIHD: Visual Intervention-based Hallucination Detection for Medical Visual Question Answering
by: Chen, Jiayi, et al.
Published: (2026)
by: Chen, Jiayi, et al.
Published: (2026)
AQAScore: Evaluating Semantic Alignment in Text-to-Audio Generation via Audio Question Answering
by: Kuan, Chun-Yi, et al.
Published: (2026)
by: Kuan, Chun-Yi, et al.
Published: (2026)
Dataset and Benchmark for Urdu Natural Scenes Text Detection, Recognition and Visual Question Answering
by: Maryam, Hiba, et al.
Published: (2024)
by: Maryam, Hiba, et al.
Published: (2024)
GraphPad: Inference-Time 3D Scene Graph Updates for Embodied Question Answering
by: Ali, Muhammad Qasim, et al.
Published: (2025)
by: Ali, Muhammad Qasim, et al.
Published: (2025)
Denoising Table-Text Retrieval for Open-Domain Question Answering
by: Kang, Deokhyung, et al.
Published: (2024)
by: Kang, Deokhyung, et al.
Published: (2024)
The Role of Exploration Modules in Small Language Models for Knowledge Graph Question Answering
by: Cheng, Yi-Jie, et al.
Published: (2025)
by: Cheng, Yi-Jie, et al.
Published: (2025)
ViConsFormer: Constituting Meaningful Phrases of Scene Texts using Transformer-based Method in Vietnamese Text-based Visual Question Answering
by: Nguyen, Nghia Hieu, et al.
Published: (2024)
by: Nguyen, Nghia Hieu, et al.
Published: (2024)
Knowledge Graph-Guided Multi-Agent Distillation for Reliable Industrial Question Answering with Datasets
by: Pan, Jiqun, et al.
Published: (2025)
by: Pan, Jiqun, et al.
Published: (2025)
Towards Top-Down Reasoning: An Explainable Multi-Agent Approach for Visual Question Answering
by: Wang, Zeqing, et al.
Published: (2023)
by: Wang, Zeqing, et al.
Published: (2023)
Question-to-Question Retrieval for Hallucination-Free Knowledge Access: An Approach for Wikipedia and Wikidata Question Answering
by: Thottingal, Santhosh
Published: (2025)
by: Thottingal, Santhosh
Published: (2025)
Subgraph Retrieval Enhanced by Graph-Text Alignment for Commonsense Question Answering
by: Peng, Boci, et al.
Published: (2024)
by: Peng, Boci, et al.
Published: (2024)
GeoSceneGraph: Geometric Scene Graph Diffusion Model for Text-guided 3D Indoor Scene Synthesis
by: Ruiz, Antonio, et al.
Published: (2025)
by: Ruiz, Antonio, et al.
Published: (2025)
Multi-Modal Scene Graph with Kolmogorov-Arnold Experts for Audio-Visual Question Answering
by: Fu, Zijian, et al.
Published: (2025)
by: Fu, Zijian, et al.
Published: (2025)
GraphEQA: Using 3D Semantic Scene Graphs for Real-time Embodied Question Answering
by: Saxena, Saumya, et al.
Published: (2024)
by: Saxena, Saumya, et al.
Published: (2024)
Similar Items
-
iDPA: Instance Decoupled Prompt Attention for Incremental Medical Object Detection
by: Yi, Huahui, et al.
Published: (2025) -
Guiding Medical Vision-Language Models with Explicit Visual Prompts: Framework Design and Comprehensive Exploration of Prompt Variations
by: Zhu, Kangyu, et al.
Published: (2025) -
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
by: Lim, Youngsun, et al.
Published: (2024) -
Improving Medical Visual Reinforcement Fine-Tuning via Perception and Reasoning Augmentation
by: Yang, Guangjing, et al.
Published: (2026) -
MLAE: Masked LoRA Experts for Visual Parameter-Efficient Fine-Tuning
by: Wang, Junjie, et al.
Published: (2024)