Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Lim, Youngsun, Shim, Hyunjung |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
Label-Augmented Dataset Distillation
di: Kang, Seoungyoon, et al.
Pubblicazione: (2024)
di: Kang, Seoungyoon, et al.
Pubblicazione: (2024)
FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models
di: Lim, Youngsun, et al.
Pubblicazione: (2026)
di: Lim, Youngsun, et al.
Pubblicazione: (2026)
Robust Driving QA through Metadata-Grounded Context and Task-Specific Prompts
di: Yu, Seungjun, et al.
Pubblicazione: (2025)
di: Yu, Seungjun, et al.
Pubblicazione: (2025)
What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging
di: Kang, Inha, et al.
Pubblicazione: (2025)
di: Kang, Inha, et al.
Pubblicazione: (2025)
Memory-Efficient Fine-Tuning for Quantized Diffusion Model
di: Ryu, Hyogon, et al.
Pubblicazione: (2024)
di: Ryu, Hyogon, et al.
Pubblicazione: (2024)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
di: Park, NaHyeon, et al.
Pubblicazione: (2024)
di: Park, NaHyeon, et al.
Pubblicazione: (2024)
Scribble-Guided Diffusion for Training-free Text-to-Image Generation
di: Lee, Seonho, et al.
Pubblicazione: (2024)
di: Lee, Seonho, et al.
Pubblicazione: (2024)
Precision matters: Precision-aware ensemble for weakly supervised semantic segmentation
di: Park, Junsung, et al.
Pubblicazione: (2024)
di: Park, Junsung, et al.
Pubblicazione: (2024)
Classifier-guided CLIP Distillation for Unsupervised Multi-label Classification
di: Kim, Dongseob, et al.
Pubblicazione: (2025)
di: Kim, Dongseob, et al.
Pubblicazione: (2025)
T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts
di: Huang, Ziwei, et al.
Pubblicazione: (2024)
di: Huang, Ziwei, et al.
Pubblicazione: (2024)
MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval
di: Park, Seojeong, et al.
Pubblicazione: (2024)
di: Park, Seojeong, et al.
Pubblicazione: (2024)
Knowledge-based learning in Text-RAG and Image-RAG
di: Shim, Alexander, et al.
Pubblicazione: (2026)
di: Shim, Alexander, et al.
Pubblicazione: (2026)
CoT-PL: Chain-of-Thought Pseudo-Labeling for Open-Vocabulary Object Detection
di: Choi, Hojun, et al.
Pubblicazione: (2025)
di: Choi, Hojun, et al.
Pubblicazione: (2025)
Clustering-based Image-Text Graph Matching for Domain Generalization
di: Park, Nokyung, et al.
Pubblicazione: (2023)
di: Park, Nokyung, et al.
Pubblicazione: (2023)
Open Multimodal Retrieval-Augmented Factual Image Generation
di: Tian, Yang, et al.
Pubblicazione: (2025)
di: Tian, Yang, et al.
Pubblicazione: (2025)
Grounding Driving VLA via Inverse Kinematics
di: Park, Junsung, et al.
Pubblicazione: (2026)
di: Park, Junsung, et al.
Pubblicazione: (2026)
GEA: Generation-Enhanced Alignment for Text-to-Image Person Retrieval
di: Zou, Hao, et al.
Pubblicazione: (2025)
di: Zou, Hao, et al.
Pubblicazione: (2025)
Spherical Linear Interpolation and Text-Anchoring for Zero-shot Composed Image Retrieval
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
Sampling Bag of Views for Open-Vocabulary Object Detection
di: Choi, Hojun, et al.
Pubblicazione: (2024)
di: Choi, Hojun, et al.
Pubblicazione: (2024)
Rethinking Data Augmentation for Robust LiDAR Semantic Segmentation in Adverse Weather
di: Park, Junsung, et al.
Pubblicazione: (2024)
di: Park, Junsung, et al.
Pubblicazione: (2024)
Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval
di: Liu, Delong, et al.
Pubblicazione: (2023)
di: Liu, Delong, et al.
Pubblicazione: (2023)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
di: Zhao, Yu, et al.
Pubblicazione: (2024)
di: Zhao, Yu, et al.
Pubblicazione: (2024)
GenEval 2: Addressing Benchmark Drift in Text-to-Image Evaluation
di: Kamath, Amita, et al.
Pubblicazione: (2025)
di: Kamath, Amita, et al.
Pubblicazione: (2025)
The Factuality Tax of Diversity-Intervened Text-to-Image Generation: Benchmark and Fact-Augmented Intervention
di: Wan, Yixin, et al.
Pubblicazione: (2024)
di: Wan, Yixin, et al.
Pubblicazione: (2024)
Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken Generation
di: Li, Yongqi, et al.
Pubblicazione: (2024)
di: Li, Yongqi, et al.
Pubblicazione: (2024)
ChainMPQ: Interleaved Text-Image Reasoning Chains for Mitigating Relation Hallucinations
di: Wu, Yike, et al.
Pubblicazione: (2025)
di: Wu, Yike, et al.
Pubblicazione: (2025)
Directional Textual Inversion for Personalized Text-to-Image Generation
di: Kim, Kunhee, et al.
Pubblicazione: (2025)
di: Kim, Kunhee, et al.
Pubblicazione: (2025)
Invisible Relevance Bias: Text-Image Retrieval Models Prefer AI-Generated Images
di: Xu, Shicheng, et al.
Pubblicazione: (2023)
di: Xu, Shicheng, et al.
Pubblicazione: (2023)
Addressing Image Authenticity When Cameras Use Generative AI
di: Masud, Umar, et al.
Pubblicazione: (2026)
di: Masud, Umar, et al.
Pubblicazione: (2026)
MUMU: Bootstrapping Multimodal Image Generation from Text-to-Image Data
di: Berman, William, et al.
Pubblicazione: (2024)
di: Berman, William, et al.
Pubblicazione: (2024)
Agentic Retoucher for Text-To-Image Generation
di: Shen, Shaocheng, et al.
Pubblicazione: (2026)
di: Shen, Shaocheng, et al.
Pubblicazione: (2026)
No Thing, Nothing: Highlighting Safety-Critical Classes for Robust LiDAR Semantic Segmentation in Adverse Weather
di: Park, Junsung, et al.
Pubblicazione: (2025)
di: Park, Junsung, et al.
Pubblicazione: (2025)
Blind to Position, Biased in Language: Probing Mid-Layer Representational Bias in Vision-Language Encoders for Zero-Shot Language-Grounded Spatial Understanding
di: An, Na Min, et al.
Pubblicazione: (2025)
di: An, Na Min, et al.
Pubblicazione: (2025)
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
di: Duan, Zhongjie, et al.
Pubblicazione: (2024)
di: Duan, Zhongjie, et al.
Pubblicazione: (2024)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
Evaluating Text-to-Image Generative Models: An Empirical Study on Human Image Synthesis
di: Chen, Muxi, et al.
Pubblicazione: (2024)
di: Chen, Muxi, et al.
Pubblicazione: (2024)
World-To-Image: Grounding Text-to-Image Generation with Agent-Driven World Knowledge
di: Son, Moo Hyun, et al.
Pubblicazione: (2025)
di: Son, Moo Hyun, et al.
Pubblicazione: (2025)
Grounding Text-to-Image Diffusion Models for Controlled High-Quality Image Generation
di: Süleyman, Ahmad, et al.
Pubblicazione: (2025)
di: Süleyman, Ahmad, et al.
Pubblicazione: (2025)
Multi-path Exploration and Feedback Adjustment for Text-to-Image Person Retrieval
di: Kang, Bin, et al.
Pubblicazione: (2024)
di: Kang, Bin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
di: Lim, Youngsun, et al.
Pubblicazione: (2024) -
Label-Augmented Dataset Distillation
di: Kang, Seoungyoon, et al.
Pubblicazione: (2024) -
FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models
di: Lim, Youngsun, et al.
Pubblicazione: (2026) -
Robust Driving QA through Metadata-Grounded Context and Task-Specific Prompts
di: Yu, Seungjun, et al.
Pubblicazione: (2025) -
What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging
di: Kang, Inha, et al.
Pubblicazione: (2025)