FilterRAG: Zero-Shot Informed Retrieval-Augmented Generation to Mitigate Hallucinations in VQA
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Sarwar, Nobin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VideoRAG: Retrieval-Augmented Generation over Video Corpus
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025)
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025)
UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities
von: Yeo, Woongyeong, et al.
Veröffentlicht: (2025)
von: Yeo, Woongyeong, et al.
Veröffentlicht: (2025)
TabRAG: Improving Tabular Document Question Answering for Retrieval Augmented Generation via Structured Representations
von: Si, Jacob, et al.
Veröffentlicht: (2025)
von: Si, Jacob, et al.
Veröffentlicht: (2025)
iRAG: Advancing RAG for Videos with an Incremental Approach
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2024)
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2024)
RAG-Check: Evaluating Multimodal Retrieval Augmented Generation Performance
von: Mortaheb, Matin, et al.
Veröffentlicht: (2025)
von: Mortaheb, Matin, et al.
Veröffentlicht: (2025)
FlowVQA: Mapping Multimodal Logic in Visual Question Answering with Flowcharts
von: Singh, Shubhankar, et al.
Veröffentlicht: (2024)
von: Singh, Shubhankar, et al.
Veröffentlicht: (2024)
SignRAG: A Retrieval-Augmented System for Scalable Zero-Shot Road Sign Recognition
von: Zhu, Minghao, et al.
Veröffentlicht: (2025)
von: Zhu, Minghao, et al.
Veröffentlicht: (2025)
DREAM: Improving Video-Text Retrieval Through Relevance-Based Augmentation Using Large Foundation Models
von: Wang, Yimu, et al.
Veröffentlicht: (2024)
von: Wang, Yimu, et al.
Veröffentlicht: (2024)
Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation
von: Luo, Weiqing, et al.
Veröffentlicht: (2026)
von: Luo, Weiqing, et al.
Veröffentlicht: (2026)
VDocRAG: Retrieval-Augmented Generation over Visually-Rich Documents
von: Tanaka, Ryota, et al.
Veröffentlicht: (2025)
von: Tanaka, Ryota, et al.
Veröffentlicht: (2025)
MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation
von: Hsiao, Chi-Hsiang, et al.
Veröffentlicht: (2025)
von: Hsiao, Chi-Hsiang, et al.
Veröffentlicht: (2025)
Multimedia-Aware Question Answering: A Review of Retrieval and Cross-Modal Reasoning Architectures
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
Re-ranking the Context for Multimodal Retrieval Augmented Generation
von: Mortaheb, Matin, et al.
Veröffentlicht: (2025)
von: Mortaheb, Matin, et al.
Veröffentlicht: (2025)
Windsock is Dancing: Adaptive Multimodal Retrieval-Augmented Generation
von: Zhao, Shu, et al.
Veröffentlicht: (2025)
von: Zhao, Shu, et al.
Veröffentlicht: (2025)
Are We on the Right Way for Assessing Document Retrieval-Augmented Generation?
von: Shen, Wenxuan, et al.
Veröffentlicht: (2025)
von: Shen, Wenxuan, et al.
Veröffentlicht: (2025)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
ViDoRAG: Visual Document Retrieval-Augmented Generation via Dynamic Iterative Reasoning Agents
von: Wang, Qiuchen, et al.
Veröffentlicht: (2025)
von: Wang, Qiuchen, et al.
Veröffentlicht: (2025)
FactCHD: Benchmarking Fact-Conflicting Hallucination Detection
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation
von: Martin, Alexander, et al.
Veröffentlicht: (2025)
von: Martin, Alexander, et al.
Veröffentlicht: (2025)
TelcoAI: Advancing 3GPP Technical Specification Search through Agentic Multi-Modal Retrieval-Augmented Generation
von: Ghosh, Rahul, et al.
Veröffentlicht: (2025)
von: Ghosh, Rahul, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Generation with Graphs (GraphRAG)
von: Han, Haoyu, et al.
Veröffentlicht: (2024)
von: Han, Haoyu, et al.
Veröffentlicht: (2024)
Online Learning via Memory: Retrieval-Augmented Detector Adaptation
von: Jian, Yanan, et al.
Veröffentlicht: (2024)
von: Jian, Yanan, et al.
Veröffentlicht: (2024)
VLM-KG: Multimodal Radiology Knowledge Graph Generation
von: Abdullah, Abdullah, et al.
Veröffentlicht: (2025)
von: Abdullah, Abdullah, et al.
Veröffentlicht: (2025)
Tile-Based ViT Inference with Visual-Cluster Priors for Zero-Shot Multi-Species Plant Identification
von: Gustineli, Murilo, et al.
Veröffentlicht: (2025)
von: Gustineli, Murilo, et al.
Veröffentlicht: (2025)
ZeShot-VQA: Zero-Shot Visual Question Answering Framework with Answer Mapping for Natural Disaster Damage Assessment
von: Karimi, Ehsan, et al.
Veröffentlicht: (2025)
von: Karimi, Ehsan, et al.
Veröffentlicht: (2025)
Towards Universal Video Retrieval: Generalizing Video Embedding via Synthesized Multimodal Pyramid Curriculum
von: Guo, Zhuoning, et al.
Veröffentlicht: (2025)
von: Guo, Zhuoning, et al.
Veröffentlicht: (2025)
Large Language Model Informed Patent Image Retrieval
von: Lo, Hao-Cheng, et al.
Veröffentlicht: (2024)
von: Lo, Hao-Cheng, et al.
Veröffentlicht: (2024)
Open Multimodal Retrieval-Augmented Factual Image Generation
von: Tian, Yang, et al.
Veröffentlicht: (2025)
von: Tian, Yang, et al.
Veröffentlicht: (2025)
Generalized Contrastive Learning for Multi-Modal Retrieval and Ranking
von: Zhu, Tianyu, et al.
Veröffentlicht: (2024)
von: Zhu, Tianyu, et al.
Veröffentlicht: (2024)
VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents
von: Yu, Shi, et al.
Veröffentlicht: (2024)
von: Yu, Shi, et al.
Veröffentlicht: (2024)
MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs
von: Lin, Sheng-Chieh, et al.
Veröffentlicht: (2024)
von: Lin, Sheng-Chieh, et al.
Veröffentlicht: (2024)
Retrieval-augmented Prompt Learning for Pre-trained Foundation Models
von: Chen, Xiang, et al.
Veröffentlicht: (2025)
von: Chen, Xiang, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Visual Question Answering via Built-in Autoregressive Search Engines
von: Long, Xinwei, et al.
Veröffentlicht: (2025)
von: Long, Xinwei, et al.
Veröffentlicht: (2025)
FM2DS: Few-Shot Multimodal Multihop Data Synthesis with Knowledge Distillation for Question Answering
von: Abaskohi, Amirhossein, et al.
Veröffentlicht: (2024)
von: Abaskohi, Amirhossein, et al.
Veröffentlicht: (2024)
Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval
von: Deanda, Demetrio, et al.
Veröffentlicht: (2025)
von: Deanda, Demetrio, et al.
Veröffentlicht: (2025)
Using Knowledge Graphs to harvest datasets for efficient CLIP model training
von: Ging, Simon, et al.
Veröffentlicht: (2025)
von: Ging, Simon, et al.
Veröffentlicht: (2025)
It's Not a Modality Gap: Characterizing and Addressing the Contrastive Gap
von: Fahim, Abrar, et al.
Veröffentlicht: (2024)
von: Fahim, Abrar, et al.
Veröffentlicht: (2024)
Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval
von: Tu, Rong-Cheng, et al.
Veröffentlicht: (2025)
von: Tu, Rong-Cheng, et al.
Veröffentlicht: (2025)
Reasoning-Augmented Representations for Multimodal Retrieval
von: Zhang, Jianrui, et al.
Veröffentlicht: (2026)
von: Zhang, Jianrui, et al.
Veröffentlicht: (2026)
Reversed in Time: A Novel Temporal-Emphasized Benchmark for Cross-Modal Video-Text Retrieval
von: Du, Yang, et al.
Veröffentlicht: (2024)
von: Du, Yang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
VideoRAG: Retrieval-Augmented Generation over Video Corpus
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025) -
UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities
von: Yeo, Woongyeong, et al.
Veröffentlicht: (2025) -
TabRAG: Improving Tabular Document Question Answering for Retrieval Augmented Generation via Structured Representations
von: Si, Jacob, et al.
Veröffentlicht: (2025) -
iRAG: Advancing RAG for Videos with an Incremental Approach
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2024) -
RAG-Check: Evaluating Multimodal Retrieval Augmented Generation Performance
von: Mortaheb, Matin, et al.
Veröffentlicht: (2025)