Gespeichert in:
| Hauptverfasser: | Rizk, Basem, Walsh, Joel, Core, Mark, Nye, Benjamin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2510.01513 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Indexing Multimodal Language Models for Large-scale Image Retrieval
von: Tharwat, Bahey, et al.
Veröffentlicht: (2026)
von: Tharwat, Bahey, et al.
Veröffentlicht: (2026)
VLM-KG: Multimodal Radiology Knowledge Graph Generation
von: Abdullah, Abdullah, et al.
Veröffentlicht: (2025)
von: Abdullah, Abdullah, et al.
Veröffentlicht: (2025)
CLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval
von: Wan, David, et al.
Veröffentlicht: (2025)
von: Wan, David, et al.
Veröffentlicht: (2025)
DSRAG: A Domain-Specific Retrieval Framework Based on Document-derived Multimodal Knowledge Graph
von: Yang, Mengzheng, et al.
Veröffentlicht: (2025)
von: Yang, Mengzheng, et al.
Veröffentlicht: (2025)
MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation
von: Hsiao, Chi-Hsiang, et al.
Veröffentlicht: (2025)
von: Hsiao, Chi-Hsiang, et al.
Veröffentlicht: (2025)
Multi-Vector Index Compression in Any Modality
von: Qin, Hanxiang, et al.
Veröffentlicht: (2026)
von: Qin, Hanxiang, et al.
Veröffentlicht: (2026)
SemRAG: Semantic Knowledge-Augmented RAG for Improved Question-Answering
von: Zhong, Kezhen, et al.
Veröffentlicht: (2025)
von: Zhong, Kezhen, et al.
Veröffentlicht: (2025)
GeoOutageKG: A Multimodal Geospatiotemporal Knowledge Graph for Multiresolution Power Outage Analysis
von: Frakes, Ethan, et al.
Veröffentlicht: (2025)
von: Frakes, Ethan, et al.
Veröffentlicht: (2025)
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
von: Guo, Minghao, et al.
Veröffentlicht: (2026)
von: Guo, Minghao, et al.
Veröffentlicht: (2026)
Few-Shot Prompting for Extractive Quranic QA with Instruction-Tuned LLMs
von: Basem, Mohamed, et al.
Veröffentlicht: (2025)
von: Basem, Mohamed, et al.
Veröffentlicht: (2025)
Cross-Language Approach for Quranic QA
von: Oshallah, Islam, et al.
Veröffentlicht: (2025)
von: Oshallah, Islam, et al.
Veröffentlicht: (2025)
Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation
von: Martin, Alexander, et al.
Veröffentlicht: (2025)
von: Martin, Alexander, et al.
Veröffentlicht: (2025)
VideoAgent: Long-form Video Understanding with Large Language Model as Agent
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
An Index-based Approach for Efficient and Effective Web Content Extraction
von: Chen, Yihan, et al.
Veröffentlicht: (2025)
von: Chen, Yihan, et al.
Veröffentlicht: (2025)
The Structure-Content Trade-off in Knowledge Graph Retrieval
von: Six, Valentin, et al.
Veröffentlicht: (2025)
von: Six, Valentin, et al.
Veröffentlicht: (2025)
Improving Content Recommendation: Knowledge Graph-Based Semantic Contrastive Learning for Diversity and Cold-Start Users
von: Kim, Yejin, et al.
Veröffentlicht: (2024)
von: Kim, Yejin, et al.
Veröffentlicht: (2024)
Optimized Quran Passage Retrieval Using an Expanded QA Dataset and Fine-Tuned Language Models
von: Basem, Mohamed, et al.
Veröffentlicht: (2024)
von: Basem, Mohamed, et al.
Veröffentlicht: (2024)
Two-Stage Quranic QA via Ensemble Retrieval and Instruction-Tuned Answer Extraction
von: Basem, Mohamed, et al.
Veröffentlicht: (2025)
von: Basem, Mohamed, et al.
Veröffentlicht: (2025)
Self Knowledge Re-expression: A Fully Local Method for Adapting LLMs to Tasks Using Intrinsic Knowledge
von: Wang, Mengyu, et al.
Veröffentlicht: (2026)
von: Wang, Mengyu, et al.
Veröffentlicht: (2026)
Index Light, Reason Deep: Deferred Visual Ingestion for Visual-Dense Document Question Answering
von: Xu, Tao
Veröffentlicht: (2026)
von: Xu, Tao
Veröffentlicht: (2026)
Using Knowledge Graphs to harvest datasets for efficient CLIP model training
von: Ging, Simon, et al.
Veröffentlicht: (2025)
von: Ging, Simon, et al.
Veröffentlicht: (2025)
From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents
von: Lian, Niu, et al.
Veröffentlicht: (2026)
von: Lian, Niu, et al.
Veröffentlicht: (2026)
KGMEL: Knowledge Graph-Enhanced Multimodal Entity Linking
von: Kim, Juyeon, et al.
Veröffentlicht: (2025)
von: Kim, Juyeon, et al.
Veröffentlicht: (2025)
Windsock is Dancing: Adaptive Multimodal Retrieval-Augmented Generation
von: Zhao, Shu, et al.
Veröffentlicht: (2025)
von: Zhao, Shu, et al.
Veröffentlicht: (2025)
X-Reflect: Cross-Reflection Prompting for Multimodal Recommendation
von: Lyu, Hanjia, et al.
Veröffentlicht: (2024)
von: Lyu, Hanjia, et al.
Veröffentlicht: (2024)
A Survey of Knowledge Graph Reasoning on Graph Types: Static, Dynamic, and Multimodal
von: Liang, Ke, et al.
Veröffentlicht: (2022)
von: Liang, Ke, et al.
Veröffentlicht: (2022)
Ontology-Based Knowledge Graph Framework for Industrial Standard Documents via Hierarchical and Propositional Structuring
von: Park, Jiin, et al.
Veröffentlicht: (2025)
von: Park, Jiin, et al.
Veröffentlicht: (2025)
InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search
von: Hou, Bohan, et al.
Veröffentlicht: (2026)
von: Hou, Bohan, et al.
Veröffentlicht: (2026)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
E5-V: Universal Embeddings with Multimodal Large Language Models
von: Jiang, Ting, et al.
Veröffentlicht: (2024)
von: Jiang, Ting, et al.
Veröffentlicht: (2024)
NativE: Multi-modal Knowledge Graph Completion in the Wild
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
ConceptFormer: Towards Efficient Use of Knowledge-Graph Embeddings in Large Language Models
von: Barmettler, Joel, et al.
Veröffentlicht: (2025)
von: Barmettler, Joel, et al.
Veröffentlicht: (2025)
Supervised Fine-Tuning or Contrastive Learning? Towards Better Multimodal LLM Reranking
von: Dai, Ziqi, et al.
Veröffentlicht: (2025)
von: Dai, Ziqi, et al.
Veröffentlicht: (2025)
MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interaction
von: Xiao, Zilin, et al.
Veröffentlicht: (2025)
von: Xiao, Zilin, et al.
Veröffentlicht: (2025)
Understanding Parametric Knowledge Injection in Retrieval-Augmented Generation
von: Tang, Minghao, et al.
Veröffentlicht: (2025)
von: Tang, Minghao, et al.
Veröffentlicht: (2025)
Do Recommender Systems Really Leverage Multimodal Content? A Comprehensive Analysis on Multimodal Representations for Recommendation
von: Pomo, Claudio, et al.
Veröffentlicht: (2025)
von: Pomo, Claudio, et al.
Veröffentlicht: (2025)
CollEX -- A Multimodal Agentic RAG System Enabling Interactive Exploration of Scientific Collections
von: Schneider, Florian, et al.
Veröffentlicht: (2025)
von: Schneider, Florian, et al.
Veröffentlicht: (2025)
WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering
von: Zhu, Yingjian, et al.
Veröffentlicht: (2026)
von: Zhu, Yingjian, et al.
Veröffentlicht: (2026)
Think When Needed: Adaptive Reasoning-Driven Multimodal Embeddings with a Dual-LoRA Architecture
von: Zhang, Longxiang, et al.
Veröffentlicht: (2026)
von: Zhang, Longxiang, et al.
Veröffentlicht: (2026)
Structurally Refined Graph Transformer for Multimodal Recommendation
von: Shi, Ke, et al.
Veröffentlicht: (2025)
von: Shi, Ke, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Indexing Multimodal Language Models for Large-scale Image Retrieval
von: Tharwat, Bahey, et al.
Veröffentlicht: (2026) -
VLM-KG: Multimodal Radiology Knowledge Graph Generation
von: Abdullah, Abdullah, et al.
Veröffentlicht: (2025) -
CLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval
von: Wan, David, et al.
Veröffentlicht: (2025) -
DSRAG: A Domain-Specific Retrieval Framework Based on Document-derived Multimodal Knowledge Graph
von: Yang, Mengzheng, et al.
Veröffentlicht: (2025) -
MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation
von: Hsiao, Chi-Hsiang, et al.
Veröffentlicht: (2025)