Guardado en:
| Autores principales: | Rizk, Basem, Walsh, Joel, Core, Mark, Nye, Benjamin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2510.01513 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Indexing Multimodal Language Models for Large-scale Image Retrieval
por: Tharwat, Bahey, et al.
Publicado: (2026)
por: Tharwat, Bahey, et al.
Publicado: (2026)
VLM-KG: Multimodal Radiology Knowledge Graph Generation
por: Abdullah, Abdullah, et al.
Publicado: (2025)
por: Abdullah, Abdullah, et al.
Publicado: (2025)
CLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval
por: Wan, David, et al.
Publicado: (2025)
por: Wan, David, et al.
Publicado: (2025)
DSRAG: A Domain-Specific Retrieval Framework Based on Document-derived Multimodal Knowledge Graph
por: Yang, Mengzheng, et al.
Publicado: (2025)
por: Yang, Mengzheng, et al.
Publicado: (2025)
MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation
por: Hsiao, Chi-Hsiang, et al.
Publicado: (2025)
por: Hsiao, Chi-Hsiang, et al.
Publicado: (2025)
Multi-Vector Index Compression in Any Modality
por: Qin, Hanxiang, et al.
Publicado: (2026)
por: Qin, Hanxiang, et al.
Publicado: (2026)
SemRAG: Semantic Knowledge-Augmented RAG for Improved Question-Answering
por: Zhong, Kezhen, et al.
Publicado: (2025)
por: Zhong, Kezhen, et al.
Publicado: (2025)
GeoOutageKG: A Multimodal Geospatiotemporal Knowledge Graph for Multiresolution Power Outage Analysis
por: Frakes, Ethan, et al.
Publicado: (2025)
por: Frakes, Ethan, et al.
Publicado: (2025)
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
por: Guo, Minghao, et al.
Publicado: (2026)
por: Guo, Minghao, et al.
Publicado: (2026)
Few-Shot Prompting for Extractive Quranic QA with Instruction-Tuned LLMs
por: Basem, Mohamed, et al.
Publicado: (2025)
por: Basem, Mohamed, et al.
Publicado: (2025)
Cross-Language Approach for Quranic QA
por: Oshallah, Islam, et al.
Publicado: (2025)
por: Oshallah, Islam, et al.
Publicado: (2025)
Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation
por: Martin, Alexander, et al.
Publicado: (2025)
por: Martin, Alexander, et al.
Publicado: (2025)
VideoAgent: Long-form Video Understanding with Large Language Model as Agent
por: Wang, Xiaohan, et al.
Publicado: (2024)
por: Wang, Xiaohan, et al.
Publicado: (2024)
An Index-based Approach for Efficient and Effective Web Content Extraction
por: Chen, Yihan, et al.
Publicado: (2025)
por: Chen, Yihan, et al.
Publicado: (2025)
The Structure-Content Trade-off in Knowledge Graph Retrieval
por: Six, Valentin, et al.
Publicado: (2025)
por: Six, Valentin, et al.
Publicado: (2025)
Improving Content Recommendation: Knowledge Graph-Based Semantic Contrastive Learning for Diversity and Cold-Start Users
por: Kim, Yejin, et al.
Publicado: (2024)
por: Kim, Yejin, et al.
Publicado: (2024)
Optimized Quran Passage Retrieval Using an Expanded QA Dataset and Fine-Tuned Language Models
por: Basem, Mohamed, et al.
Publicado: (2024)
por: Basem, Mohamed, et al.
Publicado: (2024)
Two-Stage Quranic QA via Ensemble Retrieval and Instruction-Tuned Answer Extraction
por: Basem, Mohamed, et al.
Publicado: (2025)
por: Basem, Mohamed, et al.
Publicado: (2025)
Self Knowledge Re-expression: A Fully Local Method for Adapting LLMs to Tasks Using Intrinsic Knowledge
por: Wang, Mengyu, et al.
Publicado: (2026)
por: Wang, Mengyu, et al.
Publicado: (2026)
Index Light, Reason Deep: Deferred Visual Ingestion for Visual-Dense Document Question Answering
por: Xu, Tao
Publicado: (2026)
por: Xu, Tao
Publicado: (2026)
Using Knowledge Graphs to harvest datasets for efficient CLIP model training
por: Ging, Simon, et al.
Publicado: (2025)
por: Ging, Simon, et al.
Publicado: (2025)
From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents
por: Lian, Niu, et al.
Publicado: (2026)
por: Lian, Niu, et al.
Publicado: (2026)
KGMEL: Knowledge Graph-Enhanced Multimodal Entity Linking
por: Kim, Juyeon, et al.
Publicado: (2025)
por: Kim, Juyeon, et al.
Publicado: (2025)
Windsock is Dancing: Adaptive Multimodal Retrieval-Augmented Generation
por: Zhao, Shu, et al.
Publicado: (2025)
por: Zhao, Shu, et al.
Publicado: (2025)
X-Reflect: Cross-Reflection Prompting for Multimodal Recommendation
por: Lyu, Hanjia, et al.
Publicado: (2024)
por: Lyu, Hanjia, et al.
Publicado: (2024)
A Survey of Knowledge Graph Reasoning on Graph Types: Static, Dynamic, and Multimodal
por: Liang, Ke, et al.
Publicado: (2022)
por: Liang, Ke, et al.
Publicado: (2022)
Ontology-Based Knowledge Graph Framework for Industrial Standard Documents via Hierarchical and Propositional Structuring
por: Park, Jiin, et al.
Publicado: (2025)
por: Park, Jiin, et al.
Publicado: (2025)
InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search
por: Hou, Bohan, et al.
Publicado: (2026)
por: Hou, Bohan, et al.
Publicado: (2026)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
por: Dong, Kuicai, et al.
Publicado: (2025)
por: Dong, Kuicai, et al.
Publicado: (2025)
E5-V: Universal Embeddings with Multimodal Large Language Models
por: Jiang, Ting, et al.
Publicado: (2024)
por: Jiang, Ting, et al.
Publicado: (2024)
NativE: Multi-modal Knowledge Graph Completion in the Wild
por: Zhang, Yichi, et al.
Publicado: (2024)
por: Zhang, Yichi, et al.
Publicado: (2024)
ConceptFormer: Towards Efficient Use of Knowledge-Graph Embeddings in Large Language Models
por: Barmettler, Joel, et al.
Publicado: (2025)
por: Barmettler, Joel, et al.
Publicado: (2025)
Supervised Fine-Tuning or Contrastive Learning? Towards Better Multimodal LLM Reranking
por: Dai, Ziqi, et al.
Publicado: (2025)
por: Dai, Ziqi, et al.
Publicado: (2025)
MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interaction
por: Xiao, Zilin, et al.
Publicado: (2025)
por: Xiao, Zilin, et al.
Publicado: (2025)
Understanding Parametric Knowledge Injection in Retrieval-Augmented Generation
por: Tang, Minghao, et al.
Publicado: (2025)
por: Tang, Minghao, et al.
Publicado: (2025)
Do Recommender Systems Really Leverage Multimodal Content? A Comprehensive Analysis on Multimodal Representations for Recommendation
por: Pomo, Claudio, et al.
Publicado: (2025)
por: Pomo, Claudio, et al.
Publicado: (2025)
CollEX -- A Multimodal Agentic RAG System Enabling Interactive Exploration of Scientific Collections
por: Schneider, Florian, et al.
Publicado: (2025)
por: Schneider, Florian, et al.
Publicado: (2025)
WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering
por: Zhu, Yingjian, et al.
Publicado: (2026)
por: Zhu, Yingjian, et al.
Publicado: (2026)
Think When Needed: Adaptive Reasoning-Driven Multimodal Embeddings with a Dual-LoRA Architecture
por: Zhang, Longxiang, et al.
Publicado: (2026)
por: Zhang, Longxiang, et al.
Publicado: (2026)
Structurally Refined Graph Transformer for Multimodal Recommendation
por: Shi, Ke, et al.
Publicado: (2025)
por: Shi, Ke, et al.
Publicado: (2025)
Ejemplares similares
-
Indexing Multimodal Language Models for Large-scale Image Retrieval
por: Tharwat, Bahey, et al.
Publicado: (2026) -
VLM-KG: Multimodal Radiology Knowledge Graph Generation
por: Abdullah, Abdullah, et al.
Publicado: (2025) -
CLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval
por: Wan, David, et al.
Publicado: (2025) -
DSRAG: A Domain-Specific Retrieval Framework Based on Document-derived Multimodal Knowledge Graph
por: Yang, Mengzheng, et al.
Publicado: (2025) -
MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation
por: Hsiao, Chi-Hsiang, et al.
Publicado: (2025)