RoRA-VLM: Robust Retrieval-Augmented Vision Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qi, Jingyuan, Xu, Zhiyang, Shao, Rulin, Chen, Yang, Di, Jin, Cheng, Yu, Wang, Qifan, Huang, Lifu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AR-RAG: Autoregressive Retrieval Augmentation for Image Generation
von: Qi, Jingyuan, et al.
Veröffentlicht: (2025)
von: Qi, Jingyuan, et al.
Veröffentlicht: (2025)
Vision-Flan: Scaling Human-Labeled Tasks in Visual Instruction Tuning
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024)
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024)
Modality-Specialized Synergizers for Interleaved Vision-Language Generalists
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024)
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024)
RoRA: Efficient Fine-Tuning of LLM with Reliability Optimization for Rank Adaptation
von: Liu, Jun, et al.
Veröffentlicht: (2025)
von: Liu, Jun, et al.
Veröffentlicht: (2025)
MULTISCRIPT: Multimodal Script Learning for Supporting Open Domain Everyday Tasks
von: Qi, Jingyuan, et al.
Veröffentlicht: (2023)
von: Qi, Jingyuan, et al.
Veröffentlicht: (2023)
UniHGKR: Unified Instruction-aware Heterogeneous Knowledge Retrievers
von: Min, Dehai, et al.
Veröffentlicht: (2024)
von: Min, Dehai, et al.
Veröffentlicht: (2024)
Multimodal Instruction Tuning with Conditional Mixture of LoRA
von: Shen, Ying, et al.
Veröffentlicht: (2024)
von: Shen, Ying, et al.
Veröffentlicht: (2024)
Inference Compute-Optimal Video Vision Language Models
von: Wang, Peiqi, et al.
Veröffentlicht: (2025)
von: Wang, Peiqi, et al.
Veröffentlicht: (2025)
Error-driven Data-efficient Large Multimodal Model Tuning
von: Yao, Barry Menglong, et al.
Veröffentlicht: (2024)
von: Yao, Barry Menglong, et al.
Veröffentlicht: (2024)
RoSA: Enhancing Parameter-Efficient Fine-Tuning via RoPE-aware Selective Adaptation in Large Language Models
von: Pan, Dayan, et al.
Veröffentlicht: (2025)
von: Pan, Dayan, et al.
Veröffentlicht: (2025)
AMELI: Enhancing Multimodal Entity Linking with Fine-Grained Attributes
von: Yao, Barry Menglong, et al.
Veröffentlicht: (2023)
von: Yao, Barry Menglong, et al.
Veröffentlicht: (2023)
Soft Prompt Tuning for Augmenting Dense Retrieval with Large Language Models
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2023)
Targeted Augmentation for Low-Resource Event Extraction
von: Wang, Sijia, et al.
Veröffentlicht: (2024)
von: Wang, Sijia, et al.
Veröffentlicht: (2024)
Chain-of-Note: Enhancing Robustness in Retrieval-Augmented Language Models
von: Yu, Wenhao, et al.
Veröffentlicht: (2023)
von: Yu, Wenhao, et al.
Veröffentlicht: (2023)
RoCoIns: Enhancing Robustness of Large Language Models through Code-Style Instructions
von: Zhang, Yuansen, et al.
Veröffentlicht: (2024)
von: Zhang, Yuansen, et al.
Veröffentlicht: (2024)
SparkRA: A Retrieval-Augmented Knowledge Service System Based on Spark Large Language Model
von: Wu, Dayong, et al.
Veröffentlicht: (2024)
von: Wu, Dayong, et al.
Veröffentlicht: (2024)
Debate as Optimization: Adaptive Conformal Prediction and Diverse Retrieval for Event Extraction
von: Wang, Sijia, et al.
Veröffentlicht: (2024)
von: Wang, Sijia, et al.
Veröffentlicht: (2024)
How Do Large Language Models Learn Concepts During Continual Pre-Training?
von: Yao, Barry Menglong, et al.
Veröffentlicht: (2026)
von: Yao, Barry Menglong, et al.
Veröffentlicht: (2026)
RoLoRA: Fine-tuning Rotated Outlier-free LLMs for Effective Weight-Activation Quantization
von: Huang, Xijie, et al.
Veröffentlicht: (2024)
von: Huang, Xijie, et al.
Veröffentlicht: (2024)
Benchmarking Retrieval-Augmented Large Language Models in Biomedical NLP: Application, Robustness, and Self-Awareness
von: Li, Mingchen, et al.
Veröffentlicht: (2024)
von: Li, Mingchen, et al.
Veröffentlicht: (2024)
Unleashing Implicit Rewards: Prefix-Value Learning for Distribution-Level Optimization
von: Gao, Shiping, et al.
Veröffentlicht: (2026)
von: Gao, Shiping, et al.
Veröffentlicht: (2026)
VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction
von: Fan, Zhiwen, et al.
Veröffentlicht: (2025)
von: Fan, Zhiwen, et al.
Veröffentlicht: (2025)
Training a Utility-based Retriever Through Shared Context Attribution for Retrieval-Augmented Language Models
von: Xu, Yilong, et al.
Veröffentlicht: (2025)
von: Xu, Yilong, et al.
Veröffentlicht: (2025)
InternalInspector $I^2$: Robust Confidence Estimation in LLMs through Internal States
von: Beigi, Mohammad, et al.
Veröffentlicht: (2024)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2024)
RPO: Retrieval Preference Optimization for Robust Retrieval-Augmented Generation
von: Yan, Shi-Qi, et al.
Veröffentlicht: (2025)
von: Yan, Shi-Qi, et al.
Veröffentlicht: (2025)
RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
RAGViz: Diagnose and Visualize Retrieval-Augmented Generation
von: Wang, Tevin, et al.
Veröffentlicht: (2024)
von: Wang, Tevin, et al.
Veröffentlicht: (2024)
HPE-CogVLM: Advancing Vision Language Models with a Head Pose Grounding Task
von: Tian, Yu, et al.
Veröffentlicht: (2024)
von: Tian, Yu, et al.
Veröffentlicht: (2024)
Toward Robust RALMs: Revealing the Impact of Imperfect Retrieval on Retrieval-Augmented Language Models
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
Think, Act, Build: An Agentic Framework with Vision Language Models for Zero-Shot 3D Visual Grounding
von: Wang, Haibo, et al.
Veröffentlicht: (2026)
von: Wang, Haibo, et al.
Veröffentlicht: (2026)
Frustratingly Simple Retrieval Improves Challenging, Reasoning-Intensive Benchmarks
von: Lyu, Xinxi, et al.
Veröffentlicht: (2025)
von: Lyu, Xinxi, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Language Model for Extreme Multi-Label Knowledge Graph Link Prediction
von: Lin, Yu-Hsiang, et al.
Veröffentlicht: (2024)
von: Lin, Yu-Hsiang, et al.
Veröffentlicht: (2024)
X-Eval: Generalizable Multi-aspect Text Evaluation via Augmented Instruction Tuning with Auxiliary Evaluation Aspects
von: Liu, Minqian, et al.
Veröffentlicht: (2023)
von: Liu, Minqian, et al.
Veröffentlicht: (2023)
Scaling Retrieval-Based Language Models with a Trillion-Token Datastore
von: Shao, Rulin, et al.
Veröffentlicht: (2024)
von: Shao, Rulin, et al.
Veröffentlicht: (2024)
ICONS: Influence Consensus for Vision-Language Data Selection
von: Wu, Xindi, et al.
Veröffentlicht: (2024)
von: Wu, Xindi, et al.
Veröffentlicht: (2024)
Auto-RAG: Autonomous Retrieval-Augmented Generation for Large Language Models
von: Yu, Tian, et al.
Veröffentlicht: (2024)
von: Yu, Tian, et al.
Veröffentlicht: (2024)
Retrieval Augmented Generation Evaluation in the Era of Large Language Models: A Comprehensive Survey
von: Gan, Aoran, et al.
Veröffentlicht: (2025)
von: Gan, Aoran, et al.
Veröffentlicht: (2025)
Navigating the Dual Facets: A Comprehensive Evaluation of Sequential Memory Editing in Large Language Models
von: Lin, Zihao, et al.
Veröffentlicht: (2024)
von: Lin, Zihao, et al.
Veröffentlicht: (2024)
RA-DIT: Retrieval-Augmented Dual Instruction Tuning
von: Lin, Xi Victoria, et al.
Veröffentlicht: (2023)
von: Lin, Xi Victoria, et al.
Veröffentlicht: (2023)
Making Retrieval-Augmented Language Models Robust to Irrelevant Context
von: Yoran, Ori, et al.
Veröffentlicht: (2023)
von: Yoran, Ori, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
AR-RAG: Autoregressive Retrieval Augmentation for Image Generation
von: Qi, Jingyuan, et al.
Veröffentlicht: (2025) -
Vision-Flan: Scaling Human-Labeled Tasks in Visual Instruction Tuning
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024) -
Modality-Specialized Synergizers for Interleaved Vision-Language Generalists
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024) -
RoRA: Efficient Fine-Tuning of LLM with Reliability Optimization for Rank Adaptation
von: Liu, Jun, et al.
Veröffentlicht: (2025) -
MULTISCRIPT: Multimodal Script Learning for Supporting Open Domain Everyday Tasks
von: Qi, Jingyuan, et al.
Veröffentlicht: (2023)