LIVE: Learnable In-Context Vector for Visual Question Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Yingzhe, Hao, Chenduo, Yang, Xu, Peng, Jiawei, Hu, Xinting, Geng, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mimic In-Context Learning for Multimodal Tasks
by: Jiang, Yuchu, et al.
Published: (2025)
by: Jiang, Yuchu, et al.
Published: (2025)
Rationale-guided Prompting for Knowledge-based Visual Question Answering
by: Hu, Zhongjian, et al.
Published: (2024)
by: Hu, Zhongjian, et al.
Published: (2024)
Multi-Agents Based on Large Language Models for Knowledge-based Visual Question Answering
by: Hu, Zhongjian, et al.
Published: (2024)
by: Hu, Zhongjian, et al.
Published: (2024)
Improving Retrieval Augmented Open-Domain Question-Answering with Vectorized Contexts
by: Chen, Zhuo, et al.
Published: (2024)
by: Chen, Zhuo, et al.
Published: (2024)
Overcoming Language Priors for Visual Question Answering Based on Knowledge Distillation
by: Peng, Daowan, et al.
Published: (2025)
by: Peng, Daowan, et al.
Published: (2025)
Efficient Multimodal Planning Agent for Visual Question-Answering
by: Chen, Zhuo, et al.
Published: (2026)
by: Chen, Zhuo, et al.
Published: (2026)
Exploring Diverse Methods in Visual Question Answering
by: Li, Panfeng, et al.
Published: (2024)
by: Li, Panfeng, et al.
Published: (2024)
Enhancing Question Answering Precision with Optimized Vector Retrieval and Instructions
by: Yang, Lixiao, et al.
Published: (2024)
by: Yang, Lixiao, et al.
Published: (2024)
CommVQA: Situating Visual Question Answering in Communicative Contexts
by: Naik, Nandita Shankar, et al.
Published: (2024)
by: Naik, Nandita Shankar, et al.
Published: (2024)
LIVE: LaTex Interactive Visual Editing
by: Lin, Jinwei
Published: (2024)
by: Lin, Jinwei
Published: (2024)
What Factors Affect LLMs and RLLMs in Financial Question Answering?
by: Wang, Peng, et al.
Published: (2025)
by: Wang, Peng, et al.
Published: (2025)
Condition-Gated Reasoning for Context-Dependent Biomedical Question Answering
by: Parekh, Jash Rajesh, et al.
Published: (2026)
by: Parekh, Jash Rajesh, et al.
Published: (2026)
Desiderata for the Context Use of Question Answering Systems
by: Shaier, Sagi, et al.
Published: (2024)
by: Shaier, Sagi, et al.
Published: (2024)
Context Filtering with Reward Modeling in Question Answering
by: Kim, Sangryul, et al.
Published: (2024)
by: Kim, Sangryul, et al.
Published: (2024)
Selectively Answering Visual Questions
by: Eisenschlos, Julian Martin, et al.
Published: (2024)
by: Eisenschlos, Julian Martin, et al.
Published: (2024)
Lever LM: Configuring In-Context Sequence to Lever Large Vision Language Models
by: Yang, Xu, et al.
Published: (2023)
by: Yang, Xu, et al.
Published: (2023)
Design as Desired: Utilizing Visual Question Answering for Multimodal Pre-training
by: Su, Tongkun, et al.
Published: (2024)
by: Su, Tongkun, et al.
Published: (2024)
EVJVQA Challenge: Multilingual Visual Question Answering
by: Nguyen, Ngan Luu-Thuy, et al.
Published: (2023)
by: Nguyen, Ngan Luu-Thuy, et al.
Published: (2023)
PathReasoner: Modeling Reasoning Path with Equivalent Extension for Logical Question Answering
by: Xu, Fangzhi, et al.
Published: (2024)
by: Xu, Fangzhi, et al.
Published: (2024)
Multimodal Commonsense Knowledge Distillation for Visual Question Answering
by: Yang, Shuo, et al.
Published: (2024)
by: Yang, Shuo, et al.
Published: (2024)
Open Domain Question Answering with Conflicting Contexts
by: Liu, Siyi, et al.
Published: (2024)
by: Liu, Siyi, et al.
Published: (2024)
Context Convergence Improves Answering Inferential Questions
by: Mozafari, Jamshid, et al.
Published: (2026)
by: Mozafari, Jamshid, et al.
Published: (2026)
EffiQA: Efficient Question-Answering with Strategic Multi-Model Collaboration on Knowledge Graphs
by: Dong, Zixuan, et al.
Published: (2024)
by: Dong, Zixuan, et al.
Published: (2024)
GraphWalker: Agentic Knowledge Graph Question Answering via Synthetic Trajectory Curriculum
by: Xu, Shuwen, et al.
Published: (2026)
by: Xu, Shuwen, et al.
Published: (2026)
TQA-Bench: Evaluating LLMs for Multi-Table Question Answering with Scalable Context and Symbolic Extension
by: Qiu, Zipeng, et al.
Published: (2024)
by: Qiu, Zipeng, et al.
Published: (2024)
SplaXBERT: Leveraging Mixed Precision Training and Context Splitting for Question Answering
by: Yufan, Zhu, et al.
Published: (2024)
by: Yufan, Zhu, et al.
Published: (2024)
A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts
by: Ge, Suyu, et al.
Published: (2024)
by: Ge, Suyu, et al.
Published: (2024)
MultiCube-RAG for Multi-hop Question Answering
by: Shi, Jimeng, et al.
Published: (2026)
by: Shi, Jimeng, et al.
Published: (2026)
An In-Context Schema Understanding Method for Knowledge Base Question Answering
by: Liu, Yantao, et al.
Published: (2023)
by: Liu, Yantao, et al.
Published: (2023)
Evaluating Long-Term Memory for Long-Context Question Answering
by: Terranova, Alessandra, et al.
Published: (2025)
by: Terranova, Alessandra, et al.
Published: (2025)
DARE: Diverse Visual Question Answering with Robustness Evaluation
by: Sterz, Hannah, et al.
Published: (2024)
by: Sterz, Hannah, et al.
Published: (2024)
Knowledge-Based Counterfactual Queries for Visual Question Answering
by: Stoikou, Theodoti, et al.
Published: (2023)
by: Stoikou, Theodoti, et al.
Published: (2023)
LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
by: Peng, Yingzhe, et al.
Published: (2025)
by: Peng, Yingzhe, et al.
Published: (2025)
Privacy-protected Retrieval-Augmented Generation for Knowledge Graph Question Answering
by: Ning, Yunfeng, et al.
Published: (2025)
by: Ning, Yunfeng, et al.
Published: (2025)
DragonVerseQA: Open-Domain Long-Form Context-Aware Question-Answering
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
Continual Learning for Temporal-Sensitive Question Answering
by: Yang, Wanqi, et al.
Published: (2024)
by: Yang, Wanqi, et al.
Published: (2024)
Investigating LLM Capabilities on Long Context Comprehension for Medical Question Answering
by: AlMannaa, Feras, et al.
Published: (2025)
by: AlMannaa, Feras, et al.
Published: (2025)
A Comprehensive Evaluation of GPT-4V on Knowledge-Intensive Visual Question Answering
by: Li, Yunxin, et al.
Published: (2023)
by: Li, Yunxin, et al.
Published: (2023)
Index Light, Reason Deep: Deferred Visual Ingestion for Visual-Dense Document Question Answering
by: Xu, Tao
Published: (2026)
by: Xu, Tao
Published: (2026)
Similar Items
-
Mimic In-Context Learning for Multimodal Tasks
by: Jiang, Yuchu, et al.
Published: (2025) -
Rationale-guided Prompting for Knowledge-based Visual Question Answering
by: Hu, Zhongjian, et al.
Published: (2024) -
Multi-Agents Based on Large Language Models for Knowledge-based Visual Question Answering
by: Hu, Zhongjian, et al.
Published: (2024) -
Improving Retrieval Augmented Open-Domain Question-Answering with Vectorized Contexts
by: Chen, Zhuo, et al.
Published: (2024) -
Overcoming Language Priors for Visual Question Answering Based on Knowledge Distillation
by: Peng, Daowan, et al.
Published: (2025)