Towards Efficient Visual-Language Alignment of the Q-Former for Visual Reasoning Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Sungkyung, Lee, Adam, Park, Junyoung, Chung, Andrew, Oh, Jusang, Lee, Jay-Yoon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models
von: Kim, Woojin, et al.
Veröffentlicht: (2026)
von: Kim, Woojin, et al.
Veröffentlicht: (2026)
Expanding Search Space with Diverse Prompting Agents: An Efficient Sampling Approach for LLM Mathematical Reasoning
von: Lee, Gisang, et al.
Veröffentlicht: (2024)
von: Lee, Gisang, et al.
Veröffentlicht: (2024)
Toward Robust RALMs: Revealing the Impact of Imperfect Retrieval on Retrieval-Augmented Language Models
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
Latent Self-Consistency for Reliable Majority-Set Selection in Short- and Long-Answer Reasoning
von: Oh, Jungsuk, et al.
Veröffentlicht: (2025)
von: Oh, Jungsuk, et al.
Veröffentlicht: (2025)
More Human, More Efficient: Aligning Annotations with Quantized SLMs
von: Wang, Jiayu, et al.
Veröffentlicht: (2026)
von: Wang, Jiayu, et al.
Veröffentlicht: (2026)
Selective Vision is the Challenge for Visual Reasoning: A Benchmark for Visual Argument Understanding
von: Chung, Jiwan, et al.
Veröffentlicht: (2024)
von: Chung, Jiwan, et al.
Veröffentlicht: (2024)
Case-Based Reasoning Approach for Solving Financial Question Answering
von: Kim, Yikyung, et al.
Veröffentlicht: (2024)
von: Kim, Yikyung, et al.
Veröffentlicht: (2024)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
Thanos: Enhancing Conversational Agents with Skill-of-Mind-Infused Large Language Model
von: Lee, Young-Jun, et al.
Veröffentlicht: (2024)
von: Lee, Young-Jun, et al.
Veröffentlicht: (2024)
MATA: Multi-Agent Framework for Reliable and Flexible Table Question Answering
von: Hyeon, Sieun, et al.
Veröffentlicht: (2026)
von: Hyeon, Sieun, et al.
Veröffentlicht: (2026)
SDS KoPub VDR: A Benchmark Dataset for Visual Document Retrieval in Korean Public Documents
von: Lee, Jaehoon, et al.
Veröffentlicht: (2025)
von: Lee, Jaehoon, et al.
Veröffentlicht: (2025)
RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation
von: Kim, Kiseung, et al.
Veröffentlicht: (2024)
von: Kim, Kiseung, et al.
Veröffentlicht: (2024)
RPM: Reasoning-Level Personalization for Black-Box Large Language Models
von: Kim, Jieyong, et al.
Veröffentlicht: (2025)
von: Kim, Jieyong, et al.
Veröffentlicht: (2025)
ECG-Reasoning-Benchmark: A Benchmark for Evaluating Clinical Reasoning Capabilities in ECG Interpretation
von: Oh, Jungwoo, et al.
Veröffentlicht: (2026)
von: Oh, Jungwoo, et al.
Veröffentlicht: (2026)
How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding
von: Yu, Zhuoran, et al.
Veröffentlicht: (2025)
von: Yu, Zhuoran, et al.
Veröffentlicht: (2025)
The Curious Case of Analogies: Investigating Analogical Reasoning in Large Language Models
von: Lee, Taewhoo, et al.
Veröffentlicht: (2025)
von: Lee, Taewhoo, et al.
Veröffentlicht: (2025)
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
Enhancing Robustness of Retrieval-Augmented Language Models with In-Context Learning
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
Adaptive Task Vectors for Large Language Models
von: Kang, Joonseong, et al.
Veröffentlicht: (2025)
von: Kang, Joonseong, et al.
Veröffentlicht: (2025)
Unleashing Multi-Hop Reasoning Potential in Large Language Models through Repetition of Misordered Context
von: Yu, Sangwon, et al.
Veröffentlicht: (2024)
von: Yu, Sangwon, et al.
Veröffentlicht: (2024)
Small Language Models are Equation Reasoners
von: Kim, Bumjun, et al.
Veröffentlicht: (2024)
von: Kim, Bumjun, et al.
Veröffentlicht: (2024)
Locate&Edit: Energy-based Text Editing for Efficient, Flexible, and Faithful Controlled Text Generation
von: Son, Hye Ryung, et al.
Veröffentlicht: (2024)
von: Son, Hye Ryung, et al.
Veröffentlicht: (2024)
On Efficient Language and Vision Assistants for Visually-Situated Natural Language Understanding: What Matters in Reading and Reasoning
von: Kim, Geewook, et al.
Veröffentlicht: (2024)
von: Kim, Geewook, et al.
Veröffentlicht: (2024)
RadZero: Similarity-Based Cross-Attention for Explainable Vision-Language Alignment in Chest X-ray with Zero-Shot Multi-Task Capability
von: Park, Jonggwon, et al.
Veröffentlicht: (2025)
von: Park, Jonggwon, et al.
Veröffentlicht: (2025)
Task-Aware Resolution Optimization for Visual Large Language Models
von: Luo, Weiqing, et al.
Veröffentlicht: (2025)
von: Luo, Weiqing, et al.
Veröffentlicht: (2025)
Alignment Data Map for Efficient Preference Data Selection and Diagnosis
von: Lee, Seohyeong, et al.
Veröffentlicht: (2025)
von: Lee, Seohyeong, et al.
Veröffentlicht: (2025)
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
von: Kim, Minchan, et al.
Veröffentlicht: (2024)
von: Kim, Minchan, et al.
Veröffentlicht: (2024)
ETHIC: Evaluating Large Language Models on Long-Context Tasks with High Information Coverage
von: Lee, Taewhoo, et al.
Veröffentlicht: (2024)
von: Lee, Taewhoo, et al.
Veröffentlicht: (2024)
Introducing Verification Task of Set Consistency with Set-Consistency Energy Networks
von: Song, Mooho, et al.
Veröffentlicht: (2025)
von: Song, Mooho, et al.
Veröffentlicht: (2025)
GraphCheck: Multipath Fact-Checking with Entity-Relationship Graphs
von: Jeon, Hyewon, et al.
Veröffentlicht: (2025)
von: Jeon, Hyewon, et al.
Veröffentlicht: (2025)
R1-ACT: Efficient Reasoning Model Safety Alignment by Activating Safety Knowledge
von: In, Yeonjun, et al.
Veröffentlicht: (2025)
von: In, Yeonjun, et al.
Veröffentlicht: (2025)
Safeguarding Privacy of Retrieval Data against Membership Inference Attacks: Is This Query Too Close to Home?
von: Choi, Yujin, et al.
Veröffentlicht: (2025)
von: Choi, Yujin, et al.
Veröffentlicht: (2025)
PAD: Towards Efficient Data Generation for Transfer Learning Using Phrase Alignment
von: Kim, Jong Myoung, et al.
Veröffentlicht: (2025)
von: Kim, Jong Myoung, et al.
Veröffentlicht: (2025)
TimeChara: Evaluating Point-in-Time Character Hallucination of Role-Playing Large Language Models
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2024)
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2024)
Language Model Can Do Knowledge Tracing: Simple but Effective Method to Integrate Language Model and Knowledge Tracing Task
von: Lee, Unggi, et al.
Veröffentlicht: (2024)
von: Lee, Unggi, et al.
Veröffentlicht: (2024)
LCSB: Layer-Cyclic Selective Backpropagation for Memory-Efficient On-Device LLM Fine-Tuning
von: Park, Juneyoung, et al.
Veröffentlicht: (2026)
von: Park, Juneyoung, et al.
Veröffentlicht: (2026)
Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities
von: Zhou, Ziwei, et al.
Veröffentlicht: (2025)
von: Zhou, Ziwei, et al.
Veröffentlicht: (2025)
Direct Alignment of Draft Model for Speculative Decoding with Chat-Fine-Tuned LLMs
von: Goel, Raghavv, et al.
Veröffentlicht: (2024)
von: Goel, Raghavv, et al.
Veröffentlicht: (2024)
Compressed Context Memory For Online Language Model Interaction
von: Kim, Jang-Hyun, et al.
Veröffentlicht: (2023)
von: Kim, Jang-Hyun, et al.
Veröffentlicht: (2023)
RAISE: Enhancing Scientific Reasoning in LLMs via Step-by-Step Retrieval
von: Oh, Minhae, et al.
Veröffentlicht: (2025)
von: Oh, Minhae, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models
von: Kim, Woojin, et al.
Veröffentlicht: (2026) -
Expanding Search Space with Diverse Prompting Agents: An Efficient Sampling Approach for LLM Mathematical Reasoning
von: Lee, Gisang, et al.
Veröffentlicht: (2024) -
Toward Robust RALMs: Revealing the Impact of Imperfect Retrieval on Retrieval-Augmented Language Models
von: Park, Seong-Il, et al.
Veröffentlicht: (2024) -
Latent Self-Consistency for Reliable Majority-Set Selection in Short- and Long-Answer Reasoning
von: Oh, Jungsuk, et al.
Veröffentlicht: (2025) -
More Human, More Efficient: Aligning Annotations with Quantized SLMs
von: Wang, Jiayu, et al.
Veröffentlicht: (2026)