VLind-Bench: Measuring Language Priors in Large Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Kang-il, Kim, Minbeom, Yoon, Seunghyun, Kim, Minsung, Lee, Dongryeol, Koh, Hyukhun, Jung, Kyomin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
di: Min, Kyungmin, et al.
Pubblicazione: (2024)
di: Min, Kyungmin, et al.
Pubblicazione: (2024)
Program Synthesis via Test-Time Transduction
di: Lee, Kang-il, et al.
Pubblicazione: (2025)
di: Lee, Kang-il, et al.
Pubblicazione: (2025)
Fine-grained Gender Control in Machine Translation with Large Language Models
di: Lee, Minwoo, et al.
Pubblicazione: (2024)
di: Lee, Minwoo, et al.
Pubblicazione: (2024)
ReflectCAP: Detailed Image Captioning with Reflective Memory
di: Min, Kyungmin, et al.
Pubblicazione: (2026)
di: Min, Kyungmin, et al.
Pubblicazione: (2026)
Generating Diverse Hypotheses for Inductive Reasoning
di: Lee, Kang-il, et al.
Pubblicazione: (2024)
di: Lee, Kang-il, et al.
Pubblicazione: (2024)
Fooling the LVLM Judges: Visual Biases in LVLM-Based Evaluation
di: Hwang, Yerin, et al.
Pubblicazione: (2025)
di: Hwang, Yerin, et al.
Pubblicazione: (2025)
MVMR: A New Framework for Evaluating Faithfulness of Video Moment Retrieval against Multiple Distractors
di: Yang, Nakyeong, et al.
Pubblicazione: (2023)
di: Yang, Nakyeong, et al.
Pubblicazione: (2023)
Casual as an Anchor: Resolving Supervision Misalignment in Formality Transfer Dataset
di: Yu, Hyojeong, et al.
Pubblicazione: (2026)
di: Yu, Hyojeong, et al.
Pubblicazione: (2026)
Rethinking Post-Unlearning Behavior of Large Vision-Language Models
di: Kim, Minsung, et al.
Pubblicazione: (2025)
di: Kim, Minsung, et al.
Pubblicazione: (2025)
Conditional [MASK] Discrete Diffusion Language Model
di: Koh, Hyukhun, et al.
Pubblicazione: (2024)
di: Koh, Hyukhun, et al.
Pubblicazione: (2024)
FaithUn: Toward Faithful Forgetting in Language Models by Investigating the Interconnectedness of Knowledge
di: Yang, Nakyeong, et al.
Pubblicazione: (2025)
di: Yang, Nakyeong, et al.
Pubblicazione: (2025)
Doubly-Universal Adversarial Perturbations: Deceiving Vision-Language Models Across Both Images and Text with a Single Perturbation
di: Kim, Hee-Seon, et al.
Pubblicazione: (2024)
di: Kim, Hee-Seon, et al.
Pubblicazione: (2024)
Can LLMs Recognize Toxicity? A Structured Investigation Framework and Toxicity Metric
di: Koh, Hyukhun, et al.
Pubblicazione: (2024)
di: Koh, Hyukhun, et al.
Pubblicazione: (2024)
Confidence-Guided Stepwise Model Routing for Cost-Efficient Reasoning
di: Lee, Sangmook, et al.
Pubblicazione: (2025)
di: Lee, Sangmook, et al.
Pubblicazione: (2025)
Guaranteed Generation from Large Language Models
di: Kim, Minbeom, et al.
Pubblicazione: (2024)
di: Kim, Minbeom, et al.
Pubblicazione: (2024)
Drift: Decoding-time Personalized Alignments with Implicit User Preferences
di: Kim, Minbeom, et al.
Pubblicazione: (2025)
di: Kim, Minbeom, et al.
Pubblicazione: (2025)
Benchmarking Direct Preference Optimization for Medical Large Vision-Language Models
di: Kim, Dain, et al.
Pubblicazione: (2026)
di: Kim, Dain, et al.
Pubblicazione: (2026)
How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?
di: Lee, Seongyun, et al.
Pubblicazione: (2024)
di: Lee, Seongyun, et al.
Pubblicazione: (2024)
Harmful Prompt Laundering: Jailbreaking LLMs with Abductive Styles and Symbolic Encoding
di: Joo, Seongho, et al.
Pubblicazione: (2025)
di: Joo, Seongho, et al.
Pubblicazione: (2025)
Public Data Assisted Differentially Private In-Context Learning
di: Joo, Seongho, et al.
Pubblicazione: (2025)
di: Joo, Seongho, et al.
Pubblicazione: (2025)
Intriguing Properties of Large Language and Vision Models
di: Lee, Young-Jun, et al.
Pubblicazione: (2024)
di: Lee, Young-Jun, et al.
Pubblicazione: (2024)
Better Safe Than Sorry? Overreaction Problem of Vision Language Models in Visual Emergency Recognition
di: Choi, Dasol, et al.
Pubblicazione: (2025)
di: Choi, Dasol, et al.
Pubblicazione: (2025)
KOFFVQA: An Objectively Evaluated Free-form VQA Benchmark for Large Vision-Language Models in the Korean Language
di: Kim, Yoonshik, et al.
Pubblicazione: (2025)
di: Kim, Yoonshik, et al.
Pubblicazione: (2025)
A Character-Centric Creative Story Generation via Imagination
di: Park, Kyeongman, et al.
Pubblicazione: (2024)
di: Park, Kyeongman, et al.
Pubblicazione: (2024)
Can You Trick the Grader? Adversarial Persuasion of LLM Judges
di: Hwang, Yerin, et al.
Pubblicazione: (2025)
di: Hwang, Yerin, et al.
Pubblicazione: (2025)
AdvisorQA: Towards Helpful and Harmless Advice-seeking Question Answering with Collective Intelligence
di: Kim, Minbeom, et al.
Pubblicazione: (2024)
di: Kim, Minbeom, et al.
Pubblicazione: (2024)
TroL: Traversal of Layers for Large Language and Vision Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
Vision-Language Models Do Not Understand Negation
di: Alhamoud, Kumail, et al.
Pubblicazione: (2025)
di: Alhamoud, Kumail, et al.
Pubblicazione: (2025)
Toward Interactive Regional Understanding in Vision-Large Language Models
di: Lee, Jungbeom, et al.
Pubblicazione: (2024)
di: Lee, Jungbeom, et al.
Pubblicazione: (2024)
ChatEXAONEPath: An Expert-level Multimodal Large Language Model for Histopathology Using Whole Slide Images
di: Kim, Sangwook, et al.
Pubblicazione: (2025)
di: Kim, Sangwook, et al.
Pubblicazione: (2025)
Focus Matters: Phase-Aware Suppression for Hallucination in Vision-Language Models
di: Kim, Sohyeon, et al.
Pubblicazione: (2026)
di: Kim, Sohyeon, et al.
Pubblicazione: (2026)
MFC-Bench: Benchmarking Multimodal Fact-Checking with Large Vision-Language Models
di: Wang, Shengkang, et al.
Pubblicazione: (2024)
di: Wang, Shengkang, et al.
Pubblicazione: (2024)
Expanding the Boundaries of Vision Prior Knowledge in Multi-modal Large Language Models
di: Liang, Qiao, et al.
Pubblicazione: (2025)
di: Liang, Qiao, et al.
Pubblicazione: (2025)
Are LLM-Judges Robust to Expressions of Uncertainty? Investigating the effect of Epistemic Markers on LLM-based Evaluation
di: Lee, Dongryeol, et al.
Pubblicazione: (2024)
di: Lee, Dongryeol, et al.
Pubblicazione: (2024)
Finer: Investigating and Enhancing Fine-Grained Visual Concept Recognition in Large Vision Language Models
di: Kim, Jeonghwan, et al.
Pubblicazione: (2024)
di: Kim, Jeonghwan, et al.
Pubblicazione: (2024)
Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models
di: Gröpl, Marcel, et al.
Pubblicazione: (2026)
di: Gröpl, Marcel, et al.
Pubblicazione: (2026)
HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models
di: Guan, Tianrui, et al.
Pubblicazione: (2023)
di: Guan, Tianrui, et al.
Pubblicazione: (2023)
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression
di: Kundu, Souvik, et al.
Pubblicazione: (2025)
di: Kundu, Souvik, et al.
Pubblicazione: (2025)
LifeTox: Unveiling Implicit Toxicity in Life Advice
di: Kim, Minbeom, et al.
Pubblicazione: (2023)
di: Kim, Minbeom, et al.
Pubblicazione: (2023)
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
di: Kim, Minchan, et al.
Pubblicazione: (2024)
di: Kim, Minchan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
di: Min, Kyungmin, et al.
Pubblicazione: (2024) -
Program Synthesis via Test-Time Transduction
di: Lee, Kang-il, et al.
Pubblicazione: (2025) -
Fine-grained Gender Control in Machine Translation with Large Language Models
di: Lee, Minwoo, et al.
Pubblicazione: (2024) -
ReflectCAP: Detailed Image Captioning with Reflective Memory
di: Min, Kyungmin, et al.
Pubblicazione: (2026) -
Generating Diverse Hypotheses for Inductive Reasoning
di: Lee, Kang-il, et al.
Pubblicazione: (2024)