Hospitality-VQA: Decision-Oriented Informativeness Evaluation for Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Jeongwoo, Duhyeong, Baek, Han, Eungyeol, Shin, Soyeon, han, Gukin, Kim, Seungduk, Jeon, Jaehyun, Jeong, Taewoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient and Effective Vocabulary Expansion Towards Multilingual Large Language Models
von: Kim, Seungduk, et al.
Veröffentlicht: (2024)
von: Kim, Seungduk, et al.
Veröffentlicht: (2024)
Vision-Language Models Generate More Homogeneous Stories for Phenotypically Black Individuals
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
Robust Deep Joint Source Channel Coding for Task-Oriented Semantic Communications
von: Park, Taewoo, et al.
Veröffentlicht: (2025)
von: Park, Taewoo, et al.
Veröffentlicht: (2025)
Vision-aligned Latent Reasoning for Multi-modal Large Language Model
von: Jeon, Byungwoo, et al.
Veröffentlicht: (2026)
von: Jeon, Byungwoo, et al.
Veröffentlicht: (2026)
SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?
von: Shin, Jongmin, et al.
Veröffentlicht: (2026)
von: Shin, Jongmin, et al.
Veröffentlicht: (2026)
Visual Cues of Gender and Race are Associated with Stereotyping in Vision-Language Models
von: Lee, Messi H. J., et al.
Veröffentlicht: (2025)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2025)
NEXT-EVAL: Next Evaluation of Traditional and LLM Web Data Record Extraction
von: Kim, Soyeon, et al.
Veröffentlicht: (2025)
von: Kim, Soyeon, et al.
Veröffentlicht: (2025)
What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models
von: Choi, Dasol, et al.
Veröffentlicht: (2026)
von: Choi, Dasol, et al.
Veröffentlicht: (2026)
KOFFVQA: An Objectively Evaluated Free-form VQA Benchmark for Large Vision-Language Models in the Korean Language
von: Kim, Yoonshik, et al.
Veröffentlicht: (2025)
von: Kim, Yoonshik, et al.
Veröffentlicht: (2025)
HDRSDR-VQA: A Subjective Video Quality Dataset for HDR and SDR Comparative Evaluation
von: Chen, Bowen, et al.
Veröffentlicht: (2025)
von: Chen, Bowen, et al.
Veröffentlicht: (2025)
MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
Latent Diffusion Models with Masked AutoEncoders
von: Lee, Junho, et al.
Veröffentlicht: (2025)
von: Lee, Junho, et al.
Veröffentlicht: (2025)
M3-SLU: Evaluating Speaker-Attributed Reasoning in Multimodal Large Language Models
von: Kwon, Yejin, et al.
Veröffentlicht: (2025)
von: Kwon, Yejin, et al.
Veröffentlicht: (2025)
Equivariant Latent Alignment via Flow Matching under Group Symmetries
von: Kim, Sunghyun, et al.
Veröffentlicht: (2026)
von: Kim, Sunghyun, et al.
Veröffentlicht: (2026)
Multimodal Large Language Models and Tunings: Vision, Language, Sensors, Audio, and Beyond
von: Han, Soyeon Caren, et al.
Veröffentlicht: (2024)
von: Han, Soyeon Caren, et al.
Veröffentlicht: (2024)
Diagnosing Causal Reasoning in Vision-Language Models via Structured Relevance Graphs
von: Pratama, Dhita Putri, et al.
Veröffentlicht: (2026)
von: Pratama, Dhita Putri, et al.
Veröffentlicht: (2026)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
Examining the Impact of Telehealth Stoma Care Interventions on the Ostomates: A Systematic Review and Meta‐Analysis
von: Soyeon Kim, et al.
Veröffentlicht: (2025)
von: Soyeon Kim, et al.
Veröffentlicht: (2025)
IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2026)
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2026)
Hierarchical Latent Space Item Response Model for Analyzing Mental Health Vulnerability of Elementary School Students in South Korea
von: Park, Soyeon, et al.
Veröffentlicht: (2026)
von: Park, Soyeon, et al.
Veröffentlicht: (2026)
DAM: Domain-Aware Module for Multi-Domain Dataset Condensation
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)
Cog3DMap: Multi-View Vision-Language Reasoning with 3D Cognitive Maps
von: Gwak, Chanyoung, et al.
Veröffentlicht: (2026)
von: Gwak, Chanyoung, et al.
Veröffentlicht: (2026)
HandVQA: Diagnosing and Improving Fine-Grained Spatial Reasoning about Hands in Vision-Language Models
von: Sayem, MD Khalequzzaman Chowdhury, et al.
Veröffentlicht: (2026)
von: Sayem, MD Khalequzzaman Chowdhury, et al.
Veröffentlicht: (2026)
Self-Guided Masked Autoencoder
von: Shin, Jeongwoo, et al.
Veröffentlicht: (2025)
von: Shin, Jeongwoo, et al.
Veröffentlicht: (2025)
Unlocking the Capabilities of Masked Generative Models for Image Synthesis via Self-Guidance
von: Hur, Jiwan, et al.
Veröffentlicht: (2024)
von: Hur, Jiwan, et al.
Veröffentlicht: (2024)
Naturalness-Aware Curriculum Learning with Dynamic Temperature for Speech Deepfake Detection
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
Redefining Evaluation Standards: A Unified Framework for Evaluating the Korean Capabilities of Language Models
von: Lee, Hanwool, et al.
Veröffentlicht: (2025)
von: Lee, Hanwool, et al.
Veröffentlicht: (2025)
Characterizations of smooth projective horospherical varieties of Picard number one
von: Hong, Jaehyun, et al.
Veröffentlicht: (2022)
von: Hong, Jaehyun, et al.
Veröffentlicht: (2022)
NegVQA: Can Vision Language Models Understand Negation?
von: Zhang, Yuhui, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhui, et al.
Veröffentlicht: (2025)
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning
von: Nam, Jaehyun, et al.
Veröffentlicht: (2024)
von: Nam, Jaehyun, et al.
Veröffentlicht: (2024)
Scaling Up Diffusion and Flow-based XGBoost Models
von: Cresswell, Jesse C., et al.
Veröffentlicht: (2024)
von: Cresswell, Jesse C., et al.
Veröffentlicht: (2024)
Explainable AI-Based Interface System for Weather Forecasting Model
von: Kim, Soyeon, et al.
Veröffentlicht: (2025)
von: Kim, Soyeon, et al.
Veröffentlicht: (2025)
SIA: Enhancing Safety via Intent Awareness for Vision-Language Models
von: Na, Youngjin, et al.
Veröffentlicht: (2025)
von: Na, Youngjin, et al.
Veröffentlicht: (2025)
VEHME: A Vision-Language Model For Evaluating Handwritten Mathematics Expressions
von: Nguyen, Thu Phuong, et al.
Veröffentlicht: (2025)
von: Nguyen, Thu Phuong, et al.
Veröffentlicht: (2025)
Self-Refining Language Model Anonymizers via Adversarial Distillation
von: Kim, Kyuyoung, et al.
Veröffentlicht: (2025)
von: Kim, Kyuyoung, et al.
Veröffentlicht: (2025)
A Survey of Large Language Models in Finance (FinLLMs)
von: Lee, Jean, et al.
Veröffentlicht: (2024)
von: Lee, Jean, et al.
Veröffentlicht: (2024)
Design and Evaluation of an Uncertainty-Aware Shared-Autonomy System with Hierarchical Conservative Skill Inference
von: Kim, Taewoo, et al.
Veröffentlicht: (2023)
von: Kim, Taewoo, et al.
Veröffentlicht: (2023)
Measuring Sample Importance in Data Pruning for Language Models based on Information Entropy
von: Kim, Minsang, et al.
Veröffentlicht: (2024)
von: Kim, Minsang, et al.
Veröffentlicht: (2024)
Cluster automorphism group of braid varieties
von: Kim, Soyeon
Veröffentlicht: (2025)
von: Kim, Soyeon
Veröffentlicht: (2025)
A submodular optimization approach to trustworthy loan approval automation
von: Kyungsik Lee, et al.
Veröffentlicht: (2024)
von: Kyungsik Lee, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Efficient and Effective Vocabulary Expansion Towards Multilingual Large Language Models
von: Kim, Seungduk, et al.
Veröffentlicht: (2024) -
Vision-Language Models Generate More Homogeneous Stories for Phenotypically Black Individuals
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024) -
Robust Deep Joint Source Channel Coding for Task-Oriented Semantic Communications
von: Park, Taewoo, et al.
Veröffentlicht: (2025) -
Vision-aligned Latent Reasoning for Multi-modal Large Language Model
von: Jeon, Byungwoo, et al.
Veröffentlicht: (2026) -
SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?
von: Shin, Jongmin, et al.
Veröffentlicht: (2026)