Do MLLMs Capture How Interfaces Guide User Behavior? A Benchmark for Multimodal UI/UX Design Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Jeon, Jaehyun, Kim, Min Soo, Yoon, Jang Han, Shim, Sumin, Choi, Yejin, Kim, Hanbin, Kim, Dae Hyun, Yu, Youngjae |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zero-shot Multimodal Document Retrieval via Cross-modal Question Generation
by: Choi, Yejin, et al.
Published: (2025)
by: Choi, Yejin, et al.
Published: (2025)
A11YN: aligning LLMs for accessible web UI code generation
by: Yoon, Janghan, et al.
Published: (2025)
by: Yoon, Janghan, et al.
Published: (2025)
Towards Visual Text Design Transfer Across Languages
by: Choi, Yejin, et al.
Published: (2024)
by: Choi, Yejin, et al.
Published: (2024)
A More Word-like Image Tokenization for MLLMs
by: Lee, Hyun, et al.
Published: (2026)
by: Lee, Hyun, et al.
Published: (2026)
Building UI/UX Dataset for Dark Pattern Detection and YOLOv12x-based Real-Time Object Recognition Detection System
by: Jang, Se-Young, et al.
Published: (2025)
by: Jang, Se-Young, et al.
Published: (2025)
Audio-Based Linguistic Feature Extraction for Enhancing Multi-lingual and Low-Resource Text-to-Speech
by: Kim, Youngjae, et al.
Published: (2024)
by: Kim, Youngjae, et al.
Published: (2024)
CANVAS: A Benchmark for Vision-Language Models on Tool-Based User Interface Design
by: Jeong, Daeheon, et al.
Published: (2025)
by: Jeong, Daeheon, et al.
Published: (2025)
MLLM as a UI Judge: Benchmarking Multimodal LLMs for Predicting Human Perception of User Interfaces
by: Luera, Reuben A., et al.
Published: (2025)
by: Luera, Reuben A., et al.
Published: (2025)
Multi-View Graph Convolution Network for Internal Talent Recommendation Based on Enterprise Emails
by: Kim, Soo Hyun, et al.
Published: (2025)
by: Kim, Soo Hyun, et al.
Published: (2025)
Progressive Facial Granularity Aggregation with Bilateral Attribute-based Enhancement for Face-to-Speech Synthesis
by: Jeon, Yejin, et al.
Published: (2025)
by: Jeon, Yejin, et al.
Published: (2025)
Facilitating Personalized TTS for Dysarthric Speakers Using Knowledge Anchoring and Curriculum Learning
by: Jeon, Yejin, et al.
Published: (2025)
by: Jeon, Yejin, et al.
Published: (2025)
Configurable 3D‐Printed Microstructured Stamp as a User‐Friendly Tool for Versatile Patterning of Low‐Viscosity Bioinks
by: Yejin Choi, et al.
Published: (2025)
by: Yejin Choi, et al.
Published: (2025)
Enumeration of multiplex juggling card sequences using generalized q-derivatives
by: Cho, Yumin, et al.
Published: (2024)
by: Cho, Yumin, et al.
Published: (2024)
SMILE: Multimodal Dataset for Understanding Laughter in Video with Language Models
by: Hyun, Lee, et al.
Published: (2023)
by: Hyun, Lee, et al.
Published: (2023)
Adaptive Travel Behaviors During Crisis Waves: A Co‐Occurrence Network of Social Media Discourse
by: Yejin Lee, et al.
Published: (2026)
by: Yejin Lee, et al.
Published: (2026)
Mind the Motions: Benchmarking Theory-of-Mind in Everyday Body Language
by: Lee, Seungbeen, et al.
Published: (2025)
by: Lee, Seungbeen, et al.
Published: (2025)
Unveiling Ru(bpy) 3 2+ ‐Encapsulated Zeolite Y as Photocatalyst: Harnessing Photocatalytic Singlet Oxygen Generation for Mustard Gas Simulant Detoxification
by: Sumin Kim, et al.
Published: (2024)
by: Sumin Kim, et al.
Published: (2024)
Unleashing frailty from laboratory into real world: A critical step toward frailty‐guided clinical care of older adults
by: Dae Hyun Kim
Published: (2024)
by: Dae Hyun Kim
Published: (2024)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
by: Chung, Jiwan, et al.
Published: (2025)
by: Chung, Jiwan, et al.
Published: (2025)
SMART UI/UX DESIGN EVALUATOR
by: Andry, Andry
Published: (2026)
by: Andry, Andry
Published: (2026)
Conductive Agent‐Controlled Tortuosity in Solvent‐Free Thick‐Film Electrodes for High‐Energy Lithium‐Ion Batteries
by: Byeongjin Kim, et al.
Published: (2025)
by: Byeongjin Kim, et al.
Published: (2025)
Multi‐Channel Neural Interface for Neural Recording and Neuromodulation
by: Eunmin Kim, et al.
Published: (2025)
by: Eunmin Kim, et al.
Published: (2025)
Latent Preference Modeling for Cross-Session Personalized Tool Calling
by: Yoon, Yejin, et al.
Published: (2026)
by: Yoon, Yejin, et al.
Published: (2026)
Explicit Feature Interaction-aware Graph Neural Networks
by: Kim, Minkyu, et al.
Published: (2022)
by: Kim, Minkyu, et al.
Published: (2022)
Higher-order Neural Additive Models: An Interpretable Machine Learning Model with Feature Interactions
by: Kim, Minkyu, et al.
Published: (2022)
by: Kim, Minkyu, et al.
Published: (2022)
MIRROR: Multimodal Cognitive Reframing Therapy for Rolling with Resistance
by: Kim, Subin, et al.
Published: (2025)
by: Kim, Subin, et al.
Published: (2025)
Subtle Risks, Critical Failures: A Framework for Diagnosing Physical Safety of LLMs for Embodied Decision Making
by: Son, Yejin, et al.
Published: (2025)
by: Son, Yejin, et al.
Published: (2025)
Light-Wave Engineering for Selective Polarization of a Single $\mathbf{Q}$ Valley in Transition Metal Dichalcogenides
by: Kim, Youngjae
Published: (2025)
by: Kim, Youngjae
Published: (2025)
Pseudospins revealed through the giant dynamical Franz-Keldysh effect in massless Dirac materials
by: Kim, Youngjae
Published: (2024)
by: Kim, Youngjae
Published: (2024)
Modulating Molecular Interaction of Zwitterion Toward Rational Interface Engineering of Perovskite Solar Cells
by: Hangyeol Kim, et al.
Published: (2024)
by: Hangyeol Kim, et al.
Published: (2024)
Cylinders in Du Val del Pezzo surfaces of degree one with Picard rank two
by: Kim, Jaehyun, et al.
Published: (2025)
by: Kim, Jaehyun, et al.
Published: (2025)
Antibody development for the diagnosis of Oryctes rhinoceros nudivirus
by: Hyun‐Soo Kim, et al.
Published: (2024)
by: Hyun‐Soo Kim, et al.
Published: (2024)
POaaS: Minimal-Edit Prompt Optimization as a Service to Lift Accuracy and Cut Hallucinations on On-Device sLLMs
by: Shim, Jungwoo, et al.
Published: (2026)
by: Shim, Jungwoo, et al.
Published: (2026)
Anchoring and Rescaling Attention for Semantically Coherent Inbetweening
by: Choi, Tae Eun, et al.
Published: (2026)
by: Choi, Tae Eun, et al.
Published: (2026)
Privacy Starts with UI: Privacy Patterns and Designer Perspectives in UI/UX Practice
by: Maloku, Anxhela, et al.
Published: (2026)
by: Maloku, Anxhela, et al.
Published: (2026)
Adjoint‐based observation impact on meteorological forecast errors in the Arctic
by: Dae‐Hui Kim, et al.
Published: (2024)
by: Dae‐Hui Kim, et al.
Published: (2024)
Structural Optimization of NVP/C Composites by an Advanced Two‐Step Spray Technique for High Energy Density and Long‐Life Symmetric Sodium‐Ion Batteries
by: Yejin Ra, et al.
Published: (2025)
by: Yejin Ra, et al.
Published: (2025)
Boundary-Recovering Network for Temporal Action Detection
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
Leveraging the Power of MLLMs for Gloss-Free Sign Language Translation
by: Kim, Jungeun, et al.
Published: (2024)
by: Kim, Jungeun, et al.
Published: (2024)
Heterogeneous Integration of Wide Bandgap Semiconductors and 2D Materials: Processes, Applications, and Perspectives (Adv. Mater. 12/2025)
by: Soo Ho Choi, et al.
Published: (2025)
by: Soo Ho Choi, et al.
Published: (2025)
Similar Items
-
Zero-shot Multimodal Document Retrieval via Cross-modal Question Generation
by: Choi, Yejin, et al.
Published: (2025) -
A11YN: aligning LLMs for accessible web UI code generation
by: Yoon, Janghan, et al.
Published: (2025) -
Towards Visual Text Design Transfer Across Languages
by: Choi, Yejin, et al.
Published: (2024) -
A More Word-like Image Tokenization for MLLMs
by: Lee, Hyun, et al.
Published: (2026) -
Building UI/UX Dataset for Dark Pattern Detection and YOLOv12x-based Real-Time Object Recognition Detection System
by: Jang, Se-Young, et al.
Published: (2025)