The Visual Iconicity Challenge: Evaluating Vision-Language Models on Sign Language Form-Meaning Mapping
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Keleş, Onur, Özyürek, Aslı, Ortega, Gerardo, Gökgöz, Kadir, Ghaleb, Esam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
I see what you mean: Co-Speech Gestures for Reference Resolution in Multimodal Dialogue
von: Ghaleb, Esam, et al.
Veröffentlicht: (2025)
von: Ghaleb, Esam, et al.
Veröffentlicht: (2025)
HolisticSemGes: Semantic Grounding of Holistic Co-Speech Gesture Generation with Contrastive Flow-Matching
von: Liu, Lanmiao, et al.
Veröffentlicht: (2026)
von: Liu, Lanmiao, et al.
Veröffentlicht: (2026)
SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning
von: Liu, Lanmiao, et al.
Veröffentlicht: (2025)
von: Liu, Lanmiao, et al.
Veröffentlicht: (2025)
DuoGesture: Neuro-Inspired and Biomechanically Informed Dual-Stream Co-Speech Gesture Generation
von: Paar, Ferdinand, et al.
Veröffentlicht: (2026)
von: Paar, Ferdinand, et al.
Veröffentlicht: (2026)
Iconicity and Gesture Jointly Facilitate Learning of Second Language Signs at First Exposure in Hearing Nonsigners
von: Dilay Z. Karadöller, et al.
Veröffentlicht: (2024)
von: Dilay Z. Karadöller, et al.
Veröffentlicht: (2024)
E-TSL: A Continuous Educational Turkish Sign Language Dataset with Baseline Methods
von: Öztürk, Şükrü, et al.
Veröffentlicht: (2024)
von: Öztürk, Şükrü, et al.
Veröffentlicht: (2024)
The Influence of Iconicity in Transfer Learning for Sign Language Recognition
von: Artiaga, Keren, et al.
Veröffentlicht: (2026)
von: Artiaga, Keren, et al.
Veröffentlicht: (2026)
Learning Co-Speech Gesture Representations in Dialogue through Contrastive Learning: An Intrinsic Evaluation
von: Ghaleb, Esam, et al.
Veröffentlicht: (2024)
von: Ghaleb, Esam, et al.
Veröffentlicht: (2024)
Co-Speech Gesture Detection through Multi-Phase Sequence Labeling
von: Ghaleb, Esam, et al.
Veröffentlicht: (2023)
von: Ghaleb, Esam, et al.
Veröffentlicht: (2023)
Sign-IDD: Iconicity Disentangled Diffusion for Sign Language Production
von: Tang, Shengeng, et al.
Veröffentlicht: (2024)
von: Tang, Shengeng, et al.
Veröffentlicht: (2024)
SignLLM: Sign Language Production Large Language Models
von: Fang, Sen, et al.
Veröffentlicht: (2024)
von: Fang, Sen, et al.
Veröffentlicht: (2024)
Analysing Cross-Speaker Convergence in Face-to-Face Dialogue through the Lens of Automatically Detected Shared Linguistic Constructions
von: Ghaleb, Esam, et al.
Veröffentlicht: (2024)
von: Ghaleb, Esam, et al.
Veröffentlicht: (2024)
Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector
von: Li, Sifan, et al.
Veröffentlicht: (2025)
von: Li, Sifan, et al.
Veröffentlicht: (2025)
Visual Riddles: a Commonsense and World Knowledge Challenge for Large Vision and Language Models
von: Bitton-Guetta, Nitzan, et al.
Veröffentlicht: (2024)
von: Bitton-Guetta, Nitzan, et al.
Veröffentlicht: (2024)
Do Vision-Language Models Really Understand Visual Language?
von: Hou, Yifan, et al.
Veröffentlicht: (2024)
von: Hou, Yifan, et al.
Veröffentlicht: (2024)
CartoMapQA: A Fundamental Benchmark Dataset Evaluating Vision-Language Models on Cartographic Map Understanding
von: Ung, Huy Quang, et al.
Veröffentlicht: (2025)
von: Ung, Huy Quang, et al.
Veröffentlicht: (2025)
Bangla Sign Language Translation: Dataset Creation Challenges, Benchmarking and Prospects
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025)
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025)
Evaluating Vision-Language Models as Evaluators in Path Planning
von: Aghzal, Mohamed, et al.
Veröffentlicht: (2024)
von: Aghzal, Mohamed, et al.
Veröffentlicht: (2024)
Do Vision-Language Models Understand Visual Persuasiveness?
von: Park, Gyuwon
Veröffentlicht: (2025)
von: Park, Gyuwon
Veröffentlicht: (2025)
Visual In-Context Learning for Large Vision-Language Models
von: Zhou, Yucheng, et al.
Veröffentlicht: (2024)
von: Zhou, Yucheng, et al.
Veröffentlicht: (2024)
Leveraging Speech for Gesture Detection in Multimodal Communication
von: Ghaleb, Esam, et al.
Veröffentlicht: (2024)
von: Ghaleb, Esam, et al.
Veröffentlicht: (2024)
Evaluating Vision-Language Models for Emotion Recognition
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025)
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025)
Evaluation of Cultural Competence of Vision-Language Models
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
Using Sign Language Production as Data Augmentation to enhance Sign Language Translation
von: Walsh, Harry, et al.
Veröffentlicht: (2025)
von: Walsh, Harry, et al.
Veröffentlicht: (2025)
Clean Evaluations on Contaminated Visual Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
SignDATA: Data Pipeline for Sign Language Translation
von: Chen, Kuanwei, et al.
Veröffentlicht: (2026)
von: Chen, Kuanwei, et al.
Veröffentlicht: (2026)
Can Vision-Language Models Evaluate Handwritten Math?
von: Nath, Oikantik, et al.
Veröffentlicht: (2025)
von: Nath, Oikantik, et al.
Veröffentlicht: (2025)
Instruction-Following Evaluation of Large Vision-Language Models
von: Shiono, Daiki, et al.
Veröffentlicht: (2025)
von: Shiono, Daiki, et al.
Veröffentlicht: (2025)
Gloss-Free Sign Language Translation: An Unbiased Evaluation of Progress in the Field
von: Sincan, Ozge Mercanoglu, et al.
Veröffentlicht: (2026)
von: Sincan, Ozge Mercanoglu, et al.
Veröffentlicht: (2026)
Studying and Mitigating Biases in Sign Language Understanding Models
von: Atwell, Katherine, et al.
Veröffentlicht: (2024)
von: Atwell, Katherine, et al.
Veröffentlicht: (2024)
Malayalam Sign Language Identification using Finetuned YOLOv8 and Computer Vision Techniques
von: K., Abhinand, et al.
Veröffentlicht: (2024)
von: K., Abhinand, et al.
Veröffentlicht: (2024)
Sign Stitching: A Novel Approach to Sign Language Production
von: Walsh, Harry, et al.
Veröffentlicht: (2024)
von: Walsh, Harry, et al.
Veröffentlicht: (2024)
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
von: Góral, Gracjan, et al.
Veröffentlicht: (2024)
von: Góral, Gracjan, et al.
Veröffentlicht: (2024)
Representing Signs as Signs: One-Shot ISLR to Facilitate Functional Sign Language Technologies
von: Vandendriessche, Toon, et al.
Veröffentlicht: (2025)
von: Vandendriessche, Toon, et al.
Veröffentlicht: (2025)
BabyVision: Visual Reasoning Beyond Language
von: Chen, Liang, et al.
Veröffentlicht: (2026)
von: Chen, Liang, et al.
Veröffentlicht: (2026)
Large Vision-Language Models for Remote Sensing Visual Question Answering
von: Siripong, Surasakdi, et al.
Veröffentlicht: (2024)
von: Siripong, Surasakdi, et al.
Veröffentlicht: (2024)
Vision-Language Modeling in PET/CT for Visual Grounding of Positive Findings
von: Huemann, Zachary, et al.
Veröffentlicht: (2025)
von: Huemann, Zachary, et al.
Veröffentlicht: (2025)
Teaching Vision-Language Models to Ask: Resolving Ambiguity in Visual Questions
von: Jian, Pu, et al.
Veröffentlicht: (2025)
von: Jian, Pu, et al.
Veröffentlicht: (2025)
VividMed: Vision Language Model with Versatile Visual Grounding for Medicine
von: Luo, Lingxiao, et al.
Veröffentlicht: (2024)
von: Luo, Lingxiao, et al.
Veröffentlicht: (2024)
Evaluating Vision Language Model Adaptations for Radiology Report Generation in Low-Resource Languages
von: Salmè, Marco, et al.
Veröffentlicht: (2025)
von: Salmè, Marco, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
I see what you mean: Co-Speech Gestures for Reference Resolution in Multimodal Dialogue
von: Ghaleb, Esam, et al.
Veröffentlicht: (2025) -
HolisticSemGes: Semantic Grounding of Holistic Co-Speech Gesture Generation with Contrastive Flow-Matching
von: Liu, Lanmiao, et al.
Veröffentlicht: (2026) -
SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning
von: Liu, Lanmiao, et al.
Veröffentlicht: (2025) -
DuoGesture: Neuro-Inspired and Biomechanically Informed Dual-Stream Co-Speech Gesture Generation
von: Paar, Ferdinand, et al.
Veröffentlicht: (2026) -
Iconicity and Gesture Jointly Facilitate Learning of Second Language Signs at First Exposure in Hearing Nonsigners
von: Dilay Z. Karadöller, et al.
Veröffentlicht: (2024)