Reasoning or Pattern Matching? Probing Large Vision-Language Models with Visual Puzzles
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lymperaiou, Maria, Karampinis, Vasileios, Filandrianos, Giorgos, Vlachos, Angelos, Zerva, Chrysoula, Voulodimos, Athanasios |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning
von: Vlachos, Angelos, et al.
Veröffentlicht: (2025)
von: Vlachos, Angelos, et al.
Veröffentlicht: (2025)
HalCECE: A Framework for Explainable Hallucination Detection through Conceptual Counterfactuals in Image Captioning
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2025)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2025)
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2025)
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2025)
Fast Post-Hoc Confidence Fusion for 3-Class Open-Set Aerial Object Detection
von: Loukovitis, Spyridon, et al.
Veröffentlicht: (2025)
von: Loukovitis, Spyridon, et al.
Veröffentlicht: (2025)
Model-Agnostic Open-Set Air-to-Air Visual Object Detection for Reliable UAV Perception
von: Loukovitis, Spyridon, et al.
Veröffentlicht: (2025)
von: Loukovitis, Spyridon, et al.
Veröffentlicht: (2025)
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
Counterfactual Edits for Generative Evaluation
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
Masked Generative Story Transformer with Character Guidance and Caption Augmentation
von: Papadimitriou, Christos, et al.
Veröffentlicht: (2024)
von: Papadimitriou, Christos, et al.
Veröffentlicht: (2024)
Common Corruptions for Enhancing and Evaluating Robustness in Air-to-Air Visual Object Detection
von: Arsenos, Anastasios, et al.
Veröffentlicht: (2024)
von: Arsenos, Anastasios, et al.
Veröffentlicht: (2024)
Prompt2Fashion: An automatically generated fashion dataset
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
Pitfalls of Scale: Investigating the Inverse Task of Redefinition in Large Language Models
von: Stringli, Elena, et al.
Veröffentlicht: (2025)
von: Stringli, Elena, et al.
Veröffentlicht: (2025)
Structure Your Data: Towards Semantic Graph Counterfactuals
von: Dimitriou, Angeliki, et al.
Veröffentlicht: (2024)
von: Dimitriou, Angeliki, et al.
Veröffentlicht: (2024)
U-CECE: A Universal Multi-Resolution Framework for Conceptual Counterfactual Explanations
von: Dimitriou, Angeliki, et al.
Veröffentlicht: (2026)
von: Dimitriou, Angeliki, et al.
Veröffentlicht: (2026)
ARPA: A Novel Hybrid Model for Advancing Visual Word Disambiguation Using Large Language Models and Transformers
von: Papastavrou, Aristi, et al.
Veröffentlicht: (2024)
von: Papastavrou, Aristi, et al.
Veröffentlicht: (2024)
Explaining Vision GNNs: A Semantic and Visual Analysis of Graph-based Image Classification
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025)
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025)
Puzzle Solving using Reasoning of Large Language Models: A Survey
von: Giadikiaroglou, Panagiotis, et al.
Veröffentlicht: (2024)
von: Giadikiaroglou, Panagiotis, et al.
Veröffentlicht: (2024)
Ensuring UAV Safety: A Vision-only and Real-time Framework for Collision Avoidance Through Object Detection, Tracking, and Distance Estimation
von: Karampinis, Vasileios, et al.
Veröffentlicht: (2024)
von: Karampinis, Vasileios, et al.
Veröffentlicht: (2024)
SemEval-2026 Task 6: CLARITY -- Unmasking Political Question Evasions
von: Thomas, Konstantinos, et al.
Veröffentlicht: (2026)
von: Thomas, Konstantinos, et al.
Veröffentlicht: (2026)
"I Never Said That": A dataset, taxonomy and baselines on response clarity classification
von: Thomas, Konstantinos, et al.
Veröffentlicht: (2024)
von: Thomas, Konstantinos, et al.
Veröffentlicht: (2024)
AILS-NTUA at SemEval-2026 Task 12: Graph-Based Retrieval and Reflective Prompting for Abductive Event Reasoning
von: Karafyllis, Nikolas, et al.
Veröffentlicht: (2026)
von: Karafyllis, Nikolas, et al.
Veröffentlicht: (2026)
AILS-NTUA at SemEval-2025 Task 8: Language-to-Code prompting and Error Fixing for Tabular Question Answering
von: Evangelatos, Andreas, et al.
Veröffentlicht: (2025)
von: Evangelatos, Andreas, et al.
Veröffentlicht: (2025)
CrossFlowDG: Bridging the Modality Gap with Cross-modal Flow Matching for Domain Generalization
von: Kritikos, Antonios, et al.
Veröffentlicht: (2026)
von: Kritikos, Antonios, et al.
Veröffentlicht: (2026)
AILS-NTUA at SemEval-2025 Task 3: Leveraging Large Language Models and Translation Strategies for Multilingual Hallucination Detection
von: Karkani, Dimitra, et al.
Veröffentlicht: (2025)
von: Karkani, Dimitra, et al.
Veröffentlicht: (2025)
AILS-NTUA at SemEval-2026 Task 8: Evaluating Multi-Turn RAG Conversations
von: Athanasiou, Dimosthenis, et al.
Veröffentlicht: (2026)
von: Athanasiou, Dimosthenis, et al.
Veröffentlicht: (2026)
VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models
von: Ren, Yufan, et al.
Veröffentlicht: (2025)
von: Ren, Yufan, et al.
Veröffentlicht: (2025)
AILS-NTUA at SemEval-2025 Task 4: Parameter-Efficient Unlearning for Large Language Models using Data Chunking
von: Premptis, Iraklis, et al.
Veröffentlicht: (2025)
von: Premptis, Iraklis, et al.
Veröffentlicht: (2025)
A Conformal Risk Control Framework for Granular Word Assessment and Uncertainty Calibration of CLIPScore Quality Estimates
von: Gomes, Gonçalo, et al.
Veröffentlicht: (2025)
von: Gomes, Gonçalo, et al.
Veröffentlicht: (2025)
The Contribution of Knowledge in Visiolinguistic Learning: A Survey on Tasks and Challenges
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
PuzzleVQA: Diagnosing Multimodal Reasoning Challenges of Language Models with Abstract Visual Patterns
von: Chia, Yew Ken, et al.
Veröffentlicht: (2024)
von: Chia, Yew Ken, et al.
Veröffentlicht: (2024)
SCENIR: Visual Semantic Clarity through Unsupervised Scene Graph Retrieval
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025)
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025)
AILS-NTUA at SemEval-2026 Task 10: Agentic LLMs for Psycholinguistic Marker Extraction and Conspiracy Endorsement Detection
von: Spanakis, Panagiotis Alexios, et al.
Veröffentlicht: (2026)
von: Spanakis, Panagiotis Alexios, et al.
Veröffentlicht: (2026)
Fine-Grained ImageNet Classification in the Wild
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2023)
AILS-NTUA at SemEval-2026 Task 3: Efficient Dimensional Aspect-Based Sentiment Analysis
von: Gazetas, Stavros, et al.
Veröffentlicht: (2026)
von: Gazetas, Stavros, et al.
Veröffentlicht: (2026)
SD-MVSum: Script-Driven Multimodal Video Summarization Method and Datasets
von: Mylonas, Manolis, et al.
Veröffentlicht: (2025)
von: Mylonas, Manolis, et al.
Veröffentlicht: (2025)
AILS-NTUA at SemEval-2024 Task 9: Cracking Brain Teasers: Transformer Models for Lateral Thinking Puzzles
von: Panagiotopoulos, Ioannis, et al.
Veröffentlicht: (2024)
von: Panagiotopoulos, Ioannis, et al.
Veröffentlicht: (2024)
Evaluating Counterfactual Strategic Reasoning in Large Language Models
von: Georgousis, Dimitrios, et al.
Veröffentlicht: (2026)
von: Georgousis, Dimitrios, et al.
Veröffentlicht: (2026)
U-Sketch: An Efficient Approach for Sketch to Image Diffusion Models
von: Mitsouras, Ilias, et al.
Veröffentlicht: (2024)
von: Mitsouras, Ilias, et al.
Veröffentlicht: (2024)
Probing Conceptual Understanding of Large Visual-Language Models
von: Schiappa, Madeline, et al.
Veröffentlicht: (2023)
von: Schiappa, Madeline, et al.
Veröffentlicht: (2023)
A Unified Masked Jigsaw Puzzle Framework for Vision and Language Models
von: Ye, Weixin, et al.
Veröffentlicht: (2026)
von: Ye, Weixin, et al.
Veröffentlicht: (2026)
Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning
von: Ghosal, Deepanway, et al.
Veröffentlicht: (2024)
von: Ghosal, Deepanway, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning
von: Vlachos, Angelos, et al.
Veröffentlicht: (2025) -
HalCECE: A Framework for Explainable Hallucination Detection through Conceptual Counterfactuals in Image Captioning
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2025) -
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2025) -
Fast Post-Hoc Confidence Fusion for 3-Class Open-Set Aerial Object Detection
von: Loukovitis, Spyridon, et al.
Veröffentlicht: (2025) -
Model-Agnostic Open-Set Air-to-Air Visual Object Detection for Reliable UAV Perception
von: Loukovitis, Spyridon, et al.
Veröffentlicht: (2025)