MIRAGE: The Illusion of Visual Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Asadi, Mohammad, O'Sullivan, Jack W., Cao, Fang, Nedaee, Tahoura, Rajabalifardi, Kamyar, Li, Fei-Fei, Adeli, Ehsan, Ashley, Euan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deterministic Hallucination Detection in Medical VQA via Confidence-Evidence Bayesian Gain
by: Asadi, Mohammad, et al.
Published: (2026)
by: Asadi, Mohammad, et al.
Published: (2026)
MARCUS: An agentic, multimodal vision-language model for cardiac diagnosis and management
by: O'Sullivan, Jack W, et al.
Published: (2026)
by: O'Sullivan, Jack W, et al.
Published: (2026)
VideoWeave: A Data-Centric Approach for Efficient Video Understanding
by: Durante, Zane, et al.
Published: (2026)
by: Durante, Zane, et al.
Published: (2026)
MIRAGE: Evaluating and Explaining Inductive Reasoning Process in Language Models
by: Li, Jiachun, et al.
Published: (2024)
by: Li, Jiachun, et al.
Published: (2024)
QuantiPhy: A Quantitative Benchmark Evaluating Physical Reasoning Abilities of Vision-Language Models
by: Puyin, Li, et al.
Published: (2025)
by: Puyin, Li, et al.
Published: (2025)
Few-Shot Classification of Interactive Activities of Daily Living (InteractADL)
by: Durante, Zane, et al.
Published: (2024)
by: Durante, Zane, et al.
Published: (2024)
Questionnaire meets LLM: A Benchmark and Empirical Study of Structural Skills for Understanding Questions and Responses
by: Nguyen, Duc-Hai, et al.
Published: (2025)
by: Nguyen, Duc-Hai, et al.
Published: (2025)
Reasoning Transfer for an Extremely Low-Resource and Endangered Language: Bridging Languages Through Sample-Efficient Language Understanding
by: Tran, Khanh-Tung, et al.
Published: (2025)
by: Tran, Khanh-Tung, et al.
Published: (2025)
Towards Fine-Grained Video Question Answering
by: Dai, Wei, et al.
Published: (2025)
by: Dai, Wei, et al.
Published: (2025)
Refined Iterated Pareto Greedy for Energy-aware Hybrid Flowshop Scheduling with Blocking Constraints
by: Missaoui, Ahmed, et al.
Published: (2025)
by: Missaoui, Ahmed, et al.
Published: (2025)
MIRAGE: Knowledge Graph-Guided Cross-Cohort MRI Synthesis for Alzheimer's Disease Prediction
by: Wu, Guanchen, et al.
Published: (2026)
by: Wu, Guanchen, et al.
Published: (2026)
MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them
by: Zhang, Weichen, et al.
Published: (2025)
by: Zhang, Weichen, et al.
Published: (2025)
The Initial Exploration Problem in Knowledge Graph Exploration
by: McNamara, Claire, et al.
Published: (2026)
by: McNamara, Claire, et al.
Published: (2026)
AdaVid: Adaptive Video-Language Pretraining
by: Patel, Chaitanya, et al.
Published: (2025)
by: Patel, Chaitanya, et al.
Published: (2025)
Pun Unintended: LLMs and the Illusion of Humor Understanding
by: Zangari, Alessandro, et al.
Published: (2025)
by: Zangari, Alessandro, et al.
Published: (2025)
Discovering Latent Graphs with GFlowNets for Diverse Conditional Image Generation
by: Trang, Bailey, et al.
Published: (2025)
by: Trang, Bailey, et al.
Published: (2025)
Neuro-Symbolic Decoding of Neural Activity
by: Wang, Yanchen, et al.
Published: (2026)
by: Wang, Yanchen, et al.
Published: (2026)
SRNN: Spatiotemporal Relational Neural Network for Intuitive Physics Understanding
by: Yang, Fei
Published: (2025)
by: Yang, Fei
Published: (2025)
Surrogate-Based Prevalence Measurement for Large-Scale A/B Testing
by: Xu, Zehao, et al.
Published: (2026)
by: Xu, Zehao, et al.
Published: (2026)
Spherical Leech Quantization for Visual Tokenization and Generation
by: Zhao, Yue, et al.
Published: (2025)
by: Zhao, Yue, et al.
Published: (2025)
Is Complexity an Illusion?
by: Bennett, Michael Timothy
Published: (2024)
by: Bennett, Michael Timothy
Published: (2024)
MIRAGE: Towards AI-Generated Image Detection in the Wild
by: Xia, Cheng, et al.
Published: (2025)
by: Xia, Cheng, et al.
Published: (2025)
Towards Fast Algorithms for the Preference Consistency Problem Based on Hierarchical Models
by: George, Anne-Marie, et al.
Published: (2024)
by: George, Anne-Marie, et al.
Published: (2024)
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
by: Park, Chanhee, et al.
Published: (2025)
by: Park, Chanhee, et al.
Published: (2025)
IRLBench: A Multi-modal, Culturally Grounded, Parallel Irish-English Benchmark for Open-Ended LLM Reasoning Evaluation
by: Tran, Khanh-Tung, et al.
Published: (2025)
by: Tran, Khanh-Tung, et al.
Published: (2025)
UCCIX: Irish-eXcellence Large Language Model
by: Tran, Khanh-Tung, et al.
Published: (2024)
by: Tran, Khanh-Tung, et al.
Published: (2024)
Optimising 4th-Order Runge-Kutta Methods: A Dynamic Heuristic Approach for Efficiency and Low Storage
by: Goodship, Gavin Lee, et al.
Published: (2025)
by: Goodship, Gavin Lee, et al.
Published: (2025)
GeoSAE: Geometric Prior-Guided Layer-Wise Sparse Autoencoder Annotation of Brain MRI Foundation Models
by: Nerrise, Favour, et al.
Published: (2026)
by: Nerrise, Favour, et al.
Published: (2026)
MIRAGE: Multimodal Identification and Recognition of Annotations in Indian General Prescriptions
by: Mankash, Tavish, et al.
Published: (2024)
by: Mankash, Tavish, et al.
Published: (2024)
Modality-Aware and Anatomical Vector-Quantized Autoencoding for Multimodal Brain MRI
by: Li, Mingjie, et al.
Published: (2026)
by: Li, Mingjie, et al.
Published: (2026)
Comment on Is Complexity an Illusion?
by: Simmons, Gabriel
Published: (2024)
by: Simmons, Gabriel
Published: (2024)
Rethinking the Illusion of Thinking
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
MIRAGE-Bench: Automatic Multilingual Benchmark Arena for Retrieval-Augmented Generation Systems
by: Thakur, Nandan, et al.
Published: (2024)
by: Thakur, Nandan, et al.
Published: (2024)
GAMMA-PD: Graph-based Analysis of Multi-Modal Motor Impairment Assessments in Parkinson's Disease
by: Nerrise, Favour, et al.
Published: (2024)
by: Nerrise, Favour, et al.
Published: (2024)
Qualitative Analysis of $ω$-Regular Objectives on Robust MDPs
by: Asadi, Ali, et al.
Published: (2025)
by: Asadi, Ali, et al.
Published: (2025)
Codev-Bench: How Do LLMs Understand Developer-Centric Code Completion?
by: Pan, Zhenyu, et al.
Published: (2024)
by: Pan, Zhenyu, et al.
Published: (2024)
TWIG: Towards pre-hoc Hyperparameter Optimisation and Cross-Graph Generalisation via Simulated KGE Models
by: Sardina, Jeffrey, et al.
Published: (2024)
by: Sardina, Jeffrey, et al.
Published: (2024)
MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence
by: Liu, Chonghan, et al.
Published: (2025)
by: Liu, Chonghan, et al.
Published: (2025)
Multimodal Climate Disinformation Detection: Integrating Vision-Language Models with External Knowledge Sources
by: Shamsabad, Marzieh Adeli, et al.
Published: (2026)
by: Shamsabad, Marzieh Adeli, et al.
Published: (2026)
Visual Self-supervised Learning Scheme for Dense Prediction Tasks on X-ray Images
by: Halat, Shervin, et al.
Published: (2023)
by: Halat, Shervin, et al.
Published: (2023)
Similar Items
-
Deterministic Hallucination Detection in Medical VQA via Confidence-Evidence Bayesian Gain
by: Asadi, Mohammad, et al.
Published: (2026) -
MARCUS: An agentic, multimodal vision-language model for cardiac diagnosis and management
by: O'Sullivan, Jack W, et al.
Published: (2026) -
VideoWeave: A Data-Centric Approach for Efficient Video Understanding
by: Durante, Zane, et al.
Published: (2026) -
MIRAGE: Evaluating and Explaining Inductive Reasoning Process in Language Models
by: Li, Jiachun, et al.
Published: (2024) -
QuantiPhy: A Quantitative Benchmark Evaluating Physical Reasoning Abilities of Vision-Language Models
by: Puyin, Li, et al.
Published: (2025)