The Illusion-Illusion: Vision Language Models See Illusions Where There are None
Fuente:
arXiv
Saved in:
| Main Author: | Ullman, Tomer |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Rational Analysis of the Speech-to-Song Illusion
by: Marjieh, Raja, et al.
Published: (2024)
by: Marjieh, Raja, et al.
Published: (2024)
Linton Stereo Illusion
by: Linton, Paul
Published: (2024)
by: Linton, Paul
Published: (2024)
Deciphering Functions of Neurons in Vision-Language Models
by: Xu, Jiaqi, et al.
Published: (2025)
by: Xu, Jiaqi, et al.
Published: (2025)
IllusionVQA: A Challenging Optical Illusion Dataset for Vision Language Models
by: Shahgir, Haz Sameen, et al.
Published: (2024)
by: Shahgir, Haz Sameen, et al.
Published: (2024)
Flexible Tool Selection through Low-dimensional Attribute Alignment of Vision and Language
by: Hao, Guangfu, et al.
Published: (2025)
by: Hao, Guangfu, et al.
Published: (2025)
Beginning with You: Perceptual-Initialization Improves Vision-Language Representation and Alignment
by: Hu, Yang, et al.
Published: (2025)
by: Hu, Yang, et al.
Published: (2025)
Interpreting the structure of multi-object representations in vision encoders
by: Khajuria, Tarun, et al.
Published: (2024)
by: Khajuria, Tarun, et al.
Published: (2024)
A Cognitive Architecture for Machine Consciousness and Artificial Superintelligence: Thought Is Structured by the Iterative Updating of Working Memory
by: Reser, Jared Edward
Published: (2022)
by: Reser, Jared Edward
Published: (2022)
Linton Stereo Illusion: Response on Johnston (1991)
by: Linton, Paul
Published: (2024)
by: Linton, Paul
Published: (2024)
Explicitly Modeling Subcortical Vision with a Neuro-Inspired Front-End Improves CNN Robustness
by: Piper, Lucas, et al.
Published: (2025)
by: Piper, Lucas, et al.
Published: (2025)
Explicitly Modeling Pre-Cortical Vision with a Neuro-Inspired Front-End Improves CNN Robustness
by: Piper, Lucas, et al.
Published: (2024)
by: Piper, Lucas, et al.
Published: (2024)
The Illusion of Opting in AI-Mediated Consequential Decisions
by: Ji, Eugene Yu
Published: (2026)
by: Ji, Eugene Yu
Published: (2026)
Utilizing Computer Vision for Continuous Monitoring of Vaccine Side Effects in Experimental Mice
by: Li, Chuang, et al.
Published: (2024)
by: Li, Chuang, et al.
Published: (2024)
Brain Mapping with Dense Features: Grounding Cortical Semantic Selectivity in Natural Images With Vision Transformers
by: Luo, Andrew F., et al.
Published: (2024)
by: Luo, Andrew F., et al.
Published: (2024)
Language learning shapes visual category-selectivity in deep neural networks
by: Lu, Zitong, et al.
Published: (2025)
by: Lu, Zitong, et al.
Published: (2025)
Model-Guided Microstimulation Steers Primate Visual Behavior
by: Mehrer, Johannes, et al.
Published: (2025)
by: Mehrer, Johannes, et al.
Published: (2025)
Sparse Autoencoders Bridge The Deep Learning Model and The Brain
by: Mao, Ziming, et al.
Published: (2025)
by: Mao, Ziming, et al.
Published: (2025)
Human-Aligned Evaluation of a Pixel-wise DNN Color Constancy Model
by: Heidari-Gorji, Hamed, et al.
Published: (2026)
by: Heidari-Gorji, Hamed, et al.
Published: (2026)
Simple Models, Rich Representations: Visual Decoding from Primate Intracortical Neural Signals
by: Ciferri, Matteo, et al.
Published: (2026)
by: Ciferri, Matteo, et al.
Published: (2026)
SLIM-Brain: A Data- and Training-Efficient Foundation Model for fMRI Data Analysis
by: Wang, Mo, et al.
Published: (2025)
by: Wang, Mo, et al.
Published: (2025)
Brain-DiT: A Universal Multi-state fMRI Foundation Model with Metadata-Conditioned Pretraining
by: Xia, Junfeng, et al.
Published: (2026)
by: Xia, Junfeng, et al.
Published: (2026)
Characterizing Universal Object Representations Across Vision Models
by: Mahner, Florian P., et al.
Published: (2026)
by: Mahner, Florian P., et al.
Published: (2026)
Illusion-Aware Visual Preprocessing and Anti-Illusion Prompting for Classic Illusion Understanding in Vision-Language Models
by: Zha, Junli, et al.
Published: (2026)
by: Zha, Junli, et al.
Published: (2026)
Aligned with LLM: a new multi-modal training paradigm for encoding fMRI activity in visual cortex
by: Ma, Shuxiao, et al.
Published: (2024)
by: Ma, Shuxiao, et al.
Published: (2024)
Relationships between the degrees of freedom in the affine Gaussian derivative model for visual receptive fields and 2-D affine image transformations, with application to covariance properties of simple cells in the primary visual cortex
by: Lindeberg, Tony
Published: (2024)
by: Lindeberg, Tony
Published: (2024)
Supervised Learning without Backpropagation using Spike-Timing-Dependent Plasticity for Image Recognition
by: Xie, Wei
Published: (2024)
by: Xie, Wei
Published: (2024)
Primary visual cortex contributes to color constancy by predicting rather than discounting the illuminant: evidence from a computational study
by: Gao, Shaobing, et al.
Published: (2024)
by: Gao, Shaobing, et al.
Published: (2024)
Self-Attention-Based Contextual Modulation Improves Neural System Identification
by: Lin, Isaac, et al.
Published: (2024)
by: Lin, Isaac, et al.
Published: (2024)
Correcting Biased Centered Kernel Alignment Measures in Biological and Artificial Neural Networks
by: Murphy, Alex, et al.
Published: (2024)
by: Murphy, Alex, et al.
Published: (2024)
Reliable Object Tracking by Multimodal Hybrid Feature Extraction and Transformer-Based Fusion
by: Sun, Hongze, et al.
Published: (2024)
by: Sun, Hongze, et al.
Published: (2024)
Towards understanding the nature of direct functional connectivity in visual brain network
by: Bhattacharya, Debanjali, et al.
Published: (2024)
by: Bhattacharya, Debanjali, et al.
Published: (2024)
Parametric PerceptNet: A bio-inspired deep-net trained for Image Quality Assessment
by: Vila-Tomás, Jorge, et al.
Published: (2024)
by: Vila-Tomás, Jorge, et al.
Published: (2024)
Foveated Retinotopy Improves Classification and Localization in Convolutional Neural Networks
by: Jérémie, Jean-Nicolas, et al.
Published: (2024)
by: Jérémie, Jean-Nicolas, et al.
Published: (2024)
What Makes a Face Look like a Hat: Decoupling Low-level and High-level Visual Properties with Image Triplets
by: Piriyajitakonkij, Maytus, et al.
Published: (2024)
by: Piriyajitakonkij, Maytus, et al.
Published: (2024)
Cracking the neural code for word recognition in convolutional neural networks
by: Agrawal, Aakash, et al.
Published: (2024)
by: Agrawal, Aakash, et al.
Published: (2024)
Synthesis and Perceptual Scaling of High Resolution Naturalistic Images Using Stable Diffusion
by: Pettini, Leonardo, et al.
Published: (2024)
by: Pettini, Leonardo, et al.
Published: (2024)
Universal dimensions of visual representation
by: Chen, Zirui, et al.
Published: (2024)
by: Chen, Zirui, et al.
Published: (2024)
Towards Explainable Automated Neuroanatomy
by: Qian, Kui, et al.
Published: (2024)
by: Qian, Kui, et al.
Published: (2024)
Brain Network Diffusion-Driven fMRI Connectivity Augmentation for Enhanced Autism Spectrum Disorder Diagnosis
by: Zhao, Haokai, et al.
Published: (2024)
by: Zhao, Haokai, et al.
Published: (2024)
Time Series Analysis of Spiking Neural Systems via Transfer Entropy and Directed Persistent Homology
by: Peek, Dylan, et al.
Published: (2025)
by: Peek, Dylan, et al.
Published: (2025)
Similar Items
-
A Rational Analysis of the Speech-to-Song Illusion
by: Marjieh, Raja, et al.
Published: (2024) -
Linton Stereo Illusion
by: Linton, Paul
Published: (2024) -
Deciphering Functions of Neurons in Vision-Language Models
by: Xu, Jiaqi, et al.
Published: (2025) -
IllusionVQA: A Challenging Optical Illusion Dataset for Vision Language Models
by: Shahgir, Haz Sameen, et al.
Published: (2024) -
Flexible Tool Selection through Low-dimensional Attribute Alignment of Vision and Language
by: Hao, Guangfu, et al.
Published: (2025)