Interpreting the structure of multi-object representations in vision encoders
Fuente:
arXiv
Saved in:
| Main Authors: | Khajuria, Tarun, Dias, Braian Olmiro, Domnich, Marharyta, Aru, Jaan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Aligned with LLM: a new multi-modal training paradigm for encoding fMRI activity in visual cortex
by: Ma, Shuxiao, et al.
Published: (2024)
by: Ma, Shuxiao, et al.
Published: (2024)
MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery
by: Kneeland, Reese, et al.
Published: (2026)
by: Kneeland, Reese, et al.
Published: (2026)
Animal behavioral analysis and neural encoding with transformer-based self-supervised pretraining
by: Wang, Yanchen, et al.
Published: (2025)
by: Wang, Yanchen, et al.
Published: (2025)
Universal dimensions of visual representation
by: Chen, Zirui, et al.
Published: (2024)
by: Chen, Zirui, et al.
Published: (2024)
Perceptual misalignment of texture representations in convolutional neural networks
by: de Paolis, Ludovica, et al.
Published: (2026)
by: de Paolis, Ludovica, et al.
Published: (2026)
The Illusion-Illusion: Vision Language Models See Illusions Where There are None
by: Ullman, Tomer
Published: (2024)
by: Ullman, Tomer
Published: (2024)
A Cognitive Architecture for Machine Consciousness and Artificial Superintelligence: Thought Is Structured by the Iterative Updating of Working Memory
by: Reser, Jared Edward
Published: (2022)
by: Reser, Jared Edward
Published: (2022)
Revealing the core dimensions underlying representations in brains, behavior and AI
by: Mahner, Florian P., et al.
Published: (2026)
by: Mahner, Florian P., et al.
Published: (2026)
Predicting upcoming visual features during eye movements yields scene representations aligned with human visual cortex
by: Thorat, Sushrut, et al.
Published: (2025)
by: Thorat, Sushrut, et al.
Published: (2025)
Hybrid Lie semi-group and cascade structures for the generalized Gaussian derivative model for visual receptive fields
by: Lindeberg, Tony
Published: (2025)
by: Lindeberg, Tony
Published: (2025)
Human face perception reflects inverse-generative and naturalistic discriminative objectives
by: Guo, Wenxuan, et al.
Published: (2026)
by: Guo, Wenxuan, et al.
Published: (2026)
Short-term AI literacy intervention does not reduce over-reliance on incorrect ChatGPT recommendations
by: Puppart, Brett, et al.
Published: (2025)
by: Puppart, Brett, et al.
Published: (2025)
Artificial intelligence and the internal processes of creativity
by: Aru, Jaan
Published: (2024)
by: Aru, Jaan
Published: (2024)
Discriminating image representations with principal distortions
by: Feather, Jenelle, et al.
Published: (2024)
by: Feather, Jenelle, et al.
Published: (2024)
Flexible Tool Selection through Low-dimensional Attribute Alignment of Vision and Language
by: Hao, Guangfu, et al.
Published: (2025)
by: Hao, Guangfu, et al.
Published: (2025)
Visual representations in the human brain are aligned with large language models
by: Doerig, Adrien, et al.
Published: (2022)
by: Doerig, Adrien, et al.
Published: (2022)
Representational Similarity via Interpretable Visual Concepts
by: Kondapaneni, Neehar, et al.
Published: (2025)
by: Kondapaneni, Neehar, et al.
Published: (2025)
FOVI: A biologically-inspired foveated interface for deep vision models
by: Blauch, Nicholas M., et al.
Published: (2026)
by: Blauch, Nicholas M., et al.
Published: (2026)
Relationships between the degrees of freedom in the affine Gaussian derivative model for visual receptive fields and 2-D affine image transformations, with application to covariance properties of simple cells in the primary visual cortex
by: Lindeberg, Tony
Published: (2024)
by: Lindeberg, Tony
Published: (2024)
Utilizing Computer Vision for Continuous Monitoring of Vaccine Side Effects in Experimental Mice
by: Li, Chuang, et al.
Published: (2024)
by: Li, Chuang, et al.
Published: (2024)
Supervised Learning without Backpropagation using Spike-Timing-Dependent Plasticity for Image Recognition
by: Xie, Wei
Published: (2024)
by: Xie, Wei
Published: (2024)
Primary visual cortex contributes to color constancy by predicting rather than discounting the illuminant: evidence from a computational study
by: Gao, Shaobing, et al.
Published: (2024)
by: Gao, Shaobing, et al.
Published: (2024)
Explicitly Modeling Pre-Cortical Vision with a Neuro-Inspired Front-End Improves CNN Robustness
by: Piper, Lucas, et al.
Published: (2024)
by: Piper, Lucas, et al.
Published: (2024)
Self-Attention-Based Contextual Modulation Improves Neural System Identification
by: Lin, Isaac, et al.
Published: (2024)
by: Lin, Isaac, et al.
Published: (2024)
Correcting Biased Centered Kernel Alignment Measures in Biological and Artificial Neural Networks
by: Murphy, Alex, et al.
Published: (2024)
by: Murphy, Alex, et al.
Published: (2024)
Reliable Object Tracking by Multimodal Hybrid Feature Extraction and Transformer-Based Fusion
by: Sun, Hongze, et al.
Published: (2024)
by: Sun, Hongze, et al.
Published: (2024)
Towards understanding the nature of direct functional connectivity in visual brain network
by: Bhattacharya, Debanjali, et al.
Published: (2024)
by: Bhattacharya, Debanjali, et al.
Published: (2024)
Parametric PerceptNet: A bio-inspired deep-net trained for Image Quality Assessment
by: Vila-Tomás, Jorge, et al.
Published: (2024)
by: Vila-Tomás, Jorge, et al.
Published: (2024)
Foveated Retinotopy Improves Classification and Localization in Convolutional Neural Networks
by: Jérémie, Jean-Nicolas, et al.
Published: (2024)
by: Jérémie, Jean-Nicolas, et al.
Published: (2024)
What Makes a Face Look like a Hat: Decoupling Low-level and High-level Visual Properties with Image Triplets
by: Piriyajitakonkij, Maytus, et al.
Published: (2024)
by: Piriyajitakonkij, Maytus, et al.
Published: (2024)
Cracking the neural code for word recognition in convolutional neural networks
by: Agrawal, Aakash, et al.
Published: (2024)
by: Agrawal, Aakash, et al.
Published: (2024)
Brain Mapping with Dense Features: Grounding Cortical Semantic Selectivity in Natural Images With Vision Transformers
by: Luo, Andrew F., et al.
Published: (2024)
by: Luo, Andrew F., et al.
Published: (2024)
Synthesis and Perceptual Scaling of High Resolution Naturalistic Images Using Stable Diffusion
by: Pettini, Leonardo, et al.
Published: (2024)
by: Pettini, Leonardo, et al.
Published: (2024)
Towards Explainable Automated Neuroanatomy
by: Qian, Kui, et al.
Published: (2024)
by: Qian, Kui, et al.
Published: (2024)
Brain Network Diffusion-Driven fMRI Connectivity Augmentation for Enhanced Autism Spectrum Disorder Diagnosis
by: Zhao, Haokai, et al.
Published: (2024)
by: Zhao, Haokai, et al.
Published: (2024)
Beginning with You: Perceptual-Initialization Improves Vision-Language Representation and Alignment
by: Hu, Yang, et al.
Published: (2025)
by: Hu, Yang, et al.
Published: (2025)
Time Series Analysis of Spiking Neural Systems via Transfer Entropy and Directed Persistent Homology
by: Peek, Dylan, et al.
Published: (2025)
by: Peek, Dylan, et al.
Published: (2025)
Insights from the Algonauts 2025 Winners
by: Scotti, Paul S., et al.
Published: (2025)
by: Scotti, Paul S., et al.
Published: (2025)
Error correction in multiclass image classification of facial emotion on unbalanced samples
by: Lebedev, Andrey A., et al.
Published: (2025)
by: Lebedev, Andrey A., et al.
Published: (2025)
Model-Guided Microstimulation Steers Primate Visual Behavior
by: Mehrer, Johannes, et al.
Published: (2025)
by: Mehrer, Johannes, et al.
Published: (2025)
Similar Items
-
Aligned with LLM: a new multi-modal training paradigm for encoding fMRI activity in visual cortex
by: Ma, Shuxiao, et al.
Published: (2024) -
MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery
by: Kneeland, Reese, et al.
Published: (2026) -
Animal behavioral analysis and neural encoding with transformer-based self-supervised pretraining
by: Wang, Yanchen, et al.
Published: (2025) -
Universal dimensions of visual representation
by: Chen, Zirui, et al.
Published: (2024) -
Perceptual misalignment of texture representations in convolutional neural networks
by: de Paolis, Ludovica, et al.
Published: (2026)