Human-like Object Grouping in Self-supervised Vision Transformers
Fuente:
arXiv
Guardado en:
| Autores principales: | Adeli, Hossein, Ahn, Seoyoung, Luo, Andrew, Zhang, Mengmi, Kriegeskorte, Nikolaus, Zelinsky, Gregory |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Transformer brain encoders explain human high-level visual responses
por: Adeli, Hossein, et al.
Publicado: (2025)
por: Adeli, Hossein, et al.
Publicado: (2025)
In Silico Mapping of Visual Categorical Selectivity Across the Whole Brain
por: Hwang, Ethan, et al.
Publicado: (2025)
por: Hwang, Ethan, et al.
Publicado: (2025)
Human face perception reflects inverse-generative and naturalistic discriminative objectives
por: Guo, Wenxuan, et al.
Publicado: (2026)
por: Guo, Wenxuan, et al.
Publicado: (2026)
Brain Mapping with Dense Features: Grounding Cortical Semantic Selectivity in Natural Images With Vision Transformers
por: Luo, Andrew F., et al.
Publicado: (2024)
por: Luo, Andrew F., et al.
Publicado: (2024)
Scaling Vision Transformers for Functional MRI with Flat Maps
por: Lane, Connor, et al.
Publicado: (2025)
por: Lane, Connor, et al.
Publicado: (2025)
Transformers self-organize like newborn visual systems when trained in prenatal worlds
por: Pandey, Lalit, et al.
Publicado: (2026)
por: Pandey, Lalit, et al.
Publicado: (2026)
Motion Mapping Cognition: A Nondecomposable Primary Process in Human Vision
por: Xie, Zhenping
Publicado: (2024)
por: Xie, Zhenping
Publicado: (2024)
Does Object Binding Naturally Emerge in Large Pretrained Vision Transformers?
por: Li, Yihao, et al.
Publicado: (2025)
por: Li, Yihao, et al.
Publicado: (2025)
How does the primate brain combine generative and discriminative computations in vision?
por: Peters, Benjamin, et al.
Publicado: (2024)
por: Peters, Benjamin, et al.
Publicado: (2024)
Reliable Object Tracking by Multimodal Hybrid Feature Extraction and Transformer-Based Fusion
por: Sun, Hongze, et al.
Publicado: (2024)
por: Sun, Hongze, et al.
Publicado: (2024)
Reanimating Images using Neural Representations of Dynamic Stimuli
por: Yeung, Jacob, et al.
Publicado: (2024)
por: Yeung, Jacob, et al.
Publicado: (2024)
Animal behavioral analysis and neural encoding with transformer-based self-supervised pretraining
por: Wang, Yanchen, et al.
Publicado: (2025)
por: Wang, Yanchen, et al.
Publicado: (2025)
A Robotics-Inspired Scanpath Model Reveals the Importance of Uncertainty and Semantic Object Cues for Gaze Guidance in Dynamic Scenes
por: Mengers, Vito, et al.
Publicado: (2024)
por: Mengers, Vito, et al.
Publicado: (2024)
Probing Human Visual Robustness with Neurally-Guided Deep Neural Networks
por: Shao, Zhenan, et al.
Publicado: (2024)
por: Shao, Zhenan, et al.
Publicado: (2024)
Brain-IT: Image Reconstruction from fMRI via Brain-Interaction Transformer
por: Beliy, Roman, et al.
Publicado: (2025)
por: Beliy, Roman, et al.
Publicado: (2025)
Hierarchical Mesh Transformers with Topology-Guided Pretraining for Morphometric Analysis of Brain Structures
por: Xiong, Yujian, et al.
Publicado: (2026)
por: Xiong, Yujian, et al.
Publicado: (2026)
Source Invariance and Probabilistic Transfer: A Testable Theory of Probabilistic Neural Representations
por: Lippl, Samuel, et al.
Publicado: (2024)
por: Lippl, Samuel, et al.
Publicado: (2024)
The Topology and Geometry of Neural Representations
por: Lin, Baihan, et al.
Publicado: (2023)
por: Lin, Baihan, et al.
Publicado: (2023)
NeuroPath: A Neural Pathway Transformer for Joining the Dots of Human Connectomes
por: Wei, Ziquan, et al.
Publicado: (2024)
por: Wei, Ziquan, et al.
Publicado: (2024)
Deciphering Functions of Neurons in Vision-Language Models
por: Xu, Jiaqi, et al.
Publicado: (2025)
por: Xu, Jiaqi, et al.
Publicado: (2025)
Characterizing Universal Object Representations Across Vision Models
por: Mahner, Florian P., et al.
Publicado: (2026)
por: Mahner, Florian P., et al.
Publicado: (2026)
SIMON: Saliency-aware Integrative Multi-view Object-centric Neural Decoding
por: Lin, YuSheng, et al.
Publicado: (2026)
por: Lin, YuSheng, et al.
Publicado: (2026)
Beginning with You: Perceptual-Initialization Improves Vision-Language Representation and Alignment
por: Hu, Yang, et al.
Publicado: (2025)
por: Hu, Yang, et al.
Publicado: (2025)
What Makes a Face Look like a Hat: Decoupling Low-level and High-level Visual Properties with Image Triplets
por: Piriyajitakonkij, Maytus, et al.
Publicado: (2024)
por: Piriyajitakonkij, Maytus, et al.
Publicado: (2024)
Utilizing Computer Vision for Continuous Monitoring of Vaccine Side Effects in Experimental Mice
por: Li, Chuang, et al.
Publicado: (2024)
por: Li, Chuang, et al.
Publicado: (2024)
Scalable Diffusion Transformer for Conditional 4D fMRI Synthesis
por: Seo, Jungwoo, et al.
Publicado: (2025)
por: Seo, Jungwoo, et al.
Publicado: (2025)
Explicitly Modeling Subcortical Vision with a Neuro-Inspired Front-End Improves CNN Robustness
por: Piper, Lucas, et al.
Publicado: (2025)
por: Piper, Lucas, et al.
Publicado: (2025)
Self-Attention-Based Contextual Modulation Improves Neural System Identification
por: Lin, Isaac, et al.
Publicado: (2024)
por: Lin, Isaac, et al.
Publicado: (2024)
Explicitly Modeling Pre-Cortical Vision with a Neuro-Inspired Front-End Improves CNN Robustness
por: Piper, Lucas, et al.
Publicado: (2024)
por: Piper, Lucas, et al.
Publicado: (2024)
Human-Aligned Evaluation of a Pixel-wise DNN Color Constancy Model
por: Heidari-Gorji, Hamed, et al.
Publicado: (2026)
por: Heidari-Gorji, Hamed, et al.
Publicado: (2026)
A Multimodal Seq2Seq Transformer for Predicting Brain Responses to Naturalistic Stimuli
por: He, Qianyi, et al.
Publicado: (2025)
por: He, Qianyi, et al.
Publicado: (2025)
End-to-end Topographic Auditory Models Replicate Signatures of Human Auditory Cortex
por: Al-Tahan, Haider, et al.
Publicado: (2025)
por: Al-Tahan, Haider, et al.
Publicado: (2025)
Stimulus Motion Perception Studies Imply Specific Neural Computations in Human Visual Stabilization
por: Arathorn, David W, et al.
Publicado: (2025)
por: Arathorn, David W, et al.
Publicado: (2025)
Simple 3D Pose Features Support Human and Machine Social Scene Understanding
por: Qin, Wenshuo, et al.
Publicado: (2025)
por: Qin, Wenshuo, et al.
Publicado: (2025)
Comparing supervised learning dynamics: Deep neural networks match human data efficiency but show a generalisation lag
por: Huber, Lukas S., et al.
Publicado: (2024)
por: Huber, Lukas S., et al.
Publicado: (2024)
The Multiscale Surface Vision Transformer
por: Dahan, Simon, et al.
Publicado: (2023)
por: Dahan, Simon, et al.
Publicado: (2023)
Probability-Invariant Random Walk Learning on Gyral Folding-Based Cortical Similarity Networks for Alzheimer's and Lewy Body Dementia Diagnosis
por: Chen, Minheng, et al.
Publicado: (2026)
por: Chen, Minheng, et al.
Publicado: (2026)
ConnectomeDiffuser: Generative AI Enables Brain Network Construction from Diffusion Tensor Imaging
por: Chen, Xuhang, et al.
Publicado: (2025)
por: Chen, Xuhang, et al.
Publicado: (2025)
Action Without Interaction: Probing the Physical Foundations of Video LMMs via Contact-Release Detection
por: Harari, Daniel, et al.
Publicado: (2025)
por: Harari, Daniel, et al.
Publicado: (2025)
Grounding Social Perception in Intuitive Physics
por: Ying, Lance, et al.
Publicado: (2026)
por: Ying, Lance, et al.
Publicado: (2026)
Ejemplares similares
-
Transformer brain encoders explain human high-level visual responses
por: Adeli, Hossein, et al.
Publicado: (2025) -
In Silico Mapping of Visual Categorical Selectivity Across the Whole Brain
por: Hwang, Ethan, et al.
Publicado: (2025) -
Human face perception reflects inverse-generative and naturalistic discriminative objectives
por: Guo, Wenxuan, et al.
Publicado: (2026) -
Brain Mapping with Dense Features: Grounding Cortical Semantic Selectivity in Natural Images With Vision Transformers
por: Luo, Andrew F., et al.
Publicado: (2024) -
Scaling Vision Transformers for Functional MRI with Flat Maps
por: Lane, Connor, et al.
Publicado: (2025)