An extremely coarse feedback signal is sufficient for learning human-aligned visual representations
Fuente:
arXiv
Guardado en:
| Autores principales: | Mehta, Yash, Bonner, Michael F. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Universal dimensions of visual representation
por: Chen, Zirui, et al.
Publicado: (2024)
por: Chen, Zirui, et al.
Publicado: (2024)
Efficient coding along the visual hierarchy
por: Passi, Ananya, et al.
Publicado: (2026)
por: Passi, Ananya, et al.
Publicado: (2026)
Predicting upcoming visual features during eye movements yields scene representations aligned with human visual cortex
por: Thorat, Sushrut, et al.
Publicado: (2025)
por: Thorat, Sushrut, et al.
Publicado: (2025)
Self-supervised video pretraining yields robust and more human-aligned visual representations
por: Parthasarathy, Nikhil, et al.
Publicado: (2022)
por: Parthasarathy, Nikhil, et al.
Publicado: (2022)
Evaluating alignment between humans and neural network representations in image-based learning tasks
por: Demircan, Can, et al.
Publicado: (2023)
por: Demircan, Can, et al.
Publicado: (2023)
Do text-free diffusion models learn discriminative visual representations?
por: Mukhopadhyay, Soumik, et al.
Publicado: (2023)
por: Mukhopadhyay, Soumik, et al.
Publicado: (2023)
Visual representations in the human brain are aligned with large language models
por: Doerig, Adrien, et al.
Publicado: (2022)
por: Doerig, Adrien, et al.
Publicado: (2022)
Can multimodal representation learning by alignment preserve modality-specific information?
por: Thoreau, Romain, et al.
Publicado: (2025)
por: Thoreau, Romain, et al.
Publicado: (2025)
Characterizing the visual representation of objects from the child's view
por: Yang, Jane, et al.
Publicado: (2026)
por: Yang, Jane, et al.
Publicado: (2026)
Vision Transformer attention alignment with human visual perception in aesthetic object evaluation
por: Carrasco, Miguel, et al.
Publicado: (2025)
por: Carrasco, Miguel, et al.
Publicado: (2025)
Dimensions underlying the representational alignment of deep neural networks with humans
por: Mahner, Florian P., et al.
Publicado: (2024)
por: Mahner, Florian P., et al.
Publicado: (2024)
Do computer vision foundation models learn the low-level characteristics of the human visual system?
por: Cai, Yancheng, et al.
Publicado: (2025)
por: Cai, Yancheng, et al.
Publicado: (2025)
Dilated Convolution with Learnable Spacings makes visual models more aligned with humans: a Grad-CAM study
por: Chamas, Rabih, et al.
Publicado: (2024)
por: Chamas, Rabih, et al.
Publicado: (2024)
MoDiPO: text-to-motion alignment via AI-feedback-driven Direct Preference Optimization
por: Pappa, Massimiliano, et al.
Publicado: (2024)
por: Pappa, Massimiliano, et al.
Publicado: (2024)
Closing the gap in multimodal medical representation alignment
por: Grassucci, Eleonora, et al.
Publicado: (2026)
por: Grassucci, Eleonora, et al.
Publicado: (2026)
Towards aligned body representations in vision models
por: Gizdov, Andrey, et al.
Publicado: (2025)
por: Gizdov, Andrey, et al.
Publicado: (2025)
Extending global-local view alignment for self-supervised learning with remote sensing imagery
por: Wanyan, Xinye, et al.
Publicado: (2023)
por: Wanyan, Xinye, et al.
Publicado: (2023)
Self-supervised vision-langage alignment of deep learning representations for bone X-rays analysis
por: Englebert, Alexandre, et al.
Publicado: (2024)
por: Englebert, Alexandre, et al.
Publicado: (2024)
A transition towards virtual representations of visual scenes
por: Pereira, Américo, et al.
Publicado: (2024)
por: Pereira, Américo, et al.
Publicado: (2024)
Assessing the alignment between infants' visual and linguistic experience using multimodal language models
por: Tan, Alvin Wei Ming, et al.
Publicado: (2025)
por: Tan, Alvin Wei Ming, et al.
Publicado: (2025)
AGA: An adaptive group alignment framework for structured medical cross-modal representation learning
por: Li, Wei, et al.
Publicado: (2025)
por: Li, Wei, et al.
Publicado: (2025)
Sparse components distinguish visual pathways & their alignment to neural networks
por: Marvi, Ammar I, et al.
Publicado: (2025)
por: Marvi, Ammar I, et al.
Publicado: (2025)
Learning complete and explainable visual representations from itemized text supervision
por: Lyu, Yiwei, et al.
Publicado: (2025)
por: Lyu, Yiwei, et al.
Publicado: (2025)
On the dynamic evolution of CLIP texture-shape bias and its relationship to human alignment and model robustness
por: Hernández-Cámara, Pablo, et al.
Publicado: (2025)
por: Hernández-Cámara, Pablo, et al.
Publicado: (2025)
A Self supervised learning framework for imbalanced medical imaging datasets
por: Sharma, Yash Kumar, et al.
Publicado: (2026)
por: Sharma, Yash Kumar, et al.
Publicado: (2026)
CoCoG-2: Controllable generation of visual stimuli for understanding human concept representation
por: Wei, Chen, et al.
Publicado: (2024)
por: Wei, Chen, et al.
Publicado: (2024)
Generating visual explanations from deep networks using implicit neural representations
por: Byra, Michal, et al.
Publicado: (2025)
por: Byra, Michal, et al.
Publicado: (2025)
Self-supervised structured object representation learning
por: Hadjerci, Oussama, et al.
Publicado: (2025)
por: Hadjerci, Oussama, et al.
Publicado: (2025)
Deep video representation learning: a survey
por: Ravanbakhsh, Elham, et al.
Publicado: (2024)
por: Ravanbakhsh, Elham, et al.
Publicado: (2024)
A deep multiple instance learning approach based on coarse labels for high-resolution land-cover mapping
por: Perantoni, Gianmarco, et al.
Publicado: (2025)
por: Perantoni, Gianmarco, et al.
Publicado: (2025)
Pathological Truth Bias in Vision-Language Models
por: Thube, Yash
Publicado: (2025)
por: Thube, Yash
Publicado: (2025)
Noise-aware few-shot learning through bi-directional multi-view prompt alignment
por: Niu, Lu, et al.
Publicado: (2026)
por: Niu, Lu, et al.
Publicado: (2026)
UniAR: A Unified model for predicting human Attention and Responses on visual content
por: Li, Peizhao, et al.
Publicado: (2023)
por: Li, Peizhao, et al.
Publicado: (2023)
Vision CNNs trained to estimate spatial latents learned similar ventral-stream-aligned representations
por: Xie, Yudi, et al.
Publicado: (2024)
por: Xie, Yudi, et al.
Publicado: (2024)
Affine transformation estimation improves visual self-supervised learning
por: Torpey, David, et al.
Publicado: (2024)
por: Torpey, David, et al.
Publicado: (2024)
Human alignment of neural network representations
por: Muttenthaler, Lukas, et al.
Publicado: (2022)
por: Muttenthaler, Lukas, et al.
Publicado: (2022)
CheXalign: Preference fine-tuning in chest X-ray interpretation models without human feedback
por: Hein, Dennis, et al.
Publicado: (2024)
por: Hein, Dennis, et al.
Publicado: (2024)
KBody: Towards general, robust, and aligned monocular whole-body estimation
por: Zioulis, Nikolaos, et al.
Publicado: (2023)
por: Zioulis, Nikolaos, et al.
Publicado: (2023)
Video alignment using unsupervised learning of local and global features
por: Fakhfour, Niloufar, et al.
Publicado: (2023)
por: Fakhfour, Niloufar, et al.
Publicado: (2023)
Bridging the gap to real-world language-grounded visual concept learning
por: Jung, Whie, et al.
Publicado: (2025)
por: Jung, Whie, et al.
Publicado: (2025)
Ejemplares similares
-
Universal dimensions of visual representation
por: Chen, Zirui, et al.
Publicado: (2024) -
Efficient coding along the visual hierarchy
por: Passi, Ananya, et al.
Publicado: (2026) -
Predicting upcoming visual features during eye movements yields scene representations aligned with human visual cortex
por: Thorat, Sushrut, et al.
Publicado: (2025) -
Self-supervised video pretraining yields robust and more human-aligned visual representations
por: Parthasarathy, Nikhil, et al.
Publicado: (2022) -
Evaluating alignment between humans and neural network representations in image-based learning tasks
por: Demircan, Can, et al.
Publicado: (2023)