Do computer vision foundation models learn the low-level characteristics of the human visual system?
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Yancheng, Yin, Fei, Hammou, Dounia, Mantiuk, Rafal |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating quality metrics through the lenses of psychophysical measurements of low-level vision
by: Hammou, Dounia, et al.
Published: (2025)
by: Hammou, Dounia, et al.
Published: (2025)
CameraVDP: Perceptual Display Assessment with Uncertainty Estimation via Camera and Visual Difference Prediction
by: Cai, Yancheng, et al.
Published: (2025)
by: Cai, Yancheng, et al.
Published: (2025)
elaTCSF: A Temporal Contrast Sensitivity Function for Flicker Detection and Modeling Variable Refresh Rate Flicker
by: Cai, Yancheng, et al.
Published: (2025)
by: Cai, Yancheng, et al.
Published: (2025)
Implicit neural representation of textures
by: Kwok, Albert, et al.
Published: (2026)
by: Kwok, Albert, et al.
Published: (2026)
FaceCraft4D: Animated 3D Facial Avatar Generation from a Single Image
by: Yin, Fei, et al.
Published: (2025)
by: Yin, Fei, et al.
Published: (2025)
Are foundation models for computer vision good conformal predictors?
by: Fillioux, Leo, et al.
Published: (2024)
by: Fillioux, Leo, et al.
Published: (2024)
ColorVideoVDP: A visual difference predictor for image, video and display distortions
by: Mantiuk, Rafal K., et al.
Published: (2024)
by: Mantiuk, Rafal K., et al.
Published: (2024)
Perceptual Assessment and Optimization of HDR Image Rendering
by: Cao, Peibei, et al.
Published: (2023)
by: Cao, Peibei, et al.
Published: (2023)
Intelligent bear deterrence system based on computer vision: Reducing human-bear conflicts in remote areas
by: Chen, Pengyu, et al.
Published: (2025)
by: Chen, Pengyu, et al.
Published: (2025)
A Neural Quality Metric for BRDF Models
by: Kavoosighafi, Behnaz, et al.
Published: (2025)
by: Kavoosighafi, Behnaz, et al.
Published: (2025)
Quantifying the human visual exposome with vision language models
by: Rominger, Christian, et al.
Published: (2026)
by: Rominger, Christian, et al.
Published: (2026)
Expert-level vision-language foundation model for real-world radiology and comprehensive evaluation
by: Liu, Xiaohong, et al.
Published: (2024)
by: Liu, Xiaohong, et al.
Published: (2024)
X2HDR: HDR Image Generation in a Perceptually Uniform Space
by: Wu, Ronghuan, et al.
Published: (2026)
by: Wu, Ronghuan, et al.
Published: (2026)
Do vision models perceive illusory motion in static images like humans?
by: Rosario, Isabella Elaine, et al.
Published: (2026)
by: Rosario, Isabella Elaine, et al.
Published: (2026)
Do text-free diffusion models learn discriminative visual representations?
by: Mukhopadhyay, Soumik, et al.
Published: (2023)
by: Mukhopadhyay, Soumik, et al.
Published: (2023)
Ensemble learning of foundation models for precision oncology
by: Luo, Xiangde, et al.
Published: (2025)
by: Luo, Xiangde, et al.
Published: (2025)
Rapidly deploying on-device eye tracking by distilling visual foundation models
by: Jiang, Cheng, et al.
Published: (2026)
by: Jiang, Cheng, et al.
Published: (2026)
Streaming of rendered content with adaptive frame rate and resolution
by: Liu, Yaru, et al.
Published: (2026)
by: Liu, Yaru, et al.
Published: (2026)
Fine-tuning vision foundation model for crack segmentation in civil infrastructures
by: Ge, Kang, et al.
Published: (2023)
by: Ge, Kang, et al.
Published: (2023)
ICME 2025 Generalizable HDR and SDR Video Quality Measurement Grand Challenge
by: Chen, Yixu, et al.
Published: (2025)
by: Chen, Yixu, et al.
Published: (2025)
Multi-label classification for multi-temporal, multi-spatial coral reef condition monitoring using vision foundation model with adapter learning
by: Shao, Xinlei, et al.
Published: (2025)
by: Shao, Xinlei, et al.
Published: (2025)
Enabling clinical use of foundation models for computational pathology
by: Henriksen, Audun L, et al.
Published: (2026)
by: Henriksen, Audun L, et al.
Published: (2026)
MedDINOv3: How to adapt vision foundation models for medical image segmentation?
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
Enhancing the vision-language foundation model with key semantic knowledge-emphasized report refinement
by: Huang, Weijian, et al.
Published: (2024)
by: Huang, Weijian, et al.
Published: (2024)
For a semiotic AI: Bridging computer vision and visual semiotics for computational observation of large scale facial image archives
by: Morra, Lia, et al.
Published: (2024)
by: Morra, Lia, et al.
Published: (2024)
Scaling up self-supervised learning for improved surgical foundation models
by: Jaspers, Tim J. M., et al.
Published: (2025)
by: Jaspers, Tim J. M., et al.
Published: (2025)
Thinker: A vision-language foundation model for embodied intelligence
by: Pan, Baiyu, et al.
Published: (2026)
by: Pan, Baiyu, et al.
Published: (2026)
A multimodal vision foundation model for generalizable knee pathology
by: Yu, Kang, et al.
Published: (2026)
by: Yu, Kang, et al.
Published: (2026)
Estimating the distribution of numerosity and non-numerical visual magnitudes in natural scenes using computer vision
by: Hou, Kuinan, et al.
Published: (2024)
by: Hou, Kuinan, et al.
Published: (2024)
Object segmentation in the wild with foundation models: application to vision assisted neuro-prostheses for upper limbs
by: Atoki, Bolutife, et al.
Published: (2025)
by: Atoki, Bolutife, et al.
Published: (2025)
Zero-shot segmentation of skin tumors in whole-slide images with vision-language foundation models
by: Moreno, Santiago, et al.
Published: (2025)
by: Moreno, Santiago, et al.
Published: (2025)
EXACT: an explainable anomaly-aware vision foundation model for analysis of 3D chest CT
by: Bai, Xuguang, et al.
Published: (2026)
by: Bai, Xuguang, et al.
Published: (2026)
Attend what matters: Leveraging vision foundational models for breast cancer classification using mammograms
by: Sanghvi, Samyak, et al.
Published: (2026)
by: Sanghvi, Samyak, et al.
Published: (2026)
ProFound: A moderate-sized vision foundation model for multi-task prostate imaging
by: Wang, Yipei, et al.
Published: (2026)
by: Wang, Yipei, et al.
Published: (2026)
VISTA-PATH: An interactive foundation model for pathology image segmentation and quantitative analysis in computational pathology
by: Liang, Peixian, et al.
Published: (2026)
by: Liang, Peixian, et al.
Published: (2026)
A data- and compute-efficient chest X-ray foundation model beyond aggressive scaling
by: Wang, Chong, et al.
Published: (2026)
by: Wang, Chong, et al.
Published: (2026)
Ultra-low-light computer vision using trained photon correlations
by: Sohoni, Mandar M., et al.
Published: (2026)
by: Sohoni, Mandar M., et al.
Published: (2026)
ActiveMark: on watermarking of visual foundation models via massive activations
by: Chistyakova, Anna, et al.
Published: (2025)
by: Chistyakova, Anna, et al.
Published: (2025)
Tissue Concepts: supervised foundation models in computational pathology
by: Nicke, Till, et al.
Published: (2024)
by: Nicke, Till, et al.
Published: (2024)
An extremely coarse feedback signal is sufficient for learning human-aligned visual representations
by: Mehta, Yash, et al.
Published: (2026)
by: Mehta, Yash, et al.
Published: (2026)
Similar Items
-
Evaluating quality metrics through the lenses of psychophysical measurements of low-level vision
by: Hammou, Dounia, et al.
Published: (2025) -
CameraVDP: Perceptual Display Assessment with Uncertainty Estimation via Camera and Visual Difference Prediction
by: Cai, Yancheng, et al.
Published: (2025) -
elaTCSF: A Temporal Contrast Sensitivity Function for Flicker Detection and Modeling Variable Refresh Rate Flicker
by: Cai, Yancheng, et al.
Published: (2025) -
Implicit neural representation of textures
by: Kwok, Albert, et al.
Published: (2026) -
FaceCraft4D: Animated 3D Facial Avatar Generation from a Single Image
by: Yin, Fei, et al.
Published: (2025)