Black Sheep in the Herd: Playing with Spuriously Correlated Attributes for Vision-Language Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Tian, Xinyu, Zou, Shu, Yang, Zhaoyuan, He, Mengqi, Zhang, Jing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Remote Sensing Semantic Segmentation Quality Assessment based on Vision Language Model
by: Shi, Huiying, et al.
Published: (2025)
by: Shi, Huiying, et al.
Published: (2025)
Uncovering Memorization Effect in the Presence of Spurious Correlations
by: You, Chenyu, et al.
Published: (2025)
by: You, Chenyu, et al.
Published: (2025)
VALD-MD: Visual Attribution via Latent Diffusion for Medical Diagnostics
by: Siddiqui, Ammar A., et al.
Published: (2024)
by: Siddiqui, Ammar A., et al.
Published: (2024)
End-to-end learned Lossy Dynamic Point Cloud Attribute Compression
by: Nguyen, Dat Thanh, et al.
Published: (2024)
by: Nguyen, Dat Thanh, et al.
Published: (2024)
Comparison of ConvNeXt and Vision-Language Models for Breast Density Assessment in Screening Mammography
by: Molina-Román, Yusdivia, et al.
Published: (2025)
by: Molina-Román, Yusdivia, et al.
Published: (2025)
PSI3D: Plug-and-Play 3D Stochastic Inference with Slice-wise Latent Diffusion Prior
by: Guo, Wenhan, et al.
Published: (2025)
by: Guo, Wenhan, et al.
Published: (2025)
Lossless Point Cloud Geometry and Attribute Compression Using a Learned Conditional Probability Model
by: Nguyen, Dat Thanh, et al.
Published: (2023)
by: Nguyen, Dat Thanh, et al.
Published: (2023)
Validating the Clinical Utility of CineECG 3D Reconstructions through Cross-Modal Feature Attribution
by: Dobiczek, Karol, et al.
Published: (2026)
by: Dobiczek, Karol, et al.
Published: (2026)
Scribe Verification in Chinese manuscripts using Siamese, Triplet, and Vision Transformer Neural Networks
by: Liakopoulos, Dimitrios-Chrysovalantis, et al.
Published: (2026)
by: Liakopoulos, Dimitrios-Chrysovalantis, et al.
Published: (2026)
Plug-and-Play Regularization on Magnitude with Deep Priors for 3D Near-Field MIMO Imaging
by: Oral, Okyanus, et al.
Published: (2023)
by: Oral, Okyanus, et al.
Published: (2023)
Neural Architecture Search generated Phase Retrieval Net for Real-time Off-axis Quantitative Phase Imaging
by: Shu, Xin, et al.
Published: (2022)
by: Shu, Xin, et al.
Published: (2022)
Learned Nonlinear Predictor for Critically Sampled 3D Point Cloud Attribute Compression
by: Do, Tam Thuc, et al.
Published: (2023)
by: Do, Tam Thuc, et al.
Published: (2023)
SmoothSegNet: A Global-Local Framework for Liver Tumor Segmentation with Clinical KnowledgeInformed Label Smoothing
by: Wang, Hairong, et al.
Published: (2024)
by: Wang, Hairong, et al.
Published: (2024)
A Deep Learning-Driven Pipeline for Differentiating Hypertrophic Cardiomyopathy from Cardiac Amyloidosis Using 2D Multi-View Echocardiography
by: Peng, Bo, et al.
Published: (2024)
by: Peng, Bo, et al.
Published: (2024)
SAR-GTR: Attributed Scattering Information Guided SAR Graph Transformer Recognition Algorithm
by: Xiong, Xuying, et al.
Published: (2025)
by: Xiong, Xuying, et al.
Published: (2025)
Mind Your Vision: Multimodal Estimation of Refractive Disorders Using Electrooculography and Eye Tracking
by: Wei, Xin, et al.
Published: (2025)
by: Wei, Xin, et al.
Published: (2025)
Lane-Wise Highway Anomaly Detection
by: Qiu, Mei, et al.
Published: (2025)
by: Qiu, Mei, et al.
Published: (2025)
Tiny-VBF: Resource-Efficient Vision Transformer based Lightweight Beamformer for Ultrasound Single-Angle Plane Wave Imaging
by: Rahoof, Abdul, et al.
Published: (2023)
by: Rahoof, Abdul, et al.
Published: (2023)
3D-U-SAM Network For Few-shot Tooth Segmentation in CBCT Images
by: Zhang, Yifu, et al.
Published: (2023)
by: Zhang, Yifu, et al.
Published: (2023)
RETO: A Rotary-Enhanced Transformer Operator for High-Fidelity Prediction of Automotive Aerodynamics
by: Zhang, Bojun, et al.
Published: (2026)
by: Zhang, Bojun, et al.
Published: (2026)
Incomplete Graph Learning: A Comprehensive Survey
by: Xia, Riting, et al.
Published: (2025)
by: Xia, Riting, et al.
Published: (2025)
Towards In-Vehicle Multi-Task Facial Attribute Recognition: Investigating Synthetic Data and Vision Foundation Models
by: Seraj, Esmaeil, et al.
Published: (2024)
by: Seraj, Esmaeil, et al.
Published: (2024)
Super Resolution Based on Deep Operator Networks
by: Yang, Siyuan
Published: (2024)
by: Yang, Siyuan
Published: (2024)
Comparative Analysis of ImageNet Pre-Trained Deep Learning Models and DINOv2 in Medical Imaging Classification
by: Huang, Yuning, et al.
Published: (2024)
by: Huang, Yuning, et al.
Published: (2024)
Beyond Single-Channel: Multichannel Signal Imaging for PPG-to-ECG Reconstruction with Vision Transformers
by: Li, Xiaoyan, et al.
Published: (2025)
by: Li, Xiaoyan, et al.
Published: (2025)
Learning the irreversible progression trajectory of Alzheimer's disease
by: Wang, Yipei, et al.
Published: (2024)
by: Wang, Yipei, et al.
Published: (2024)
Plug-and-Play Diffusion Meets ADMM: Dual-Variable Coupling for Robust Medical Image Reconstruction
by: Du, Chenhe, et al.
Published: (2026)
by: Du, Chenhe, et al.
Published: (2026)
NeuroMambaLLM: Dynamic Graph Learning of fMRI Functional Connectivity in Autistic Brains Using Mamba and Language Model Reasoning
by: Torabi, Yasaman, et al.
Published: (2026)
by: Torabi, Yasaman, et al.
Published: (2026)
Dynamic PET Image Prediction Using a Network Combining Reversible and Irreversible Modules
by: Sun, Jie, et al.
Published: (2024)
by: Sun, Jie, et al.
Published: (2024)
Beyond Manual Annotation: A Human-AI Collaborative Framework for Medical Image Segmentation Using Only "Better or Worse" Expert Feedback
by: Zhang, Yizhe
Published: (2025)
by: Zhang, Yizhe
Published: (2025)
Deep Unrolling of Sparsity-Induced RDO for 3D Point Cloud Attribute Coding
by: Do, Tam Thuc, et al.
Published: (2025)
by: Do, Tam Thuc, et al.
Published: (2025)
Deep Separable Spatiotemporal Learning for Fast Dynamic Cardiac MRI
by: Wang, Zi, et al.
Published: (2024)
by: Wang, Zi, et al.
Published: (2024)
Efficient Test-Time Adaptation through Latent Subspace Coefficients Search
by: Luo, Xinyu, et al.
Published: (2025)
by: Luo, Xinyu, et al.
Published: (2025)
Shrinking Your TimeStep: Towards Low-Latency Neuromorphic Object Recognition with Spiking Neural Network
by: Ding, Yongqi, et al.
Published: (2024)
by: Ding, Yongqi, et al.
Published: (2024)
Edge-boosted graph learning for functional brain connectivity analysis
by: Yang, David, et al.
Published: (2025)
by: Yang, David, et al.
Published: (2025)
Leveraging Second-Order Curvature for Efficient Learned Image Compression: Theory and Empirical Evidence
by: Zhang, Yichi, et al.
Published: (2026)
by: Zhang, Yichi, et al.
Published: (2026)
A Literature Review on Fetus Brain Motion Correction in MRI
by: Zhang, Haoran, et al.
Published: (2024)
by: Zhang, Haoran, et al.
Published: (2024)
DiSC-Med: Diffusion-based Semantic Communications for Robust Medical Image Transmission
by: Guo, Fupei, et al.
Published: (2025)
by: Guo, Fupei, et al.
Published: (2025)
Physics-Aware Neural Operators for Direct Inversion in 3D Photoacoustic Tomography
by: Wang, Jiayun, et al.
Published: (2025)
by: Wang, Jiayun, et al.
Published: (2025)
On Privacy-Preserving Image Transmission in Low-Altitude Networks: A Swin Transformer-Based Framework with Federated Learning
by: Zhang, Kexin, et al.
Published: (2026)
by: Zhang, Kexin, et al.
Published: (2026)
Similar Items
-
Remote Sensing Semantic Segmentation Quality Assessment based on Vision Language Model
by: Shi, Huiying, et al.
Published: (2025) -
Uncovering Memorization Effect in the Presence of Spurious Correlations
by: You, Chenyu, et al.
Published: (2025) -
VALD-MD: Visual Attribution via Latent Diffusion for Medical Diagnostics
by: Siddiqui, Ammar A., et al.
Published: (2024) -
End-to-end learned Lossy Dynamic Point Cloud Attribute Compression
by: Nguyen, Dat Thanh, et al.
Published: (2024) -
Comparison of ConvNeXt and Vision-Language Models for Breast Density Assessment in Screening Mammography
by: Molina-Román, Yusdivia, et al.
Published: (2025)