Cross-Modal Clustering-Guided Negative Sampling for Self-Supervised Joint Learning from Medical Images and Reports
Fuente:
arXiv
Saved in:
| Main Authors: | Lan, Libin, Li, Hongxing, Xia, Zunhui, Zhou, Juan, Zhu, Xiaofei, Li, Yongmei, Zhang, Yudong, Luo, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DMAF-Net: An Effective Modality Rebalancing Framework for Incomplete Multi-Modal Medical Image Segmentation
by: Lan, Libin, et al.
Published: (2025)
by: Lan, Libin, et al.
Published: (2025)
TCSAFormer: Efficient Vision Transformer with Token Compression and Sparse Attention for Medical Image Segmentation
by: Xia, Zunhui, et al.
Published: (2025)
by: Xia, Zunhui, et al.
Published: (2025)
TCSAFormer : Efficient Vision Transformer With Token Compression and Sparse Attention for Medical Image Segmentation
by: Zunhui Xia, et al.
Published: (2026)
by: Zunhui Xia, et al.
Published: (2026)
MedFormer: Hierarchical Medical Vision Transformer with Content-Aware Dual Sparse Selection Attention
by: Xia, Zunhui, et al.
Published: (2025)
by: Xia, Zunhui, et al.
Published: (2025)
DSSAU-Net:U-Shaped Hybrid Network for Pubic Symphysis and Fetal Head Segmentation
by: Xia, Zunhui, et al.
Published: (2025)
by: Xia, Zunhui, et al.
Published: (2025)
BRAU-Net++: U-Shaped Hybrid CNN-Transformer Network for Medical Image Segmentation
by: Lan, Libin, et al.
Published: (2024)
by: Lan, Libin, et al.
Published: (2024)
DCAU-Net: Differential Cross Attention and Channel-Spatial Feature Fusion for Medical Image Segmentation
by: Li, Yanxin, et al.
Published: (2026)
by: Li, Yanxin, et al.
Published: (2026)
MSLAU-Net: A Hybrid CNN-Transformer Network for Medical Image Segmentation
by: Lan, Libin, et al.
Published: (2025)
by: Lan, Libin, et al.
Published: (2025)
Self-Enhanced Image Clustering with Cross-Modal Semantic Consistency
by: Li, Zihan, et al.
Published: (2025)
by: Li, Zihan, et al.
Published: (2025)
Self-Supervised Learning of Spatial Acoustic Representation with Cross-Channel Signal Reconstruction and Multi-Channel Conformer
by: Yang, Bing, et al.
Published: (2023)
by: Yang, Bing, et al.
Published: (2023)
Cross-Modal Causal Intervention for Medical Report Generation
by: Chen, Weixing, et al.
Published: (2023)
by: Chen, Weixing, et al.
Published: (2023)
Semi-MedRef: Semi-Supervised Medical Referring Image Segmentation with Cross-Modal Alignment
by: Li, Yuchen, et al.
Published: (2026)
by: Li, Yuchen, et al.
Published: (2026)
Self‐Supervised Transfer Learning of Cross‐Domains Histopathological Images for Cancer Diagnosis
by: Jianbo Zhu, et al.
Published: (2026)
by: Jianbo Zhu, et al.
Published: (2026)
Self-Supervised Graph Embedding Clustering
by: Li, Fangfang, et al.
Published: (2024)
by: Li, Fangfang, et al.
Published: (2024)
Beyond Instance-Level Self-Supervision in 3D Multi-Modal Medical Imaging
by: Pan, Tan, et al.
Published: (2026)
by: Pan, Tan, et al.
Published: (2026)
Teacher-Guided Pseudo Supervision and Cross-Modal Alignment for Audio-Visual Video Parsing
by: Chen, Yaru, et al.
Published: (2025)
by: Chen, Yaru, et al.
Published: (2025)
Multi-Level Sequence Denoising with Cross-Signal Contrastive Learning for Sequential Recommendation
by: Zhu, Xiaofei, et al.
Published: (2024)
by: Zhu, Xiaofei, et al.
Published: (2024)
Self-Supervised Cross-Modal Learning for Image-to-Point Cloud Registration
by: Wang, Xingmei, et al.
Published: (2025)
by: Wang, Xingmei, et al.
Published: (2025)
PRIOR: Prototype Representation Joint Learning from Medical Images and Reports
by: Cheng, Pujin, et al.
Published: (2023)
by: Cheng, Pujin, et al.
Published: (2023)
Unsupervised and Supervised Algorithms for Identification of Sample Pixels in FTIR Images
by: Zhao, Xiangyu, et al.
Published: (2025)
by: Zhao, Xiangyu, et al.
Published: (2025)
Self-Paced Sample Selection for Barely-Supervised Medical Image Segmentation
by: Su, Junming, et al.
Published: (2024)
by: Su, Junming, et al.
Published: (2024)
Self-Supervised Alignment Learning for Medical Image Segmentation
by: Li, Haofeng, et al.
Published: (2024)
by: Li, Haofeng, et al.
Published: (2024)
Depth-Guided Self-Supervised Human Keypoint Detection via Cross-Modal Distillation
by: Anand, Aman, et al.
Published: (2024)
by: Anand, Aman, et al.
Published: (2024)
Semi-Supervised Multi-Modal Medical Image Segmentation for Complex Situations
by: Meng, Dongdong, et al.
Published: (2025)
by: Meng, Dongdong, et al.
Published: (2025)
Modality-Guided Dynamic Graph Fusion and Temporal Diffusion for Self-Supervised RGB-T Tracking
by: Li, Shenglan, et al.
Published: (2025)
by: Li, Shenglan, et al.
Published: (2025)
Federated Self-Supervised Learning for One-Shot Cross-Modal and Cross-Imaging Technique Segmentation
by: Manna, Siladittya, et al.
Published: (2025)
by: Manna, Siladittya, et al.
Published: (2025)
TAP-SLF: Parameter-Efficient Adaptation of Vision Foundation Models for Multi-Task Ultrasound Image Analysis
by: Wan, Hui, et al.
Published: (2026)
by: Wan, Hui, et al.
Published: (2026)
Mediffusion: Joint Diffusion for Self-Explainable Semi-Supervised Classification and Medical Image Generation
by: Kaleta, Joanna, et al.
Published: (2024)
by: Kaleta, Joanna, et al.
Published: (2024)
Positive2Negative: Breaking the Information-Lossy Barrier in Self-Supervised Single Image Denoising
by: Li, Tong, et al.
Published: (2024)
by: Li, Tong, et al.
Published: (2024)
MIMNet: Multi-Interest Meta Network with Multi-Granularity Target-Guided Attention for Cross-domain Recommendation
by: Zhu, Xiaofei, et al.
Published: (2024)
by: Zhu, Xiaofei, et al.
Published: (2024)
Principle-Guided Supervision for Interpretable Uncertainty in Medical Image Segmentation
by: Sui, An, et al.
Published: (2026)
by: Sui, An, et al.
Published: (2026)
Generating Negative Samples for Multi-Modal Recommendation
by: Ji, Yanbiao, et al.
Published: (2025)
by: Ji, Yanbiao, et al.
Published: (2025)
GraphCL: Graph-based Clustering for Semi-Supervised Medical Image Segmentation
by: Wang, Mengzhu, et al.
Published: (2024)
by: Wang, Mengzhu, et al.
Published: (2024)
Prototype-Enhanced Confidence Modeling for Cross-Modal Medical Image-Report Retrieval
by: Gowda, Shreyank N, et al.
Published: (2025)
by: Gowda, Shreyank N, et al.
Published: (2025)
Self-Supervised Cross-Modal Text-Image Time Series Retrieval in Remote Sensing
by: Hoxha, Genc, et al.
Published: (2025)
by: Hoxha, Genc, et al.
Published: (2025)
Learning to Attend to Depression-Related Patterns: An Adaptive Cross-Modal Gating Network for Depression Detection
by: Yu, Hangbin, et al.
Published: (2026)
by: Yu, Hangbin, et al.
Published: (2026)
Unsupervised Hyperspectral Image Super-Resolution via Self-Supervised Modality Decoupling
by: Du, Songcheng, et al.
Published: (2024)
by: Du, Songcheng, et al.
Published: (2024)
Cross-Modal Conditioned Reconstruction for Language-guided Medical Image Segmentation
by: Huang, Xiaoshuang, et al.
Published: (2024)
by: Huang, Xiaoshuang, et al.
Published: (2024)
CIG-MAE: Cross-Modal Information-Guided Masked Autoencoder for Self-Supervised WiFi Sensing
by: Liu, Gang, et al.
Published: (2025)
by: Liu, Gang, et al.
Published: (2025)
Joint Superpixel and Self-Representation Learning for Scalable Hyperspectral Image Clustering
by: Li, Xianlu, et al.
Published: (2025)
by: Li, Xianlu, et al.
Published: (2025)
Similar Items
-
DMAF-Net: An Effective Modality Rebalancing Framework for Incomplete Multi-Modal Medical Image Segmentation
by: Lan, Libin, et al.
Published: (2025) -
TCSAFormer: Efficient Vision Transformer with Token Compression and Sparse Attention for Medical Image Segmentation
by: Xia, Zunhui, et al.
Published: (2025) -
TCSAFormer : Efficient Vision Transformer With Token Compression and Sparse Attention for Medical Image Segmentation
by: Zunhui Xia, et al.
Published: (2026) -
MedFormer: Hierarchical Medical Vision Transformer with Content-Aware Dual Sparse Selection Attention
by: Xia, Zunhui, et al.
Published: (2025) -
DSSAU-Net:U-Shaped Hybrid Network for Pubic Symphysis and Fetal Head Segmentation
by: Xia, Zunhui, et al.
Published: (2025)