Cross-Modality Clustering-based Self-Labeling for Multimodal Data Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Zyblewski, Paweł, Minku, Leandro L. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multimodal Classification via Modal-Aware Interactive Enhancement
by: Jiang, Qing-Yuan, et al.
Published: (2024)
by: Jiang, Qing-Yuan, et al.
Published: (2024)
On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
by: Cai, Yichao, et al.
Published: (2025)
by: Cai, Yichao, et al.
Published: (2025)
Enhancing Multi-Modal Video Sentiment Classification Through Semi-Supervised Clustering
by: Saadatinia, Mehrshad, et al.
Published: (2025)
by: Saadatinia, Mehrshad, et al.
Published: (2025)
Robust Classification by Coupling Data Mollification with Label Smoothing
by: Heinonen, Markus, et al.
Published: (2024)
by: Heinonen, Markus, et al.
Published: (2024)
Gradient based Severity Labeling for Biomarker Classification in OCT
by: Kokilepersaud, Kiran, et al.
Published: (2026)
by: Kokilepersaud, Kiran, et al.
Published: (2026)
Score-based Conditional Generation with Fewer Labeled Data by Self-calibrating Classifier Guidance
by: Huang, Paul Kuo-Ming, et al.
Published: (2023)
by: Huang, Paul Kuo-Ming, et al.
Published: (2023)
CROME: Cross-Modal Adapters for Efficient Multimodal LLM
by: Ebrahimi, Sayna, et al.
Published: (2024)
by: Ebrahimi, Sayna, et al.
Published: (2024)
Mediffusion: Joint Diffusion for Self-Explainable Semi-Supervised Classification and Medical Image Generation
by: Kaleta, Joanna, et al.
Published: (2024)
by: Kaleta, Joanna, et al.
Published: (2024)
Contrastive Learning for Multimodal Human Activity Recognition with Limited Labeled Data
by: Jing, Long, et al.
Published: (2026)
by: Jing, Long, et al.
Published: (2026)
Multimodal Structure Learning: Disentangling Shared and Specific Topology via Cross-Modal Graphical Lasso
by: Wang, Fei, et al.
Published: (2026)
by: Wang, Fei, et al.
Published: (2026)
Radiation-Preserving Selective Imaging for Pediatric Hip Dysplasia: A Cross-Modal Ultrasound-Xray Policy with Limited Labels
by: Stothers, Duncan, et al.
Published: (2025)
by: Stothers, Duncan, et al.
Published: (2025)
Leveraging Semi-Supervised Learning to Enhance Data Mining for Image Classification under Limited Labeled Data
by: Shen, Aoran, et al.
Published: (2024)
by: Shen, Aoran, et al.
Published: (2024)
SelfFed: Self-Supervised Federated Learning for Data Heterogeneity and Label Scarcity in Medical Images
by: Khowaja, Sunder Ali, et al.
Published: (2023)
by: Khowaja, Sunder Ali, et al.
Published: (2023)
Employing Two-Dimensional Word Embedding for Difficult Tabular Data Stream Classification
by: Zyblewski, Paweł
Published: (2024)
by: Zyblewski, Paweł
Published: (2024)
Multimodal LLMs under Pairwise Modalities
by: Li, Yan, et al.
Published: (2026)
by: Li, Yan, et al.
Published: (2026)
Pairwise Similarity Distribution Clustering for Noisy Label Learning
by: Bai, Sihan
Published: (2024)
by: Bai, Sihan
Published: (2024)
Cooperative Pseudo Labeling for Unsupervised Federated Classification
by: Guo, Kuangpu, et al.
Published: (2025)
by: Guo, Kuangpu, et al.
Published: (2025)
Robust Multimodal Learning via Cross-Modal Proxy Tokens
by: Reza, Md Kaykobad, et al.
Published: (2025)
by: Reza, Md Kaykobad, et al.
Published: (2025)
Enhancing Cross-Modal Fine-Tuning with Gradually Intermediate Modality Generation
by: Cai, Lincan, et al.
Published: (2024)
by: Cai, Lincan, et al.
Published: (2024)
Conformal Cross-Modal Active Learning
by: Nguyen, Huy Hoang, et al.
Published: (2026)
by: Nguyen, Huy Hoang, et al.
Published: (2026)
ViTaPEs: Visuotactile Position Encodings for Cross-Modal Alignment in Multimodal Transformers
by: Lygerakis, Fotios, et al.
Published: (2025)
by: Lygerakis, Fotios, et al.
Published: (2025)
Data Selection for Fine-tuning Vision Language Models via Cross Modal Alignment Trajectories
by: Naharas, Nilay, et al.
Published: (2025)
by: Naharas, Nilay, et al.
Published: (2025)
Mind the Gap: Learning Modality-Agnostic Representations with a Cross-Modality UNet
by: Niu, Xin, et al.
Published: (2026)
by: Niu, Xin, et al.
Published: (2026)
Dynamic Cross-Modal Prompt Generation for Multimodal Continual Instruction Tuning
by: Hu, Tao, et al.
Published: (2026)
by: Hu, Tao, et al.
Published: (2026)
Multimodal Foundation Model for Cross-Modal Retrieval and Activity Recognition Tasks
by: Matsuishi, Koki, et al.
Published: (2025)
by: Matsuishi, Koki, et al.
Published: (2025)
Multi-Label Plant Species Classification with Self-Supervised Vision Transformers
by: Gustineli, Murilo, et al.
Published: (2024)
by: Gustineli, Murilo, et al.
Published: (2024)
No Labels Needed: Zero-Shot Image Classification with Collaborative Self-Learning
by: Todescato, Matheus Vinícius, et al.
Published: (2025)
by: Todescato, Matheus Vinícius, et al.
Published: (2025)
Enhancing Cognition and Explainability of Multimodal Foundation Models with Self-Synthesized Data
by: Shi, Yucheng, et al.
Published: (2025)
by: Shi, Yucheng, et al.
Published: (2025)
Cross-Validation Is All You Need: A Statistical Approach To Label Noise Estimation
by: Chen, Jianan, et al.
Published: (2023)
by: Chen, Jianan, et al.
Published: (2023)
Multimodal Learning Without Labeled Multimodal Data: Guarantees and Applications
by: Liang, Paul Pu, et al.
Published: (2023)
by: Liang, Paul Pu, et al.
Published: (2023)
Label Propagation for Zero-shot Classification with Vision-Language Models
by: Stojnić, Vladan, et al.
Published: (2024)
by: Stojnić, Vladan, et al.
Published: (2024)
A Re-node Self-training Approach for Deep Graph-based Semi-supervised Classification on Multi-view Image Data
by: Bi, Jingjun, et al.
Published: (2025)
by: Bi, Jingjun, et al.
Published: (2025)
Active Label Refinement for Robust Training of Imbalanced Medical Image Classification Tasks in the Presence of High Label Noise
by: Khanal, Bidur, et al.
Published: (2024)
by: Khanal, Bidur, et al.
Published: (2024)
LEMoN: Label Error Detection using Multimodal Neighbors
by: Zhang, Haoran, et al.
Published: (2024)
by: Zhang, Haoran, et al.
Published: (2024)
PROMISE: Prompt-Attentive Hierarchical Contrastive Learning for Robust Cross-Modal Representation with Missing Modalities
by: Chen, Jiajun, et al.
Published: (2025)
by: Chen, Jiajun, et al.
Published: (2025)
Cross-Modal Redundancy and the Geometry of Vision-Language Embeddings
by: Dhimoïla, Grégoire, et al.
Published: (2026)
by: Dhimoïla, Grégoire, et al.
Published: (2026)
Quantifying Cross-Modality Memorization in Vision-Language Models
by: Wen, Yuxin, et al.
Published: (2025)
by: Wen, Yuxin, et al.
Published: (2025)
Mitigating Instance-Dependent Label Noise: Integrating Self-Supervised Pretraining with Pseudo-Label Refinement
by: Bala, Gouranga, et al.
Published: (2024)
by: Bala, Gouranga, et al.
Published: (2024)
Multimodal Unsupervised Domain Generalization by Retrieving Across the Modality Gap
by: Liao, Christopher, et al.
Published: (2024)
by: Liao, Christopher, et al.
Published: (2024)
Multimodal Guidance Network for Missing-Modality Inference in Content Moderation
by: Zhao, Zhuokai, et al.
Published: (2023)
by: Zhao, Zhuokai, et al.
Published: (2023)
Similar Items
-
Multimodal Classification via Modal-Aware Interactive Enhancement
by: Jiang, Qing-Yuan, et al.
Published: (2024) -
On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
by: Cai, Yichao, et al.
Published: (2025) -
Enhancing Multi-Modal Video Sentiment Classification Through Semi-Supervised Clustering
by: Saadatinia, Mehrshad, et al.
Published: (2025) -
Robust Classification by Coupling Data Mollification with Label Smoothing
by: Heinonen, Markus, et al.
Published: (2024) -
Gradient based Severity Labeling for Biomarker Classification in OCT
by: Kokilepersaud, Kiran, et al.
Published: (2026)