Google is all you need: Semi-Supervised Transfer Learning Strategy For Light Multimodal Multi-Task Classification Model
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Haixu, Jiang, Penghao, Tao, Zerui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tighnari: Multi-modal Plant Species Prediction Based on Hierarchical Cross-Attention Using Graph-Based and Vision Backbone-Extracted Features
di: Liu, Haixu, et al.
Pubblicazione: (2025)
di: Liu, Haixu, et al.
Pubblicazione: (2025)
Weak to Strong: VLM-Based Pseudo-Labeling as a Weakly Supervised Training Strategy in Multimodal Video-based Hidden Emotion Understanding Tasks
di: Wang, Yufei, et al.
Pubblicazione: (2026)
di: Wang, Yufei, et al.
Pubblicazione: (2026)
Multi-Modal Video Feature Extraction for Popularity Prediction
di: Liu, Haixu, et al.
Pubblicazione: (2025)
di: Liu, Haixu, et al.
Pubblicazione: (2025)
Unreal is all you need: Multimodal ISAC Data Simulation with Only One Engine
di: Huang, Kongwu, et al.
Pubblicazione: (2025)
di: Huang, Kongwu, et al.
Pubblicazione: (2025)
Tighnari v2: Mitigating Label Noise and Distribution Shift in Multimodal Plant Distribution Prediction via Mixture of Experts and Weakly Supervised Learning
di: Liu, Haixu, et al.
Pubblicazione: (2026)
di: Liu, Haixu, et al.
Pubblicazione: (2026)
nnY-Net: Swin-NeXt with Cross-Attention for 3D Medical Images Segmentation
di: Liu, Haixu, et al.
Pubblicazione: (2025)
di: Liu, Haixu, et al.
Pubblicazione: (2025)
Regression is all you need for medical image translation
di: Rassmann, Sebastian, et al.
Pubblicazione: (2025)
di: Rassmann, Sebastian, et al.
Pubblicazione: (2025)
Semi-Supervised Multi-Task Learning for Interpretable Quality As- sessment of Fundus Images
di: Telesco, Lucas Gabriel, et al.
Pubblicazione: (2025)
di: Telesco, Lucas Gabriel, et al.
Pubblicazione: (2025)
Exploring Facial Expression Recognition through Semi-Supervised Pretraining and Temporal Modeling
di: Yu, Jun, et al.
Pubblicazione: (2024)
di: Yu, Jun, et al.
Pubblicazione: (2024)
CRTrack: Low-Light Semi-Supervised Multi-object Tracking Based on Consistency Regularization
di: Zhao, Zijing, et al.
Pubblicazione: (2025)
di: Zhao, Zijing, et al.
Pubblicazione: (2025)
Revisiting Network Perturbation for Semi-Supervised Semantic Segmentation
di: Li, Sien, et al.
Pubblicazione: (2024)
di: Li, Sien, et al.
Pubblicazione: (2024)
Shifting to Machine Supervision: Annotation-Efficient Semi and Self-Supervised Learning for Automatic Medical Image Segmentation and Classification
di: Singh, Pranav, et al.
Pubblicazione: (2023)
di: Singh, Pranav, et al.
Pubblicazione: (2023)
Is attention all you need in medical image analysis? A review
di: Papanastasiou, Giorgos, et al.
Pubblicazione: (2023)
di: Papanastasiou, Giorgos, et al.
Pubblicazione: (2023)
STAA: Spatio-Temporal Attention Attribution for Real-Time Interpreting Transformer-based Video Models
di: Wang, Zerui, et al.
Pubblicazione: (2024)
di: Wang, Zerui, et al.
Pubblicazione: (2024)
SIAVC: Semi-Supervised Framework for Industrial Accident Video Classification
di: Li, Zuoyong, et al.
Pubblicazione: (2024)
di: Li, Zuoyong, et al.
Pubblicazione: (2024)
Exploring Beyond Logits: Hierarchical Dynamic Labeling Based on Embeddings for Semi-Supervised Classification
di: Ma, Yanbiao, et al.
Pubblicazione: (2024)
di: Ma, Yanbiao, et al.
Pubblicazione: (2024)
Lance: Unified Multimodal Modeling by Multi-Task Synergy
di: Fu, Fengyi, et al.
Pubblicazione: (2026)
di: Fu, Fengyi, et al.
Pubblicazione: (2026)
GUI-Reflection: Empowering Multimodal GUI Models with Self-Reflection Behavior
di: Wu, Penghao, et al.
Pubblicazione: (2025)
di: Wu, Penghao, et al.
Pubblicazione: (2025)
Image compositing is all you need for data augmentation
di: Shermaine, Ang Jia Ning, et al.
Pubblicazione: (2025)
di: Shermaine, Ang Jia Ning, et al.
Pubblicazione: (2025)
Patho-R1: A Multimodal Reinforcement Learning-Based Pathology Expert Reasoner
di: Zhang, Wenchuan, et al.
Pubblicazione: (2025)
di: Zhang, Wenchuan, et al.
Pubblicazione: (2025)
MiSuRe is all you need to explain your image segmentation
di: Hasany, Syed Nouman, et al.
Pubblicazione: (2024)
di: Hasany, Syed Nouman, et al.
Pubblicazione: (2024)
Are CLIP features all you need for Universal Synthetic Image Origin Attribution?
di: Cioni, Dario, et al.
Pubblicazione: (2024)
di: Cioni, Dario, et al.
Pubblicazione: (2024)
Integrating Semi-Supervised and Active Learning for Semantic Segmentation
di: Ma, Wanli, et al.
Pubblicazione: (2025)
di: Ma, Wanli, et al.
Pubblicazione: (2025)
FixCLR: Negative-Class Contrastive Learning for Semi-Supervised Domain Generalization
di: Son, Ha Min, et al.
Pubblicazione: (2025)
di: Son, Ha Min, et al.
Pubblicazione: (2025)
PathM3: A Multimodal Multi-Task Multiple Instance Learning Framework for Whole Slide Image Classification and Captioning
di: Zhou, Qifeng, et al.
Pubblicazione: (2024)
di: Zhou, Qifeng, et al.
Pubblicazione: (2024)
Semi-Supervised Learning for Deep Causal Generative Models
di: Ibrahim, Yasin, et al.
Pubblicazione: (2024)
di: Ibrahim, Yasin, et al.
Pubblicazione: (2024)
Adversarial Guided Diffusion Models for Adversarial Purification
di: Lin, Guang, et al.
Pubblicazione: (2024)
di: Lin, Guang, et al.
Pubblicazione: (2024)
Semi-Supervised Facial Expression Recognition based on Dynamic Threshold and Negative Learning
di: Cai, Zhongpeng, et al.
Pubblicazione: (2026)
di: Cai, Zhongpeng, et al.
Pubblicazione: (2026)
Multimodal Medical Image Classification via Synergistic Learning Pre-training
di: Lin, Qinghua, et al.
Pubblicazione: (2025)
di: Lin, Qinghua, et al.
Pubblicazione: (2025)
Learning from Noisy Preferences: A Semi-Supervised Learning Approach to Direct Preference Optimization
di: Liu, Xinxin, et al.
Pubblicazione: (2026)
di: Liu, Xinxin, et al.
Pubblicazione: (2026)
Semi-Supervised Coupled Thin-Plate Spline Model for Rotation Correction and Beyond
di: Nie, Lang, et al.
Pubblicazione: (2024)
di: Nie, Lang, et al.
Pubblicazione: (2024)
Pose is all you need: The pose only group activity recognition system (POGARS)
di: Thilakarathne, Haritha, et al.
Pubblicazione: (2021)
di: Thilakarathne, Haritha, et al.
Pubblicazione: (2021)
An accurate detection is not all you need to combat label noise in web-noisy datasets
di: Albert, Paul, et al.
Pubblicazione: (2024)
di: Albert, Paul, et al.
Pubblicazione: (2024)
Fusion is all you need: Face Fusion for Customized Identity-Preserving Image Synthesis
di: Mohamed, Salaheldin, et al.
Pubblicazione: (2024)
di: Mohamed, Salaheldin, et al.
Pubblicazione: (2024)
OCR is All you need: Importing Multi-Modality into Image-based Defect Detection System
di: Hsu, Chih-Chung, et al.
Pubblicazione: (2024)
di: Hsu, Chih-Chung, et al.
Pubblicazione: (2024)
Learning Disentangled Stain and Structural Representations for Semi-Supervised Histopathology Segmentation
di: Pham, Ha-Hieu, et al.
Pubblicazione: (2025)
di: Pham, Ha-Hieu, et al.
Pubblicazione: (2025)
Practical Transferability Estimation for Image Classification Tasks
di: Tan, Yang, et al.
Pubblicazione: (2021)
di: Tan, Yang, et al.
Pubblicazione: (2021)
Universal Semi-Supervised Learning for Medical Image Classification
di: Ju, Lie, et al.
Pubblicazione: (2023)
di: Ju, Lie, et al.
Pubblicazione: (2023)
An Analysis of Multi-Task Architectures for the Hierarchic Multi-Label Problem of Vehicle Model and Make Classification
di: Manole, Alexandru, et al.
Pubblicazione: (2026)
di: Manole, Alexandru, et al.
Pubblicazione: (2026)
Decouple-Then-Merge: Finetune Diffusion Models as Multi-Task Learning
di: Ma, Qianli, et al.
Pubblicazione: (2024)
di: Ma, Qianli, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Tighnari: Multi-modal Plant Species Prediction Based on Hierarchical Cross-Attention Using Graph-Based and Vision Backbone-Extracted Features
di: Liu, Haixu, et al.
Pubblicazione: (2025) -
Weak to Strong: VLM-Based Pseudo-Labeling as a Weakly Supervised Training Strategy in Multimodal Video-based Hidden Emotion Understanding Tasks
di: Wang, Yufei, et al.
Pubblicazione: (2026) -
Multi-Modal Video Feature Extraction for Popularity Prediction
di: Liu, Haixu, et al.
Pubblicazione: (2025) -
Unreal is all you need: Multimodal ISAC Data Simulation with Only One Engine
di: Huang, Kongwu, et al.
Pubblicazione: (2025) -
Tighnari v2: Mitigating Label Noise and Distribution Shift in Multimodal Plant Distribution Prediction via Mixture of Experts and Weakly Supervised Learning
di: Liu, Haixu, et al.
Pubblicazione: (2026)