Saved in:
| Main Authors: | Liu, Haixu, Jiang, Penghao, Tao, Zerui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.01611 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tighnari: Multi-modal Plant Species Prediction Based on Hierarchical Cross-Attention Using Graph-Based and Vision Backbone-Extracted Features
by: Liu, Haixu, et al.
Published: (2025)
by: Liu, Haixu, et al.
Published: (2025)
Weak to Strong: VLM-Based Pseudo-Labeling as a Weakly Supervised Training Strategy in Multimodal Video-based Hidden Emotion Understanding Tasks
by: Wang, Yufei, et al.
Published: (2026)
by: Wang, Yufei, et al.
Published: (2026)
Multi-Modal Video Feature Extraction for Popularity Prediction
by: Liu, Haixu, et al.
Published: (2025)
by: Liu, Haixu, et al.
Published: (2025)
nnY-Net: Swin-NeXt with Cross-Attention for 3D Medical Images Segmentation
by: Liu, Haixu, et al.
Published: (2025)
by: Liu, Haixu, et al.
Published: (2025)
Regression is all you need for medical image translation
by: Rassmann, Sebastian, et al.
Published: (2025)
by: Rassmann, Sebastian, et al.
Published: (2025)
Tighnari v2: Mitigating Label Noise and Distribution Shift in Multimodal Plant Distribution Prediction via Mixture of Experts and Weakly Supervised Learning
by: Liu, Haixu, et al.
Published: (2026)
by: Liu, Haixu, et al.
Published: (2026)
Unreal is all you need: Multimodal ISAC Data Simulation with Only One Engine
by: Huang, Kongwu, et al.
Published: (2025)
by: Huang, Kongwu, et al.
Published: (2025)
Semi-Supervised Multi-Task Learning for Interpretable Quality As- sessment of Fundus Images
by: Telesco, Lucas Gabriel, et al.
Published: (2025)
by: Telesco, Lucas Gabriel, et al.
Published: (2025)
Exploring Facial Expression Recognition through Semi-Supervised Pretraining and Temporal Modeling
by: Yu, Jun, et al.
Published: (2024)
by: Yu, Jun, et al.
Published: (2024)
Is attention all you need in medical image analysis? A review
by: Papanastasiou, Giorgos, et al.
Published: (2023)
by: Papanastasiou, Giorgos, et al.
Published: (2023)
CRTrack: Low-Light Semi-Supervised Multi-object Tracking Based on Consistency Regularization
by: Zhao, Zijing, et al.
Published: (2025)
by: Zhao, Zijing, et al.
Published: (2025)
STAA: Spatio-Temporal Attention Attribution for Real-Time Interpreting Transformer-based Video Models
by: Wang, Zerui, et al.
Published: (2024)
by: Wang, Zerui, et al.
Published: (2024)
Revisiting Network Perturbation for Semi-Supervised Semantic Segmentation
by: Li, Sien, et al.
Published: (2024)
by: Li, Sien, et al.
Published: (2024)
Shifting to Machine Supervision: Annotation-Efficient Semi and Self-Supervised Learning for Automatic Medical Image Segmentation and Classification
by: Singh, Pranav, et al.
Published: (2023)
by: Singh, Pranav, et al.
Published: (2023)
GUI-Reflection: Empowering Multimodal GUI Models with Self-Reflection Behavior
by: Wu, Penghao, et al.
Published: (2025)
by: Wu, Penghao, et al.
Published: (2025)
SIAVC: Semi-Supervised Framework for Industrial Accident Video Classification
by: Li, Zuoyong, et al.
Published: (2024)
by: Li, Zuoyong, et al.
Published: (2024)
Exploring Beyond Logits: Hierarchical Dynamic Labeling Based on Embeddings for Semi-Supervised Classification
by: Ma, Yanbiao, et al.
Published: (2024)
by: Ma, Yanbiao, et al.
Published: (2024)
Lance: Unified Multimodal Modeling by Multi-Task Synergy
by: Fu, Fengyi, et al.
Published: (2026)
by: Fu, Fengyi, et al.
Published: (2026)
Image compositing is all you need for data augmentation
by: Shermaine, Ang Jia Ning, et al.
Published: (2025)
by: Shermaine, Ang Jia Ning, et al.
Published: (2025)
Patho-R1: A Multimodal Reinforcement Learning-Based Pathology Expert Reasoner
by: Zhang, Wenchuan, et al.
Published: (2025)
by: Zhang, Wenchuan, et al.
Published: (2025)
Adversarial Guided Diffusion Models for Adversarial Purification
by: Lin, Guang, et al.
Published: (2024)
by: Lin, Guang, et al.
Published: (2024)
MiSuRe is all you need to explain your image segmentation
by: Hasany, Syed Nouman, et al.
Published: (2024)
by: Hasany, Syed Nouman, et al.
Published: (2024)
Are CLIP features all you need for Universal Synthetic Image Origin Attribution?
by: Cioni, Dario, et al.
Published: (2024)
by: Cioni, Dario, et al.
Published: (2024)
OCR is All you need: Importing Multi-Modality into Image-based Defect Detection System
by: Hsu, Chih-Chung, et al.
Published: (2024)
by: Hsu, Chih-Chung, et al.
Published: (2024)
Semi-Supervised Learning for Deep Causal Generative Models
by: Ibrahim, Yasin, et al.
Published: (2024)
by: Ibrahim, Yasin, et al.
Published: (2024)
Integrating Semi-Supervised and Active Learning for Semantic Segmentation
by: Ma, Wanli, et al.
Published: (2025)
by: Ma, Wanli, et al.
Published: (2025)
FixCLR: Negative-Class Contrastive Learning for Semi-Supervised Domain Generalization
by: Son, Ha Min, et al.
Published: (2025)
by: Son, Ha Min, et al.
Published: (2025)
PathM3: A Multimodal Multi-Task Multiple Instance Learning Framework for Whole Slide Image Classification and Captioning
by: Zhou, Qifeng, et al.
Published: (2024)
by: Zhou, Qifeng, et al.
Published: (2024)
Multimodal Medical Image Classification via Synergistic Learning Pre-training
by: Lin, Qinghua, et al.
Published: (2025)
by: Lin, Qinghua, et al.
Published: (2025)
Semi-Supervised Facial Expression Recognition based on Dynamic Threshold and Negative Learning
by: Cai, Zhongpeng, et al.
Published: (2026)
by: Cai, Zhongpeng, et al.
Published: (2026)
Learning from Noisy Preferences: A Semi-Supervised Learning Approach to Direct Preference Optimization
by: Liu, Xinxin, et al.
Published: (2026)
by: Liu, Xinxin, et al.
Published: (2026)
Pose is all you need: The pose only group activity recognition system (POGARS)
by: Thilakarathne, Haritha, et al.
Published: (2021)
by: Thilakarathne, Haritha, et al.
Published: (2021)
An accurate detection is not all you need to combat label noise in web-noisy datasets
by: Albert, Paul, et al.
Published: (2024)
by: Albert, Paul, et al.
Published: (2024)
Fusion is all you need: Face Fusion for Customized Identity-Preserving Image Synthesis
by: Mohamed, Salaheldin, et al.
Published: (2024)
by: Mohamed, Salaheldin, et al.
Published: (2024)
Semi-Supervised Coupled Thin-Plate Spline Model for Rotation Correction and Beyond
by: Nie, Lang, et al.
Published: (2024)
by: Nie, Lang, et al.
Published: (2024)
Practical Transferability Estimation for Image Classification Tasks
by: Tan, Yang, et al.
Published: (2021)
by: Tan, Yang, et al.
Published: (2021)
SG-OIF: A Stability-Guided Online Influence Framework for Reliable Vision Data
by: Rao, Penghao, et al.
Published: (2025)
by: Rao, Penghao, et al.
Published: (2025)
Learning Disentangled Stain and Structural Representations for Semi-Supervised Histopathology Segmentation
by: Pham, Ha-Hieu, et al.
Published: (2025)
by: Pham, Ha-Hieu, et al.
Published: (2025)
SS-ADA: A Semi-Supervised Active Domain Adaptation Framework for Semantic Segmentation
by: Yan, Weihao, et al.
Published: (2024)
by: Yan, Weihao, et al.
Published: (2024)
CAST: Contrastive Adaptation and Distillation for Semi-Supervised Instance Segmentation
by: Taghavi, Pardis, et al.
Published: (2025)
by: Taghavi, Pardis, et al.
Published: (2025)
Similar Items
-
Tighnari: Multi-modal Plant Species Prediction Based on Hierarchical Cross-Attention Using Graph-Based and Vision Backbone-Extracted Features
by: Liu, Haixu, et al.
Published: (2025) -
Weak to Strong: VLM-Based Pseudo-Labeling as a Weakly Supervised Training Strategy in Multimodal Video-based Hidden Emotion Understanding Tasks
by: Wang, Yufei, et al.
Published: (2026) -
Multi-Modal Video Feature Extraction for Popularity Prediction
by: Liu, Haixu, et al.
Published: (2025) -
nnY-Net: Swin-NeXt with Cross-Attention for 3D Medical Images Segmentation
by: Liu, Haixu, et al.
Published: (2025) -
Regression is all you need for medical image translation
by: Rassmann, Sebastian, et al.
Published: (2025)