Cross-modal learning for plankton recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Kareinen, Joona, Immonen, Veikka, Eerola, Tuomas, Haraguchi, Lumi, Lensu, Lasse, Kraft, Kaisa, Suikkanen, Sanna, Kälviäinen, Heikki |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self-Supervised Pretraining for Fine-Grained Plankton Recognition
by: Kareinen, Joona, et al.
Published: (2025)
by: Kareinen, Joona, et al.
Published: (2025)
Open-Set Plankton Recognition
by: Kareinen, Joona, et al.
Published: (2025)
by: Kareinen, Joona, et al.
Published: (2025)
DAPlankton: Benchmark Dataset for Multi-instrument Plankton Recognition via Fine-grained Domain Adaptation
by: Batrakhanov, Daniel, et al.
Published: (2024)
by: Batrakhanov, Daniel, et al.
Published: (2024)
Towards Phytoplankton Parasite Detection Using Autoencoders
by: Bilik, Simon, et al.
Published: (2023)
by: Bilik, Simon, et al.
Published: (2023)
Unsupervised Pelage Pattern Unwrapping for Animal Re-identification
by: Algasov, Aleksandr, et al.
Published: (2025)
by: Algasov, Aleksandr, et al.
Published: (2025)
Understanding the Impact of Training Set Size on Animal Re-identification
by: Algasov, Aleksandr, et al.
Published: (2024)
by: Algasov, Aleksandr, et al.
Published: (2024)
Deep Unsupervised Segmentation of Log Point Clouds
by: Zolotarev, Fedor, et al.
Published: (2025)
by: Zolotarev, Fedor, et al.
Published: (2025)
EALD-MLLM: Emotion Analysis in Long-sequential and De-identity videos with Multi-modal Large Language Model
by: Li, Deng, et al.
Published: (2024)
by: Li, Deng, et al.
Published: (2024)
Insights from Visual Cognition: Understanding Human Action Dynamics with Overall Glance and Refined Gaze Transformer
by: Xing, Bohao, et al.
Published: (2026)
by: Xing, Bohao, et al.
Published: (2026)
AU-TTT: Vision Test-Time Training model for Facial Action Unit Detection
by: Xing, Bohao, et al.
Published: (2025)
by: Xing, Bohao, et al.
Published: (2025)
Towards synthetic generation of realistic wooden logs
by: Zolotarev, Fedor, et al.
Published: (2025)
by: Zolotarev, Fedor, et al.
Published: (2025)
Mineral segmentation using electron microscope images and spectral sampling through multimodal graph neural networks
by: Repka, Samuel, et al.
Published: (2025)
by: Repka, Samuel, et al.
Published: (2025)
DEEMO: De-identity Multimodal Emotion Recognition and Reasoning
by: Li, Deng, et al.
Published: (2025)
by: Li, Deng, et al.
Published: (2025)
Multimodal surface defect detection from wooden logs for sawing optimization
by: Reich, Bořek, et al.
Published: (2025)
by: Reich, Bořek, et al.
Published: (2025)
EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models
by: Xing, Bohao, et al.
Published: (2025)
by: Xing, Bohao, et al.
Published: (2025)
DiffFAS: Face Anti-Spoofing via Generative Diffusion Models
by: Ge, Xinxu, et al.
Published: (2024)
by: Ge, Xinxu, et al.
Published: (2024)
FSBench: A Figure Skating Benchmark for Advancing Artistic Sports Understanding
by: Gao, Rong, et al.
Published: (2025)
by: Gao, Rong, et al.
Published: (2025)
MSF-Mamba: Motion-aware State Fusion Mamba for Efficient Micro-Gesture Recognition
by: Li, Deng, et al.
Published: (2025)
by: Li, Deng, et al.
Published: (2025)
A cross-modal network for facial expression recognition
by: Tian, Chunwei, et al.
Published: (2026)
by: Tian, Chunwei, et al.
Published: (2026)
EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning
by: Xing, Bohao, et al.
Published: (2024)
by: Xing, Bohao, et al.
Published: (2024)
Cross-Domain Image Conversion by CycleDM
by: Shimotsumagari, Sho, et al.
Published: (2024)
by: Shimotsumagari, Sho, et al.
Published: (2024)
Identity-free Artificial Emotional Intelligence via Micro-Gesture Understanding
by: Gao, Rong, et al.
Published: (2024)
by: Gao, Rong, et al.
Published: (2024)
Cross-modal ultra-scale learning with tri-modalities of renal biopsy images for glomerular multi-disease auxiliary diagnosis
by: Long, Kaixing, et al.
Published: (2025)
by: Long, Kaixing, et al.
Published: (2025)
Shape vs. Context: Examining Human--AI Gaps in Ambiguous Japanese Character Recognition
by: Haraguchi, Daichi
Published: (2026)
by: Haraguchi, Daichi
Published: (2026)
Community-aware evaluation and threshold calibration for open-set plankton image recognition
by: Chen, Xi, et al.
Published: (2026)
by: Chen, Xi, et al.
Published: (2026)
A Study on the Matching Rate of Dance Movements Using 2D Skeleton Detection and 3D Pose Estimation: Why Is SEVENTEEN's Performance So Bita-Zoroi (Perfectly Synchronized)?
by: Simojo, Atsushi, et al.
Published: (2025)
by: Simojo, Atsushi, et al.
Published: (2025)
Deeply Dual Supervised learning for melanoma recognition
by: Polma, Rujosh, et al.
Published: (2025)
by: Polma, Rujosh, et al.
Published: (2025)
Multi-modality action recognition based on dual feature shift in vehicle cabin monitoring
by: Lin, Dan, et al.
Published: (2024)
by: Lin, Dan, et al.
Published: (2024)
Cross-modal Prompting for Balanced Incomplete Multi-modal Emotion Recognition
by: He, Wen-Jue, et al.
Published: (2025)
by: He, Wen-Jue, et al.
Published: (2025)
Enhanced Cross-modal 3D Retrieval via Tri-modal Reconstruction
by: Ren, Junlong, et al.
Published: (2025)
by: Ren, Junlong, et al.
Published: (2025)
CarGait: Cross-Attention based Re-ranking for Gait recognition
by: Habib, Gavriel, et al.
Published: (2025)
by: Habib, Gavriel, et al.
Published: (2025)
Impression-CLIP: Contrastive Shape-Impression Embedding for Fonts
by: Kubota, Yugo, et al.
Published: (2024)
by: Kubota, Yugo, et al.
Published: (2024)
Font Impression Estimation in the Wild
by: Kitajima, Kazuki, et al.
Published: (2024)
by: Kitajima, Kazuki, et al.
Published: (2024)
CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation
by: Zhang, Zelin, et al.
Published: (2026)
by: Zhang, Zelin, et al.
Published: (2026)
Multi-model learning by sequential reading of untrimmed videos for action recognition
by: Kamiya, Kodai, et al.
Published: (2024)
by: Kamiya, Kodai, et al.
Published: (2024)
Multi-modal Semantic Understanding with Contrastive Cross-modal Feature Alignment
by: Zhang, Ming, et al.
Published: (2024)
by: Zhang, Ming, et al.
Published: (2024)
SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning
by: Zhang, Runmin, et al.
Published: (2024)
by: Zhang, Runmin, et al.
Published: (2024)
Camera-LiDAR Cross-modality Gait Recognition
by: Guo, Wenxuan, et al.
Published: (2024)
by: Guo, Wenxuan, et al.
Published: (2024)
XoFTR: Cross-modal Feature Matching Transformer
by: Tuzcuoğlu, Önder, et al.
Published: (2024)
by: Tuzcuoğlu, Önder, et al.
Published: (2024)
A self-supervised learning approach to deep filter banks for texture recognition
by: Florindo, Joao B., et al.
Published: (2026)
by: Florindo, Joao B., et al.
Published: (2026)
Similar Items
-
Self-Supervised Pretraining for Fine-Grained Plankton Recognition
by: Kareinen, Joona, et al.
Published: (2025) -
Open-Set Plankton Recognition
by: Kareinen, Joona, et al.
Published: (2025) -
DAPlankton: Benchmark Dataset for Multi-instrument Plankton Recognition via Fine-grained Domain Adaptation
by: Batrakhanov, Daniel, et al.
Published: (2024) -
Towards Phytoplankton Parasite Detection Using Autoencoders
by: Bilik, Simon, et al.
Published: (2023) -
Unsupervised Pelage Pattern Unwrapping for Animal Re-identification
by: Algasov, Aleksandr, et al.
Published: (2025)