ShellfishNet: A Domain-Specific Benchmark for Visual Recognition of Marine Molluscs
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Ziheng, Wang, Yang, Wang, Nan, Wu, Chengliang, Yan, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
InfoSyncNet: Information Synchronization Temporal Convolutional Network for Visual Speech Recognition
by: Xue, Junxiao, et al.
Published: (2025)
by: Xue, Junxiao, et al.
Published: (2025)
VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition
by: Yadav, Tanush, et al.
Published: (2026)
by: Yadav, Tanush, et al.
Published: (2026)
OptiSAR-Net++: A Large-Scale Benchmark and Transformer-Free Framework for Cross-Domain Remote Sensing Visual Grounding
by: Tang, Xiaoyu, et al.
Published: (2026)
by: Tang, Xiaoyu, et al.
Published: (2026)
AVAR-Net: A Lightweight Audio-Visual Anomaly Recognition Framework with a Benchmark Dataset
by: Ali, Amjid, et al.
Published: (2025)
by: Ali, Amjid, et al.
Published: (2025)
Visual Instruction Pretraining for Domain-Specific Foundation Models
by: Li, Yuxuan, et al.
Published: (2025)
by: Li, Yuxuan, et al.
Published: (2025)
Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks
by: Lee, Jusung, et al.
Published: (2024)
by: Lee, Jusung, et al.
Published: (2024)
MFH: Marrying Frequency Domain with Handwritten Mathematical Expression Recognition
by: Yang, Huanxin, et al.
Published: (2025)
by: Yang, Huanxin, et al.
Published: (2025)
MMDG-Bench: A Benchmark for Multimodal Domain Generalization
by: Zhan, Qianshan, et al.
Published: (2026)
by: Zhan, Qianshan, et al.
Published: (2026)
4Seasons: Benchmarking Visual SLAM and Long-Term Localization for Autonomous Driving in Challenging Conditions
by: Wenzel, Patrick, et al.
Published: (2022)
by: Wenzel, Patrick, et al.
Published: (2022)
MolParser: End-to-end Visual Recognition of Molecule Structures in the Wild
by: Fang, Xi, et al.
Published: (2024)
by: Fang, Xi, et al.
Published: (2024)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
by: Zhang, Xiaoman, et al.
Published: (2023)
by: Zhang, Xiaoman, et al.
Published: (2023)
Virtual Classification: Modulating Domain-Specific Knowledge for Multidomain Crowd Counting
by: Guo, Mingyue, et al.
Published: (2024)
by: Guo, Mingyue, et al.
Published: (2024)
A$^{3}$lign-DFER: Pioneering Comprehensive Dynamic Affective Alignment for Dynamic Facial Expression Recognition with CLIP
by: Tao, Zeng, et al.
Published: (2024)
by: Tao, Zeng, et al.
Published: (2024)
OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
Unsupervised Domain Adaptation with Dynamic Clustering and Contrastive Refinement for Gait Recognition
by: Liu, Xiaolei, et al.
Published: (2025)
by: Liu, Xiaolei, et al.
Published: (2025)
BizGenEval: A Systematic Benchmark for Commercial Visual Content Generation
by: Li, Yan, et al.
Published: (2026)
by: Li, Yan, et al.
Published: (2026)
Benchmarking Micro-action Recognition: Dataset, Methods, and Applications
by: Guo, Dan, et al.
Published: (2024)
by: Guo, Dan, et al.
Published: (2024)
U-Bench: A Comprehensive Understanding of U-Net through 100-Variant Benchmarking
by: Tang, Fenghe, et al.
Published: (2025)
by: Tang, Fenghe, et al.
Published: (2025)
MoiréNet: A Compact Dual-Domain Network for Image Demoiréing
by: Guo, Shuwei, et al.
Published: (2025)
by: Guo, Shuwei, et al.
Published: (2025)
From Generalist to Specialist: Adapting Vision Language Models via Task-Specific Visual Instruction Tuning
by: Bai, Yang, et al.
Published: (2024)
by: Bai, Yang, et al.
Published: (2024)
VicKAM: Visual Conceptual Knowledge Guided Action Map for Weakly Supervised Group Activity Recognition
by: Wang, Zhuming, et al.
Published: (2025)
by: Wang, Zhuming, et al.
Published: (2025)
On the Estimation of Image-matching Uncertainty in Visual Place Recognition
by: Zaffar, Mubariz, et al.
Published: (2024)
by: Zaffar, Mubariz, et al.
Published: (2024)
ActionAtlas: A VideoQA Benchmark for Domain-specialized Action Recognition
by: Salehi, Mohammadreza, et al.
Published: (2024)
by: Salehi, Mohammadreza, et al.
Published: (2024)
Cascade-Free Mandarin Visual Speech Recognition via Semantic-Guided Cross-Representation Alignment
by: Yang, Lei, et al.
Published: (2026)
by: Yang, Lei, et al.
Published: (2026)
Region-aware Spatiotemporal Modeling with Collaborative Domain Generalization for Cross-Subject EEG Emotion Recognition
by: Wu, Weiwei, et al.
Published: (2026)
by: Wu, Weiwei, et al.
Published: (2026)
Ghost-dil-NetVLAD: A Lightweight Neural Network for Visual Place Recognition
by: Gong, Qingyuan, et al.
Published: (2021)
by: Gong, Qingyuan, et al.
Published: (2021)
A Survey on Interpretability in Visual Recognition
by: Wan, Qiyang, et al.
Published: (2025)
by: Wan, Qiyang, et al.
Published: (2025)
Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations
by: Huang, Hai, et al.
Published: (2025)
by: Huang, Hai, et al.
Published: (2025)
Chimera: Improving Generalist Model with Domain-Specific Experts
by: Peng, Tianshuo, et al.
Published: (2024)
by: Peng, Tianshuo, et al.
Published: (2024)
Knowledge-Aligned Counterfactual-Enhancement Diffusion Perception for Unsupervised Cross-Domain Visual Emotion Recognition
by: Yin, Wen, et al.
Published: (2025)
by: Yin, Wen, et al.
Published: (2025)
Deep Learning for Cross-Domain Few-Shot Visual Recognition: A Survey
by: Xu, Huali, et al.
Published: (2023)
by: Xu, Huali, et al.
Published: (2023)
OmniEval: A Benchmark for Evaluating Omni-modal Models with Visual, Auditory, and Textual Inputs
by: Zhang, Yiman, et al.
Published: (2025)
by: Zhang, Yiman, et al.
Published: (2025)
Certified L2-Norm Robustness of 3D Point Cloud Recognition in the Frequency Domain
by: Zhou, Liang, et al.
Published: (2025)
by: Zhou, Liang, et al.
Published: (2025)
Camera-LiDAR Cross-modality Gait Recognition
by: Guo, Wenxuan, et al.
Published: (2024)
by: Guo, Wenxuan, et al.
Published: (2024)
ChartAct: A Benchmark for Dynamic Chart Understanding
by: Huang, Muye, et al.
Published: (2026)
by: Huang, Muye, et al.
Published: (2026)
Optimizing Domain-Specific Image Retrieval: A Benchmark of FAISS and Annoy with Fine-Tuned Features
by: Rahman, MD Shaikh, et al.
Published: (2024)
by: Rahman, MD Shaikh, et al.
Published: (2024)
Understanding Dataset Bias in Medical Imaging: A Case Study on Chest X-rays
by: Dack, Ethan, et al.
Published: (2025)
by: Dack, Ethan, et al.
Published: (2025)
Evaluating Few-Shot Pill Recognition Under Visual Domain Shift
by: Chu, W. I., et al.
Published: (2026)
by: Chu, W. I., et al.
Published: (2026)
A Visual Self-attention Mechanism Facial Expression Recognition Network beyond Convnext
by: Nan, Bingyu, et al.
Published: (2025)
by: Nan, Bingyu, et al.
Published: (2025)
EraW-Net: Enhance-Refine-Align W-Net for Scene-Associated Driver Attention Estimation
by: Zhou, Jun, et al.
Published: (2024)
by: Zhou, Jun, et al.
Published: (2024)
Similar Items
-
InfoSyncNet: Information Synchronization Temporal Convolutional Network for Visual Speech Recognition
by: Xue, Junxiao, et al.
Published: (2025) -
VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition
by: Yadav, Tanush, et al.
Published: (2026) -
OptiSAR-Net++: A Large-Scale Benchmark and Transformer-Free Framework for Cross-Domain Remote Sensing Visual Grounding
by: Tang, Xiaoyu, et al.
Published: (2026) -
AVAR-Net: A Lightweight Audio-Visual Anomaly Recognition Framework with a Benchmark Dataset
by: Ali, Amjid, et al.
Published: (2025) -
Visual Instruction Pretraining for Domain-Specific Foundation Models
by: Li, Yuxuan, et al.
Published: (2025)