Progressive Multimodal Interaction Network for Reliable Quantification of Fish Feeding Intensity in Aquaculture
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Shulong, Yao, Mingyuan, Zhao, Jiayin, Li, Daoliang, Chen, Yingyi, Wang, Haihua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Comprehensive Review of Fish Feeding Behavior Analysis in Aquaculture: Tasks, Techniques, and Applications
by: Zhang, Shulong, et al.
Published: (2025)
by: Zhang, Shulong, et al.
Published: (2025)
Prompt to Protection: A Comparative Study of Multimodal LLMs in Construction Hazard Recognition
by: Chaudhary, Nishi, et al.
Published: (2025)
by: Chaudhary, Nishi, et al.
Published: (2025)
Fourier-based Action Recognition for Wildlife Behavior Quantification with Event Cameras
by: Hamann, Friedhelm, et al.
Published: (2024)
by: Hamann, Friedhelm, et al.
Published: (2024)
V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?
by: Jongwiriyanurak, Natchapon, et al.
Published: (2024)
by: Jongwiriyanurak, Natchapon, et al.
Published: (2024)
FollowGen: A Scaled Noise Conditional Diffusion Model for Car-Following Trajectory Prediction
by: You, Junwei, et al.
Published: (2024)
by: You, Junwei, et al.
Published: (2024)
PhytoSynth: Leveraging Multi-modal Generative Models for Crop Disease Data Generation with Novel Benchmarking and Prompt Engineering Approach
by: Rai, Nitin, et al.
Published: (2025)
by: Rai, Nitin, et al.
Published: (2025)
Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems
by: Kyem, Blessing Agyei, et al.
Published: (2026)
by: Kyem, Blessing Agyei, et al.
Published: (2026)
Improving Object Detector Training on Synthetic Data by Starting With a Strong Baseline Methodology
by: Ruis, Frank A., et al.
Published: (2024)
by: Ruis, Frank A., et al.
Published: (2024)
Detecting Multiple Diseases in Multiple Crops Using Deep Learning
by: Yadav, Vivek, et al.
Published: (2025)
by: Yadav, Vivek, et al.
Published: (2025)
From Images to Insights: Explainable Biodiversity Monitoring with Plain Language Habitat Explanations
by: Zhou, Yutong, et al.
Published: (2025)
by: Zhou, Yutong, et al.
Published: (2025)
Improving watermelon (Citrullus lanatus) disease classification with generative artificial intelligence (GenAI)-based synthetic and real-field images via a custom EfficientNetV2-L model
by: Rai, Nitin, et al.
Published: (2025)
by: Rai, Nitin, et al.
Published: (2025)
Labits: Layered Bidirectional Time Surfaces Representation for Event Camera-based Continuous Dense Trajectory Estimation
by: Zhang, Zhongyang, et al.
Published: (2024)
by: Zhang, Zhongyang, et al.
Published: (2024)
Automated Facility Enumeration for Building Compliance Checking using Door Detection and Large Language Models
by: Zhang, Licheng, et al.
Published: (2025)
by: Zhang, Licheng, et al.
Published: (2025)
Human-in-the-Loop: Quantitative Evaluation of 3D Models Generation by Large Language Models
by: Sadik, Ahmed R., et al.
Published: (2025)
by: Sadik, Ahmed R., et al.
Published: (2025)
QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing
by: Mittal, Garvit Kumar, et al.
Published: (2026)
by: Mittal, Garvit Kumar, et al.
Published: (2026)
DoorDet: Semi-Automated Multi-Class Door Detection Dataset via Object Detection and Large Language Models
by: Zhang, Licheng, et al.
Published: (2025)
by: Zhang, Licheng, et al.
Published: (2025)
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
by: Goodge, Adam, et al.
Published: (2025)
by: Goodge, Adam, et al.
Published: (2025)
SAM-SP: Self-Prompting Makes SAM Great Again
by: Zhou, Chunpeng, et al.
Published: (2024)
by: Zhou, Chunpeng, et al.
Published: (2024)
Toward Accountable AI-Generated Content on Social Platforms: Steganographic Attribution and Multimodal Harm Detection
by: Guan, Xinlei, et al.
Published: (2026)
by: Guan, Xinlei, et al.
Published: (2026)
All-Optical Segmentation via Diffractive Neural Networks for Autonomous Driving
by: Li, Yingjie, et al.
Published: (2026)
by: Li, Yingjie, et al.
Published: (2026)
Can Foundation Models Revolutionize Mobile AR Sparse Sensing?
by: Zhao, Yiqin, et al.
Published: (2025)
by: Zhao, Yiqin, et al.
Published: (2025)
Human Cognition in Machines: A Unified Perspective of World Models
by: Rupprecht, Timothy, et al.
Published: (2026)
by: Rupprecht, Timothy, et al.
Published: (2026)
Fast Quantum Convolutional Neural Networks for Low-Complexity Object Detection in Autonomous Driving Applications
by: Baek, Hankyul, et al.
Published: (2023)
by: Baek, Hankyul, et al.
Published: (2023)
Learned Display Radiance Fields with Lensless Cameras
by: Chen, Ziyang, et al.
Published: (2025)
by: Chen, Ziyang, et al.
Published: (2025)
DeepSeek-Inspired Exploration of RL-based LLMs and Synergy with Wireless Networks: A Survey
by: Qiao, Yu, et al.
Published: (2025)
by: Qiao, Yu, et al.
Published: (2025)
SpecTrack: Learned Multi-Rotation Tracking via Speckle Imaging
by: Chen, Ziyang, et al.
Published: (2024)
by: Chen, Ziyang, et al.
Published: (2024)
MM-Telco: Benchmarks and Multimodal Large Language Models for Telecom Applications
by: Kumar, Anshul, et al.
Published: (2025)
by: Kumar, Anshul, et al.
Published: (2025)
From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage
by: Ruan, Cihan, et al.
Published: (2026)
by: Ruan, Cihan, et al.
Published: (2026)
MicroCrackAttentionNeXt: Advancing Microcrack Detection in Wave Field Analysis Using Deep Neural Networks through Feature Visualization
by: Moreh, Fatahlla, et al.
Published: (2024)
by: Moreh, Fatahlla, et al.
Published: (2024)
Enhancing Autism Spectrum Disorder Early Detection with the Parent-Child Dyads Block-Play Protocol and an Attention-enhanced GCN-xLSTM Hybrid Deep Learning Framework
by: Li, Xiang, et al.
Published: (2024)
by: Li, Xiang, et al.
Published: (2024)
VOLMO: Versatile and Open Large Models for Ophthalmology
by: Qin, Zhenyue, et al.
Published: (2026)
by: Qin, Zhenyue, et al.
Published: (2026)
Semi-Supervised Multimodal Multi-Instance Learning for Aortic Stenosis Diagnosis
by: Huang, Zhe, et al.
Published: (2024)
by: Huang, Zhe, et al.
Published: (2024)
Multi-Image Super Resolution Framework for Detection and Analysis of Plant Roots
by: Agarwal, Shubham, et al.
Published: (2026)
by: Agarwal, Shubham, et al.
Published: (2026)
Evaluating and Enhancing Trustworthiness of LLMs in Perception Tasks
by: Dona, Malsha Ashani Mahawatta, et al.
Published: (2024)
by: Dona, Malsha Ashani Mahawatta, et al.
Published: (2024)
Introducing Nylon Face Mask Attacks: A Dataset for Evaluating Generalised Face Presentation Attack Detection
by: Manasa, et al.
Published: (2025)
by: Manasa, et al.
Published: (2025)
SynSpill: Improved Industrial Spill Detection With Synthetic Data
by: Baranwal, Aaditya, et al.
Published: (2025)
by: Baranwal, Aaditya, et al.
Published: (2025)
Scrutinizing Data from Sky: An Examination of Its Veracity in Area Based Traffic Contexts
by: Ali, Yawar, et al.
Published: (2024)
by: Ali, Yawar, et al.
Published: (2024)
x-RAGE: eXtended Reality -- Action & Gesture Events Dataset
by: Parmar, Vivek, et al.
Published: (2024)
by: Parmar, Vivek, et al.
Published: (2024)
Towards Railway Domain Adaptation for LiDAR-based 3D Detection: Road-to-Rail and Sim-to-Real via SynDRA-BBox
by: Diaz, Xavier, et al.
Published: (2025)
by: Diaz, Xavier, et al.
Published: (2025)
DashCam Video: A complementary low-cost data stream for on-demand forest-infrastructure system monitoring
by: Joshi, Durga, et al.
Published: (2025)
by: Joshi, Durga, et al.
Published: (2025)
Similar Items
-
A Comprehensive Review of Fish Feeding Behavior Analysis in Aquaculture: Tasks, Techniques, and Applications
by: Zhang, Shulong, et al.
Published: (2025) -
Prompt to Protection: A Comparative Study of Multimodal LLMs in Construction Hazard Recognition
by: Chaudhary, Nishi, et al.
Published: (2025) -
Fourier-based Action Recognition for Wildlife Behavior Quantification with Event Cameras
by: Hamann, Friedhelm, et al.
Published: (2024) -
V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?
by: Jongwiriyanurak, Natchapon, et al.
Published: (2024) -
FollowGen: A Scaled Noise Conditional Diffusion Model for Car-Following Trajectory Prediction
by: You, Junwei, et al.
Published: (2024)