MultiFoodhat: A potential new paradigm for intelligent food quality inspection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Yue, Zhuang, Guohang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GCAM: Gaussian and causal-attention model of food fine-grained recognition
von: Zhuang, Guohang, et al.
Veröffentlicht: (2024)
von: Zhuang, Guohang, et al.
Veröffentlicht: (2024)
STAR: A Benchmark for Astronomical Star Fields Super-Resolution
von: Wu, Kuo-Cheng, et al.
Veröffentlicht: (2025)
von: Wu, Kuo-Cheng, et al.
Veröffentlicht: (2025)
SR$^{2}$-Net: A General Plug-and-Play Model for Spectral Refinement in Hyperspectral Image Super-Resolution
von: He, Ji-Xuan, et al.
Veröffentlicht: (2026)
von: He, Ji-Xuan, et al.
Veröffentlicht: (2026)
Convolution-based Probability Gradient Loss for Semantic Segmentation
von: Shan, Guohang, et al.
Veröffentlicht: (2024)
von: Shan, Guohang, et al.
Veröffentlicht: (2024)
An annotated grain kernel image database for visual quality inspection
von: Fan, Lei, et al.
Veröffentlicht: (2023)
von: Fan, Lei, et al.
Veröffentlicht: (2023)
High Performance Space Debris Tracking in Complex Skylight Backgrounds with a Large-Scale Dataset
von: Zhuang, Guohang, et al.
Veröffentlicht: (2025)
von: Zhuang, Guohang, et al.
Veröffentlicht: (2025)
Multi-spectral Class Center Network for Face Manipulation Detection and Localization
von: Miao, Changtao, et al.
Veröffentlicht: (2023)
von: Miao, Changtao, et al.
Veröffentlicht: (2023)
Evidential Calibrated Uncertainty-Guided Interactive Segmentation paradigm for Ultrasound Images
von: Shang, Jiang, et al.
Veröffentlicht: (2025)
von: Shang, Jiang, et al.
Veröffentlicht: (2025)
ContextMix: A context-aware data augmentation method for industrial visual inspection systems
von: Kim, Hyungmin, et al.
Veröffentlicht: (2024)
von: Kim, Hyungmin, et al.
Veröffentlicht: (2024)
Visual inspection for illicit items in X-ray images using Deep Learning
von: Mademlis, Ioannis, et al.
Veröffentlicht: (2023)
von: Mademlis, Ioannis, et al.
Veröffentlicht: (2023)
Online LiDAR-Camera Extrinsic Parameters Self-checking
von: Wei, Pengjin, et al.
Veröffentlicht: (2022)
von: Wei, Pengjin, et al.
Veröffentlicht: (2022)
Communication-Efficient Multi-Agent 3D Detection via Hybrid Collaboration
von: Hu, Yue, et al.
Veröffentlicht: (2025)
von: Hu, Yue, et al.
Veröffentlicht: (2025)
Aligned with LLM: a new multi-modal training paradigm for encoding fMRI activity in visual cortex
von: Ma, Shuxiao, et al.
Veröffentlicht: (2024)
von: Ma, Shuxiao, et al.
Veröffentlicht: (2024)
Robust infrared small target detection using self-supervised and a contrario paradigms
von: Ciocarlan, Alina, et al.
Veröffentlicht: (2024)
von: Ciocarlan, Alina, et al.
Veröffentlicht: (2024)
Defect Segmentation in OCT scans of ceramic parts for non-destructive inspection using deep learning
von: Laveda-Martínez, Andrés, et al.
Veröffentlicht: (2025)
von: Laveda-Martínez, Andrés, et al.
Veröffentlicht: (2025)
Image class translation: visual inspection of class-specific hypotheticals and classification based on translation distance
von: Bowen, Mikyla K., et al.
Veröffentlicht: (2024)
von: Bowen, Mikyla K., et al.
Veröffentlicht: (2024)
Mastering Collaborative Multi-modal Data Selection: A Focus on Informativeness, Uniqueness, and Representativeness
von: Yu, Qifan, et al.
Veröffentlicht: (2024)
von: Yu, Qifan, et al.
Veröffentlicht: (2024)
Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function
von: Zhuang, Chenyi, et al.
Veröffentlicht: (2024)
von: Zhuang, Chenyi, et al.
Veröffentlicht: (2024)
EarthGPT: A Universal Multi-modal Large Language Model for Multi-sensor Image Comprehension in Remote Sensing Domain
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
Rendering Multi-Human and Multi-Object with 3D Gaussian Splatting
von: Wang, Weiquan, et al.
Veröffentlicht: (2026)
von: Wang, Weiquan, et al.
Veröffentlicht: (2026)
Multi-organ segmentation: a progressive exploration of learning paradigms under scarce annotation
von: Li, Shiman, et al.
Veröffentlicht: (2023)
von: Li, Shiman, et al.
Veröffentlicht: (2023)
A holistic perception system of internal and external monitoring for ground autonomous vehicles: AutoTRUST paradigm
von: Gkillas, Alexandros, et al.
Veröffentlicht: (2025)
von: Gkillas, Alexandros, et al.
Veröffentlicht: (2025)
Pragmatic Communication in Multi-Agent Collaborative Perception
von: Hu, Yue, et al.
Veröffentlicht: (2024)
von: Hu, Yue, et al.
Veröffentlicht: (2024)
Mixture-of-Noises Enhanced Forgery-Aware Predictor for Multi-Face Manipulation Detection and Localization
von: Miao, Changtao, et al.
Veröffentlicht: (2024)
von: Miao, Changtao, et al.
Veröffentlicht: (2024)
Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model
von: Liu, Ting, et al.
Veröffentlicht: (2024)
von: Liu, Ting, et al.
Veröffentlicht: (2024)
A Benchmark and Knowledge-Grounded Framework for Advanced Multimodal Personalization Study
von: Hu, Xia, et al.
Veröffentlicht: (2026)
von: Hu, Xia, et al.
Veröffentlicht: (2026)
MELLA: Bridging Linguistic Capability and Cultural Groundedness for Low-Resource Language MLLMs
von: Gao, Yufei, et al.
Veröffentlicht: (2025)
von: Gao, Yufei, et al.
Veröffentlicht: (2025)
Computer vision tasks for intelligent aerospace missions: An overview
von: Chen, Huilin, et al.
Veröffentlicht: (2024)
von: Chen, Huilin, et al.
Veröffentlicht: (2024)
Hierarchical Context Transformer for Multi-level Semantic Scene Understanding
von: Hao, Luoying, et al.
Veröffentlicht: (2025)
von: Hao, Luoying, et al.
Veröffentlicht: (2025)
EarthMarker: A Visual Prompting Multi-modal Large Language Model for Remote Sensing
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
Conditional Diffusion Model with Anatomical-Dose Dual Constraints for End-to-End Multi-Tumor Dose Prediction
von: Xie, Hui, et al.
Veröffentlicht: (2025)
von: Xie, Hui, et al.
Veröffentlicht: (2025)
APEX: Learning Adaptive Priorities for Multi-Objective Alignment in Vision-Language Generation
von: Chen, Dongliang, et al.
Veröffentlicht: (2026)
von: Chen, Dongliang, et al.
Veröffentlicht: (2026)
FALIP: Visual Prompt as Foveal Attention Boosts CLIP Zero-Shot Performance
von: Zhuang, Jiedong, et al.
Veröffentlicht: (2024)
von: Zhuang, Jiedong, et al.
Veröffentlicht: (2024)
DiffuseST: Unleashing the Capability of the Diffusion Model for Style Transfer
von: Hu, Ying, et al.
Veröffentlicht: (2024)
von: Hu, Ying, et al.
Veröffentlicht: (2024)
Thermo-LIO: A Novel Multi-Sensor Integrated System for Structural Health Monitoring
von: Yang, Chao, et al.
Veröffentlicht: (2026)
von: Yang, Chao, et al.
Veröffentlicht: (2026)
Bias-constrained multimodal intelligence for equitable and reliable clinical AI
von: Li, Cheng, et al.
Veröffentlicht: (2026)
von: Li, Cheng, et al.
Veröffentlicht: (2026)
MUSES: 3D-Controllable Image Generation via Multi-Modal Agent Collaboration
von: Ding, Yanbo, et al.
Veröffentlicht: (2024)
von: Ding, Yanbo, et al.
Veröffentlicht: (2024)
Popeye: A Unified Visual-Language Model for Multi-Source Ship Detection from Remote Sensing Imagery
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
EarthGPT-X: A Spatial MLLM for Multi-level Multi-Source Remote Sensing Imagery Understanding with Visual Prompting
von: Zhang, Wei, et al.
Veröffentlicht: (2025)
von: Zhang, Wei, et al.
Veröffentlicht: (2025)
MUVR: A Multi-Modal Untrimmed Video Retrieval Benchmark with Multi-Level Visual Correspondence
von: Feng, Yue, et al.
Veröffentlicht: (2025)
von: Feng, Yue, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GCAM: Gaussian and causal-attention model of food fine-grained recognition
von: Zhuang, Guohang, et al.
Veröffentlicht: (2024) -
STAR: A Benchmark for Astronomical Star Fields Super-Resolution
von: Wu, Kuo-Cheng, et al.
Veröffentlicht: (2025) -
SR$^{2}$-Net: A General Plug-and-Play Model for Spectral Refinement in Hyperspectral Image Super-Resolution
von: He, Ji-Xuan, et al.
Veröffentlicht: (2026) -
Convolution-based Probability Gradient Loss for Semantic Segmentation
von: Shan, Guohang, et al.
Veröffentlicht: (2024) -
An annotated grain kernel image database for visual quality inspection
von: Fan, Lei, et al.
Veröffentlicht: (2023)