GCAM: Gaussian and causal-attention model of food fine-grained recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Zhuang, Guohang, Hu, Yue, Yan, Tianxing, Gao, JiaZhan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Demographic-aware fine-grained visual recognition of pediatric wrist pathologies
by: Ahmed, Ammar, et al.
Published: (2025)
by: Ahmed, Ammar, et al.
Published: (2025)
MultiFoodhat: A potential new paradigm for intelligent food quality inspection
by: Hu, Yue, et al.
Published: (2025)
by: Hu, Yue, et al.
Published: (2025)
Convolution-based Probability Gradient Loss for Semantic Segmentation
by: Shan, Guohang, et al.
Published: (2024)
by: Shan, Guohang, et al.
Published: (2024)
The devil is in the fine-grained details: Evaluating open-vocabulary object detectors for fine-grained understanding
by: Bianchi, Lorenzo, et al.
Published: (2023)
by: Bianchi, Lorenzo, et al.
Published: (2023)
Understanding attention-based encoder-decoder networks: a case study with chess scoresheet recognition
by: Hayashi, Sergio Y., et al.
Published: (2024)
by: Hayashi, Sergio Y., et al.
Published: (2024)
Improving fine-grained understanding in image-text pre-training
by: Bica, Ioana, et al.
Published: (2024)
by: Bica, Ioana, et al.
Published: (2024)
High Performance Space Debris Tracking in Complex Skylight Backgrounds with a Large-Scale Dataset
by: Zhuang, Guohang, et al.
Published: (2025)
by: Zhuang, Guohang, et al.
Published: (2025)
Attributing Data for Sharpness-Aware Minimization
by: Ren, Chenyang, et al.
Published: (2025)
by: Ren, Chenyang, et al.
Published: (2025)
Uncertainty modeling for fine-tuned implicit functions
by: Susmelj, Anna, et al.
Published: (2024)
by: Susmelj, Anna, et al.
Published: (2024)
QUEST: A robust attention formulation using query-modulated spherical attention
by: Govindarajan, Hariprasath, et al.
Published: (2026)
by: Govindarajan, Hariprasath, et al.
Published: (2026)
Intelligent recognition of GPR road hidden defect images based on feature fusion and attention mechanism
by: Lv, Haotian, et al.
Published: (2025)
by: Lv, Haotian, et al.
Published: (2025)
PixelGaussian: Generalizable 3D Gaussian Reconstruction from Arbitrary Views
by: Fei, Xin, et al.
Published: (2024)
by: Fei, Xin, et al.
Published: (2024)
Adaptive receptive field-based spatial-frequency feature reconstruction network for few-shot fine-grained image classification
by: Zhang, Linyue, et al.
Published: (2026)
by: Zhang, Linyue, et al.
Published: (2026)
This changes to that : Combining causal and non-causal explanations to generate disease progression in capsule endoscopy
by: Vats, Anuja, et al.
Published: (2022)
by: Vats, Anuja, et al.
Published: (2022)
Online LiDAR-Camera Extrinsic Parameters Self-checking
by: Wei, Pengjin, et al.
Published: (2022)
by: Wei, Pengjin, et al.
Published: (2022)
Pig behavior dataset and Spatial-temporal perception and enhancement networks based on the attention mechanism for pig behavior recognition
by: Qi, Fangzheng, et al.
Published: (2025)
by: Qi, Fangzheng, et al.
Published: (2025)
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
by: Gu, Difei, et al.
Published: (2025)
by: Gu, Difei, et al.
Published: (2025)
Multi-scale Quaternion CNN and BiGRU with Cross Self-attention Feature Fusion for Fault Diagnosis of Bearing
by: Liu, Huanbai, et al.
Published: (2024)
by: Liu, Huanbai, et al.
Published: (2024)
Classification for everyone : Building geography agnostic models for fairer recognition
by: Jindal, Akshat, et al.
Published: (2023)
by: Jindal, Akshat, et al.
Published: (2023)
Developing emotion recognition for video conference software to support people with autism
by: Franzen, Marc, et al.
Published: (2021)
by: Franzen, Marc, et al.
Published: (2021)
Color histogram equalization and fine-tuning to improve expression recognition of (partially occluded) faces on sign language datasets
by: Nunnari, Fabrizio, et al.
Published: (2025)
by: Nunnari, Fabrizio, et al.
Published: (2025)
Eliminating the Language Bias for Visual Question Answering with fine-grained Causal Intervention
by: Liu, Ying, et al.
Published: (2024)
by: Liu, Ying, et al.
Published: (2024)
Sensing technologies and machine learning methods for emotion recognition in autism: Systematic review
by: Banos, Oresti, et al.
Published: (2024)
by: Banos, Oresti, et al.
Published: (2024)
ProtoGS: Efficient and High-Quality Rendering with 3D Gaussian Prototypes
by: Gao, Zhengqing, et al.
Published: (2025)
by: Gao, Zhengqing, et al.
Published: (2025)
CDAN: Convolutional dense attention-guided network for low-light image enhancement
by: Shakibania, Hossein, et al.
Published: (2023)
by: Shakibania, Hossein, et al.
Published: (2023)
Moving object detection from multi-depth images with an attention-enhanced CNN
by: Shibukawa, Masato, et al.
Published: (2025)
by: Shibukawa, Masato, et al.
Published: (2025)
Vision Transformer attention alignment with human visual perception in aesthetic object evaluation
by: Carrasco, Miguel, et al.
Published: (2025)
by: Carrasco, Miguel, et al.
Published: (2025)
EPBC-YOLOv8: An efficient and accurate improved YOLOv8 underwater detector based on an attention mechanism
by: Jiang, Xing, et al.
Published: (2025)
by: Jiang, Xing, et al.
Published: (2025)
Sign language recognition from skeletal data using graph and recurrent neural networks
by: Mederos, B., et al.
Published: (2025)
by: Mederos, B., et al.
Published: (2025)
Performance of computer vision algorithms for fine-grained classification using crowdsourced insect images
by: Pucci, Rita, et al.
Published: (2024)
by: Pucci, Rita, et al.
Published: (2024)
MELLA: Bridging Linguistic Capability and Cultural Groundedness for Low-Resource Language MLLMs
by: Gao, Yufei, et al.
Published: (2025)
by: Gao, Yufei, et al.
Published: (2025)
Investigating Redundancy in Multimodal Large Language Models with Multiple Vision Encoders
by: Wang, Yizhou, et al.
Published: (2025)
by: Wang, Yizhou, et al.
Published: (2025)
Aligning Vision to Language: Annotation-Free Multimodal Knowledge Graph Construction for Enhanced LLMs Reasoning
by: Liu, Junming, et al.
Published: (2025)
by: Liu, Junming, et al.
Published: (2025)
HCR-Net: A deep learning based script independent handwritten character recognition network
by: Chauhan, Vinod Kumar, et al.
Published: (2021)
by: Chauhan, Vinod Kumar, et al.
Published: (2021)
Towards Black-Box Membership Inference Attack for Diffusion Models
by: Li, Jingwei, et al.
Published: (2024)
by: Li, Jingwei, et al.
Published: (2024)
GaussianAD: Gaussian-Centric End-to-End Autonomous Driving
by: Zheng, Wenzhao, et al.
Published: (2024)
by: Zheng, Wenzhao, et al.
Published: (2024)
Open Your Eyes: Vision Enhances Message Passing Neural Networks in Link Prediction
by: Wei, Yanbin, et al.
Published: (2025)
by: Wei, Yanbin, et al.
Published: (2025)
BLADE: Block-Sparse Attention Meets Step Distillation for Efficient Video Generation
by: Gu, Youping, et al.
Published: (2025)
by: Gu, Youping, et al.
Published: (2025)
Visual RAG: Expanding MLLM visual knowledge without fine-tuning
by: Bonomo, Mirco, et al.
Published: (2025)
by: Bonomo, Mirco, et al.
Published: (2025)
GCA-ResUNet:Image segmentation in medical images using grouped coordinate attention
by: Ding, Jun, et al.
Published: (2025)
by: Ding, Jun, et al.
Published: (2025)
Similar Items
-
Demographic-aware fine-grained visual recognition of pediatric wrist pathologies
by: Ahmed, Ammar, et al.
Published: (2025) -
MultiFoodhat: A potential new paradigm for intelligent food quality inspection
by: Hu, Yue, et al.
Published: (2025) -
Convolution-based Probability Gradient Loss for Semantic Segmentation
by: Shan, Guohang, et al.
Published: (2024) -
The devil is in the fine-grained details: Evaluating open-vocabulary object detectors for fine-grained understanding
by: Bianchi, Lorenzo, et al.
Published: (2023) -
Understanding attention-based encoder-decoder networks: a case study with chess scoresheet recognition
by: Hayashi, Sergio Y., et al.
Published: (2024)