Tighnari: Multi-modal Plant Species Prediction Based on Hierarchical Cross-Attention Using Graph-Based and Vision Backbone-Extracted Features
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Haixu, Jiang, Penghao, Tao, Zerui, Wan, Muyan, Sun, Qiuzhuang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
nnY-Net: Swin-NeXt with Cross-Attention for 3D Medical Images Segmentation
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
Google is all you need: Semi-Supervised Transfer Learning Strategy For Light Multimodal Multi-Task Classification Model
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
Multi-Modal Video Feature Extraction for Popularity Prediction
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
Tighnari v2: Mitigating Label Noise and Distribution Shift in Multimodal Plant Distribution Prediction via Mixture of Experts and Weakly Supervised Learning
von: Liu, Haixu, et al.
Veröffentlicht: (2026)
von: Liu, Haixu, et al.
Veröffentlicht: (2026)
SARIMAX-Based Power Outage Prediction During Extreme Weather Events
von: Ye, Haoran, et al.
Veröffentlicht: (2025)
von: Ye, Haoran, et al.
Veröffentlicht: (2025)
Molecular Odor Prediction Based on Multi-Feature Graph Attention Networks
von: Xie, HongXin, et al.
Veröffentlicht: (2025)
von: Xie, HongXin, et al.
Veröffentlicht: (2025)
Unveiling and Controlling Anomalous Attention Distribution in Transformers
von: Yan, Ruiqing, et al.
Veröffentlicht: (2024)
von: Yan, Ruiqing, et al.
Veröffentlicht: (2024)
Multi-View Hierarchical Graph Neural Network for Sketch-Based 3D Shape Retrieval
von: Cheng, Hang, et al.
Veröffentlicht: (2026)
von: Cheng, Hang, et al.
Veröffentlicht: (2026)
Best Transition Matrix Esitimation or Best Label Noise Robustness Classifier? Two Possible Methods to Enhance the Performance of T-revision
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2023)
von: Dutta, Soumya, et al.
Veröffentlicht: (2023)
Hierarchical Multi-modal Transformer for Cross-modal Long Document Classification
von: Liu, Tengfei, et al.
Veröffentlicht: (2024)
von: Liu, Tengfei, et al.
Veröffentlicht: (2024)
A Graph-Based Approach to Spectrum Demand Prediction Using Hierarchical Attention Networks
von: Alkadamani, Mohamad, et al.
Veröffentlicht: (2026)
von: Alkadamani, Mohamad, et al.
Veröffentlicht: (2026)
Disparity-in-Differences: Extracting Hierarchical Backbones of Weighted Directed Networks
von: Kim, Hyunuk
Veröffentlicht: (2025)
von: Kim, Hyunuk
Veröffentlicht: (2025)
Adapting Frozen Mono-modal Backbones for Multi-modal Registration via Contrast-Agnostic Instance Optimization
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
EfficientViT: Multi-Scale Linear Attention for High-Resolution Dense Prediction
von: Cai, Han, et al.
Veröffentlicht: (2022)
von: Cai, Han, et al.
Veröffentlicht: (2022)
Revisiting the Integration of Convolution and Attention for Vision Backbone
von: Zhu, Lei, et al.
Veröffentlicht: (2024)
von: Zhu, Lei, et al.
Veröffentlicht: (2024)
Insecure Despite Proven Updated: Extracting the Root VCEK Seed on EPYC Milan via a Software-Only Attack
von: Shen, Muyan, et al.
Veröffentlicht: (2026)
von: Shen, Muyan, et al.
Veröffentlicht: (2026)
Cross-modal Medical Image Generation Based on Pyramid Convolutional Attention Network
von: Mao, Fuyou, et al.
Veröffentlicht: (2024)
von: Mao, Fuyou, et al.
Veröffentlicht: (2024)
A Depression Detection Method Based on Multi-Modal Feature Fusion Using Cross-Attention
von: Li, Shengjie, et al.
Veröffentlicht: (2024)
von: Li, Shengjie, et al.
Veröffentlicht: (2024)
Hierarchical Attention Models for Multi-Relational Graphs
von: Iyer, Roshni G., et al.
Veröffentlicht: (2024)
von: Iyer, Roshni G., et al.
Veröffentlicht: (2024)
Multi-Label Plant Species Prediction with Metadata-Enhanced Multi-Head Vision Transformers
von: Herasimchyk, Hanna, et al.
Veröffentlicht: (2025)
von: Herasimchyk, Hanna, et al.
Veröffentlicht: (2025)
Hierarchical Cross-modal Prompt Learning for Vision-Language Models
von: Zheng, Hao, et al.
Veröffentlicht: (2025)
von: Zheng, Hao, et al.
Veröffentlicht: (2025)
SIGMMA: Hierarchical Graph-Based Multi-Scale Multi-modal Contrastive Alignment of Histopathology Image and Spatial Transcriptome
von: Jeong, Dabin, et al.
Veröffentlicht: (2025)
von: Jeong, Dabin, et al.
Veröffentlicht: (2025)
Multi-modal Semantic Understanding with Contrastive Cross-modal Feature Alignment
von: Zhang, Ming, et al.
Veröffentlicht: (2024)
von: Zhang, Ming, et al.
Veröffentlicht: (2024)
Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition
von: Zhao, Ruoyu, et al.
Veröffentlicht: (2025)
von: Zhao, Ruoyu, et al.
Veröffentlicht: (2025)
Stacked Cross-modal Feature Consolidation Attention Networks for Image Captioning
von: Pourkeshavarz, Mozhgan, et al.
Veröffentlicht: (2023)
von: Pourkeshavarz, Mozhgan, et al.
Veröffentlicht: (2023)
Vision KAN: Towards an Attention-Free Backbone for Vision with Kolmogorov-Arnold Networks
von: Yang, Zhuoqin, et al.
Veröffentlicht: (2026)
von: Yang, Zhuoqin, et al.
Veröffentlicht: (2026)
DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models
von: Zhang, Yudong, et al.
Veröffentlicht: (2024)
von: Zhang, Yudong, et al.
Veröffentlicht: (2024)
Optimal Abort Policy for Mission-Critical Systems under Imperfect Condition Monitoring
von: Sun, Qiuzhuang, et al.
Veröffentlicht: (2025)
von: Sun, Qiuzhuang, et al.
Veröffentlicht: (2025)
Multi-Task Semantic Communication With Graph Attention-Based Feature Correlation Extraction
von: Yu, Xi, et al.
Veröffentlicht: (2025)
von: Yu, Xi, et al.
Veröffentlicht: (2025)
Attack Detection in Dynamic Games with Quadratic Measurements
von: Jiang, Muyan, et al.
Veröffentlicht: (2025)
von: Jiang, Muyan, et al.
Veröffentlicht: (2025)
Online Planning of Power Flows for Power Systems Against Bushfires Using Spatial Context
von: Xu, Jianyu, et al.
Veröffentlicht: (2024)
von: Xu, Jianyu, et al.
Veröffentlicht: (2024)
Vision Transformers with Hierarchical Attention
von: Liu, Yun, et al.
Veröffentlicht: (2021)
von: Liu, Yun, et al.
Veröffentlicht: (2021)
FSR-VLN: Fast and Slow Reasoning for Vision-Language Navigation with Hierarchical Multi-modal Scene Graph
von: Zhou, Xiaolin, et al.
Veröffentlicht: (2025)
von: Zhou, Xiaolin, et al.
Veröffentlicht: (2025)
Modeling Multi-modal Cross-interaction for Multi-label Few-shot Image Classification Based on Local Feature Selection
von: Yan, Kun, et al.
Veröffentlicht: (2024)
von: Yan, Kun, et al.
Veröffentlicht: (2024)
Cross-Paradigm Evaluation of Gaze-Based Semantic Object Identification for Intelligent Vehicles
von: Deng, Penghao, et al.
Veröffentlicht: (2026)
von: Deng, Penghao, et al.
Veröffentlicht: (2026)
VisionTS++: Cross-Modal Time Series Foundation Model with Continual Pre-trained Vision Backbones
von: Shen, Lefei, et al.
Veröffentlicht: (2025)
von: Shen, Lefei, et al.
Veröffentlicht: (2025)
CollabOD: Collaborative Multi-Backbone with Cross-scale Vision for UAV Small Object Detection
von: Bai, Xuecheng, et al.
Veröffentlicht: (2026)
von: Bai, Xuecheng, et al.
Veröffentlicht: (2026)
Extracting the Multiscale Causal Backbone of Brain Dynamics
von: D'Acunto, Gabriele, et al.
Veröffentlicht: (2023)
von: D'Acunto, Gabriele, et al.
Veröffentlicht: (2023)
Clinical Multi-modal Fusion with Heterogeneous Graph and Disease Correlation Learning for Multi-Disease Prediction
von: Jiang, Yueheng, et al.
Veröffentlicht: (2025)
von: Jiang, Yueheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
nnY-Net: Swin-NeXt with Cross-Attention for 3D Medical Images Segmentation
von: Liu, Haixu, et al.
Veröffentlicht: (2025) -
Google is all you need: Semi-Supervised Transfer Learning Strategy For Light Multimodal Multi-Task Classification Model
von: Liu, Haixu, et al.
Veröffentlicht: (2025) -
Multi-Modal Video Feature Extraction for Popularity Prediction
von: Liu, Haixu, et al.
Veröffentlicht: (2025) -
Tighnari v2: Mitigating Label Noise and Distribution Shift in Multimodal Plant Distribution Prediction via Mixture of Experts and Weakly Supervised Learning
von: Liu, Haixu, et al.
Veröffentlicht: (2026) -
SARIMAX-Based Power Outage Prediction During Extreme Weather Events
von: Ye, Haoran, et al.
Veröffentlicht: (2025)