Local-to-Global Cross-Modal Attention-Aware Fusion for HSI-X Semantic Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Xuming, Yokoya, Naoto, Gu, Xingfa, Tian, Qingjiu, Bruzzone, Lorenzo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CoMiX: Cross-Modal Fusion with Deformable Convolutions for HSI-X Semantic Segmentation
by: Zhang, Xuming, et al.
Published: (2024)
by: Zhang, Xuming, et al.
Published: (2024)
Segment Anything with Multiple Modalities
by: Xiao, Aoran, et al.
Published: (2024)
by: Xiao, Aoran, et al.
Published: (2024)
Robust Self-Supervised Cross-Modal Super-Resolution against Real-World Misaligned Observations
by: Dong, Xiaoyu, et al.
Published: (2026)
by: Dong, Xiaoyu, et al.
Published: (2026)
Enhancing 3D LiDAR Segmentation by Shaping Dense and Accurate 2D Semantic Predictions
by: Dong, Xiaoyu, et al.
Published: (2026)
by: Dong, Xiaoyu, et al.
Published: (2026)
MM-OVSeg:Multimodal Optical-SAR Fusion for Open-Vocabulary Segmentation in Remote Sensing
by: Wei, Yimin, et al.
Published: (2026)
by: Wei, Yimin, et al.
Published: (2026)
Spectral-Aware Global Fusion for RGB-Thermal Semantic Segmentation
by: Zhang, Ce, et al.
Published: (2025)
by: Zhang, Ce, et al.
Published: (2025)
Generalized Few-Shot Semantic Segmentation in Remote Sensing: Challenge and Benchmark
by: Broni-Bediako, Clifford, et al.
Published: (2024)
by: Broni-Bediako, Clifford, et al.
Published: (2024)
SynRS3D: A Synthetic Dataset for Global 3D Semantic Understanding from Monocular Remote Sensing Imagery
by: Song, Jian, et al.
Published: (2024)
by: Song, Jian, et al.
Published: (2024)
Bidirectional Cross-Attention Fusion of High-Res RGB and Low-Res HSI for Multimodal Automated Waste Sorting
by: Funk, Jonas V., et al.
Published: (2026)
by: Funk, Jonas V., et al.
Published: (2026)
A Survey of Sample-Efficient Deep Learning for Change Detection in Remote Sensing: Tasks, Strategies, and Challenges
by: Ding, Lei, et al.
Published: (2025)
by: Ding, Lei, et al.
Published: (2025)
MAGIC++: Efficient and Resilient Modality-Agnostic Semantic Segmentation via Hierarchical Modality Selection
by: Zheng, Xu, et al.
Published: (2024)
by: Zheng, Xu, et al.
Published: (2024)
MemorySAM: Memorize Modalities and Semantics with Segment Anything Model 2 for Multi-modal Semantic Segmentation
by: Liao, Chenfei, et al.
Published: (2025)
by: Liao, Chenfei, et al.
Published: (2025)
CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation
by: Zhang, Zelin, et al.
Published: (2026)
by: Zhang, Zelin, et al.
Published: (2026)
TUNI: Real-time RGB-T Semantic Segmentation with Unified Multi-Modal Feature Extraction and Cross-Modal Feature Fusion
by: Guo, Xiaodong, et al.
Published: (2025)
by: Guo, Xiaodong, et al.
Published: (2025)
CrossEarth: Geospatial Vision Foundation Model for Domain Generalizable Remote Sensing Semantic Segmentation
by: Gong, Ziyang, et al.
Published: (2024)
by: Gong, Ziyang, et al.
Published: (2024)
A Novel Technique for Robust Training of Deep Networks With Multisource Weak Labeled Remote Sensing Data
by: Perantoni, Gianmarco, et al.
Published: (2025)
by: Perantoni, Gianmarco, et al.
Published: (2025)
A deep multiple instance learning approach based on coarse labels for high-resolution land-cover mapping
by: Perantoni, Gianmarco, et al.
Published: (2025)
by: Perantoni, Gianmarco, et al.
Published: (2025)
Hyperspectral data augmentation with transformer-based diffusion models
by: Ferrari, Mattia, et al.
Published: (2025)
by: Ferrari, Mattia, et al.
Published: (2025)
SWARD: Stochastic Window-Attention-Based Relational Distillation for Cross-Architectural Semantic Segmentation
by: Makineni, Aditya, et al.
Published: (2026)
by: Makineni, Aditya, et al.
Published: (2026)
MambaX: Image Super-Resolution with State Predictive Control
by: Li, Chenyu, et al.
Published: (2025)
by: Li, Chenyu, et al.
Published: (2025)
AffordGrasp: Cross-Modal Diffusion for Affordance-Aware Grasp Synthesis
by: Wu, Xiaofei, et al.
Published: (2026)
by: Wu, Xiaofei, et al.
Published: (2026)
Balanced Diffusion-Guided Fusion for Multimodal Remote Sensing Classification
by: Liu, Hao, et al.
Published: (2025)
by: Liu, Hao, et al.
Published: (2025)
Reducing Unimodal Bias in Multi-Modal Semantic Segmentation with Multi-Scale Functional Entropy Regularization
by: Zheng, Xu, et al.
Published: (2025)
by: Zheng, Xu, et al.
Published: (2025)
Joint Spatio-Temporal Modeling for the Semantic Change Detection in Remote Sensing Images
by: Ding, Lei, et al.
Published: (2022)
by: Ding, Lei, et al.
Published: (2022)
Joint Super-Resolution and Segmentation for 1-m Impervious Surface Area Mapping in China's Yangtze River Economic Belt
by: Deng, Jie, et al.
Published: (2025)
by: Deng, Jie, et al.
Published: (2025)
U3M: Unbiased Multiscale Modal Fusion Model for Multimodal Semantic Segmentation
by: Li, Bingyu, et al.
Published: (2024)
by: Li, Bingyu, et al.
Published: (2024)
SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation
by: Xu, Guoan, et al.
Published: (2024)
by: Xu, Guoan, et al.
Published: (2024)
Cross-Stage Attention Propagation for Efficient Semantic Segmentation
by: Kang, Beoungwoo
Published: (2026)
by: Kang, Beoungwoo
Published: (2026)
A Global-Local Cross-Attention Network for Ultra-high Resolution Remote Sensing Image Semantic Segmentation
by: Yi, Chen, et al.
Published: (2025)
by: Yi, Chen, et al.
Published: (2025)
Geo3DVQA: Evaluating Vision-Language Models for 3D Geospatial Reasoning from Aerial Imagery
by: Tsujimoto, Mai, et al.
Published: (2025)
by: Tsujimoto, Mai, et al.
Published: (2025)
MP-HSIR: A Multi-Prompt Framework for Universal Hyperspectral Image Restoration
by: Wu, Zhehui, et al.
Published: (2025)
by: Wu, Zhehui, et al.
Published: (2025)
Change Detection Between Optical Remote Sensing Imagery and Map Data via Segment Anything Model (SAM)
by: Chen, Hongruixuan, et al.
Published: (2024)
by: Chen, Hongruixuan, et al.
Published: (2024)
Sub-Region-Aware Modality Fusion and Adaptive Prompting for Multi-Modal Brain Tumor Segmentation
by: Alijani, Shadi, et al.
Published: (2026)
by: Alijani, Shadi, et al.
Published: (2026)
Boosting Few-Shot Segmentation via Instance-Aware Data Augmentation and Local Consensus Guided Cross Attention
by: Guo, Li, et al.
Published: (2024)
by: Guo, Li, et al.
Published: (2024)
Part-Aware Open-Vocabulary 3D Affordance Grounding via Prototypical Semantic and Geometric Alignment
by: Gou, Dongqiang, et al.
Published: (2026)
by: Gou, Dongqiang, et al.
Published: (2026)
Cross-Modal Fusion and Attention Mechanism for Weakly Supervised Video Anomaly Detection
by: Ghadiya, Ayush, et al.
Published: (2024)
by: Ghadiya, Ayush, et al.
Published: (2024)
Context-Aware Semantic Segmentation via Stage-Wise Attention
by: Carreaud, Antoine, et al.
Published: (2026)
by: Carreaud, Antoine, et al.
Published: (2026)
Geometry-Aware Cross Modal Alignment for Light Field-LiDAR Semantic Segmentation
by: Luo, Jie, et al.
Published: (2025)
by: Luo, Jie, et al.
Published: (2025)
Bayesian Modelling of Multi-Year Crop Type Classification Using Deep Neural Networks and Hidden Markov Models
by: Perantoni, Gianmarco, et al.
Published: (2025)
by: Perantoni, Gianmarco, et al.
Published: (2025)
A class-driven hierarchical ResNet for classification of multispectral remote sensing images
by: Weikmann, Giulio, et al.
Published: (2025)
by: Weikmann, Giulio, et al.
Published: (2025)
Similar Items
-
CoMiX: Cross-Modal Fusion with Deformable Convolutions for HSI-X Semantic Segmentation
by: Zhang, Xuming, et al.
Published: (2024) -
Segment Anything with Multiple Modalities
by: Xiao, Aoran, et al.
Published: (2024) -
Robust Self-Supervised Cross-Modal Super-Resolution against Real-World Misaligned Observations
by: Dong, Xiaoyu, et al.
Published: (2026) -
Enhancing 3D LiDAR Segmentation by Shaping Dense and Accurate 2D Semantic Predictions
by: Dong, Xiaoyu, et al.
Published: (2026) -
MM-OVSeg:Multimodal Optical-SAR Fusion for Open-Vocabulary Segmentation in Remote Sensing
by: Wei, Yimin, et al.
Published: (2026)