Saved in:
| Main Authors: | Wang, Wei, Ly, Quoc-Toan, Yu, Chong, Bai, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.13331 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SIGMMA: Hierarchical Graph-Based Multi-Scale Multi-modal Contrastive Alignment of Histopathology Image and Spatial Transcriptome
by: Jeong, Dabin, et al.
Published: (2025)
by: Jeong, Dabin, et al.
Published: (2025)
MEATRD: Multimodal Anomalous Tissue Region Detection Enhanced with Spatial Transcriptomics
by: Xu, Kaichen, et al.
Published: (2024)
by: Xu, Kaichen, et al.
Published: (2024)
A Patch-based Cross-view Regularized Framework for Backdoor Defense in Multimodal Large Language Models
by: Fang, Tianmeng, et al.
Published: (2026)
by: Fang, Tianmeng, et al.
Published: (2026)
SENCA-st: Integrating Spatial Transcriptomics and Histopathology with Cross Attention Shared Encoder for Region Identification in Cancer Pathology
by: Liyanaarachchi, Shanaka, et al.
Published: (2025)
by: Liyanaarachchi, Shanaka, et al.
Published: (2025)
Cross-modal Diffusion Modelling for Super-resolved Spatial Transcriptomics
by: Wang, Xiaofei, et al.
Published: (2024)
by: Wang, Xiaofei, et al.
Published: (2024)
Leveraging Spatial Transcriptomics as Alternative to Manual Annotations for Deep Learning-Based Nuclei Analysis
by: Nishimura, Kazuya, et al.
Published: (2026)
by: Nishimura, Kazuya, et al.
Published: (2026)
ST-Align: A Multimodal Foundation Model for Image-Gene Alignment in Spatial Transcriptomics
by: Lin, Yuxiang, et al.
Published: (2024)
by: Lin, Yuxiang, et al.
Published: (2024)
M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception
by: Udugama, U. V. B. L, et al.
Published: (2025)
by: Udugama, U. V. B. L, et al.
Published: (2025)
Global Context-aware Representation Learning for Spatially Resolved Transcriptomics
by: Oh, Yunhak, et al.
Published: (2025)
by: Oh, Yunhak, et al.
Published: (2025)
Coarse Correspondences Boost Spatial-Temporal Reasoning in Multimodal Language Model
by: Liu, Benlin, et al.
Published: (2024)
by: Liu, Benlin, et al.
Published: (2024)
ST-GDance++: A Scalable Spatial-Temporal Diffusion for Long-Duration Group Choreography
by: Xu, Jing, et al.
Published: (2026)
by: Xu, Jing, et al.
Published: (2026)
Enhancing Parameter-Efficient Fine-Tuning of Vision Transformers through Frequency-Based Adaptation
by: Ly, Son Thai, et al.
Published: (2024)
by: Ly, Son Thai, et al.
Published: (2024)
Multi-Slice Spatial Transcriptomics Data Integration Analysis with STG3Net
by: Fang, Donghai, et al.
Published: (2024)
by: Fang, Donghai, et al.
Published: (2024)
HEXST: Hexagonal Shifted-Window Transformer for Spatial Transcriptomics Gene Expression Prediction
by: Byeon, Keunho, et al.
Published: (2026)
by: Byeon, Keunho, et al.
Published: (2026)
HyperST: Hierarchical Hyperbolic Learning for Spatial Transcriptomics Prediction
by: Zhang, Chen, et al.
Published: (2025)
by: Zhang, Chen, et al.
Published: (2025)
Dual-Model Defense: Safeguarding Diffusion Models from Membership Inference Attacks through Disjoint Data Splitting
by: Tran, Bao Q., et al.
Published: (2024)
by: Tran, Bao Q., et al.
Published: (2024)
GC-GAT: Multimodal Vehicular Trajectory Prediction using Graph Goal Conditioning and Cross-context Attention
by: Gulzar, Mahir, et al.
Published: (2025)
by: Gulzar, Mahir, et al.
Published: (2025)
Feature Alignment Determines Fusion Strategy: A Comparative Study of Cross-Attention and Concatenation in Multimodal Learning
by: Zhou, Zhiqiang, et al.
Published: (2026)
by: Zhou, Zhiqiang, et al.
Published: (2026)
Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models
by: Moon, Saemi, et al.
Published: (2026)
by: Moon, Saemi, et al.
Published: (2026)
$MV_{Hybrid}$: Improving Spatial Transcriptomics Prediction with Hybrid State Space-Vision Transformer Backbone in Pathology Vision Foundation Models
by: Cho, Won June, et al.
Published: (2025)
by: Cho, Won June, et al.
Published: (2025)
CSA-Net: Channel-wise Spatially Autocorrelated Attention Networks
by: Nikzad, Nick, et al.
Published: (2024)
by: Nikzad, Nick, et al.
Published: (2024)
UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation
by: Wang, Jiayun, et al.
Published: (2026)
by: Wang, Jiayun, et al.
Published: (2026)
Multimodal Language Models Cannot Spot Spatial Inconsistencies
by: Khangaonkar, Om, et al.
Published: (2026)
by: Khangaonkar, Om, et al.
Published: (2026)
Imaging-anchored Multiomics in Cardiovascular Disease: Integrating Cardiac Imaging, Bulk, Single-cell, and Spatial Transcriptomics
by: Le, Minh H. N., et al.
Published: (2026)
by: Le, Minh H. N., et al.
Published: (2026)
Attention Frequency Modulation: Training-Free Spectral Modulation of Diffusion Cross-Attention
by: Oh, Seunghun, et al.
Published: (2026)
by: Oh, Seunghun, et al.
Published: (2026)
Attention Deep Model with Multi-Scale Deep Supervision for Person Re-Identification
by: Wu, Di, et al.
Published: (2019)
by: Wu, Di, et al.
Published: (2019)
Origins of Creativity in Attention-Based Diffusion Models
by: Finn, Emma, et al.
Published: (2025)
by: Finn, Emma, et al.
Published: (2025)
Boomda: Balanced Multi-objective Optimization for Multimodal Domain Adaptation
by: Sun, Jun, et al.
Published: (2025)
by: Sun, Jun, et al.
Published: (2025)
Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction?
by: Yan, Xinchen, et al.
Published: (2025)
by: Yan, Xinchen, et al.
Published: (2025)
HaloQuest: A Visual Hallucination Dataset for Advancing Multimodal Reasoning
by: Wang, Zhecan, et al.
Published: (2024)
by: Wang, Zhecan, et al.
Published: (2024)
A Survey of Vision Transformers in Autonomous Driving: Current Trends and Future Directions
by: Lai-Dang, Quoc-Vinh
Published: (2024)
by: Lai-Dang, Quoc-Vinh
Published: (2024)
Concepts or Skills? Rethinking Instruction Selection for Multi-modal Models
by: Bai, Andrew, et al.
Published: (2025)
by: Bai, Andrew, et al.
Published: (2025)
Latent Danger Zone: Distilling Unified Attention for Cross-Architecture Black-box Attacks
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
Gradient-Aligned Calibration for Post-Training Quantization of Diffusion Models
by: Hoang, Dung Anh, et al.
Published: (2026)
by: Hoang, Dung Anh, et al.
Published: (2026)
Privacy Protection in Personalized Diffusion Models via Targeted Cross-Attention Adversarial Attack
by: Xu, Xide, et al.
Published: (2024)
by: Xu, Xide, et al.
Published: (2024)
Introducing Visual Perception Token into Multimodal Large Language Model
by: Yu, Runpeng, et al.
Published: (2025)
by: Yu, Runpeng, et al.
Published: (2025)
Towards Understanding Multimodal Fine-Tuning: Spatial Features
by: Naghashyar, Lachin, et al.
Published: (2026)
by: Naghashyar, Lachin, et al.
Published: (2026)
MolFM-Lite: Multi-Modal Molecular Property Prediction with Conformer Ensemble Attention and Cross-Modal Fusion
by: Shah, Syed Omer, et al.
Published: (2026)
by: Shah, Syed Omer, et al.
Published: (2026)
FSTA-SNN:Frequency-based Spatial-Temporal Attention Module for Spiking Neural Networks
by: Yu, Kairong, et al.
Published: (2024)
by: Yu, Kairong, et al.
Published: (2024)
Dual-Path Knowledge-Augmented Contrastive Alignment Network for Spatially Resolved Transcriptomics
by: Zhang, Wei, et al.
Published: (2025)
by: Zhang, Wei, et al.
Published: (2025)
Similar Items
-
SIGMMA: Hierarchical Graph-Based Multi-Scale Multi-modal Contrastive Alignment of Histopathology Image and Spatial Transcriptome
by: Jeong, Dabin, et al.
Published: (2025) -
MEATRD: Multimodal Anomalous Tissue Region Detection Enhanced with Spatial Transcriptomics
by: Xu, Kaichen, et al.
Published: (2024) -
A Patch-based Cross-view Regularized Framework for Backdoor Defense in Multimodal Large Language Models
by: Fang, Tianmeng, et al.
Published: (2026) -
SENCA-st: Integrating Spatial Transcriptomics and Histopathology with Cross Attention Shared Encoder for Region Identification in Cancer Pathology
by: Liyanaarachchi, Shanaka, et al.
Published: (2025) -
Cross-modal Diffusion Modelling for Super-resolved Spatial Transcriptomics
by: Wang, Xiaofei, et al.
Published: (2024)