Spatial Transcriptomics as Images for Large-Scale Pretraining
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Yishun, Qi, Jiaxin, Wang, Jian, Zheng, Yuhua, Huang, Jianqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DELST: Dual Entailment Learning for Hyperbolic Image-Gene Pretraining in Spatial Transcriptomics
von: Chen, Xulin, et al.
Veröffentlicht: (2025)
von: Chen, Xulin, et al.
Veröffentlicht: (2025)
Learning from Gene Names, Expression Values and Images: Contrastive Masked Text-Image Pretraining for Spatial Transcriptomics Representation Learning
von: Qian, Jiahe, et al.
Veröffentlicht: (2025)
von: Qian, Jiahe, et al.
Veröffentlicht: (2025)
Long-CODE: Isolating Pure Long-Context as an Orthogonal Dimension in Video Evaluation
von: Tang, Zhijiang, et al.
Veröffentlicht: (2026)
von: Tang, Zhijiang, et al.
Veröffentlicht: (2026)
CCCaption: Dual-Reward Reinforcement Learning for Complete and Correct Image Captioning
von: Tang, Zhijiang, et al.
Veröffentlicht: (2026)
von: Tang, Zhijiang, et al.
Veröffentlicht: (2026)
Scaling Test-Time Robustness of Vision-Language Models via Self-Critical Inference Framework
von: Tang, Kaihua, et al.
Veröffentlicht: (2026)
von: Tang, Kaihua, et al.
Veröffentlicht: (2026)
Second-Order Multi-Level Variance Correction for Modality Competition in Multimodal Models
von: Lu, Yishun, et al.
Veröffentlicht: (2026)
von: Lu, Yishun, et al.
Veröffentlicht: (2026)
D-CoDe: Scaling Image-Pretrained VLMs to Video via Dynamic Compression and Question Decomposition
von: Huang, Yiyang, et al.
Veröffentlicht: (2025)
von: Huang, Yiyang, et al.
Veröffentlicht: (2025)
Large-Scale 3D Medical Image Pre-training with Geometric Context Priors
von: Wu, Linshan, et al.
Veröffentlicht: (2024)
von: Wu, Linshan, et al.
Veröffentlicht: (2024)
M2-Encoder: Advancing Bilingual Image-Text Understanding by Large-scale Efficient Pretraining
von: Guo, Qingpei, et al.
Veröffentlicht: (2024)
von: Guo, Qingpei, et al.
Veröffentlicht: (2024)
Offline Semantic Guidance for Efficient Vision-Language-Action Policy Distillation
von: Shi, Jin, et al.
Veröffentlicht: (2026)
von: Shi, Jin, et al.
Veröffentlicht: (2026)
Towards Understanding Deep Learning Model in Image Recognition via Coverage Test
von: Li, Wenkai, et al.
Veröffentlicht: (2025)
von: Li, Wenkai, et al.
Veröffentlicht: (2025)
M2OST: Many-to-one Regression for Predicting Spatial Transcriptomics from Digital Pathology Images
von: Wang, Hongyi, et al.
Veröffentlicht: (2024)
von: Wang, Hongyi, et al.
Veröffentlicht: (2024)
Sparser2Sparse: Single-shot Sparser-to-Sparse Learning for Spatial Transcriptomics Imputation with Natural Image Co-learning
von: Fang, Yaoyu, et al.
Veröffentlicht: (2025)
von: Fang, Yaoyu, et al.
Veröffentlicht: (2025)
Towards Spatial Transcriptomics-driven Pathology Foundation Models
von: Hemker, Konstantin, et al.
Veröffentlicht: (2026)
von: Hemker, Konstantin, et al.
Veröffentlicht: (2026)
FEAST: Fully Connected Expressive Attention for Spatial Transcriptomics
von: Jeong, Taejin, et al.
Veröffentlicht: (2026)
von: Jeong, Taejin, et al.
Veröffentlicht: (2026)
Text-to-Image GAN with Pretrained Representations
von: You, Xiaozhou, et al.
Veröffentlicht: (2024)
von: You, Xiaozhou, et al.
Veröffentlicht: (2024)
ScaleNet: Scaling up Pretrained Neural Networks with Incremental Parameters
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025)
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025)
SASP: Strip-Aware Spatial Perception for Fine-Grained Bird Image Classification
von: Wang, Zheng
Veröffentlicht: (2025)
von: Wang, Zheng
Veröffentlicht: (2025)
High-Resolution Spatial Transcriptomics from Histology Images using HisToSGE
von: Shi, Zhiceng, et al.
Veröffentlicht: (2024)
von: Shi, Zhiceng, et al.
Veröffentlicht: (2024)
Multi-Modal Feature Fusion for Spatial Morphology Analysis of Traditional Villages via Hierarchical Graph Neural Networks
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2025)
Unposed Sparse Views Room Layout Reconstruction in the Age of Pretrain Model
von: Huang, Yaxuan, et al.
Veröffentlicht: (2025)
von: Huang, Yaxuan, et al.
Veröffentlicht: (2025)
Scale-Aware Relay and Scale-Adaptive Loss for Tiny Object Detection in Aerial Images
von: Li, Jinfu, et al.
Veröffentlicht: (2025)
von: Li, Jinfu, et al.
Veröffentlicht: (2025)
SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images
von: Liu, Zishan, et al.
Veröffentlicht: (2026)
von: Liu, Zishan, et al.
Veröffentlicht: (2026)
UniCMs: A Unified Consistency Model For Efficient Multimodal Generation and Understanding
von: Xu, Chenkai, et al.
Veröffentlicht: (2025)
von: Xu, Chenkai, et al.
Veröffentlicht: (2025)
MLIP: Efficient Multi-Perspective Language-Image Pretraining with Exhaustive Data Utilization
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
Spatial Transcriptomics Expression Prediction from Histopathology Based on Cross-Modal Mask Reconstruction and Contrastive Learning
von: Liu, Junzhuo, et al.
Veröffentlicht: (2025)
von: Liu, Junzhuo, et al.
Veröffentlicht: (2025)
MIT-10M: A Large Scale Parallel Corpus of Multilingual Image Translation
von: Li, Bo, et al.
Veröffentlicht: (2024)
von: Li, Bo, et al.
Veröffentlicht: (2024)
Multi-Slice Spatial Transcriptomics Data Integration Analysis with STG3Net
von: Fang, Donghai, et al.
Veröffentlicht: (2024)
von: Fang, Donghai, et al.
Veröffentlicht: (2024)
UniEmoX: Cross-modal Semantic-Guided Large-Scale Pretraining for Universal Scene Emotion Perception
von: Chen, Chuang, et al.
Veröffentlicht: (2024)
von: Chen, Chuang, et al.
Veröffentlicht: (2024)
C3-Diff: Super-resolving Spatial Transcriptomics via Cross-modal Cross-content Contrastive Diffusion Modelling
von: Wang, Xiaofei, et al.
Veröffentlicht: (2025)
von: Wang, Xiaofei, et al.
Veröffentlicht: (2025)
TokenUnify: Scaling Up Autoregressive Pretraining for Neuron Segmentation
von: Chen, Yinda, et al.
Veröffentlicht: (2024)
von: Chen, Yinda, et al.
Veröffentlicht: (2024)
DenseFormer: Learning Dense Depth Map from Sparse Depth and Image via Conditional Diffusion Model
von: Yuan, Ming, et al.
Veröffentlicht: (2025)
von: Yuan, Ming, et al.
Veröffentlicht: (2025)
DST-Net: A Dual-Stream Transformer with Illumination-Independent Feature Guidance and Multi-Scale Spatial Convolution for Low-Light Image Enhancement
von: Shi, Yicui, et al.
Veröffentlicht: (2026)
von: Shi, Yicui, et al.
Veröffentlicht: (2026)
SpatialReward: Verifiable Spatial Reward Modeling for Fine-Grained Spatial Consistency in Text-to-Image Generation
von: Zhou, Sashuai, et al.
Veröffentlicht: (2026)
von: Zhou, Sashuai, et al.
Veröffentlicht: (2026)
Object-centric Binding in Contrastive Language-Image Pretraining
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources
von: Ma, Baorui, et al.
Veröffentlicht: (2026)
von: Ma, Baorui, et al.
Veröffentlicht: (2026)
LLMTrack: Semantic Multi-Object Tracking with Multi-modal Large Language Models
von: Liao, Pan, et al.
Veröffentlicht: (2026)
von: Liao, Pan, et al.
Veröffentlicht: (2026)
Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models
von: Li, You, et al.
Veröffentlicht: (2025)
von: Li, You, et al.
Veröffentlicht: (2025)
3D MRI Image Pretraining via Controllable 2D Slice Navigation Task
von: Wang, Yu, et al.
Veröffentlicht: (2026)
von: Wang, Yu, et al.
Veröffentlicht: (2026)
LLaVA-VSD: Large Language-and-Vision Assistant for Visual Spatial Description
von: Jin, Yizhang, et al.
Veröffentlicht: (2024)
von: Jin, Yizhang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DELST: Dual Entailment Learning for Hyperbolic Image-Gene Pretraining in Spatial Transcriptomics
von: Chen, Xulin, et al.
Veröffentlicht: (2025) -
Learning from Gene Names, Expression Values and Images: Contrastive Masked Text-Image Pretraining for Spatial Transcriptomics Representation Learning
von: Qian, Jiahe, et al.
Veröffentlicht: (2025) -
Long-CODE: Isolating Pure Long-Context as an Orthogonal Dimension in Video Evaluation
von: Tang, Zhijiang, et al.
Veröffentlicht: (2026) -
CCCaption: Dual-Reward Reinforcement Learning for Complete and Correct Image Captioning
von: Tang, Zhijiang, et al.
Veröffentlicht: (2026) -
Scaling Test-Time Robustness of Vision-Language Models via Self-Critical Inference Framework
von: Tang, Kaihua, et al.
Veröffentlicht: (2026)