HOMIE: Histopathology Omni-modal Embedding for Pathology Composed Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Qifeng, Zhong, Wenliang, Dang, Thao M., Ma, Hehuan, Na, Saiyang, Guo, Yuzhi, Huang, Junzhou |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Segment Any Cell: A SAM-based Auto-prompting Fine-tuning Framework for Nuclei Segmentation
by: Na, Saiyang, et al.
Published: (2024)
by: Na, Saiyang, et al.
Published: (2024)
Compositional Image Retrieval via Instruction-Aware Contrastive Learning
by: Zhong, Wenliang, et al.
Published: (2024)
by: Zhong, Wenliang, et al.
Published: (2024)
PathM3: A Multimodal Multi-Task Multiple Instance Learning Framework for Whole Slide Image Classification and Captioning
by: Zhou, Qifeng, et al.
Published: (2024)
by: Zhou, Qifeng, et al.
Published: (2024)
Text-Guided Multi-Instance Learning for Scoliosis Screening via Gait Video Analysis
by: Li, Haiqing, et al.
Published: (2025)
by: Li, Haiqing, et al.
Published: (2025)
Leveraging Gait Patterns as Biomarkers: An attention-guided Deep Multiple Instance Learning Network for Scoliosis Classification
by: Li, Haiqing, et al.
Published: (2025)
by: Li, Haiqing, et al.
Published: (2025)
SAIL-Embedding Technical Report: Omni-modal Embedding Foundation Model
by: Lin, Lin, et al.
Published: (2025)
by: Lin, Lin, et al.
Published: (2025)
InteractiveOmni: A Unified Omni-modal Model for Audio-Visual Multi-turn Dialogue
by: Tong, Wenwen, et al.
Published: (2025)
by: Tong, Wenwen, et al.
Published: (2025)
Composed Multi-modal Retrieval: A Survey of Approaches and Applications
by: Zhang, Kun, et al.
Published: (2025)
by: Zhang, Kun, et al.
Published: (2025)
Composed Object Retrieval: Object-level Retrieval via Composed Expressions
by: Wang, Tong, et al.
Published: (2025)
by: Wang, Tong, et al.
Published: (2025)
Composed Video Retrieval via Enriched Context and Discriminative Embeddings
by: Thawakar, Omkar, et al.
Published: (2024)
by: Thawakar, Omkar, et al.
Published: (2024)
Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction
by: He, Chaoqun, et al.
Published: (2026)
by: He, Chaoqun, et al.
Published: (2026)
From Mapping to Composing: A Two-Stage Framework for Zero-shot Composed Image Retrieval
by: Wang, Yabing, et al.
Published: (2025)
by: Wang, Yabing, et al.
Published: (2025)
scpFormer: A Foundation Model for Unified Representation and Integration of the Single-Cell Proteomics
by: Zhou, Qifeng, et al.
Published: (2026)
by: Zhou, Qifeng, et al.
Published: (2026)
e5-omni: Explicit Cross-modal Alignment for Omni-modal Embeddings
by: Chen, Haonan, et al.
Published: (2026)
by: Chen, Haonan, et al.
Published: (2026)
Sandboxed Coding Agents are Competitive Omni-modal Task Solvers
by: Chen, Dongping, et al.
Published: (2026)
by: Chen, Dongping, et al.
Published: (2026)
Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing
by: Tian, Zeyue, et al.
Published: (2026)
by: Tian, Zeyue, et al.
Published: (2026)
ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding
by: Guan, Yiran, et al.
Published: (2026)
by: Guan, Yiran, et al.
Published: (2026)
Enhancing Multimodal Large Language Models with Multi-instance Visual Prompt Generator for Visual Representation Enrichment
by: Zhong, Wenliang, et al.
Published: (2024)
by: Zhong, Wenliang, et al.
Published: (2024)
Imagine and Seek: Improving Composed Image Retrieval with an Imagined Proxy
by: Li, You, et al.
Published: (2024)
by: Li, You, et al.
Published: (2024)
RoboOmni: Proactive Robot Manipulation in Omni-modal Context
by: Wang, Siyin, et al.
Published: (2025)
by: Wang, Siyin, et al.
Published: (2025)
MCoT-MVS: Multi-level Vision Selection by Multi-modal Chain-of-Thought Reasoning for Composed Image Retrieval
by: Ge, Xuri, et al.
Published: (2026)
by: Ge, Xuri, et al.
Published: (2026)
OmniSelect: Dynamic Modality-Aware Token Compression for Efficient Omni-modal Large Language Models
by: Yang, Morunliu, et al.
Published: (2026)
by: Yang, Morunliu, et al.
Published: (2026)
Instance-Level Composed Image Retrieval
by: Psomas, Bill, et al.
Published: (2025)
by: Psomas, Bill, et al.
Published: (2025)
Zero Shot Composed Image Retrieval
by: Kakarla, Santhosh, et al.
Published: (2025)
by: Kakarla, Santhosh, et al.
Published: (2025)
Composed Image Retrieval for Remote Sensing
by: Psomas, Bill, et al.
Published: (2024)
by: Psomas, Bill, et al.
Published: (2024)
Knowledge-Enhanced Dual-stream Zero-shot Composed Image Retrieval
by: Suo, Yucheng, et al.
Published: (2024)
by: Suo, Yucheng, et al.
Published: (2024)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
OmniBooth: Learning Latent Control for Image Synthesis with Multi-modal Instruction
by: Li, Leheng, et al.
Published: (2024)
by: Li, Leheng, et al.
Published: (2024)
THIR: Topological Histopathological Image Retrieval
by: Tabatabaei, Zahra, et al.
Published: (2025)
by: Tabatabaei, Zahra, et al.
Published: (2025)
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation
by: Liao, Chao, et al.
Published: (2025)
by: Liao, Chao, et al.
Published: (2025)
OmniRetriever: Any-to-Any Audio-Video-Text Retrieval via Fusion-as-Teacher Distillation
by: Liu, Yunze, et al.
Published: (2026)
by: Liu, Yunze, et al.
Published: (2026)
Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval
by: Yang, Yuxin, et al.
Published: (2026)
by: Yang, Yuxin, et al.
Published: (2026)
ReCALL: Recalibrating Capability Degradation for MLLM-based Composed Image Retrieval
by: Yang, Tianyu, et al.
Published: (2026)
by: Yang, Tianyu, et al.
Published: (2026)
INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval
by: Chen, Zhiwei, et al.
Published: (2026)
by: Chen, Zhiwei, et al.
Published: (2026)
DELST: Dual Entailment Learning for Hyperbolic Image-Gene Pretraining in Spatial Transcriptomics
by: Chen, Xulin, et al.
Published: (2025)
by: Chen, Xulin, et al.
Published: (2025)
A Sanity Check on Composed Image Retrieval
by: Liu, Yikun, et al.
Published: (2026)
by: Liu, Yikun, et al.
Published: (2026)
Zero-shot Composed Text-Image Retrieval
by: Liu, Yikun, et al.
Published: (2023)
by: Liu, Yikun, et al.
Published: (2023)
ViT-Lens: Towards Omni-modal Representations
by: Lei, Weixian, et al.
Published: (2023)
by: Lei, Weixian, et al.
Published: (2023)
Feature Re-Embedding: Towards Foundation Model-Level Performance in Computational Pathology
by: Tang, Wenhao, et al.
Published: (2024)
by: Tang, Wenhao, et al.
Published: (2024)
A Comprehensive Benchmark of Histopathology Foundation Models for Kidney Digital Pathology Images
by: Kasireddy, Harishwar Reddy, et al.
Published: (2026)
by: Kasireddy, Harishwar Reddy, et al.
Published: (2026)
Similar Items
-
Segment Any Cell: A SAM-based Auto-prompting Fine-tuning Framework for Nuclei Segmentation
by: Na, Saiyang, et al.
Published: (2024) -
Compositional Image Retrieval via Instruction-Aware Contrastive Learning
by: Zhong, Wenliang, et al.
Published: (2024) -
PathM3: A Multimodal Multi-Task Multiple Instance Learning Framework for Whole Slide Image Classification and Captioning
by: Zhou, Qifeng, et al.
Published: (2024) -
Text-Guided Multi-Instance Learning for Scoliosis Screening via Gait Video Analysis
by: Li, Haiqing, et al.
Published: (2025) -
Leveraging Gait Patterns as Biomarkers: An attention-guided Deep Multiple Instance Learning Network for Scoliosis Classification
by: Li, Haiqing, et al.
Published: (2025)