Saved in:
| Main Authors: | Liu, Chengzhi, Huang, Zile, Chen, Zhe, Tang, Feilong, Tian, Yu, Xu, Zhongxing, Luo, Zihong, Zheng, Yalin, Meng, Yanda |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2502.11724 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Incomplete-Modality Alignment for Ophthalmic Disease Grading and Diagnosis via Labeled Optimal Transport
by: Yu, Qinkai, et al.
Published: (2025)
by: Yu, Qinkai, et al.
Published: (2025)
Robust Multimodal Learning for Ophthalmic Disease Grading via Disentangled Representation
by: Wang, Xinkun, et al.
Published: (2025)
by: Wang, Xinkun, et al.
Published: (2025)
Confidence-Aware Self-Distillation for Multimodal Sentiment Analysis with Incomplete Modalities
by: Luo, Yanxi, et al.
Published: (2025)
by: Luo, Yanxi, et al.
Published: (2025)
Parameterized Diffusion Optimization enabled Autoregressive Ordinal Regression for Diabetic Retinopathy Grading
by: Yu, Qinkai, et al.
Published: (2025)
by: Yu, Qinkai, et al.
Published: (2025)
Toward Modality Gap: Vision Prototype Learning for Weakly-supervised Semantic Segmentation with CLIP
by: Xu, Zhongxing, et al.
Published: (2024)
by: Xu, Zhongxing, et al.
Published: (2024)
MTSA-SNN: A Multi-modal Time Series Analysis Model Based on Spiking Neural Network
by: Liu, Chengzhi, et al.
Published: (2024)
by: Liu, Chengzhi, et al.
Published: (2024)
OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
CLIP-DR: Textual Knowledge-Guided Diabetic Retinopathy Grading with Ranking-aware Prompting
by: Yu, Qinkai, et al.
Published: (2024)
by: Yu, Qinkai, et al.
Published: (2024)
TARDis: Time Attenuated Representation Disentanglement for Incomplete Multi-Modal Tumor Segmentation and Classification
by: Wan, Zishuo, et al.
Published: (2025)
by: Wan, Zishuo, et al.
Published: (2025)
PG-SAM: Prior-Guided SAM with Medical for Multi-organ Segmentation
by: Zhong, Yiheng, et al.
Published: (2025)
by: Zhong, Yiheng, et al.
Published: (2025)
Adaptive Disentangled Representation Learning for Incomplete Multi-View Multi-Label Classification
by: Li, Quanjiang, et al.
Published: (2026)
by: Li, Quanjiang, et al.
Published: (2026)
Are Spatial-Temporal Graph Convolution Networks for Human Action Recognition Over-Parameterized?
by: Xie, Jianyang, et al.
Published: (2025)
by: Xie, Jianyang, et al.
Published: (2025)
MC-DBN: A Deep Belief Network-Based Model for Modality Completion
by: Luo, Zihong, et al.
Published: (2024)
by: Luo, Zihong, et al.
Published: (2024)
Generalizing to Unseen Domains in Diabetic Retinopathy with Disentangled Representations
by: Xia, Peng, et al.
Published: (2024)
by: Xia, Peng, et al.
Published: (2024)
OphCLIP: Hierarchical Retrieval-Augmented Learning for Ophthalmic Surgical Video-Language Pretraining
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
Hunting Attributes: Context Prototype-Aware Learning for Weakly Supervised Semantic Segmentation
by: Tang, Feilong, et al.
Published: (2024)
by: Tang, Feilong, et al.
Published: (2024)
Benchmarking Large Multimodal Models for Ophthalmic Visual Question Answering with OphthalWeChat
by: Xu, Pusheng, et al.
Published: (2025)
by: Xu, Pusheng, et al.
Published: (2025)
Towards Dynamic 3D Reconstruction of Hand-Instrument Interaction in Ophthalmic Surgery
by: Hu, Ming, et al.
Published: (2025)
by: Hu, Ming, et al.
Published: (2025)
Better Sampling, towards Better End-to-end Small Object Detection
by: Huang, Zile, et al.
Published: (2024)
by: Huang, Zile, et al.
Published: (2024)
FusionFM: Fusing Eye-specific Foundational Models for Optimized Ophthalmic Diagnosis
by: Zou, Ke, et al.
Published: (2025)
by: Zou, Ke, et al.
Published: (2025)
Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding
by: Tang, Feilong, et al.
Published: (2025)
by: Tang, Feilong, et al.
Published: (2025)
One-stage Modality Distillation for Incomplete Multimodal Learning
by: Wei, Shicai, et al.
Published: (2023)
by: Wei, Shicai, et al.
Published: (2023)
A Clinician-Friendly Platform for Ophthalmic Image Analysis Without Technical Barriers
by: Wang, Meng, et al.
Published: (2025)
by: Wang, Meng, et al.
Published: (2025)
Progressive Representation Learning for Multimodal Sentiment Analysis with Incomplete Modalities
by: Bao, Jindi, et al.
Published: (2026)
by: Bao, Jindi, et al.
Published: (2026)
GLCP: Global-to-Local Connectivity Preservation for Tubular Structure Segmentation
by: Zhou, Feixiang, et al.
Published: (2025)
by: Zhou, Feixiang, et al.
Published: (2025)
PoCo: A Self-Supervised Approach via Polar Transformation Based Progressive Contrastive Learning for Ophthalmic Disease Diagnosis
by: Wang, Jinhong, et al.
Published: (2024)
by: Wang, Jinhong, et al.
Published: (2024)
AlphaVAE: Unified End-to-End RGBA Image Reconstruction and Generation with Alpha-Aware Representation Learning
by: Wang, Zile, et al.
Published: (2025)
by: Wang, Zile, et al.
Published: (2025)
Neighbor Does Matter: Density-Aware Contrastive Learning for Medical Semi-supervised Segmentation
by: Tang, Feilong, et al.
Published: (2024)
by: Tang, Feilong, et al.
Published: (2024)
Inference-Time Dynamic Modality Selection for Incomplete Multimodal Classification
by: Du, Siyi, et al.
Published: (2026)
by: Du, Siyi, et al.
Published: (2026)
PSScreen V2: Partially Supervised Multiple Retinal Disease Screening
by: Zheng, Boyi, et al.
Published: (2025)
by: Zheng, Boyi, et al.
Published: (2025)
Distilling Cross-Modal Knowledge via Feature Disentanglement
by: Liu, Junhong, et al.
Published: (2025)
by: Liu, Junhong, et al.
Published: (2025)
Understanding Temporal Logic Consistency in Video-Language Models through Cross-Modal Attention Discriminability
by: Li, Chengzhi, et al.
Published: (2025)
by: Li, Chengzhi, et al.
Published: (2025)
Thinking in Uncertainty: Mitigating Hallucinations in MLRMs with Latent Entropy-Aware Decoding
by: Xu, Zhongxing, et al.
Published: (2026)
by: Xu, Zhongxing, et al.
Published: (2026)
Self-Supervised Generative-Contrastive Learning of Multi-Modal Euclidean Input for 3D Shape Latent Representations: A Dynamic Switching Approach
by: Wu, Chengzhi, et al.
Published: (2023)
by: Wu, Chengzhi, et al.
Published: (2023)
Learning Modality-agnostic Representation for Semantic Segmentation from Any Modalities
by: Zheng, Xu, et al.
Published: (2024)
by: Zheng, Xu, et al.
Published: (2024)
X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic Diagnosis
by: Wang, Gui, et al.
Published: (2026)
by: Wang, Gui, et al.
Published: (2026)
Decoupling Feature Representations of Ego and Other Modalities for Incomplete Multi-modal Brain Tumor Segmentation
by: Yang, Kaixiang, et al.
Published: (2024)
by: Yang, Kaixiang, et al.
Published: (2024)
OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models
by: Dong, Xuanzhao, et al.
Published: (2026)
by: Dong, Xuanzhao, et al.
Published: (2026)
A Survey of Multimodal Ophthalmic Diagnostics: From Task-Specific Approaches to Foundational Models
by: Luo, Xiaoling, et al.
Published: (2025)
by: Luo, Xiaoling, et al.
Published: (2025)
Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation
by: Liu, Ruiping, et al.
Published: (2024)
by: Liu, Ruiping, et al.
Published: (2024)
Similar Items
-
Robust Incomplete-Modality Alignment for Ophthalmic Disease Grading and Diagnosis via Labeled Optimal Transport
by: Yu, Qinkai, et al.
Published: (2025) -
Robust Multimodal Learning for Ophthalmic Disease Grading via Disentangled Representation
by: Wang, Xinkun, et al.
Published: (2025) -
Confidence-Aware Self-Distillation for Multimodal Sentiment Analysis with Incomplete Modalities
by: Luo, Yanxi, et al.
Published: (2025) -
Parameterized Diffusion Optimization enabled Autoregressive Ordinal Regression for Diabetic Retinopathy Grading
by: Yu, Qinkai, et al.
Published: (2025) -
Toward Modality Gap: Vision Prototype Learning for Weakly-supervised Semantic Segmentation with CLIP
by: Xu, Zhongxing, et al.
Published: (2024)