Gespeichert in:
| Hauptverfasser: | Liu, Zhihua, Tong, Lei, He, Xilin, Liu, Che, Arcucci, Rossella, Jin, Chen, Zhou, Huiyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2505.18052 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Segment Anyword: Mask Prompt Inversion for Open-Set Grounded Segmentation
von: Liu, Zhihua, et al.
Veröffentlicht: (2025)
von: Liu, Zhihua, et al.
Veröffentlicht: (2025)
How Does Diverse Interpretability of Textual Prompts Impact Medical Vision-Language Zero-Shot Tasks?
von: Wang, Sicheng, et al.
Veröffentlicht: (2024)
von: Wang, Sicheng, et al.
Veröffentlicht: (2024)
Utilizing Synthetic Data for Medical Vision-Language Pre-training: Bypassing the Need for Real Images
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
BIMCV-R: A Landmark Dataset for 3D CT Text-Image Retrieval
von: Chen, Yinda, et al.
Veröffentlicht: (2024)
von: Chen, Yinda, et al.
Veröffentlicht: (2024)
G2D: From Global to Dense Radiography Representation Learning via Vision-Language Pre-training
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
FMBench: Benchmarking Fairness in Multimodal Large Language Models on Medical Tasks
von: Wu, Peiran, et al.
Veröffentlicht: (2024)
von: Wu, Peiran, et al.
Veröffentlicht: (2024)
How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study
von: Liu, Che, et al.
Veröffentlicht: (2025)
von: Liu, Che, et al.
Veröffentlicht: (2025)
Freeze the backbones: A Parameter-Efficient Contrastive Approach to Robust Medical Vision-Language Pre-training
von: Qin, Jiuming, et al.
Veröffentlicht: (2024)
von: Qin, Jiuming, et al.
Veröffentlicht: (2024)
Knowledge to Sight: Reasoning over Visual Attributes via Knowledge Decomposition for Abnormality Grounding
von: Li, Jun, et al.
Veröffentlicht: (2025)
von: Li, Jun, et al.
Veröffentlicht: (2025)
IMITATE: Clinical Prior Guided Hierarchical Vision-Language Pre-training
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?
von: Liu, Che, et al.
Veröffentlicht: (2024)
von: Liu, Che, et al.
Veröffentlicht: (2024)
Enhancing Abnormality Grounding for Vision Language Models with Knowledge Descriptions
von: Li, Jun, et al.
Veröffentlicht: (2025)
von: Li, Jun, et al.
Veröffentlicht: (2025)
Noise2Noise Denoising of CRISM Hyperspectral Data
von: Platt, Robert, et al.
Veröffentlicht: (2024)
von: Platt, Robert, et al.
Veröffentlicht: (2024)
DomainForensics: Exposing Face Forgery across Domains via Bi-directional Adaptation
von: Lv, Qingxuan, et al.
Veröffentlicht: (2023)
von: Lv, Qingxuan, et al.
Veröffentlicht: (2023)
Argus: Benchmarking and Enhancing Vision-Language Models for 3D Radiology Report Generation
von: Liu, Che, et al.
Veröffentlicht: (2024)
von: Liu, Che, et al.
Veröffentlicht: (2024)
Med-UniC: Unifying Cross-Lingual Medical Vision-Language Pre-Training by Diminishing Bias
von: Wan, Zhongwei, et al.
Veröffentlicht: (2023)
von: Wan, Zhongwei, et al.
Veröffentlicht: (2023)
OT-Drive: Out-of-Distribution Off-Road Traversable Area Segmentation via Optimal Transport
von: Zhao, Zhihua, et al.
Veröffentlicht: (2026)
von: Zhao, Zhihua, et al.
Veröffentlicht: (2026)
TokenSeg: Efficient 3D Medical Image Segmentation via Hierarchical Visual Token Compression
von: Zeng, Sen, et al.
Veröffentlicht: (2026)
von: Zeng, Sen, et al.
Veröffentlicht: (2026)
OMH: Structured Sparsity via Optimally Matched Hierarchy for Unsupervised Semantic Segmentation
von: Ozaydin, Baran, et al.
Veröffentlicht: (2024)
von: Ozaydin, Baran, et al.
Veröffentlicht: (2024)
One Token to Seg Them All: Language Instructed Reasoning Segmentation in Videos
von: Bai, Zechen, et al.
Veröffentlicht: (2024)
von: Bai, Zechen, et al.
Veröffentlicht: (2024)
Neural B-frame Video Compression with Bi-directional Reference Harmonization
von: Liu, Yuxi, et al.
Veröffentlicht: (2025)
von: Liu, Yuxi, et al.
Veröffentlicht: (2025)
T3D: Advancing 3D Medical Vision-Language Pre-training by Learning Multi-View Visual Consistency
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
DTBS: Dual-Teacher Bi-directional Self-training for Domain Adaptation in Nighttime Semantic Segmentation
von: Huang, Fanding, et al.
Veröffentlicht: (2024)
von: Huang, Fanding, et al.
Veröffentlicht: (2024)
Hierarchical Spatio-temporal Segmentation Network for Ejection Fraction Estimation in Echocardiography Videos
von: Wang, Dongfang, et al.
Veröffentlicht: (2025)
von: Wang, Dongfang, et al.
Veröffentlicht: (2025)
DOMR: Establishing Cross-View Segmentation via Dense Object Matching
von: Liao, Jitong, et al.
Veröffentlicht: (2025)
von: Liao, Jitong, et al.
Veröffentlicht: (2025)
A Semi-Supervised Approach with Error Reflection for Echocardiography Segmentation
von: Han, Xiaoxiang, et al.
Veröffentlicht: (2024)
von: Han, Xiaoxiang, et al.
Veröffentlicht: (2024)
Cross Fusion RGB-T Tracking with Bi-directional Adapter
von: Zeng, Zhirong, et al.
Veröffentlicht: (2024)
von: Zeng, Zhirong, et al.
Veröffentlicht: (2024)
Towards Robust Semantic Segmentation against Patch-based Attack via Attention Refinement
von: Yuan, Zheng, et al.
Veröffentlicht: (2024)
von: Yuan, Zheng, et al.
Veröffentlicht: (2024)
SMC-NCA: Semantic-guided Multi-level Contrast for Semi-supervised Temporal Action Segmentation
von: Zhou, Feixiang, et al.
Veröffentlicht: (2023)
von: Zhou, Feixiang, et al.
Veröffentlicht: (2023)
SPARNet: Continual Test-Time Adaptation via Sample Partitioning Strategy and Anti-Forgetting Regularization
von: Meng, Xinru, et al.
Veröffentlicht: (2025)
von: Meng, Xinru, et al.
Veröffentlicht: (2025)
Mask-adaptive Gated Convolution and Bi-directional Progressive Fusion Network for Depth Completion
von: Huang, Tingxuan, et al.
Veröffentlicht: (2024)
von: Huang, Tingxuan, et al.
Veröffentlicht: (2024)
Efficient Point Clouds Upsampling via Flow Matching
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2025)
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2025)
Scaling Mesh Generation via Compressive Tokenization
von: Weng, Haohan, et al.
Veröffentlicht: (2024)
von: Weng, Haohan, et al.
Veröffentlicht: (2024)
Fuse & Calibrate: A bi-directional Vision-Language Guided Framework for Referring Image Segmentation
von: Yan, Yichen, et al.
Veröffentlicht: (2024)
von: Yan, Yichen, et al.
Veröffentlicht: (2024)
GDKVM: Echocardiography Video Segmentation via Spatiotemporal Key-Value Memory with Gated Delta Rule
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
Does DINOv3 Set a New Medical Vision Standard? Benchmarking 2D and 3D Classification, Segmentation, and Registration
von: Liu, Che, et al.
Veröffentlicht: (2025)
von: Liu, Che, et al.
Veröffentlicht: (2025)
SimToken: A Simple Baseline for Referring Audio-Visual Segmentation
von: Jin, Dian, et al.
Veröffentlicht: (2025)
von: Jin, Dian, et al.
Veröffentlicht: (2025)
Medical Referring Image Segmentation via Next-Token Mask Prediction
von: Chen, Xinyu, et al.
Veröffentlicht: (2025)
von: Chen, Xinyu, et al.
Veröffentlicht: (2025)
DCFS: Continual Test-Time Adaptation via Dual Consistency of Feature and Sample
von: Yin, Wenting, et al.
Veröffentlicht: (2025)
von: Yin, Wenting, et al.
Veröffentlicht: (2025)
OSA: Echocardiography Video Segmentation via Orthogonalized State Update and Anatomical Prior-aware Feature Enhancement
von: Wang, Rui, et al.
Veröffentlicht: (2026)
von: Wang, Rui, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Segment Anyword: Mask Prompt Inversion for Open-Set Grounded Segmentation
von: Liu, Zhihua, et al.
Veröffentlicht: (2025) -
How Does Diverse Interpretability of Textual Prompts Impact Medical Vision-Language Zero-Shot Tasks?
von: Wang, Sicheng, et al.
Veröffentlicht: (2024) -
Utilizing Synthetic Data for Medical Vision-Language Pre-training: Bypassing the Need for Real Images
von: Liu, Che, et al.
Veröffentlicht: (2023) -
BIMCV-R: A Landmark Dataset for 3D CT Text-Image Retrieval
von: Chen, Yinda, et al.
Veröffentlicht: (2024) -
G2D: From Global to Dense Radiography Representation Learning via Vision-Language Pre-training
von: Liu, Che, et al.
Veröffentlicht: (2023)