DocSAM: Unified Document Image Segmentation via Query Decomposition and Heterogeneous Mixed Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Xiao-Hui, Yin, Fei, Liu, Cheng-Lin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PGP-SAM: Prototype-Guided Prompt Learning for Efficient Few-Shot Medical Image Segmentation
by: Yan, Zhonghao, et al.
Published: (2025)
by: Yan, Zhonghao, et al.
Published: (2025)
PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction
by: Sun, Ting, et al.
Published: (2025)
by: Sun, Ting, et al.
Published: (2025)
SAM-Fed: SAM-Guided Federated Semi-Supervised Learning for Medical Image Segmentation
by: Nasirihaghighi, Sahar, et al.
Published: (2025)
by: Nasirihaghighi, Sahar, et al.
Published: (2025)
X2SAM: Any Segmentation in Images and Videos
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Beyond Adapting SAM: Towards End-to-End Ultrasound Image Segmentation via Auto Prompting
by: Lin, Xian, et al.
Published: (2023)
by: Lin, Xian, et al.
Published: (2023)
DocShaDiffusion: Diffusion Model in Latent Space for Document Image Shadow Removal
by: Liu, Wenjie, et al.
Published: (2025)
by: Liu, Wenjie, et al.
Published: (2025)
TextSAM-EUS: Text Prompt Learning for SAM to Accurately Segment Pancreatic Tumor in Endoscopic Ultrasound
by: Spiegler, Pascal, et al.
Published: (2025)
by: Spiegler, Pascal, et al.
Published: (2025)
BALR-SAM: Boundary-Aware Low-Rank Adaptation of SAM for Resource-Efficient Medical Image Segmentation
by: Liu, Zelin, et al.
Published: (2025)
by: Liu, Zelin, et al.
Published: (2025)
GeoSAM: Fine-tuning SAM with Multi-Modal Prompts for Mobility Infrastructure Segmentation
by: Sultan, Rafi Ibn, et al.
Published: (2023)
by: Sultan, Rafi Ibn, et al.
Published: (2023)
RefSAM: Efficiently Adapting Segmenting Anything Model for Referring Video Object Segmentation
by: Li, Yonglin, et al.
Published: (2023)
by: Li, Yonglin, et al.
Published: (2023)
X-SAM: From Segment Anything to Any Segmentation
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
PG-SAM: Prior-Guided SAM with Medical for Multi-organ Segmentation
by: Zhong, Yiheng, et al.
Published: (2025)
by: Zhong, Yiheng, et al.
Published: (2025)
SAM2 for Image and Video Segmentation: A Comprehensive Survey
by: Jiaxing, Zhang, et al.
Published: (2025)
by: Jiaxing, Zhang, et al.
Published: (2025)
ProMISe: Promptable Medical Image Segmentation using SAM
by: Wang, Jinfeng, et al.
Published: (2024)
by: Wang, Jinfeng, et al.
Published: (2024)
MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement Learning
by: Liu, Shengyuan, et al.
Published: (2026)
by: Liu, Shengyuan, et al.
Published: (2026)
Compress Any Segment Anything Model (SAM)
by: Fan, Juntong, et al.
Published: (2025)
by: Fan, Juntong, et al.
Published: (2025)
GeoSAM2: Unleashing the Power of SAM2 for 3D Part Segmentation
by: Deng, Ken, et al.
Published: (2025)
by: Deng, Ken, et al.
Published: (2025)
MM-UNet: A Mixed MLP Architecture for Improved Ophthalmic Image Segmentation
by: Xiao, Zunjie, et al.
Published: (2024)
by: Xiao, Zunjie, et al.
Published: (2024)
The SAM2-to-SAM3 Gap in the Segment Anything Model Family: Why Prompt-Based Expertise Fails in Concept-Driven Image Segmentation
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
ProtoSAM: One-Shot Medical Image Segmentation With Foundational Models
by: Ayzenberg, Lev, et al.
Published: (2024)
by: Ayzenberg, Lev, et al.
Published: (2024)
SAM 3: Segment Anything with Concepts
by: Carion, Nicolas, et al.
Published: (2025)
by: Carion, Nicolas, et al.
Published: (2025)
ClipSAM: CLIP and SAM Collaboration for Zero-Shot Anomaly Segmentation
by: Li, Shengze, et al.
Published: (2024)
by: Li, Shengze, et al.
Published: (2024)
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation
by: Ma, Chaofan, et al.
Published: (2023)
by: Ma, Chaofan, et al.
Published: (2023)
SAM 2: Segment Anything in Images and Videos
by: Ravi, Nikhila, et al.
Published: (2024)
by: Ravi, Nikhila, et al.
Published: (2024)
AerOSeg: Harnessing SAM for Open-Vocabulary Segmentation in Remote Sensing Images
by: Dutta, Saikat, et al.
Published: (2025)
by: Dutta, Saikat, et al.
Published: (2025)
Multimodal SAM-adapter for Semantic Segmentation
by: Curti, Iacopo, et al.
Published: (2025)
by: Curti, Iacopo, et al.
Published: (2025)
SAM-COD: SAM-guided Unified Framework for Weakly-Supervised Camouflaged Object Detection
by: Chen, Huafeng, et al.
Published: (2024)
by: Chen, Huafeng, et al.
Published: (2024)
DocPedia: Unleashing the Power of Large Multimodal Model in the Frequency Domain for Versatile Document Understanding
by: Feng, Hao, et al.
Published: (2023)
by: Feng, Hao, et al.
Published: (2023)
DocPTBench: Benchmarking End-to-End Photographed Document Parsing and Translation
by: Du, Yongkun, et al.
Published: (2025)
by: Du, Yongkun, et al.
Published: (2025)
Segmenting Visuals With Querying Words: Language Anchors For Semi-Supervised Image Segmentation
by: Nadeem, Numair, et al.
Published: (2025)
by: Nadeem, Numair, et al.
Published: (2025)
SeqSAM: Autoregressive Multiple Hypothesis Prediction for Medical Image Segmentation using SAM
by: Towle, Benjamin, et al.
Published: (2025)
by: Towle, Benjamin, et al.
Published: (2025)
QTSeg: A Query Token-Based Dual-Mix Attention Framework with Multi-Level Feature Distribution for Medical Image Segmentation
by: Tran, Phuong-Nam, et al.
Published: (2024)
by: Tran, Phuong-Nam, et al.
Published: (2024)
FedIA: Federated Medical Image Segmentation with Heterogeneous Annotation Completeness
by: Xiang, Yangyang, et al.
Published: (2024)
by: Xiang, Yangyang, et al.
Published: (2024)
SimSAM: Zero-shot Medical Image Segmentation via Simulated Interaction
by: Towle, Benjamin, et al.
Published: (2024)
by: Towle, Benjamin, et al.
Published: (2024)
LENS: Learning to Segment Anything with Unified Reinforced Reasoning
by: Zhu, Lianghui, et al.
Published: (2025)
by: Zhu, Lianghui, et al.
Published: (2025)
DescriptorMedSAM: Language-Image Fusion with Multi-Aspect Text Guidance for Medical Image Segmentation
by: Zhang, Wenjie, et al.
Published: (2025)
by: Zhang, Wenjie, et al.
Published: (2025)
AutoProSAM: Automated Prompting SAM for 3D Multi-Organ Segmentation
by: Li, Chengyin, et al.
Published: (2023)
by: Li, Chengyin, et al.
Published: (2023)
Re-purposing SAM into Efficient Visual Projectors for MLLM-Based Referring Image Segmentation
by: Yang, Xiaobo, et al.
Published: (2025)
by: Yang, Xiaobo, et al.
Published: (2025)
Mixed Prototype Consistency Learning for Semi-supervised Medical Image Segmentation
by: Li, Lijian
Published: (2024)
by: Li, Lijian
Published: (2024)
Improving Bird's Eye View Semantic Segmentation by Task Decomposition
by: Zhao, Tianhao, et al.
Published: (2024)
by: Zhao, Tianhao, et al.
Published: (2024)
Similar Items
-
PGP-SAM: Prototype-Guided Prompt Learning for Efficient Few-Shot Medical Image Segmentation
by: Yan, Zhonghao, et al.
Published: (2025) -
PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction
by: Sun, Ting, et al.
Published: (2025) -
SAM-Fed: SAM-Guided Federated Semi-Supervised Learning for Medical Image Segmentation
by: Nasirihaghighi, Sahar, et al.
Published: (2025) -
X2SAM: Any Segmentation in Images and Videos
by: Wang, Hao, et al.
Published: (2026) -
Beyond Adapting SAM: Towards End-to-End Ultrasound Image Segmentation via Auto Prompting
by: Lin, Xian, et al.
Published: (2023)