ET-SAM: Efficient Point Prompt Prediction in SAM for Unified Scene Text Detection and Layout Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Xike, Ye, Maoyuan, Liu, Juhua, Du, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hi-SAM: Marrying Segment Anything Model for Hierarchical Text Segmentation
by: Ye, Maoyuan, et al.
Published: (2024)
by: Ye, Maoyuan, et al.
Published: (2024)
VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA
by: He, Haibin, et al.
Published: (2026)
by: He, Haibin, et al.
Published: (2026)
GoMatching++: Parameter- and Data-Efficient Arbitrary-Shaped Video Text Spotting and Benchmarking
by: He, Haibin, et al.
Published: (2025)
by: He, Haibin, et al.
Published: (2025)
DeepSolo++: Let Transformer Decoder with Explicit Points Solo for Multilingual Text Spotting
by: Ye, Maoyuan, et al.
Published: (2023)
by: Ye, Maoyuan, et al.
Published: (2023)
GoMatching: A Simple Baseline for Video Text Spotting via Long and Short Term Matching
by: He, Haibin, et al.
Published: (2024)
by: He, Haibin, et al.
Published: (2024)
EdgeSAM: Prompt-In-the-Loop Distillation for SAM
by: Zhou, Chong, et al.
Published: (2023)
by: Zhou, Chong, et al.
Published: (2023)
LogicOCR: Do Your Large Multimodal Models Excel at Logical Reasoning on Text-Rich Images?
by: Ye, Maoyuan, et al.
Published: (2025)
by: Ye, Maoyuan, et al.
Published: (2025)
VRP-SAM: SAM with Visual Reference Prompt
by: Sun, Yanpeng, et al.
Published: (2024)
by: Sun, Yanpeng, et al.
Published: (2024)
HFP-SAM: Hierarchical Frequency Prompted SAM for Efficient Marine Animal Segmentation
by: Zhang, Pingping, et al.
Published: (2026)
by: Zhang, Pingping, et al.
Published: (2026)
Evaluating SAM2's Role in Camouflaged Object Detection: From SAM to SAM2
by: Tang, Lv, et al.
Published: (2024)
by: Tang, Lv, et al.
Published: (2024)
Crowd-SAM: SAM as a Smart Annotator for Object Detection in Crowded Scenes
by: Cai, Zhi, et al.
Published: (2024)
by: Cai, Zhi, et al.
Published: (2024)
P2RBox: Point Prompt Oriented Object Detection with SAM
by: Cao, Guangming, et al.
Published: (2023)
by: Cao, Guangming, et al.
Published: (2023)
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization
by: Yang, Xi, et al.
Published: (2025)
by: Yang, Xi, et al.
Published: (2025)
SAM-CP: Marrying SAM with Composable Prompts for Versatile Segmentation
by: Chen, Pengfei, et al.
Published: (2024)
by: Chen, Pengfei, et al.
Published: (2024)
IP-SAM: Prompt-Space Conditioning for Prompt-Absent Camouflaged Object Detection
by: Zhang, Huiyao, et al.
Published: (2026)
by: Zhang, Huiyao, et al.
Published: (2026)
SSP-SAM: SAM with Semantic-Spatial Prompt for Referring Expression Segmentation
by: Tang, Wei, et al.
Published: (2026)
by: Tang, Wei, et al.
Published: (2026)
Adapting Segment Anything Model for Power Transmission Corridor Hazard Segmentation
by: Chen, Hang, et al.
Published: (2025)
by: Chen, Hang, et al.
Published: (2025)
Semantic-aware SAM for Point-Prompted Instance Segmentation
by: Wei, Zhaoyang, et al.
Published: (2023)
by: Wei, Zhaoyang, et al.
Published: (2023)
OFL-SAM2: Prompt SAM2 with Online Few-shot Learner for Efficient Medical Image Segmentation
by: Lan, Meng, et al.
Published: (2025)
by: Lan, Meng, et al.
Published: (2025)
Self-Prompt SAM: Medical Image Segmentation via Automatic Prompt SAM Adaptation
by: Xie, Bin, et al.
Published: (2025)
by: Xie, Bin, et al.
Published: (2025)
BiPrompt-SAM: Enhancing Image Segmentation via Explicit Selection between Point and Text Prompts
by: Xu, Suzhe, et al.
Published: (2025)
by: Xu, Suzhe, et al.
Published: (2025)
TextSAM-EUS: Text Prompt Learning for SAM to Accurately Segment Pancreatic Tumor in Endoscopic Ultrasound
by: Spiegler, Pascal, et al.
Published: (2025)
by: Spiegler, Pascal, et al.
Published: (2025)
Reasoning-OCR: Can Large Multimodal Models Solve Complex Logical Reasoning Problems from OCR Cues?
by: He, Haibin, et al.
Published: (2025)
by: He, Haibin, et al.
Published: (2025)
ProSAM: Enhancing the Robustness of SAM-based Visual Reference Segmentation with Probabilistic Prompts
by: Wang, Xiaoqi, et al.
Published: (2025)
by: Wang, Xiaoqi, et al.
Published: (2025)
SEG-SAM: Semantic-Guided SAM for Unified Medical Image Segmentation
by: Huang, Shuangping, et al.
Published: (2024)
by: Huang, Shuangping, et al.
Published: (2024)
SAM-PD: How Far Can SAM Take Us in Tracking and Segmenting Anything in Videos by Prompt Denoising
by: Zhou, Tao, et al.
Published: (2024)
by: Zhou, Tao, et al.
Published: (2024)
EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model
by: Zhang, Yuxuan, et al.
Published: (2024)
by: Zhang, Yuxuan, et al.
Published: (2024)
Inspiring the Next Generation of Segment Anything Models: Comprehensively Evaluate SAM and SAM 2 with Diverse Prompts Towards Context-Dependent Concepts under Different Scenes
by: Zhao, Xiaoqi, et al.
Published: (2024)
by: Zhao, Xiaoqi, et al.
Published: (2024)
Char-SAM: Turning Segment Anything Model into Scene Text Segmentation Annotator with Character-level Visual Prompts
by: Xie, Enze, et al.
Published: (2024)
by: Xie, Enze, et al.
Published: (2024)
SAM-PTx: Text-Guided Fine-Tuning of SAM with Parameter-Efficient, Parallel-Text Adapters
by: Jalilian, Shayan, et al.
Published: (2025)
by: Jalilian, Shayan, et al.
Published: (2025)
Efficient-SAM2: Accelerating SAM2 with Object-Aware Visual Encoding and Memory Retrieval
by: Zhang, Jing, et al.
Published: (2026)
by: Zhang, Jing, et al.
Published: (2026)
SAM-COD: SAM-guided Unified Framework for Weakly-Supervised Camouflaged Object Detection
by: Chen, Huafeng, et al.
Published: (2024)
by: Chen, Huafeng, et al.
Published: (2024)
MedSAM-U: Uncertainty-Guided Auto Multi-Prompt Adaptation for Reliable MedSAM
by: Zhou, Nan, et al.
Published: (2024)
by: Zhou, Nan, et al.
Published: (2024)
VesSAM: Efficient Multi-Prompting for Segmenting Complex Vessel
by: Fu, Suzhong, et al.
Published: (2025)
by: Fu, Suzhong, et al.
Published: (2025)
SSS: Semi-Supervised SAM-2 with Efficient Prompting for Medical Imaging Segmentation
by: Zhu, Hongjie, et al.
Published: (2025)
by: Zhu, Hongjie, et al.
Published: (2025)
OpenWorldSAM: Extending SAM2 for Universal Image Segmentation with Language Prompts
by: Xiao, Shiting, et al.
Published: (2025)
by: Xiao, Shiting, et al.
Published: (2025)
ST-SAM: SAM-Driven Self-Training Framework for Semi-Supervised Camouflaged Object Detection
by: Hu, Xihang, et al.
Published: (2025)
by: Hu, Xihang, et al.
Published: (2025)
SAM-Guided Masked Token Prediction for 3D Scene Understanding
by: Chen, Zhimin, et al.
Published: (2024)
by: Chen, Zhimin, et al.
Published: (2024)
AuralSAM2: Enabling SAM2 Hear Through Pyramid Audio-Visual Feature Prompting
by: Liu, Yuyuan, et al.
Published: (2025)
by: Liu, Yuyuan, et al.
Published: (2025)
Ref-SAM3D: Bridging SAM3D with Text for Reference 3D Reconstruction
by: Zhou, Yun, et al.
Published: (2025)
by: Zhou, Yun, et al.
Published: (2025)
Similar Items
-
Hi-SAM: Marrying Segment Anything Model for Hierarchical Text Segmentation
by: Ye, Maoyuan, et al.
Published: (2024) -
VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA
by: He, Haibin, et al.
Published: (2026) -
GoMatching++: Parameter- and Data-Efficient Arbitrary-Shaped Video Text Spotting and Benchmarking
by: He, Haibin, et al.
Published: (2025) -
DeepSolo++: Let Transformer Decoder with Explicit Points Solo for Multilingual Text Spotting
by: Ye, Maoyuan, et al.
Published: (2023) -
GoMatching: A Simple Baseline for Video Text Spotting via Long and Short Term Matching
by: He, Haibin, et al.
Published: (2024)