Edge Approximation Text Detector
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Chuang, Han, Xu, Han, Tao, Han, Han, Zhao, Bingxuan, Wang, Qi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TextMamba: Scene Text Detector with Mamba
by: Zhao, Qiyan, et al.
Published: (2025)
by: Zhao, Qiyan, et al.
Published: (2025)
Text-Pass Filter: An Efficient Scene Text Detector
by: Yang, Chuang, et al.
Published: (2026)
by: Yang, Chuang, et al.
Published: (2026)
Spotlight Text Detector: Spotlight on Candidate Regions Like a Camera
by: Han, Xu, et al.
Published: (2024)
by: Han, Xu, et al.
Published: (2024)
GloTSFormer: Global Video Text Spotting Transformer
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
Cross-Axis Feature Fusion with Joint-Wise Motion Difference Prediction for Text-Based 3D Human Motion Editing
by: Han, Gyojin, et al.
Published: (2026)
by: Han, Gyojin, et al.
Published: (2026)
Frame-Difference Guided Dynamic Region Perception for CLIP Adaptation in Text-Video Retrieval
by: Yu, Jiaao, et al.
Published: (2025)
by: Yu, Jiaao, et al.
Published: (2025)
Gold-YOLO: Efficient Object Detector via Gather-and-Distribute Mechanism
by: Wang, Chengcheng, et al.
Published: (2023)
by: Wang, Chengcheng, et al.
Published: (2023)
When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models
by: Yakun, Cui, et al.
Published: (2026)
by: Yakun, Cui, et al.
Published: (2026)
Integrated Image-Text Based on Semi-supervised Learning for Small Sample Instance Segmentation
by: Chi, Ruting, et al.
Published: (2024)
by: Chi, Ruting, et al.
Published: (2024)
MoMBS: Mixed-order minibatch sampling enhances model training from diverse-quality images
by: Li, Han, et al.
Published: (2025)
by: Li, Han, et al.
Published: (2025)
A Dynamic Knowledge Distillation Method Based on the Gompertz Curve
by: Yang, Han, et al.
Published: (2025)
by: Yang, Han, et al.
Published: (2025)
Instant Preference Alignment for Text-to-Image Diffusion Models
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
TextEditBench: Evaluating Reasoning-aware Text Editing Beyond Rendering
by: Gui, Rui, et al.
Published: (2025)
by: Gui, Rui, et al.
Published: (2025)
FIFO-Diffusion: Generating Infinite Videos from Text without Training
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
Decompose the model: Mechanistic interpretability in image models with Generalized Integrated Gradients (GIG)
by: Kim, Yearim, et al.
Published: (2024)
by: Kim, Yearim, et al.
Published: (2024)
Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
by: Han, Jiaming, et al.
Published: (2025)
by: Han, Jiaming, et al.
Published: (2025)
FlexGen: Flexible Multi-View Generation from Text and Image Inputs
by: Xu, Xinli, et al.
Published: (2024)
by: Xu, Xinli, et al.
Published: (2024)
FAIRT2V: Training-Free Debiasing for Text-to-Video Diffusion Models
by: Zhong, Haonan, et al.
Published: (2026)
by: Zhong, Haonan, et al.
Published: (2026)
SINE: SINgle Image Editing with Text-to-Image Diffusion Models
by: Zhang, Zhixing, et al.
Published: (2022)
by: Zhang, Zhixing, et al.
Published: (2022)
Feature-EndoGaussian: Feature Distilled Gaussian Splatting in Surgical Deformable Scene Reconstruction
by: Li, Kai, et al.
Published: (2025)
by: Li, Kai, et al.
Published: (2025)
Projecting Gaussian Ellipsoids While Avoiding Affine Projection Approximation
by: Qi, Han, et al.
Published: (2024)
by: Qi, Han, et al.
Published: (2024)
PATIMT-Bench: A Multi-Scenario Benchmark for Position-Aware Text Image Machine Translation in Large Vision-Language Models
by: Zhuang, Wanru, et al.
Published: (2025)
by: Zhuang, Wanru, et al.
Published: (2025)
HoneyImage: Verifiable, Harmless, and Stealthy Dataset Ownership Verification for Image Models
by: Zhu, Zhihao, et al.
Published: (2025)
by: Zhu, Zhihao, et al.
Published: (2025)
Memory-based Cross-modal Semantic Alignment Network for Radiology Report Generation
by: Tao, Yitian, et al.
Published: (2024)
by: Tao, Yitian, et al.
Published: (2024)
Retrieval-Augmented Prompt for OOD Detection
by: Han, Ruisong, et al.
Published: (2025)
by: Han, Ruisong, et al.
Published: (2025)
Can Diffusion Models Learn Hidden Inter-Feature Rules Behind Images?
by: Han, Yujin, et al.
Published: (2025)
by: Han, Yujin, et al.
Published: (2025)
DAPE: Dynamic Non-uniform Alignment and Progressive Detail Enhancement Techniques for Improving the Performance of Efficient Visual Language Models
by: Tian, Mengyuan, et al.
Published: (2026)
by: Tian, Mengyuan, et al.
Published: (2026)
A Novel Spike Transformer Network for Depth Estimation from Event Cameras via Cross-modality Knowledge Distillation
by: Zhang, Xin, et al.
Published: (2024)
by: Zhang, Xin, et al.
Published: (2024)
Steering Away from Harm: An Adaptive Approach to Defending Vision Language Model Against Jailbreaks
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
CLIP-Guided Unsupervised Semantic-Aware Exposure Correction
by: Wu, Puzhen, et al.
Published: (2026)
by: Wu, Puzhen, et al.
Published: (2026)
Salience Adjustment for Context-Based Emotion Recognition
by: Han, Bin, et al.
Published: (2025)
by: Han, Bin, et al.
Published: (2025)
MetaSpatial: Reinforcing 3D Spatial Reasoning in VLMs for the Metaverse
by: Pan, Zhenyu, et al.
Published: (2025)
by: Pan, Zhenyu, et al.
Published: (2025)
Multi-Source Collaborative Gradient Discrepancy Minimization for Federated Domain Generalization
by: Wei, Yikang, et al.
Published: (2024)
by: Wei, Yikang, et al.
Published: (2024)
Focus Entirety and Perceive Environment for Arbitrary-Shaped Text Detection
by: Han, Xu, et al.
Published: (2024)
by: Han, Xu, et al.
Published: (2024)
3D-Layout-R1: Structured Reasoning for Language-Instructed Spatial Editing
by: Zhen, Haoyu, et al.
Published: (2026)
by: Zhen, Haoyu, et al.
Published: (2026)
Dynamic Double Space Tower
by: Sun, Weikai, et al.
Published: (2025)
by: Sun, Weikai, et al.
Published: (2025)
Dissecting Out-of-Distribution Detection and Open-Set Recognition: A Critical Analysis of Methods and Benchmarks
by: Wang, Hongjun, et al.
Published: (2024)
by: Wang, Hongjun, et al.
Published: (2024)
HiLo: A Learning Framework for Generalized Category Discovery Robust to Domain Shifts
by: Wang, Hongjun, et al.
Published: (2024)
by: Wang, Hongjun, et al.
Published: (2024)
SPTNet: An Efficient Alternative Framework for Generalized Category Discovery with Spatial Prompt Tuning
by: Wang, Hongjun, et al.
Published: (2024)
by: Wang, Hongjun, et al.
Published: (2024)
PanGu-Draw: Advancing Resource-Efficient Text-to-Image Synthesis with Time-Decoupled Training and Reusable Coop-Diffusion
by: Lu, Guansong, et al.
Published: (2023)
by: Lu, Guansong, et al.
Published: (2023)
Similar Items
-
TextMamba: Scene Text Detector with Mamba
by: Zhao, Qiyan, et al.
Published: (2025) -
Text-Pass Filter: An Efficient Scene Text Detector
by: Yang, Chuang, et al.
Published: (2026) -
Spotlight Text Detector: Spotlight on Candidate Regions Like a Camera
by: Han, Xu, et al.
Published: (2024) -
GloTSFormer: Global Video Text Spotting Transformer
by: Wang, Han, et al.
Published: (2024) -
Cross-Axis Feature Fusion with Joint-Wise Motion Difference Prediction for Text-Based 3D Human Motion Editing
by: Han, Gyojin, et al.
Published: (2026)