Focus Entirety and Perceive Environment for Arbitrary-Shaped Text Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Han, Xu, Gao, Junyu, Yang, Chuang, Yuan, Yuan, Wang, Qi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Real-Time Text Detection with Similar Mask in Traffic, Industrial, and Natural Scenes
by: Han, Xu, et al.
Published: (2024)
by: Han, Xu, et al.
Published: (2024)
Spotlight Text Detector: Spotlight on Candidate Regions Like a Camera
by: Han, Xu, et al.
Published: (2024)
by: Han, Xu, et al.
Published: (2024)
Text-Pass Filter: An Efficient Scene Text Detector
by: Yang, Chuang, et al.
Published: (2026)
by: Yang, Chuang, et al.
Published: (2026)
SamLP: A Customized Segment Anything Model for License Plate Detection
by: Ding, Haoxuan, et al.
Published: (2024)
by: Ding, Haoxuan, et al.
Published: (2024)
BPDO:Boundary Points Dynamic Optimization for Arbitrary Shape Scene Text Detection
by: Zheng, Jinzhi, et al.
Published: (2024)
by: Zheng, Jinzhi, et al.
Published: (2024)
Monge-Ampere Regularization for Learning Arbitrary Shapes from Point Clouds
by: Yang, Chuanxiang, et al.
Published: (2024)
by: Yang, Chuanxiang, et al.
Published: (2024)
Edge Approximation Text Detector
by: Yang, Chuang, et al.
Published: (2025)
by: Yang, Chuang, et al.
Published: (2025)
FocusDiffuser: Perceiving Local Disparities for Camouflaged Object Detection
by: Zhao, Jianwei, et al.
Published: (2024)
by: Zhao, Jianwei, et al.
Published: (2024)
Focus-to-Perceive Representation Learning: A Cognition-Inspired Hierarchical Framework for Endoscopic Video Analysis
by: Zhang, Yuan, et al.
Published: (2026)
by: Zhang, Yuan, et al.
Published: (2026)
MorphText: Deep Morphology Regularized Arbitrary-shape Scene Text Detection
by: Xu, Chengpei, et al.
Published: (2024)
by: Xu, Chengpei, et al.
Published: (2024)
SparseFlex: High-Resolution and Arbitrary-Topology 3D Shape Modeling
by: He, Xianglong, et al.
Published: (2025)
by: He, Xianglong, et al.
Published: (2025)
NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments
by: Yao, Xuan, et al.
Published: (2025)
by: Yao, Xuan, et al.
Published: (2025)
SignEye: Traffic Sign Interpretation from Vehicle First-Person View
by: Yang, Chuang, et al.
Published: (2024)
by: Yang, Chuang, et al.
Published: (2024)
STAR-IOD: Scale-decoupled Topology Alignment with Pseudo-label Refinement for Remote Sensing Incremental Object Detection
by: Zhang, Yaoteng, et al.
Published: (2026)
by: Zhang, Yaoteng, et al.
Published: (2026)
Salient Object Detection From Arbitrary Modalities
by: Huang, Nianchang, et al.
Published: (2024)
by: Huang, Nianchang, et al.
Published: (2024)
Modality Prompts for Arbitrary Modality Salient Object Detection
by: Huang, Nianchang, et al.
Published: (2024)
by: Huang, Nianchang, et al.
Published: (2024)
Particle-Based Shape Modeling for Arbitrary Regions-of-Interest
by: Xu, Hong, et al.
Published: (2023)
by: Xu, Hong, et al.
Published: (2023)
GoMatching++: Parameter- and Data-Efficient Arbitrary-Shaped Video Text Spotting and Benchmarking
by: He, Haibin, et al.
Published: (2025)
by: He, Haibin, et al.
Published: (2025)
Arbitrary Generative Video Interpolation
by: Zhang, Guozhen, et al.
Published: (2025)
by: Zhang, Guozhen, et al.
Published: (2025)
Quantum-inspired Interpretable Deep Learning Architecture for Text Sentiment Analysis
by: Li, Bingyu, et al.
Published: (2024)
by: Li, Bingyu, et al.
Published: (2024)
DNTextSpotter: Arbitrary-Shaped Scene Text Spotting via Improved Denoising Training
by: Xie, Yu, et al.
Published: (2024)
by: Xie, Yu, et al.
Published: (2024)
RRNet: Configurable Real-Time Video Enhancement with Arbitrary Local Lighting Variations
by: Yang, Wenlong, et al.
Published: (2026)
by: Yang, Wenlong, et al.
Published: (2026)
Focus, Distinguish, and Prompt: Unleashing CLIP for Efficient and Flexible Scene Text Retrieval
by: Zeng, Gangyan, et al.
Published: (2024)
by: Zeng, Gangyan, et al.
Published: (2024)
Text-only Synthesis for Image Captioning
by: Zhou, Qing, et al.
Published: (2024)
by: Zhou, Qing, et al.
Published: (2024)
You Think, You ACT: The New Task of Arbitrary Text to Motion Generation
by: Wang, Runqi, et al.
Published: (2024)
by: Wang, Runqi, et al.
Published: (2024)
Arbitrary Reading Order Scene Text Spotter with Local Semantics Guidance
by: Lyu, Jiahao, et al.
Published: (2024)
by: Lyu, Jiahao, et al.
Published: (2024)
Generative 3D Gaussian Splatting for Arbitrary-ResolutionAtmospheric Downscaling and Forecasting
by: Han, Tao, et al.
Published: (2026)
by: Han, Tao, et al.
Published: (2026)
Like Humans to Few-Shot Learning through Knowledge Permeation of Vision and Text
by: Jia, Yuyu, et al.
Published: (2024)
by: Jia, Yuyu, et al.
Published: (2024)
Beyond Prompt Degradation: Prototype-guided Dual-pool Prompting for Incremental Object Detection
by: Zhang, Yaoteng, et al.
Published: (2026)
by: Zhang, Yaoteng, et al.
Published: (2026)
UNICBench: UNIfied Counting Benchmark for MLLM
by: Rong, Chenggang, et al.
Published: (2026)
by: Rong, Chenggang, et al.
Published: (2026)
Open-Text Aerial Detection: A Unified Framework For Aerial Visual Grounding And Detection
by: Wei, Guoting, et al.
Published: (2026)
by: Wei, Guoting, et al.
Published: (2026)
Visual Semantic Description Generation with MLLMs for Image-Text Matching
by: Chen, Junyu, et al.
Published: (2025)
by: Chen, Junyu, et al.
Published: (2025)
A Training-Free Framework for Video License Plate Tracking and Recognition with Only One-Shot
by: Ding, Haoxuan, et al.
Published: (2024)
by: Ding, Haoxuan, et al.
Published: (2024)
Text-to-3D Shape Generation
by: Lee, Han-Hung, et al.
Published: (2024)
by: Lee, Han-Hung, et al.
Published: (2024)
CritiFusion: Semantic Critique and Spectral Alignment for Faithful Text-to-Image Generation
by: Chen, ZhenQi, et al.
Published: (2025)
by: Chen, ZhenQi, et al.
Published: (2025)
Towards Arbitrary Motion Completing via Hierarchical Continuous Representation
by: Xu, Chenghao, et al.
Published: (2025)
by: Xu, Chenghao, et al.
Published: (2025)
VideoUFO: A Million-Scale User-Focused Dataset for Text-to-Video Generation
by: Wang, Wenhao, et al.
Published: (2025)
by: Wang, Wenhao, et al.
Published: (2025)
Conjugated Semantic Pool Improves OOD Detection with Pre-trained Vision-Language Models
by: Chen, Mengyuan, et al.
Published: (2024)
by: Chen, Mengyuan, et al.
Published: (2024)
Image Collage on Arbitrary Shape via Shape-Aware Slicing and Optimization
by: Wu, Dong-Yi, et al.
Published: (2023)
by: Wu, Dong-Yi, et al.
Published: (2023)
Generative Region-Language Pretraining for Open-Ended Object Detection
by: Lin, Chuang, et al.
Published: (2024)
by: Lin, Chuang, et al.
Published: (2024)
Similar Items
-
Real-Time Text Detection with Similar Mask in Traffic, Industrial, and Natural Scenes
by: Han, Xu, et al.
Published: (2024) -
Spotlight Text Detector: Spotlight on Candidate Regions Like a Camera
by: Han, Xu, et al.
Published: (2024) -
Text-Pass Filter: An Efficient Scene Text Detector
by: Yang, Chuang, et al.
Published: (2026) -
SamLP: A Customized Segment Anything Model for License Plate Detection
by: Ding, Haoxuan, et al.
Published: (2024) -
BPDO:Boundary Points Dynamic Optimization for Arbitrary Shape Scene Text Detection
by: Zheng, Jinzhi, et al.
Published: (2024)