Hyper-Local Deformable Transformers for Text Spotting on Historical Maps
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Yijun, Chiang, Yao-Yi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LIGHT: Multi-Modal Text Linking on Historical Maps
by: Lin, Yijun, et al.
Published: (2025)
by: Lin, Yijun, et al.
Published: (2025)
TiCLS : Tightly Coupled Language Text Spotter
by: Jang, Leeje, et al.
Published: (2026)
by: Jang, Leeje, et al.
Published: (2026)
VGTS: Visually Guided Text Spotting for Novel Categories in Historical Manuscripts
by: Hu, Wenbo, et al.
Published: (2023)
by: Hu, Wenbo, et al.
Published: (2023)
EdgeSpotter: Multi-Scale Dense Text Spotting for Industrial Panel Monitoring
by: Fu, Changhong, et al.
Published: (2025)
by: Fu, Changhong, et al.
Published: (2025)
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
by: Das, Alloy, et al.
Published: (2024)
by: Das, Alloy, et al.
Published: (2024)
Fine-Scale Soil Mapping in Alaska with Multimodal Machine Learning
by: Lin, Yijun, et al.
Published: (2025)
by: Lin, Yijun, et al.
Published: (2025)
GloTSFormer: Global Video Text Spotting Transformer
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
DeepSolo++: Let Transformer Decoder with Explicit Points Solo for Multilingual Text Spotting
by: Ye, Maoyuan, et al.
Published: (2023)
by: Ye, Maoyuan, et al.
Published: (2023)
Hyper-Transformer for Amodal Completion
by: Gao, Jianxiong, et al.
Published: (2024)
by: Gao, Jianxiong, et al.
Published: (2024)
LOGO: Video Text Spotting with Language Collaboration and Glyph Perception Model
by: Liu, Hongen, et al.
Published: (2024)
by: Liu, Hongen, et al.
Published: (2024)
Block-level Text Spotting with LLMs
by: Bannur, Ganesh, et al.
Published: (2024)
by: Bannur, Ganesh, et al.
Published: (2024)
SwinTextSpotter v2: Towards Better Synergy for Scene Text Spotting
by: Huang, Mingxin, et al.
Published: (2024)
by: Huang, Mingxin, et al.
Published: (2024)
HyperDiT: Hyper-Connected Transformers for High-Fidelity Pixel-Space Diffusion
by: He, Yu, et al.
Published: (2026)
by: He, Yu, et al.
Published: (2026)
VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation
by: Zhuo, Wenjie, et al.
Published: (2024)
by: Zhuo, Wenjie, et al.
Published: (2024)
DIGMAPPER: A Modular System for Automated Geologic Map Digitization
by: Duan, Weiwei, et al.
Published: (2025)
by: Duan, Weiwei, et al.
Published: (2025)
Progressive Retinal Image Registration via Global and Local Deformable Transformations
by: Liu, Yepeng, et al.
Published: (2024)
by: Liu, Yepeng, et al.
Published: (2024)
SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting
by: Deng, Yicheng, et al.
Published: (2024)
by: Deng, Yicheng, et al.
Published: (2024)
Watermark Text Pattern Spotting in Document Images
by: Krubiński, Mateusz, et al.
Published: (2024)
by: Krubiński, Mateusz, et al.
Published: (2024)
Hear the Scene: Audio-Enhanced Text Spotting
by: Li, Jing, et al.
Published: (2024)
by: Li, Jing, et al.
Published: (2024)
MiTREE: Multi-input Transformer Ecoregion Encoder for Species Distribution Modelling
by: Chen, Theresa, et al.
Published: (2024)
by: Chen, Theresa, et al.
Published: (2024)
CLIDD: Cross-Layer Independent Deformable Description for Efficient and Discriminative Local Feature Representation
by: Yao, Haodi, et al.
Published: (2026)
by: Yao, Haodi, et al.
Published: (2026)
HyperPredict: Estimating Hyperparameter Effects for Instance-Specific Regularization in Deformable Image Registration
by: Shuaibu, Aisha L., et al.
Published: (2024)
by: Shuaibu, Aisha L., et al.
Published: (2024)
Parrot Captions Teach CLIP to Spot Text
by: Lin, Yiqi, et al.
Published: (2023)
by: Lin, Yiqi, et al.
Published: (2023)
Efficiently Leveraging Linguistic Priors for Scene Text Spotting
by: Nguyen, Nguyen, et al.
Published: (2024)
by: Nguyen, Nguyen, et al.
Published: (2024)
ModeT: Learning Deformable Image Registration via Motion Decomposition Transformer
by: Wang, Haiqiao, et al.
Published: (2023)
by: Wang, Haiqiao, et al.
Published: (2023)
Uni-Fusion: Universal Continuous Mapping
by: Yuan, Yijun, et al.
Published: (2023)
by: Yuan, Yijun, et al.
Published: (2023)
OmniParser: A Unified Framework for Text Spotting, Key Information Extraction and Table Recognition
by: Wan, Jianqiang, et al.
Published: (2024)
by: Wan, Jianqiang, et al.
Published: (2024)
HATFormer: Historic Handwritten Arabic Text Recognition with Transformers
by: Chan, Adrian, et al.
Published: (2024)
by: Chan, Adrian, et al.
Published: (2024)
Diving into the Depths of Spotting Text in Multi-Domain Noisy Scenes
by: Das, Alloy, et al.
Published: (2023)
by: Das, Alloy, et al.
Published: (2023)
SSR: A Generic Framework for Text-Aided Map Compression for Localization
by: Omama, Mohammad, et al.
Published: (2026)
by: Omama, Mohammad, et al.
Published: (2026)
FlexID: Training-Free Flexible Identity Injection via Intent-Aware Modulation for Text-to-Image Generation
by: Li, Guandong, et al.
Published: (2026)
by: Li, Guandong, et al.
Published: (2026)
HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion
by: Liu, Xian, et al.
Published: (2023)
by: Liu, Xian, et al.
Published: (2023)
HGFormer: Topology-Aware Vision Transformer with HyperGraph Learning
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
Mapping the Vanishing and Transformation of Urban Villages in China
by: Zhang, Wenyu, et al.
Published: (2025)
by: Zhang, Wenyu, et al.
Published: (2025)
Weak Supervision with Arbitrary Single Frame for Micro- and Macro-expression Spotting
by: Yu, Wang-Wang, et al.
Published: (2024)
by: Yu, Wang-Wang, et al.
Published: (2024)
LRANet++: Low-Rank Approximation Network for Accurate and Efficient Text Spotting
by: Su, Yuchen, et al.
Published: (2025)
by: Su, Yuchen, et al.
Published: (2025)
Bridging the Gap Between End-to-End and Two-Step Text Spotting
by: Huang, Mingxin, et al.
Published: (2024)
by: Huang, Mingxin, et al.
Published: (2024)
HAViT: Historical Attention Vision Transformer
by: Banik, Swarnendu, et al.
Published: (2026)
by: Banik, Swarnendu, et al.
Published: (2026)
Semantic Segmentation for Sequential Historical Maps by Learning from Only One Map
by: Yuan, Yunshuang, et al.
Published: (2025)
by: Yuan, Yunshuang, et al.
Published: (2025)
ODM: A Text-Image Further Alignment Pre-training Approach for Scene Text Detection and Spotting
by: Duan, Chen, et al.
Published: (2024)
by: Duan, Chen, et al.
Published: (2024)
Similar Items
-
LIGHT: Multi-Modal Text Linking on Historical Maps
by: Lin, Yijun, et al.
Published: (2025) -
TiCLS : Tightly Coupled Language Text Spotter
by: Jang, Leeje, et al.
Published: (2026) -
VGTS: Visually Guided Text Spotting for Novel Categories in Historical Manuscripts
by: Hu, Wenbo, et al.
Published: (2023) -
EdgeSpotter: Multi-Scale Dense Text Spotting for Industrial Panel Monitoring
by: Fu, Changhong, et al.
Published: (2025) -
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
by: Das, Alloy, et al.
Published: (2024)