Diving into the Depths of Spotting Text in Multi-Domain Noisy Scenes
Fuente:
arXiv
Salvato in:
| Autori principali: | Das, Alloy, Biswas, Sanket, Pal, Umapada, Lladós, Josep |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
di: Das, Alloy, et al.
Pubblicazione: (2024)
di: Das, Alloy, et al.
Pubblicazione: (2024)
FASTER: A Font-Agnostic Scene Text Editing and Rendering Framework
di: Das, Alloy, et al.
Pubblicazione: (2023)
di: Das, Alloy, et al.
Pubblicazione: (2023)
GraphKD: Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation
di: Banerjee, Ayan, et al.
Pubblicazione: (2024)
di: Banerjee, Ayan, et al.
Pubblicazione: (2024)
DocRevive: A Unified Pipeline for Document Text Restoration
di: Purkayastha, Kunal, et al.
Pubblicazione: (2026)
di: Purkayastha, Kunal, et al.
Pubblicazione: (2026)
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)
SketchGPT: Autoregressive Modeling for Sketch Generation and Recognition
di: Tiwari, Adarsh, et al.
Pubblicazione: (2024)
di: Tiwari, Adarsh, et al.
Pubblicazione: (2024)
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
di: Banerjee, Ayan, et al.
Pubblicazione: (2024)
di: Banerjee, Ayan, et al.
Pubblicazione: (2024)
GeoContrastNet: Contrastive Key-Value Edge Learning for Language-Agnostic Document Understanding
di: Biescas, Nil, et al.
Pubblicazione: (2024)
di: Biescas, Nil, et al.
Pubblicazione: (2024)
Towards Generative Class Prompt Learning for Fine-grained Visual Recognition
di: Chattopadhyay, Soumitri, et al.
Pubblicazione: (2024)
di: Chattopadhyay, Soumitri, et al.
Pubblicazione: (2024)
NoTeS-Bank: Benchmarking Neural Transcription and Search for Scientific Notes Understanding
di: Pal, Aniket, et al.
Pubblicazione: (2025)
di: Pal, Aniket, et al.
Pubblicazione: (2025)
LayeredDoc: Domain Adaptive Document Restoration with a Layer Separation Approach
di: Pilligua, Maria, et al.
Pubblicazione: (2024)
di: Pilligua, Maria, et al.
Pubblicazione: (2024)
STEFANN: Scene Text Editor using Font Adaptive Neural Network
di: Roy, Prasun, et al.
Pubblicazione: (2019)
di: Roy, Prasun, et al.
Pubblicazione: (2019)
GlobalDoc: A Cross-Modal Vision-Language Framework for Real-World Document Image Retrieval and Classification
di: Bakkali, Souhail, et al.
Pubblicazione: (2023)
di: Bakkali, Souhail, et al.
Pubblicazione: (2023)
A Lightweight Context-Driven Training-Free Network for Scene Text Segmentation and Recognition
di: Chakraborty, Ritabrata, et al.
Pubblicazione: (2025)
di: Chakraborty, Ritabrata, et al.
Pubblicazione: (2025)
CraftGraffiti: Exploring Human Identity with Custom Graffiti Art via Facial-Preserving Diffusion Models
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)
Scene Aware Person Image Generation through Global Contextual Conditioning
di: Roy, Prasun, et al.
Pubblicazione: (2022)
di: Roy, Prasun, et al.
Pubblicazione: (2022)
Correlation Weighted Prototype-based Self-Supervised One-Shot Segmentation of Medical Images
di: Manna, Siladittya, et al.
Pubblicazione: (2024)
di: Manna, Siladittya, et al.
Pubblicazione: (2024)
TIPS: Text-Induced Pose Synthesis
di: Roy, Prasun, et al.
Pubblicazione: (2022)
di: Roy, Prasun, et al.
Pubblicazione: (2022)
Towards Robust Cross-Dataset Object Detection Generalization under Domain Specificity
di: Chakraborty, Ritabrata, et al.
Pubblicazione: (2026)
di: Chakraborty, Ritabrata, et al.
Pubblicazione: (2026)
MDIW-13: a New Multi-Lingual and Multi-Script Database and Benchmark for Script Identification
di: Ferrer, Miguel A., et al.
Pubblicazione: (2024)
di: Ferrer, Miguel A., et al.
Pubblicazione: (2024)
Multi-scale Attention Guided Pose Transfer
di: Roy, Prasun, et al.
Pubblicazione: (2022)
di: Roy, Prasun, et al.
Pubblicazione: (2022)
Hear the Scene: Audio-Enhanced Text Spotting
di: Li, Jing, et al.
Pubblicazione: (2024)
di: Li, Jing, et al.
Pubblicazione: (2024)
DistilDoc: Knowledge Distillation for Visually-Rich Document Applications
di: Van Landeghem, Jordy, et al.
Pubblicazione: (2024)
di: Van Landeghem, Jordy, et al.
Pubblicazione: (2024)
Efficiently Leveraging Linguistic Priors for Scene Text Spotting
di: Nguyen, Nguyen, et al.
Pubblicazione: (2024)
di: Nguyen, Nguyen, et al.
Pubblicazione: (2024)
DocSynthv2: A Practical Autoregressive Modeling for Document Generation
di: Biswas, Sanket, et al.
Pubblicazione: (2024)
di: Biswas, Sanket, et al.
Pubblicazione: (2024)
A CNN Based Framework for Unistroke Numeral Recognition in Air-Writing
di: Roy, Prasun, et al.
Pubblicazione: (2023)
di: Roy, Prasun, et al.
Pubblicazione: (2023)
DRG-Font: Dynamic Reference-Guided Few-shot Font Generation via Contrastive Style-Content Disentanglement
di: Chakraborty, Rejoy, et al.
Pubblicazione: (2026)
di: Chakraborty, Rejoy, et al.
Pubblicazione: (2026)
MIO : Mutual Information Optimization using Self-Supervised Binary Contrastive Learning
di: Manna, Siladittya, et al.
Pubblicazione: (2021)
di: Manna, Siladittya, et al.
Pubblicazione: (2021)
SwinTextSpotter v2: Towards Better Synergy for Scene Text Spotting
di: Huang, Mingxin, et al.
Pubblicazione: (2024)
di: Huang, Mingxin, et al.
Pubblicazione: (2024)
Fetch-A-Set: A Large-Scale OCR-Free Benchmark for Historical Document Retrieval
di: Molina, Adrià, et al.
Pubblicazione: (2024)
di: Molina, Adrià, et al.
Pubblicazione: (2024)
Reliability-Aware Weighted Multi-Scale Spatio-Temporal Maps for Heart Rate Monitoring
di: Bairagi, Arpan, et al.
Pubblicazione: (2026)
di: Bairagi, Arpan, et al.
Pubblicazione: (2026)
InstructOCR: Instruction Boosting Scene Text Spotting
di: Duan, Chen, et al.
Pubblicazione: (2024)
di: Duan, Chen, et al.
Pubblicazione: (2024)
The OCR Quest for Generalization: Learning to recognize low-resource alphabets with model editing
di: Rodríguez, Adrià Molina, et al.
Pubblicazione: (2025)
di: Rodríguez, Adrià Molina, et al.
Pubblicazione: (2025)
ODM: A Text-Image Further Alignment Pre-training Approach for Scene Text Detection and Spotting
di: Duan, Chen, et al.
Pubblicazione: (2024)
di: Duan, Chen, et al.
Pubblicazione: (2024)
Decorrelation-based Self-Supervised Visual Representation Learning for Writer Identification
di: Maitra, Arkadip, et al.
Pubblicazione: (2024)
di: Maitra, Arkadip, et al.
Pubblicazione: (2024)
Conformal uncertainty quantification to evaluate predictive fairness of foundation AI model for skin lesion classes across patient demographics
di: Bhattacharyya, Swarnava, et al.
Pubblicazione: (2025)
di: Bhattacharyya, Swarnava, et al.
Pubblicazione: (2025)
Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting
di: Colombo, Antonio, et al.
Pubblicazione: (2026)
di: Colombo, Antonio, et al.
Pubblicazione: (2026)
Ensemble Learning for Vietnamese Scene Text Spotting in Urban Environments
di: Nguyen, Hieu, et al.
Pubblicazione: (2024)
di: Nguyen, Hieu, et al.
Pubblicazione: (2024)
TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification
di: Tao, Huaqi, et al.
Pubblicazione: (2025)
di: Tao, Huaqi, et al.
Pubblicazione: (2025)
A Transformer Based Handwriting Recognition System Jointly Using Online and Offline Features
di: Lodh, Ayush, et al.
Pubblicazione: (2025)
di: Lodh, Ayush, et al.
Pubblicazione: (2025)
Documenti analoghi
-
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
di: Das, Alloy, et al.
Pubblicazione: (2024) -
FASTER: A Font-Agnostic Scene Text Editing and Rendering Framework
di: Das, Alloy, et al.
Pubblicazione: (2023) -
GraphKD: Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation
di: Banerjee, Ayan, et al.
Pubblicazione: (2024) -
DocRevive: A Unified Pipeline for Document Text Restoration
di: Purkayastha, Kunal, et al.
Pubblicazione: (2026) -
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)