Masked Self-Supervised Pre-Training for Text Recognition Transformers on Large-Scale Datasets
Fuente:
arXiv
Salvato in:
| Autori principali: | Kišš, Martin, Hradiš, Michal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Self-supervised Pre-training of Text Recognizers
di: Kišš, Martin, et al.
Pubblicazione: (2024)
di: Kišš, Martin, et al.
Pubblicazione: (2024)
AnnoPage Dataset: Dataset of Non-Textual Elements in Documents with Fine-Grained Categorization
di: Kišš, Martin, et al.
Pubblicazione: (2025)
di: Kišš, Martin, et al.
Pubblicazione: (2025)
BiblioPage: A Dataset of Scanned Title Pages for Bibliographic Metadata Extraction
di: Kohút, Jan, et al.
Pubblicazione: (2025)
di: Kohút, Jan, et al.
Pubblicazione: (2025)
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
di: Cazenavette, George, et al.
Pubblicazione: (2025)
di: Cazenavette, George, et al.
Pubblicazione: (2025)
DiNO-Diffusion. Scaling Medical Diffusion via Self-Supervised Pre-Training
di: Jimenez-Perez, Guillermo, et al.
Pubblicazione: (2024)
di: Jimenez-Perez, Guillermo, et al.
Pubblicazione: (2024)
Scale Efficient Training for Large Datasets
di: Zhou, Qing, et al.
Pubblicazione: (2025)
di: Zhou, Qing, et al.
Pubblicazione: (2025)
Bridging Diversity and Uncertainty in Active learning with Self-Supervised Pre-Training
di: Doucet, Paul, et al.
Pubblicazione: (2024)
di: Doucet, Paul, et al.
Pubblicazione: (2024)
MaskOpt: A Large-Scale Mask Optimization Dataset to Advance AI in Integrated Circuit Manufacturing
di: Hu, Yuting, et al.
Pubblicazione: (2025)
di: Hu, Yuting, et al.
Pubblicazione: (2025)
Towards Writing Style Adaptation in Handwriting Recognition
di: Kohút, Jan, et al.
Pubblicazione: (2023)
di: Kohút, Jan, et al.
Pubblicazione: (2023)
Fast Training of Diffusion Models with Masked Transformers
di: Zheng, Hongkai, et al.
Pubblicazione: (2023)
di: Zheng, Hongkai, et al.
Pubblicazione: (2023)
Advancing Comprehensive Aesthetic Insight with Multi-Scale Text-Guided Self-Supervised Learning
di: Liu, Yuti, et al.
Pubblicazione: (2024)
di: Liu, Yuti, et al.
Pubblicazione: (2024)
Integration of Self-Supervised BYOL in Semi-Supervised Medical Image Recognition
di: Feng, Hao, et al.
Pubblicazione: (2024)
di: Feng, Hao, et al.
Pubblicazione: (2024)
Masked Generative Nested Transformers with Decode Time Scaling
di: Goyal, Sahil, et al.
Pubblicazione: (2025)
di: Goyal, Sahil, et al.
Pubblicazione: (2025)
Gradient-Sign Masking for Task Vector Transport Across Pre-Trained Models
di: Rinaldi, Filippo, et al.
Pubblicazione: (2025)
di: Rinaldi, Filippo, et al.
Pubblicazione: (2025)
Blockwise Self-Supervised Learning at Scale
di: Siddiqui, Shoaib Ahmed, et al.
Pubblicazione: (2023)
di: Siddiqui, Shoaib Ahmed, et al.
Pubblicazione: (2023)
Pre-training Vision Transformers with Formula-driven Supervised Learning
di: Kataoka, Hirokatsu, et al.
Pubblicazione: (2022)
di: Kataoka, Hirokatsu, et al.
Pubblicazione: (2022)
Semi-Supervised Masked Autoencoders: Unlocking Vision Transformer Potential with Limited Data
di: Faysal, Atik, et al.
Pubblicazione: (2026)
di: Faysal, Atik, et al.
Pubblicazione: (2026)
D$^3$epth: Self-Supervised Depth Estimation with Dynamic Mask in Dynamic Scenes
di: Chen, Siyu, et al.
Pubblicazione: (2024)
di: Chen, Siyu, et al.
Pubblicazione: (2024)
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation
di: Kang, Wonjun, et al.
Pubblicazione: (2025)
di: Kang, Wonjun, et al.
Pubblicazione: (2025)
DohaScript: A Large-Scale Multi-Writer Dataset for Continuous Handwritten Hindi Text
di: Singh, Kunwar Arpit, et al.
Pubblicazione: (2026)
di: Singh, Kunwar Arpit, et al.
Pubblicazione: (2026)
An Empirical Study into Clustering of Unseen Datasets with Self-Supervised Encoders
di: Lowe, Scott C., et al.
Pubblicazione: (2024)
di: Lowe, Scott C., et al.
Pubblicazione: (2024)
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
di: Chin, Zhi-Yi, et al.
Pubblicazione: (2023)
di: Chin, Zhi-Yi, et al.
Pubblicazione: (2023)
Intelligent Anomaly Detection for Lane Rendering Using Transformer with Self-Supervised Pre-Training and Customized Fine-Tuning
di: Dong, Yongqi, et al.
Pubblicazione: (2023)
di: Dong, Yongqi, et al.
Pubblicazione: (2023)
A Survey of the Self Supervised Learning Mechanisms for Vision Transformers
di: Khan, Asifullah, et al.
Pubblicazione: (2024)
di: Khan, Asifullah, et al.
Pubblicazione: (2024)
Look Through Masks: Towards Masked Face Recognition with De-Occlusion Distillation
di: Li, Chenyu, et al.
Pubblicazione: (2024)
di: Li, Chenyu, et al.
Pubblicazione: (2024)
Masked Face Recognition with Generative-to-Discriminative Representations
di: Ge, Shiming, et al.
Pubblicazione: (2024)
di: Ge, Shiming, et al.
Pubblicazione: (2024)
Beyond Labels: A Self-Supervised Framework with Masked Autoencoders and Random Cropping for Breast Cancer Subtype Classification
di: Chiocchetti, Annalisa, et al.
Pubblicazione: (2024)
di: Chiocchetti, Annalisa, et al.
Pubblicazione: (2024)
Erasing Self-Supervised Learning Backdoor by Cluster Activation Masking
di: Qian, Shengsheng, et al.
Pubblicazione: (2023)
di: Qian, Shengsheng, et al.
Pubblicazione: (2023)
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
di: Irshad, Muhammad Zubair, et al.
Pubblicazione: (2024)
di: Irshad, Muhammad Zubair, et al.
Pubblicazione: (2024)
EUDA: An Efficient Unsupervised Domain Adaptation via Self-Supervised Vision Transformer
di: Abedi, Ali, et al.
Pubblicazione: (2024)
di: Abedi, Ali, et al.
Pubblicazione: (2024)
OCT-SelfNet: A Self-Supervised Framework with Multi-Modal Datasets for Generalized and Robust Retinal Disease Detection
di: Jannat, Fatema-E, et al.
Pubblicazione: (2024)
di: Jannat, Fatema-E, et al.
Pubblicazione: (2024)
SegGen: Supercharging Segmentation Models with Text2Mask and Mask2Img Synthesis
di: Ye, Hanrong, et al.
Pubblicazione: (2023)
di: Ye, Hanrong, et al.
Pubblicazione: (2023)
Training-Only Heterogeneous Image-Patch-Text Graph Supervision for Advancing Few-Shot Learning Adapters
di: Mohammad, Mohammed Rahman Sherif Khan, et al.
Pubblicazione: (2026)
di: Mohammad, Mohammed Rahman Sherif Khan, et al.
Pubblicazione: (2026)
Kaputt: A Large-Scale Dataset for Visual Defect Detection
di: Höfer, Sebastian, et al.
Pubblicazione: (2025)
di: Höfer, Sebastian, et al.
Pubblicazione: (2025)
Soft Label Pruning and Quantization for Large-Scale Dataset Distillation
di: Lingao, Xiao, et al.
Pubblicazione: (2026)
di: Lingao, Xiao, et al.
Pubblicazione: (2026)
Diversify, Don't Fine-Tune: Scaling Up Visual Recognition Training with Synthetic Images
di: Yu, Zhuoran, et al.
Pubblicazione: (2023)
di: Yu, Zhuoran, et al.
Pubblicazione: (2023)
Transformer-Based Self-Supervised Learning for Histopathological Classification of Ischemic Stroke Clot Origin
di: Yeh, K., et al.
Pubblicazione: (2024)
di: Yeh, K., et al.
Pubblicazione: (2024)
Robust Pre-Training of Medical Vision-and-Language Models with Domain-Invariant Multi-Modal Masked Reconstruction
di: Filvantorkaman, Melika, et al.
Pubblicazione: (2026)
di: Filvantorkaman, Melika, et al.
Pubblicazione: (2026)
DeiT-LT Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasets
di: Rangwani, Harsh, et al.
Pubblicazione: (2024)
di: Rangwani, Harsh, et al.
Pubblicazione: (2024)
Open-Vocabulary Panoptic Segmentation Using BERT Pre-Training of Vision-Language Multiway Transformer Model
di: Chen, Yi-Chia, et al.
Pubblicazione: (2024)
di: Chen, Yi-Chia, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Self-supervised Pre-training of Text Recognizers
di: Kišš, Martin, et al.
Pubblicazione: (2024) -
AnnoPage Dataset: Dataset of Non-Textual Elements in Documents with Fine-Grained Categorization
di: Kišš, Martin, et al.
Pubblicazione: (2025) -
BiblioPage: A Dataset of Scanned Title Pages for Bibliographic Metadata Extraction
di: Kohút, Jan, et al.
Pubblicazione: (2025) -
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
di: Cazenavette, George, et al.
Pubblicazione: (2025) -
DiNO-Diffusion. Scaling Medical Diffusion via Self-Supervised Pre-Training
di: Jimenez-Perez, Guillermo, et al.
Pubblicazione: (2024)