Saved in:
| Main Authors: | Fleischhacker, David, Goederle, Wolfgang, Kern, Roman |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2401.07787 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Benchmark of State-Space Models vs. Transformers and BiLSTM-based Models for Historical Newspaper OCR
by: Agbeti-messan, Merveilles, et al.
Published: (2026)
by: Agbeti-messan, Merveilles, et al.
Published: (2026)
Towards Accessible Learning: Deep Learning-Based Potential Dysgraphia Detection and OCR for Potentially Dysgraphic Handwriting
by: D, Vydeki, et al.
Published: (2024)
by: D, Vydeki, et al.
Published: (2024)
Evaluation of Ensemble Learning Techniques for handwritten OCR Improvement
by: Preiß, Martin
Published: (2025)
by: Preiß, Martin
Published: (2025)
Improving OCR using internal document redundancy
by: Belzarena, Diego, et al.
Published: (2025)
by: Belzarena, Diego, et al.
Published: (2025)
The OCR Quest for Generalization: Learning to recognize low-resource alphabets with model editing
by: Rodríguez, Adrià Molina, et al.
Published: (2025)
by: Rodríguez, Adrià Molina, et al.
Published: (2025)
Enhancement of Bengali OCR by Specialized Models and Advanced Techniques for Diverse Document Types
by: Rabby, AKM Shahariar Azad, et al.
Published: (2024)
by: Rabby, AKM Shahariar Azad, et al.
Published: (2024)
Machine Learning in Industrial Quality Control of Glass Bottle Prints
by: Bundscherer, Maximilian, et al.
Published: (2024)
by: Bundscherer, Maximilian, et al.
Published: (2024)
SimpleOCR: Rendering Visualized Questions to Teach MLLMs to Read
by: Peng, Yibo, et al.
Published: (2026)
by: Peng, Yibo, et al.
Published: (2026)
The Character Error Vector: Decomposable errors for page-level OCR evaluation
by: Bourne, Jonathan, et al.
Published: (2026)
by: Bourne, Jonathan, et al.
Published: (2026)
Optimizing Stroke Risk Prediction: A Machine Learning Pipeline Combining ROS-Balanced Ensembles and XAI
by: Akib, A S M Ahsanul Sarkar, et al.
Published: (2025)
by: Akib, A S M Ahsanul Sarkar, et al.
Published: (2025)
GutenOCR: A Grounded Vision-Language Front-End for Documents
by: Heidenreich, Hunter, et al.
Published: (2026)
by: Heidenreich, Hunter, et al.
Published: (2026)
PubMed-OCR: PMC Open Access OCR Annotations
by: Heidenreich, Hunter, et al.
Published: (2026)
by: Heidenreich, Hunter, et al.
Published: (2026)
Benchmarking OCR Pipelines with Adaptive Enhancement for Multi-Domain Retail Bill Digitization
by: Gaikwad, Vijaysinh
Published: (2026)
by: Gaikwad, Vijaysinh
Published: (2026)
Deep Companion Learning: Enhancing Generalization Through Historical Consistency
by: Zhu, Ruizhao, et al.
Published: (2024)
by: Zhu, Ruizhao, et al.
Published: (2024)
HQColon: A Hybrid Interactive Machine Learning Pipeline for High Quality Colon Labeling and Segmentation
by: Finocchiaro, Martina, et al.
Published: (2025)
by: Finocchiaro, Martina, et al.
Published: (2025)
A Low-Cost Machine Learning Approach for Timber Diameter Estimation
by: Fard, Fatemeh Hasanzadeh, et al.
Published: (2025)
by: Fard, Fatemeh Hasanzadeh, et al.
Published: (2025)
Machine Learning Approaches on Crop Pattern Recognition a Comparative Analysis
by: Kabir, Kazi Hasibul, et al.
Published: (2024)
by: Kabir, Kazi Hasibul, et al.
Published: (2024)
Using Machine Learning for move sequence visualization and generation in climbing
by: Rimbot, Thomas, et al.
Published: (2025)
by: Rimbot, Thomas, et al.
Published: (2025)
On Using Quasirandom Sequences in Machine Learning for Model Weight Initialization
by: Miranskyy, Andriy, et al.
Published: (2024)
by: Miranskyy, Andriy, et al.
Published: (2024)
Bengali Document Layout Analysis -- A YOLOV8 Based Ensembling Approach
by: Ahmed, Nazmus Sakib, et al.
Published: (2023)
by: Ahmed, Nazmus Sakib, et al.
Published: (2023)
Improving MLLM Historical Record Extraction with Test-Time Image
by: Archibald, Taylor, et al.
Published: (2025)
by: Archibald, Taylor, et al.
Published: (2025)
Unlocking the Archives: Using Large Language Models to Transcribe Handwritten Historical Documents
by: Humphries, Mark, et al.
Published: (2024)
by: Humphries, Mark, et al.
Published: (2024)
Diabetic Retinopathy Classification from Retinal Images using Machine Learning Approaches
by: Bhattacharjee, Indronil, et al.
Published: (2024)
by: Bhattacharjee, Indronil, et al.
Published: (2024)
Optimizing Urban Critical Green Space Development Using Machine Learning
by: Ganjirad, Mohammad, et al.
Published: (2025)
by: Ganjirad, Mohammad, et al.
Published: (2025)
Deep Learning-Based Approach for Identification of Potato Leaf Diseases Using Wrapper Feature Selection and Feature Concatenation
by: Naeem, Muhammad Ahtsam, et al.
Published: (2025)
by: Naeem, Muhammad Ahtsam, et al.
Published: (2025)
Improved Classification of Nitrogen Stress Severity in Plants Under Combined Stress Conditions Using Spatio-Temporal Deep Learning Framework
by: Patra, Aswini Kumar, et al.
Published: (2025)
by: Patra, Aswini Kumar, et al.
Published: (2025)
Does Combining Parameter-efficient Modules Improve Few-shot Transfer Accuracy?
by: Asadi, Nader, et al.
Published: (2024)
by: Asadi, Nader, et al.
Published: (2024)
Improving OCR for Historical Texts of Multiple Languages
by: Westerdijk, Hylke, et al.
Published: (2025)
by: Westerdijk, Hylke, et al.
Published: (2025)
Few-Shot Connectivity-Aware Text Line Segmentation in Historical Documents
by: Sterzinger, Rafael, et al.
Published: (2025)
by: Sterzinger, Rafael, et al.
Published: (2025)
Seeing Justice Clearly: Handwritten Legal Document Translation with OCR and Vision-Language Models
by: Nigam, Shubham Kumar, et al.
Published: (2025)
by: Nigam, Shubham Kumar, et al.
Published: (2025)
Context Sensitivity Improves Human-Machine Visual Alignment
by: Born, Frieda, et al.
Published: (2026)
by: Born, Frieda, et al.
Published: (2026)
Quality at the Tail of Machine Learning Inference
by: Yang, Zhengxin, et al.
Published: (2022)
by: Yang, Zhengxin, et al.
Published: (2022)
Multimodal Machine Learning in Image-Based and Clinical Biomedicine: Survey and Prospects
by: Warner, Elisa, et al.
Published: (2023)
by: Warner, Elisa, et al.
Published: (2023)
Handwritten Text Recognition of Historical Manuscripts Using Transformer-Based Models
by: Meoded, Erez
Published: (2025)
by: Meoded, Erez
Published: (2025)
Learning Vision-Based Omnidirectional Navigation: A Teacher-Student Approach Using Monocular Depth Estimation
by: Finke, Jan, et al.
Published: (2026)
by: Finke, Jan, et al.
Published: (2026)
Improving Prediction Accuracy of Semantic Segmentation Methods Using Convolutional Autoencoder Based Pre-processing Layers
by: Shimodaira, Hisashi
Published: (2024)
by: Shimodaira, Hisashi
Published: (2024)
Comparative Analysis of Machine Learning Approaches for Bone Age Assessment: A Comprehensive Study on Three Distinct Models
by: R., Nandavardhan, et al.
Published: (2024)
by: R., Nandavardhan, et al.
Published: (2024)
Using Skew to Assess the Quality of GAN-generated Image Features
by: Luzi, Lorenzo, et al.
Published: (2023)
by: Luzi, Lorenzo, et al.
Published: (2023)
Combining Radiomics and Machine Learning Approaches for Objective ASD Diagnosis: Verifying White Matter Associations with ASD
by: Song, Junlin, et al.
Published: (2024)
by: Song, Junlin, et al.
Published: (2024)
Mathematical Morphology in Machine Learning
by: Rodrigues, Erick Oliveira, et al.
Published: (2026)
by: Rodrigues, Erick Oliveira, et al.
Published: (2026)
Similar Items
-
A Benchmark of State-Space Models vs. Transformers and BiLSTM-based Models for Historical Newspaper OCR
by: Agbeti-messan, Merveilles, et al.
Published: (2026) -
Towards Accessible Learning: Deep Learning-Based Potential Dysgraphia Detection and OCR for Potentially Dysgraphic Handwriting
by: D, Vydeki, et al.
Published: (2024) -
Evaluation of Ensemble Learning Techniques for handwritten OCR Improvement
by: Preiß, Martin
Published: (2025) -
Improving OCR using internal document redundancy
by: Belzarena, Diego, et al.
Published: (2025) -
The OCR Quest for Generalization: Learning to recognize low-resource alphabets with model editing
by: Rodríguez, Adrià Molina, et al.
Published: (2025)