PubMed-OCR: PMC Open Access OCR Annotations
Fuente:
arXiv
Saved in:
| Main Authors: | Heidenreich, Hunter, Getachew, Yosheb, Dinica, Olivia, Elliott, Ben |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GutenOCR: A Grounded Vision-Language Front-End for Documents
by: Heidenreich, Hunter, et al.
Published: (2026)
by: Heidenreich, Hunter, et al.
Published: (2026)
Reading or Guessing? Visual Grounding Failures of Vision-Language Models for OCR in Ancient Greek Editions
by: Karamolegkou, Antonia, et al.
Published: (2026)
by: Karamolegkou, Antonia, et al.
Published: (2026)
A Survey of OCR Evaluation Methods and Metrics and the Invisibility of Historical Documents
by: Beyene, Fitsum Sileshi, et al.
Published: (2026)
by: Beyene, Fitsum Sileshi, et al.
Published: (2026)
Post-OCR Text Correction for Bulgarian Historical Documents
by: Beshirov, Angel, et al.
Published: (2024)
by: Beshirov, Angel, et al.
Published: (2024)
Unlocking the Archives: Using Large Language Models to Transcribe Handwritten Historical Documents
by: Humphries, Mark, et al.
Published: (2024)
by: Humphries, Mark, et al.
Published: (2024)
Efficient OCR for Building a Diverse Digital History
by: Carlson, Jacob, et al.
Published: (2023)
by: Carlson, Jacob, et al.
Published: (2023)
Dating ancient manuscripts using radiocarbon and AI-based writing style analysis
by: Popović, Mladen, et al.
Published: (2024)
by: Popović, Mladen, et al.
Published: (2024)
NLLG Quarterly arXiv Report 09/24: What are the most influential current AI Papers?
by: Leiter, Christoph, et al.
Published: (2024)
by: Leiter, Christoph, et al.
Published: (2024)
Callico: a Versatile Open-Source Document Image Annotation Platform
by: Kermorvant, Christopher, et al.
Published: (2024)
by: Kermorvant, Christopher, et al.
Published: (2024)
PubMed-Ophtha: An open resource for training ophthalmology vision-language models on scientific literature
by: Hallitschke, Verena Jasmin, et al.
Published: (2026)
by: Hallitschke, Verena Jasmin, et al.
Published: (2026)
[Re] Network Deconvolution
by: Obadage, Rochana R., et al.
Published: (2024)
by: Obadage, Rochana R., et al.
Published: (2024)
Object Recognition from Scientific Document based on Compartment Refinement Framework
by: Li, Jinghong, et al.
Published: (2023)
by: Li, Jinghong, et al.
Published: (2023)
Multimodal LLMs for OCR, OCR Post-Correction, and Named Entity Recognition in Historical Documents
by: Greif, Gavin, et al.
Published: (2025)
by: Greif, Gavin, et al.
Published: (2025)
An HTR-LLM Workflow for High-Accuracy Transcription and Analysis of Abbreviated Latin Court Hand
by: Isom, Joshua D.
Published: (2025)
by: Isom, Joshua D.
Published: (2025)
Copycats: the many lives of a publicly available medical imaging dataset
by: Jiménez-Sánchez, Amelia, et al.
Published: (2024)
by: Jiménez-Sánchez, Amelia, et al.
Published: (2024)
Qibitz: Mining PubMed for Repurposable Drugs
by: Massart, David, et al.
Published: (2024)
by: Massart, David, et al.
Published: (2024)
Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models
by: Xu, Qinwu, et al.
Published: (2026)
by: Xu, Qinwu, et al.
Published: (2026)
JaPOC: Japanese Post-OCR Correction Benchmark using Vouchers
by: Fujitake, Masato
Published: (2024)
by: Fujitake, Masato
Published: (2024)
CLOCR-C: Context Leveraging OCR Correction with Pre-trained Language Models
by: Bourne, Jonathan
Published: (2024)
by: Bourne, Jonathan
Published: (2024)
olmOCR 2: Unit Test Rewards for Document OCR
by: Poznanski, Jake, et al.
Published: (2025)
by: Poznanski, Jake, et al.
Published: (2025)
PMOA-TTS: Introducing the PubMed Open Access Textual Times Series Corpus
by: Noroozizadeh, Shahriar, et al.
Published: (2025)
by: Noroozizadeh, Shahriar, et al.
Published: (2025)
Improving OCR for Historical Texts of Multiple Languages
by: Westerdijk, Hylke, et al.
Published: (2025)
by: Westerdijk, Hylke, et al.
Published: (2025)
Why Stop at Words? Unveiling the Bigger Picture through Line-Level OCR
by: Vempati, Shashank, et al.
Published: (2025)
by: Vempati, Shashank, et al.
Published: (2025)
Position: AI/ML Influencers Have a Place in the Academic Process
by: Weissburg, Iain Xie, et al.
Published: (2024)
by: Weissburg, Iain Xie, et al.
Published: (2024)
Layout-Aware Text Editing for Efficient Transformation of Academic PDFs to Markdown
by: Duan, Changxu
Published: (2025)
by: Duan, Changxu
Published: (2025)
Seeing Justice Clearly: Handwritten Legal Document Translation with OCR and Vision-Language Models
by: Nigam, Shubham Kumar, et al.
Published: (2025)
by: Nigam, Shubham Kumar, et al.
Published: (2025)
LEMUR Neural Network Dataset: Towards Seamless AutoML
by: Goodarzi, Arash Torabi, et al.
Published: (2025)
by: Goodarzi, Arash Torabi, et al.
Published: (2025)
Handwritten Text Recognition of Historical Manuscripts Using Transformer-Based Models
by: Meoded, Erez
Published: (2025)
by: Meoded, Erez
Published: (2025)
Paper Copilot: Tracking the Evolution of Peer Review in AI Conferences
by: Yang, Jing, et al.
Published: (2025)
by: Yang, Jing, et al.
Published: (2025)
DharmaOCR: Specialized Small Language Models for Structured OCR that outperform Open-Source and Commercial Baselines
by: Cardoso, Gabriel Pimenta de Freitas, et al.
Published: (2026)
by: Cardoso, Gabriel Pimenta de Freitas, et al.
Published: (2026)
PreP-OCR: A Complete Pipeline for Document Image Restoration and Enhanced OCR Accuracy
by: Guan, Shuhao, et al.
Published: (2025)
by: Guan, Shuhao, et al.
Published: (2025)
GlotOCR Bench: OCR Models Still Struggle Beyond a Handful of Unicode Scripts
by: Kargaran, Amir Hossein, et al.
Published: (2026)
by: Kargaran, Amir Hossein, et al.
Published: (2026)
Detecting Latin in Historical Books with Large Language Models: A Multimodal Benchmark
by: Wu, Yu, et al.
Published: (2025)
by: Wu, Yu, et al.
Published: (2025)
TRIDIS: A Comprehensive Medieval and Early Modern Corpus for HTR and NER
by: Aguilar, Sergio Torres
Published: (2025)
by: Aguilar, Sergio Torres
Published: (2025)
AutoFigure: Generating and Refining Publication-Ready Scientific Illustrations
by: Zhu, Minjun, et al.
Published: (2026)
by: Zhu, Minjun, et al.
Published: (2026)
KazakhOCR: A Synthetic Benchmark for Evaluating Multimodal Models in Low-Resource Kazakh Script OCR
by: Gagnier, Henry, et al.
Published: (2026)
by: Gagnier, Henry, et al.
Published: (2026)
Evaluation of Ensemble Learning Techniques for handwritten OCR Improvement
by: Preiß, Martin
Published: (2025)
by: Preiß, Martin
Published: (2025)
Advances and Limitations in Open Source Arabic-Script OCR: A Case Study
by: Kiessling, Benjamin, et al.
Published: (2024)
by: Kiessling, Benjamin, et al.
Published: (2024)
KITAB-Bench: A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understanding
by: Heakl, Ahmed, et al.
Published: (2025)
by: Heakl, Ahmed, et al.
Published: (2025)
SimpleOCR: Rendering Visualized Questions to Teach MLLMs to Read
by: Peng, Yibo, et al.
Published: (2026)
by: Peng, Yibo, et al.
Published: (2026)
Similar Items
-
GutenOCR: A Grounded Vision-Language Front-End for Documents
by: Heidenreich, Hunter, et al.
Published: (2026) -
Reading or Guessing? Visual Grounding Failures of Vision-Language Models for OCR in Ancient Greek Editions
by: Karamolegkou, Antonia, et al.
Published: (2026) -
A Survey of OCR Evaluation Methods and Metrics and the Invisibility of Historical Documents
by: Beyene, Fitsum Sileshi, et al.
Published: (2026) -
Post-OCR Text Correction for Bulgarian Historical Documents
by: Beshirov, Angel, et al.
Published: (2024) -
Unlocking the Archives: Using Large Language Models to Transcribe Handwritten Historical Documents
by: Humphries, Mark, et al.
Published: (2024)