Layout-Aware OCR for Black Digital Archives with Unsupervised Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Beyene, Fitsum Sileshi, Dancy, Christopher L. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Survey of OCR Evaluation Methods and Metrics and the Invisibility of Historical Documents
by: Beyene, Fitsum Sileshi, et al.
Published: (2026)
by: Beyene, Fitsum Sileshi, et al.
Published: (2026)
Multimodal LLMs for OCR, OCR Post-Correction, and Named Entity Recognition in Historical Documents
by: Greif, Gavin, et al.
Published: (2025)
by: Greif, Gavin, et al.
Published: (2025)
Web Archives Metadata Generation with GPT-4o: Challenges and Insights
by: Nair, Ashwin, et al.
Published: (2024)
by: Nair, Ashwin, et al.
Published: (2024)
ASTRA: Mapping Art-Technology Institutions via Conceptual Axes, Text Embeddings, and Unsupervised Clustering
by: Bae, Joonhyung
Published: (2026)
by: Bae, Joonhyung
Published: (2026)
AI Literacy in UAE Libraries: Assessing Competencies, Training Needs, and Ethical Considerations for the Digital Age
by: Khan, Zafar Imam
Published: (2025)
by: Khan, Zafar Imam
Published: (2025)
An Axiomatic Benchmark for Evaluation of Scientific Novelty Metrics
by: Liu, Miri, et al.
Published: (2026)
by: Liu, Miri, et al.
Published: (2026)
Evaluating the quality of published medical research with ChatGPT
by: Thelwall, Mike, et al.
Published: (2024)
by: Thelwall, Mike, et al.
Published: (2024)
AI Blob! LLM-Driven Recontextualization of Italian Television Archives
by: Balestri, Roberto
Published: (2025)
by: Balestri, Roberto
Published: (2025)
Can Smaller Large Language Models Evaluate Research Quality?
by: Thelwall, Mike
Published: (2025)
by: Thelwall, Mike
Published: (2025)
Evolving Roles of LLMs in Scientific Innovation: Assistant, Collaborator, Scientist, and Evaluator
by: Zhang, Haoxuan, et al.
Published: (2025)
by: Zhang, Haoxuan, et al.
Published: (2025)
Evaluating Research Quality with Large Language Models: An Analysis of ChatGPT's Effectiveness with Different Settings and Inputs
by: Thelwall, Mike
Published: (2024)
by: Thelwall, Mike
Published: (2024)
Authority Signals in AI Cited Health Sources: A Framework for Evaluating Source Credibility in ChatGPT Responses
by: Jacques, Erin, et al.
Published: (2026)
by: Jacques, Erin, et al.
Published: (2026)
Comparing OCR Pipelines for Folkloristic Text Digitization
by: Machidon, Octavian M., et al.
Published: (2025)
by: Machidon, Octavian M., et al.
Published: (2025)
Optical Music Recognition in Manuscripts from the Ricordi Archive
by: Simonetta, Federico, et al.
Published: (2024)
by: Simonetta, Federico, et al.
Published: (2024)
Towards Large Language Models for Lunar Mission Planning and In Situ Resource Utilization
by: Pekala, Michael, et al.
Published: (2025)
by: Pekala, Michael, et al.
Published: (2025)
ArchiveGPT: A human-centered evaluation of using a vision language model for image cataloguing
by: Abele, Line, et al.
Published: (2025)
by: Abele, Line, et al.
Published: (2025)
Generative Agents Navigating Digital Libraries
by: Zerhoudi, Saber, et al.
Published: (2026)
by: Zerhoudi, Saber, et al.
Published: (2026)
Bridging the Evaluation Gap: Leveraging Large Language Models for Topic Model Evaluation
by: Tan, Zhiyin, et al.
Published: (2025)
by: Tan, Zhiyin, et al.
Published: (2025)
Causal foundations of bias, disparity and fairness
by: Traag, V. A., et al.
Published: (2022)
by: Traag, V. A., et al.
Published: (2022)
Matching Game Preferences Through Dialogical Large Language Models: A Perspective
by: Fabre, Renaud, et al.
Published: (2025)
by: Fabre, Renaud, et al.
Published: (2025)
Ontology Creation and Management Tools: the Case of Anatomical Connectivity
by: Kokash, Natallia, et al.
Published: (2025)
by: Kokash, Natallia, et al.
Published: (2025)
Can Small and Reasoning Large Language Models Score Journal Articles for Research Quality and Do Averaging and Few-shot Help?
by: Thelwall, Mike, et al.
Published: (2025)
by: Thelwall, Mike, et al.
Published: (2025)
Interpretable Link Prediction in AI-Driven Cancer Research: Uncovering Co-Authorship Patterns
by: Mosallaie, Shahab, et al.
Published: (2025)
by: Mosallaie, Shahab, et al.
Published: (2025)
The Statistical Validation of Innovation Lens
by: Radaelli, Giacomo, et al.
Published: (2025)
by: Radaelli, Giacomo, et al.
Published: (2025)
Information Ecosystem Reengineering via Public Sector Knowledge Representation
by: Bagchi, Mayukh
Published: (2025)
by: Bagchi, Mayukh
Published: (2025)
FIRESPARQL: A LLM-based Framework for SPARQL Query Generation over Scholarly Knowledge Graphs
by: Pan, Xueli, et al.
Published: (2025)
by: Pan, Xueli, et al.
Published: (2025)
In-depth Research Impact Summarization through Fine-Grained Temporal Citation Analysis
by: Arnaout, Hiba, et al.
Published: (2025)
by: Arnaout, Hiba, et al.
Published: (2025)
A Domain Ontology for Modeling the Book of Purification in Islam
by: Alawwad, Hessa
Published: (2025)
by: Alawwad, Hessa
Published: (2025)
The Rapid Growth of AI Foundation Model Usage in Science
by: Trišović, Ana, et al.
Published: (2025)
by: Trišović, Ana, et al.
Published: (2025)
Animer une base de connaissance: des ontologies aux mod{è}les d'I.A. g{é}n{é}rative
by: Stockinger, Peter
Published: (2025)
by: Stockinger, Peter
Published: (2025)
The More You Automate, the Less You See: Hidden Pitfalls of AI Scientist Systems
by: Luo, Ziming, et al.
Published: (2025)
by: Luo, Ziming, et al.
Published: (2025)
Chatting with Papers: A Hybrid Approach Using LLMs and Knowledge Graphs
by: Tykhonov, Vyacheslav, et al.
Published: (2025)
by: Tykhonov, Vyacheslav, et al.
Published: (2025)
Review of Passenger Flow Modelling Approaches Based on a Bibliometric Analysis
by: Hecht, Jonathan, et al.
Published: (2025)
by: Hecht, Jonathan, et al.
Published: (2025)
Towards a Knowledge Graph for Models and Algorithms in Applied Mathematics
by: Schembera, Björn, et al.
Published: (2024)
by: Schembera, Björn, et al.
Published: (2024)
The Artificial Intelligence Disclosure (AID) Framework: An Introduction
by: Weaver, Kari D.
Published: (2024)
by: Weaver, Kari D.
Published: (2024)
PaperX: A Unified Framework for Multimodal Academic Presentation Generation with Scholar DAG
by: Yu, Tao, et al.
Published: (2026)
by: Yu, Tao, et al.
Published: (2026)
Documentation Practices of Artificial Intelligence
by: Arnold, Stefan, et al.
Published: (2024)
by: Arnold, Stefan, et al.
Published: (2024)
Can ChatGPT evaluate research quality?
by: Thelwall, Mike
Published: (2024)
by: Thelwall, Mike
Published: (2024)
Patent Value Characterization -- An Empirical Analysis of Elevator Industry Patents
by: Guan, Yuhang, et al.
Published: (2024)
by: Guan, Yuhang, et al.
Published: (2024)
pyBibX -- A Python Library for Bibliometric and Scientometric Analysis Powered with Artificial Intelligence Tools
by: Pereira, Valdecy, et al.
Published: (2023)
by: Pereira, Valdecy, et al.
Published: (2023)
Similar Items
-
A Survey of OCR Evaluation Methods and Metrics and the Invisibility of Historical Documents
by: Beyene, Fitsum Sileshi, et al.
Published: (2026) -
Multimodal LLMs for OCR, OCR Post-Correction, and Named Entity Recognition in Historical Documents
by: Greif, Gavin, et al.
Published: (2025) -
Web Archives Metadata Generation with GPT-4o: Challenges and Insights
by: Nair, Ashwin, et al.
Published: (2024) -
ASTRA: Mapping Art-Technology Institutions via Conceptual Axes, Text Embeddings, and Unsupervised Clustering
by: Bae, Joonhyung
Published: (2026) -
AI Literacy in UAE Libraries: Assessing Competencies, Training Needs, and Ethical Considerations for the Digital Age
by: Khan, Zafar Imam
Published: (2025)