TRIDIS: A Comprehensive Medieval and Early Modern Corpus for HTR and NER
Fuente:
arXiv
Saved in:
| Main Author: | Aguilar, Sergio Torres |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An HTR-LLM Workflow for High-Accuracy Transcription and Analysis of Abbreviated Latin Court Hand
by: Isom, Joshua D.
Published: (2025)
by: Isom, Joshua D.
Published: (2025)
Detecting Latin in Historical Books with Large Language Models: A Multimodal Benchmark
by: Wu, Yu, et al.
Published: (2025)
by: Wu, Yu, et al.
Published: (2025)
Reading or Guessing? Visual Grounding Failures of Vision-Language Models for OCR in Ancient Greek Editions
by: Karamolegkou, Antonia, et al.
Published: (2026)
by: Karamolegkou, Antonia, et al.
Published: (2026)
AutoFigure: Generating and Refining Publication-Ready Scientific Illustrations
by: Zhu, Minjun, et al.
Published: (2026)
by: Zhu, Minjun, et al.
Published: (2026)
Every Part Matters: Integrity Verification of Scientific Figures Based on Multimodal Large Language Models
by: Shi, Xiang, et al.
Published: (2024)
by: Shi, Xiang, et al.
Published: (2024)
Towards Making Flowchart Images Machine Interpretable
by: Shukla, Shreya, et al.
Published: (2025)
by: Shukla, Shreya, et al.
Published: (2025)
Dating ancient manuscripts using radiocarbon and AI-based writing style analysis
by: Popović, Mladen, et al.
Published: (2024)
by: Popović, Mladen, et al.
Published: (2024)
PathoScribe: Transforming Pathology Data into a Living Library with a Unified LLM-Driven Framework for Semantic Retrieval and Clinical Integration
by: Akbar, Abdul Rehman, et al.
Published: (2026)
by: Akbar, Abdul Rehman, et al.
Published: (2026)
NLLG Quarterly arXiv Report 09/24: What are the most influential current AI Papers?
by: Leiter, Christoph, et al.
Published: (2024)
by: Leiter, Christoph, et al.
Published: (2024)
A Global Atlas of Digital Dermatology to Map Innovation and Disparities
by: Gröger, Fabian, et al.
Published: (2025)
by: Gröger, Fabian, et al.
Published: (2025)
A Literature Review of Literature Reviews in Pattern Analysis and Machine Intelligence
by: Zhao, Penghai, et al.
Published: (2024)
by: Zhao, Penghai, et al.
Published: (2024)
Optical Music Recognition in Manuscripts from the Ricordi Archive
by: Simonetta, Federico, et al.
Published: (2024)
by: Simonetta, Federico, et al.
Published: (2024)
Transfer Learning Approach for Railway Technical Map (RTM) Component Identification
by: Rumalshan, Obadage Rochana, et al.
Published: (2024)
by: Rumalshan, Obadage Rochana, et al.
Published: (2024)
Position: The Artificial Intelligence and Machine Learning Community Should Adopt a More Transparent and Regulated Peer Review Process
by: Yang, Jing
Published: (2025)
by: Yang, Jing
Published: (2025)
Position: AI/ML Influencers Have a Place in the Academic Process
by: Weissburg, Iain Xie, et al.
Published: (2024)
by: Weissburg, Iain Xie, et al.
Published: (2024)
Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus
by: Ortega, John E., et al.
Published: (2026)
by: Ortega, John E., et al.
Published: (2026)
Unfolding the Past: A Comprehensive Deep Learning Approach to Analyzing Incunabula Pages
by: Ropel, Klaudia, et al.
Published: (2025)
by: Ropel, Klaudia, et al.
Published: (2025)
LEMUR Neural Network Dataset: Towards Seamless AutoML
by: Goodarzi, Arash Torabi, et al.
Published: (2025)
by: Goodarzi, Arash Torabi, et al.
Published: (2025)
Handwritten Text Recognition of Historical Manuscripts Using Transformer-Based Models
by: Meoded, Erez
Published: (2025)
by: Meoded, Erez
Published: (2025)
Paper Copilot: Tracking the Evolution of Peer Review in AI Conferences
by: Yang, Jing, et al.
Published: (2025)
by: Yang, Jing, et al.
Published: (2025)
Iconographic Classification and Content-Based Recommendation for Digitized Artworks
by: Kutt, Krzysztof, et al.
Published: (2026)
by: Kutt, Krzysztof, et al.
Published: (2026)
Layout-Aware Text Editing for Efficient Transformation of Academic PDFs to Markdown
by: Duan, Changxu
Published: (2025)
by: Duan, Changxu
Published: (2025)
The Role of Language Models in Modern Healthcare: A Comprehensive Review
by: Khalid, Amna, et al.
Published: (2024)
by: Khalid, Amna, et al.
Published: (2024)
PubMed-OCR: PMC Open Access OCR Annotations
by: Heidenreich, Hunter, et al.
Published: (2026)
by: Heidenreich, Hunter, et al.
Published: (2026)
Unlocking the Archives: Using Large Language Models to Transcribe Handwritten Historical Documents
by: Humphries, Mark, et al.
Published: (2024)
by: Humphries, Mark, et al.
Published: (2024)
"Don't Teach Minerva": Guiding LLMs Through Complex Syntax for Faithful Latin Translation with RAG
by: Aguilar, Sergio Torres
Published: (2025)
by: Aguilar, Sergio Torres
Published: (2025)
In the Picture: Medical Imaging Datasets, Artifacts, and their Living Review
by: Jiménez-Sánchez, Amelia, et al.
Published: (2025)
by: Jiménez-Sánchez, Amelia, et al.
Published: (2025)
Style-based Composer Identification and Attribution of Symbolic Music Scores: a Systematic Survey
by: Simonetta, Federico
Published: (2025)
by: Simonetta, Federico
Published: (2025)
Automatic Modeling of Social Concepts Evoked by Art Images as Multimodal Frames
by: Pandiani, Delfina Sol Martinez, et al.
Published: (2021)
by: Pandiani, Delfina Sol Martinez, et al.
Published: (2021)
A Survey of OCR Evaluation Methods and Metrics and the Invisibility of Historical Documents
by: Beyene, Fitsum Sileshi, et al.
Published: (2026)
by: Beyene, Fitsum Sileshi, et al.
Published: (2026)
NeuroPapyri: A Deep Attention Embedding Network for Handwritten Papyri Retrieval
by: De Gregorio, Giuseppe, et al.
Published: (2024)
by: De Gregorio, Giuseppe, et al.
Published: (2024)
What Lies Beneath: A Call for Distribution-based Visual Question & Answer Datasets
by: Naiman, Jill P., et al.
Published: (2026)
by: Naiman, Jill P., et al.
Published: (2026)
WildDepth: A Multimodal Dataset for 3D Wildlife Perception and Depth Estimation
by: Aamir, Muhammad, et al.
Published: (2026)
by: Aamir, Muhammad, et al.
Published: (2026)
Evolving Thematic Map Design in Academic Cartography: A Thirty-Year Study Based on Multilingual Journals
by: Wei, Zhiwei, et al.
Published: (2026)
by: Wei, Zhiwei, et al.
Published: (2026)
Large language models for automated scholarly paper review: A survey
by: Zhuang, Zhenzhen, et al.
Published: (2025)
by: Zhuang, Zhenzhen, et al.
Published: (2025)
Hybrid Retrieval-Augmented Generation for Robust Multilingual Document Question Answering
by: Mudet, Anthony, et al.
Published: (2025)
by: Mudet, Anthony, et al.
Published: (2025)
LEGATO: Large-scale End-to-end Generalizable Approach to Typeset OMR
by: Yang, Guang, et al.
Published: (2025)
by: Yang, Guang, et al.
Published: (2025)
Quo Vadis Handwritten Text Generation for Handwritten Text Recognition?
by: Pippi, Vittorio, et al.
Published: (2025)
by: Pippi, Vittorio, et al.
Published: (2025)
Knowledge Graphs for Digitized Manuscripts in Jagiellonian Digital Library Application
by: Ignatowicz, Jan, et al.
Published: (2025)
by: Ignatowicz, Jan, et al.
Published: (2025)
Automatic Reviewers Assignment to a Research Paper Based on Allied References and Publications Weight
by: Mahmud, Tamim Al, et al.
Published: (2025)
by: Mahmud, Tamim Al, et al.
Published: (2025)
Similar Items
-
An HTR-LLM Workflow for High-Accuracy Transcription and Analysis of Abbreviated Latin Court Hand
by: Isom, Joshua D.
Published: (2025) -
Detecting Latin in Historical Books with Large Language Models: A Multimodal Benchmark
by: Wu, Yu, et al.
Published: (2025) -
Reading or Guessing? Visual Grounding Failures of Vision-Language Models for OCR in Ancient Greek Editions
by: Karamolegkou, Antonia, et al.
Published: (2026) -
AutoFigure: Generating and Refining Publication-Ready Scientific Illustrations
by: Zhu, Minjun, et al.
Published: (2026) -
Every Part Matters: Integrity Verification of Scientific Figures Based on Multimodal Large Language Models
by: Shi, Xiang, et al.
Published: (2024)