Digitizing Nepal's Written Heritage: A Comprehensive HTR Pipeline for Old Nepali Manuscripts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sarawgi, Anjali, Arias, Esteban Garces, Zotter, Christof |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lead Zirconate Titanate Reservoir Computing for Classification of Written and Spoken Digits
von: Buckley, Thomas, et al.
Veröffentlicht: (2026)
von: Buckley, Thomas, et al.
Veröffentlicht: (2026)
The Truncation Blind Spot: How Decoding Strategies Systematically Exclude Human-Like Token Choices
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2026)
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2026)
Decoding Decoded: Understanding Hyperparameter Effects in Open-Ended Text Generation
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2024)
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2024)
Development of Pre-Trained Transformer-based Models for the Nepali Language
von: Thapa, Prajwal, et al.
Veröffentlicht: (2024)
von: Thapa, Prajwal, et al.
Veröffentlicht: (2024)
Beyond Temperature: Hyperfitting as a Late-Stage Geometric Expansion
von: Li, Meimingwei, et al.
Veröffentlicht: (2026)
von: Li, Meimingwei, et al.
Veröffentlicht: (2026)
Domain-adaptative Continual Learning for Low-resource Tasks: Evaluation on Nepali
von: Duwal, Sharad, et al.
Veröffentlicht: (2024)
von: Duwal, Sharad, et al.
Veröffentlicht: (2024)
Mitigating Structural Noise in Low-Resource S2TT: An Optimized Cascaded Nepali-English Pipeline with Punctuation Restoration
von: Chongbang, Tangsang, et al.
Veröffentlicht: (2026)
von: Chongbang, Tangsang, et al.
Veröffentlicht: (2026)
Studies in Historical Documents from Nepal and India
von: Cubelic, Simon, et al.
Veröffentlicht: (2018)
von: Cubelic, Simon, et al.
Veröffentlicht: (2018)
Nepali Passport Question Answering: A Low-Resource Dataset for Public Service Applications
von: Begha, Funghang Limbu, et al.
Veröffentlicht: (2026)
von: Begha, Funghang Limbu, et al.
Veröffentlicht: (2026)
Benchmarking BERT-based Models for Sentence-level Topic Classification in Nepali Language
von: Karki, Nischal, et al.
Veröffentlicht: (2026)
von: Karki, Nischal, et al.
Veröffentlicht: (2026)
Towards Better Open-Ended Text Generation: A Multicriteria Evaluation Framework
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2024)
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2024)
Adaptive Contrastive Search: Uncertainty-Guided Decoding for Open-Ended Text Generation
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2024)
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2024)
HTR-ConvText: Leveraging Convolution and Textual Information for Handwritten Text Recognition
von: Truc, Pham Thach Thanh, et al.
Veröffentlicht: (2025)
von: Truc, Pham Thach Thanh, et al.
Veröffentlicht: (2025)
NepTam: A Nepali-Tamang Parallel Corpus and Baseline Machine Translation Experiments
von: Ghimire, Rupak Raj, et al.
Veröffentlicht: (2026)
von: Ghimire, Rupak Raj, et al.
Veröffentlicht: (2026)
A Statistical Case Against Empirical Human-AI Alignment
von: Rodemann, Julian, et al.
Veröffentlicht: (2025)
von: Rodemann, Julian, et al.
Veröffentlicht: (2025)
Unveiling Factors for Enhanced POS Tagging: A Study of Low-Resource Medieval Romance Languages
von: Schöffel, Matthias, et al.
Veröffentlicht: (2025)
von: Schöffel, Matthias, et al.
Veröffentlicht: (2025)
Can Perplexity Predict Fine-tuning Performance? An Investigation of Tokenization Effects on Sequential Language Models for Nepali
von: Luitel, Nishant, et al.
Veröffentlicht: (2024)
von: Luitel, Nishant, et al.
Veröffentlicht: (2024)
The Last Human-Written Paper: Agent-Native Research Artifacts
von: Liu, Jiachen, et al.
Veröffentlicht: (2026)
von: Liu, Jiachen, et al.
Veröffentlicht: (2026)
Uncovering Latent Connections in Indigenous Heritage: Semantic Pipelines for Cultural Preservation in Brazil
von: Zerkowski, Luis Vitor, et al.
Veröffentlicht: (2025)
von: Zerkowski, Luis Vitor, et al.
Veröffentlicht: (2025)
Unraveling Spatio-Temporal Foundation Models via the Pipeline Lens: A Comprehensive Review
von: Fang, Yuchen, et al.
Veröffentlicht: (2025)
von: Fang, Yuchen, et al.
Veröffentlicht: (2025)
Min-$k$ Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics
von: Ding, Yuanhao, et al.
Veröffentlicht: (2026)
von: Ding, Yuanhao, et al.
Veröffentlicht: (2026)
Early evidence of how LLMs outperform traditional systems on OCR/HTR tasks for historical records
von: Kim, Seorin, et al.
Veröffentlicht: (2025)
von: Kim, Seorin, et al.
Veröffentlicht: (2025)
Inter and Intra-Annual Spatio-Temporal Variability of Habitat Suitability for Asian Elephants in India: A Random Forest Model-based Analysis
von: Anjali, P., et al.
Veröffentlicht: (2021)
von: Anjali, P., et al.
Veröffentlicht: (2021)
Unsupervised Model Tree Heritage Recovery
von: Horwitz, Eliahu, et al.
Veröffentlicht: (2024)
von: Horwitz, Eliahu, et al.
Veröffentlicht: (2024)
A Comprehensive Dataset and Automated Pipeline for Nailfold Capillary Analysis
von: Zhao, Linxi, et al.
Veröffentlicht: (2023)
von: Zhao, Linxi, et al.
Veröffentlicht: (2023)
ALPACA -- Adaptive Learning Pipeline for Comprehensive AI
von: Torka, Simon, et al.
Veröffentlicht: (2024)
von: Torka, Simon, et al.
Veröffentlicht: (2024)
The Challenges of HTR Model Training: Feedback from the Project Donner le gout de l'archive a l'ere numerique
von: Couture, Beatrice, et al.
Veröffentlicht: (2022)
von: Couture, Beatrice, et al.
Veröffentlicht: (2022)
Function+Data Flow: A Framework to Specify Machine Learning Pipelines for Digital Twinning
von: de Conto, Eduardo, et al.
Veröffentlicht: (2024)
von: de Conto, Eduardo, et al.
Veröffentlicht: (2024)
Mathematical Derivation Graphs: A Relation Extraction Task in STEM Manuscripts
von: Prasad, Vishesh, et al.
Veröffentlicht: (2024)
von: Prasad, Vishesh, et al.
Veröffentlicht: (2024)
SCADE: Scalable Framework for Anomaly Detection in High-Performance System
von: Vinay, Vaishali, et al.
Veröffentlicht: (2024)
von: Vinay, Vaishali, et al.
Veröffentlicht: (2024)
Pragmatic Curiosity: A Unified Framework for Hybrid Learning and Optimization via Active Inference
von: Li, Yingke, et al.
Veröffentlicht: (2026)
von: Li, Yingke, et al.
Veröffentlicht: (2026)
Imbalanced Regression Pipeline Recommendation
von: Avelino, Juscimara G., et al.
Veröffentlicht: (2025)
von: Avelino, Juscimara G., et al.
Veröffentlicht: (2025)
Benchmarking OCR Pipelines with Adaptive Enhancement for Multi-Domain Retail Bill Digitization
von: Gaikwad, Vijaysinh
Veröffentlicht: (2026)
von: Gaikwad, Vijaysinh
Veröffentlicht: (2026)
Muharaf: Manuscripts of Handwritten Arabic Dataset for Cursive Text Recognition
von: Saeed, Mehreen, et al.
Veröffentlicht: (2024)
von: Saeed, Mehreen, et al.
Veröffentlicht: (2024)
PIPES: A Meta-dataset of Machine Learning Pipelines
von: Maia, Cynthia Moreira, et al.
Veröffentlicht: (2025)
von: Maia, Cynthia Moreira, et al.
Veröffentlicht: (2025)
Ultraspherical/Gegenbauer polynomials to unify 2D/3D Ambisonic directivity designs
von: Zotter, Franz
Veröffentlicht: (2024)
von: Zotter, Franz
Veröffentlicht: (2024)
Doubly Non-Central Beta Matrix Factorization for Stable Dimensionality Reduction of Bounded Support Matrix Data
von: Albert, Anjali N., et al.
Veröffentlicht: (2024)
von: Albert, Anjali N., et al.
Veröffentlicht: (2024)
Self-Reinforcing Controllable Synthesis of Rare Relational Data via Bayesian Calibration
von: Zhang, Chongsheng, et al.
Veröffentlicht: (2026)
von: Zhang, Chongsheng, et al.
Veröffentlicht: (2026)
Predicting Anemia Among Under-Five Children in Nepal Using Machine Learning and Deep Learning
von: Bastola, Deepak, et al.
Veröffentlicht: (2026)
von: Bastola, Deepak, et al.
Veröffentlicht: (2026)
Distinguishing AI-Generated and Human-Written Text Through Psycholinguistic Analysis
von: Opara, Chidimma
Veröffentlicht: (2025)
von: Opara, Chidimma
Veröffentlicht: (2025)
Ähnliche Einträge
-
Lead Zirconate Titanate Reservoir Computing for Classification of Written and Spoken Digits
von: Buckley, Thomas, et al.
Veröffentlicht: (2026) -
The Truncation Blind Spot: How Decoding Strategies Systematically Exclude Human-Like Token Choices
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2026) -
Decoding Decoded: Understanding Hyperparameter Effects in Open-Ended Text Generation
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2024) -
Development of Pre-Trained Transformer-based Models for the Nepali Language
von: Thapa, Prajwal, et al.
Veröffentlicht: (2024) -
Beyond Temperature: Hyperfitting as a Late-Stage Geometric Expansion
von: Li, Meimingwei, et al.
Veröffentlicht: (2026)