Devanagari Handwritten Character Recognition using Convolutional Neural Network
Fuente:
arXiv
Saved in:
| Main Authors: | Mehta, Diksha, Mehta, Prateek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An accurate and revised version of optical character recognition-based speech synthesis using LabVIEW
by: Mehta, Prateek, et al.
Published: (2025)
by: Mehta, Prateek, et al.
Published: (2025)
BabelDOC: Better Layout-Preserving PDF Translation via Intermediate Representation
by: Yang, Qi, et al.
Published: (2026)
by: Yang, Qi, et al.
Published: (2026)
HalalBench: A Multilingual OCR Benchmark for Food Packaging Ingredient Extraction
by: Arief, Hasan
Published: (2026)
by: Arief, Hasan
Published: (2026)
A Multimodal Pipeline for Clinical Data Extraction: Applying Vision-Language Models to Scans of Transfusion Reaction Reports
by: Schäfer, Henning, et al.
Published: (2025)
by: Schäfer, Henning, et al.
Published: (2025)
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
by: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Published: (2025)
by: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Published: (2025)
A Sociolinguistic Analysis of Automatic Speech Recognition Bias in Newcastle English
by: Serditova, Dana, et al.
Published: (2026)
by: Serditova, Dana, et al.
Published: (2026)
Qwerty AI: Explainable Automated Age Rating and Content Safety Assessment for Russian-Language Screenplays
by: Zmanovskii, Nikita
Published: (2025)
by: Zmanovskii, Nikita
Published: (2025)
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
by: Whittaker, Edward, et al.
Published: (2024)
by: Whittaker, Edward, et al.
Published: (2024)
UVDoc: Neural Grid-based Document Unwarping
by: Verhoeven, Floor, et al.
Published: (2023)
by: Verhoeven, Floor, et al.
Published: (2023)
Enhancing OCR for Sino-Vietnamese Language Processing via Fine-tuned PaddleOCRv5
by: Nguyen, Minh Hoang, et al.
Published: (2025)
by: Nguyen, Minh Hoang, et al.
Published: (2025)
Fine-Tuning Vision-Language Models for Understanding Current Damage and Scoring Priority with Quality Guard Agent
by: Yasuno, Takato
Published: (2026)
by: Yasuno, Takato
Published: (2026)
Accelerating Language Model Workflows with Prompt Choreography
by: Bai, TJ, et al.
Published: (2025)
by: Bai, TJ, et al.
Published: (2025)
Document Understanding for Healthcare Referrals
by: Mistry, Jimit, et al.
Published: (2023)
by: Mistry, Jimit, et al.
Published: (2023)
Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis
by: Heyne, Catyana, et al.
Published: (2026)
by: Heyne, Catyana, et al.
Published: (2026)
Language Predicts Identity Fusion Across Cultures and Reveals Divergent Pathways to Violence
by: Wright, Devin R., et al.
Published: (2026)
by: Wright, Devin R., et al.
Published: (2026)
Proximity QA: Unleashing the Power of Multi-Modal Large Language Models for Spatial Proximity Analysis
by: Li, Jianing, et al.
Published: (2024)
by: Li, Jianing, et al.
Published: (2024)
Cytoarchitecture in Words: Weakly Supervised Vision-Language Modeling for Human Brain Microscopy
by: Sutton, Matthew, et al.
Published: (2026)
by: Sutton, Matthew, et al.
Published: (2026)
GAEA: A Geolocation Aware Conversational Assistant
by: Campos, Ron, et al.
Published: (2025)
by: Campos, Ron, et al.
Published: (2025)
Chronic pain patient narratives allow for the estimation of current pain intensity
by: Nunes, Diogo A. P., et al.
Published: (2022)
by: Nunes, Diogo A. P., et al.
Published: (2022)
The Table of Media Bias Elements: A sentence-level taxonomy of media bias types and propaganda techniques
by: Menzner, Tim, et al.
Published: (2026)
by: Menzner, Tim, et al.
Published: (2026)
Pipeline and Dataset Generation for Automated Fact-checking in Almost Any Language
by: Drchal, Jan, et al.
Published: (2023)
by: Drchal, Jan, et al.
Published: (2023)
Automating Clinical Information Retrieval from Finnish Electronic Health Records Using Large Language Models
by: Saukkoriipi, Mikko, et al.
Published: (2026)
by: Saukkoriipi, Mikko, et al.
Published: (2026)
On the development of an AI performance and behavioural measures for teaching and classroom management
by: Niculescu, Andreea I., et al.
Published: (2025)
by: Niculescu, Andreea I., et al.
Published: (2025)
FS-DAG: Few Shot Domain Adapting Graph Networks for Visually Rich Document Understanding
by: Agarwal, Amit, et al.
Published: (2025)
by: Agarwal, Amit, et al.
Published: (2025)
Self-Supervised Borrowing Detection on Multilingual Wordlists
by: Wientzek, Tim
Published: (2025)
by: Wientzek, Tim
Published: (2025)
Deep Learning based Key Information Extraction from Business Documents: Systematic Literature Review
by: Rombach, Alexander Michael, et al.
Published: (2024)
by: Rombach, Alexander Michael, et al.
Published: (2024)
Transfer-learning for video classification: Video Swin Transformer on multiple domains
by: Oliveira, Daniel A. P., et al.
Published: (2022)
by: Oliveira, Daniel A. P., et al.
Published: (2022)
Towards a Robust Framework for Multimodal Hate Detection: A Study on Video vs. Image-based Content
by: Koushik, Girish A., et al.
Published: (2025)
by: Koushik, Girish A., et al.
Published: (2025)
Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning
by: Ji, Binbin, et al.
Published: (2025)
by: Ji, Binbin, et al.
Published: (2025)
Leveraging Large Language Models for Semantic Query Processing in a Scholarly Knowledge Graph
by: Jia, Runsong, et al.
Published: (2024)
by: Jia, Runsong, et al.
Published: (2024)
PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues
by: Kajare, Prajwal Vijay, et al.
Published: (2026)
by: Kajare, Prajwal Vijay, et al.
Published: (2026)
propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale
by: Idahl, Maximilian, et al.
Published: (2026)
by: Idahl, Maximilian, et al.
Published: (2026)
LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization
by: Nguyen, Huyen, et al.
Published: (2026)
by: Nguyen, Huyen, et al.
Published: (2026)
Evaluation Before Generation: A Paradigm for Robust Multimodal Sentiment Analysis with Missing Modalities
by: Chen, Rongfei, et al.
Published: (2026)
by: Chen, Rongfei, et al.
Published: (2026)
Entity Re-identification in Visual Storytelling via Contrastive Reinforcement Learning
by: Oliveira, Daniel A. P., et al.
Published: (2025)
by: Oliveira, Daniel A. P., et al.
Published: (2025)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
by: Hua, Wenjie, et al.
Published: (2025)
by: Hua, Wenjie, et al.
Published: (2025)
GeoVision Labeler: Zero-Shot Geospatial Classification with Vision and Language Models
by: Hacheme, Gilles Quentin, et al.
Published: (2025)
by: Hacheme, Gilles Quentin, et al.
Published: (2025)
Motion-Based Sign Language Video Summarization using Curvature and Torsion
by: Sartinas, Evangelos G., et al.
Published: (2023)
by: Sartinas, Evangelos G., et al.
Published: (2023)
myMNIST: Benchmark of PETNN, KAN, and Classical Deep Learning Models for Burmese Handwritten Digit Recognition
by: Thu, Ye Kyaw, et al.
Published: (2026)
by: Thu, Ye Kyaw, et al.
Published: (2026)
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
by: Banthia, Saumya, et al.
Published: (2020)
by: Banthia, Saumya, et al.
Published: (2020)
Similar Items
-
An accurate and revised version of optical character recognition-based speech synthesis using LabVIEW
by: Mehta, Prateek, et al.
Published: (2025) -
BabelDOC: Better Layout-Preserving PDF Translation via Intermediate Representation
by: Yang, Qi, et al.
Published: (2026) -
HalalBench: A Multilingual OCR Benchmark for Food Packaging Ingredient Extraction
by: Arief, Hasan
Published: (2026) -
A Multimodal Pipeline for Clinical Data Extraction: Applying Vision-Language Models to Scans of Transfusion Reaction Reports
by: Schäfer, Henning, et al.
Published: (2025) -
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
by: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Published: (2025)