Understanding Cross-Language Transfer Improvements in Low-Resource HTR: The Role of Sequence Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Al-azzawi, Sana, Liu, Chang, Habib, Nudrat, Barney, Elisa, Liwicki, Marcus |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-Language Learning within Arabic Script for Low-Resource HTR
by: Al-azzawi, Sana, et al.
Published: (2026)
by: Al-azzawi, Sana, et al.
Published: (2026)
CER-HV: A Human-in-the-Loop Framework for Cleaning Datasets Applied to Arabic-Script HTR
by: Al-azzawi, Sana, et al.
Published: (2026)
by: Al-azzawi, Sana, et al.
Published: (2026)
Instruction Makes a Difference
by: Adewumi, Tosin, et al.
Published: (2024)
by: Adewumi, Tosin, et al.
Published: (2024)
Trends and Challenges in Authorship Analysis: A Review of ML, DL, and LLM Approaches
by: Habib, Nudrat, et al.
Published: (2025)
by: Habib, Nudrat, et al.
Published: (2025)
MOoSE: Multi-Orientation Sharing Experts for Open-set Scene Text Recognition
by: Liu, Chang, et al.
Published: (2024)
by: Liu, Chang, et al.
Published: (2024)
Giving each task what it needs -- leveraging structured sparsity for tailored multi-task learning
by: Upadhyay, Richa, et al.
Published: (2024)
by: Upadhyay, Richa, et al.
Published: (2024)
Rethinking HTG Evaluation: Bridging Generation and Recognition
by: Nikolaidou, Konstantina, et al.
Published: (2024)
by: Nikolaidou, Konstantina, et al.
Published: (2024)
DiffusionPen: Towards Controlling the Style of Handwritten Text Generation
by: Nikolaidou, Konstantina, et al.
Published: (2024)
by: Nikolaidou, Konstantina, et al.
Published: (2024)
SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion
by: Nazir, Danish, et al.
Published: (2022)
by: Nazir, Danish, et al.
Published: (2022)
HTR-VT: Handwritten Text Recognition with Vision Transformer
by: Li, Yuting, et al.
Published: (2024)
by: Li, Yuting, et al.
Published: (2024)
Meta-Sparsity: Learning Optimal Sparse Structures in Multi-task Networks through Meta-learning
by: Upadhyay, Richa, et al.
Published: (2025)
by: Upadhyay, Richa, et al.
Published: (2025)
Reducing Annotation Burden for Femoral Cartilage Segmentation in Knee MRI via Cross-Sequence Transfer Learning
by: Chiumento, Francesco, et al.
Published: (2026)
by: Chiumento, Francesco, et al.
Published: (2026)
Training-Free Multi-Concept Image Editing
by: Foteinopoulou, Niki, et al.
Published: (2026)
by: Foteinopoulou, Niki, et al.
Published: (2026)
LoRAtorio: An intrinsic approach to LoRA Skill Composition
by: Foteinopoulou, Niki, et al.
Published: (2025)
by: Foteinopoulou, Niki, et al.
Published: (2025)
Transfer Learning for Cross-dataset Isolated Sign Language Recognition in Under-Resourced Datasets
by: Kindiroglu, Ahmet Alp, et al.
Published: (2024)
by: Kindiroglu, Ahmet Alp, et al.
Published: (2024)
A Systematic Performance Analysis of Deep Perceptual Loss Networks: Breaking Transfer Learning Conventions
by: Pihlgren, Gustav Grund, et al.
Published: (2023)
by: Pihlgren, Gustav Grund, et al.
Published: (2023)
Open-Vocabulary Object Detectors: Robustness Challenges under Distribution Shifts
by: Chhipa, Prakash Chandra, et al.
Published: (2024)
by: Chhipa, Prakash Chandra, et al.
Published: (2024)
Dual Orthogonal Guidance for Robust Diffusion-based Handwritten Text Generation
by: Nikolaidou, Konstantina, et al.
Published: (2025)
by: Nikolaidou, Konstantina, et al.
Published: (2025)
LCM: Log Conformal Maps for Robust Representation Learning to Mitigate Perspective Distortion
by: Chippa, Meenakshi Subhash, et al.
Published: (2024)
by: Chippa, Meenakshi Subhash, et al.
Published: (2024)
GraDeT-HTR: A Resource-Efficient Bengali Handwritten Text Recognition System utilizing Grapheme-based Tokenizer and Decoder-only Transformer
by: Hasan, Md. Mahmudul, et al.
Published: (2025)
by: Hasan, Md. Mahmudul, et al.
Published: (2025)
HTR-JAND: Handwritten Text Recognition with Joint Attention Network and Knowledge Distillation
by: Hamdan, Mohammed, et al.
Published: (2024)
by: Hamdan, Mohammed, et al.
Published: (2024)
On the Limitations of Large Language Models (LLMs): False Attribution
by: Adewumi, Tosin, et al.
Published: (2024)
by: Adewumi, Tosin, et al.
Published: (2024)
Shape2.5D: A Dataset of Texture-less Surfaces for Depth and Normals Estimation
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
DRetHTR: Linear-Time Decoder-Only Retentive Network for Handwritten Text Recognition
by: Kim, Changhun, et al.
Published: (2026)
by: Kim, Changhun, et al.
Published: (2026)
Möbius Transform for Mitigating Perspective Distortions in Representation Learning
by: Chhipa, Prakash Chandra, et al.
Published: (2024)
by: Chhipa, Prakash Chandra, et al.
Published: (2024)
Online In-Context Distillation for Low-Resource Vision Language Models
by: Kang, Zhiqi, et al.
Published: (2025)
by: Kang, Zhiqi, et al.
Published: (2025)
Understanding Information Storage and Transfer in Multi-modal Large Language Models
by: Basu, Samyadeep, et al.
Published: (2024)
by: Basu, Samyadeep, et al.
Published: (2024)
Understanding Task Transfer in Vision-Language Models
by: Sachdeva, Bhuvan, et al.
Published: (2025)
by: Sachdeva, Bhuvan, et al.
Published: (2025)
LUCES-MV: A Multi-View Dataset for Near-Field Point Light Source Photometric Stereo
by: Logothetis, Fotios, et al.
Published: (2024)
by: Logothetis, Fotios, et al.
Published: (2024)
DiaLoc: An Iterative Approach to Embodied Dialog Localization
by: Zhang, Chao, et al.
Published: (2024)
by: Zhang, Chao, et al.
Published: (2024)
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models
by: Yadav, Ankit, et al.
Published: (2025)
by: Yadav, Ankit, et al.
Published: (2025)
Temporal-Oriented Recipe for Transferring Large Vision-Language Model to Video Understanding
by: Nguyen, Thong, et al.
Published: (2025)
by: Nguyen, Thong, et al.
Published: (2025)
READ: Recurrent Adapter with Partial Video-Language Alignment for Parameter-Efficient Transfer Learning in Low-Resource Video-Language Modeling
by: Nguyen, Thong, et al.
Published: (2023)
by: Nguyen, Thong, et al.
Published: (2023)
Handwritten Text Recognition for Low Resource Languages
by: Dey, Sayantan, et al.
Published: (2025)
by: Dey, Sayantan, et al.
Published: (2025)
HTR-ConvText: Leveraging Convolution and Textual Information for Handwritten Text Recognition
by: Truc, Pham Thach Thanh, et al.
Published: (2025)
by: Truc, Pham Thach Thanh, et al.
Published: (2025)
FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding
by: Liu, Zheng, et al.
Published: (2025)
by: Liu, Zheng, et al.
Published: (2025)
Low-Resource Heuristics for Bahnaric Optical Character Recognition Improvement
by: Tran, Phat, et al.
Published: (2026)
by: Tran, Phat, et al.
Published: (2026)
Finetuning Vision-Language Models as OCR Systems for Low-Resource Languages: A Case Study of Manchu
by: Chung, Yan Hon Michael, et al.
Published: (2025)
by: Chung, Yan Hon Michael, et al.
Published: (2025)
Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video Understanding
by: Li, Yun, et al.
Published: (2025)
by: Li, Yun, et al.
Published: (2025)
Compact Model Training by Low-Rank Projection with Energy Transfer
by: Guo, Kailing, et al.
Published: (2022)
by: Guo, Kailing, et al.
Published: (2022)
Similar Items
-
Cross-Language Learning within Arabic Script for Low-Resource HTR
by: Al-azzawi, Sana, et al.
Published: (2026) -
CER-HV: A Human-in-the-Loop Framework for Cleaning Datasets Applied to Arabic-Script HTR
by: Al-azzawi, Sana, et al.
Published: (2026) -
Instruction Makes a Difference
by: Adewumi, Tosin, et al.
Published: (2024) -
Trends and Challenges in Authorship Analysis: A Review of ML, DL, and LLM Approaches
by: Habib, Nudrat, et al.
Published: (2025) -
MOoSE: Multi-Orientation Sharing Experts for Open-set Scene Text Recognition
by: Liu, Chang, et al.
Published: (2024)