A Benchmark of State-Space Models vs. Transformers and BiLSTM-based Models for Historical Newspaper OCR
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Agbeti-messan, Merveilles, Paquet, Thierry, Chatelain, Clément, Tranouez, Pierrick, Nicolas, Stéphane |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Few-shot Writer Adaptation via Multimodal In-Context Learning
von: Simon, Tom, et al.
Veröffentlicht: (2026)
von: Simon, Tom, et al.
Veröffentlicht: (2026)
Classifying the Unknown: In-Context Learning for Open-Vocabulary Text and Symbol Recognition
von: Simon, Tom, et al.
Veröffentlicht: (2025)
von: Simon, Tom, et al.
Veröffentlicht: (2025)
DANIEL: A fast Document Attention Network for Information Extraction and Labelling of handwritten documents
von: Constum, Thomas, et al.
Veröffentlicht: (2024)
von: Constum, Thomas, et al.
Veröffentlicht: (2024)
End-to-end information extraction in handwritten documents: Understanding Paris marriage records from 1880 to 1940
von: Constum, Thomas, et al.
Veröffentlicht: (2024)
von: Constum, Thomas, et al.
Veröffentlicht: (2024)
PILOT: A Promptable Interleaved Layout-aware OCR Transformer
von: Hamdi, Laziz, et al.
Veröffentlicht: (2025)
von: Hamdi, Laziz, et al.
Veröffentlicht: (2025)
BiLSTM-VHP: BiLSTM-Powered Network for Viral Host Prediction
von: Efat, Azher Ahmed, et al.
Veröffentlicht: (2025)
von: Efat, Azher Ahmed, et al.
Veröffentlicht: (2025)
IBIS: A Hybrid Inception-BiLSTM and SVM Ensemble for Robust Doppler-based Human Activity Recognition
von: Fernandes, Alison M., et al.
Veröffentlicht: (2025)
von: Fernandes, Alison M., et al.
Veröffentlicht: (2025)
ETLNet: An Efficient TCN-BiLSTM Network for Road Anomaly Detection Using Smartphone Sensors
von: Ansari, Mohd Faiz, et al.
Veröffentlicht: (2024)
von: Ansari, Mohd Faiz, et al.
Veröffentlicht: (2024)
Inverse Modeling of S‐Parameters for GaAs pHEMT Based on BiLSTM ‐ LSTM
von: Qian Lin, et al.
Veröffentlicht: (2026)
von: Qian Lin, et al.
Veröffentlicht: (2026)
Multi Class Parkinson Disease Detection Based on Finger Tapping Using Attention Enhanced CNN BiLSTM
von: Miah, Abu Saleh Musa, et al.
Veröffentlicht: (2025)
von: Miah, Abu Saleh Musa, et al.
Veröffentlicht: (2025)
Parallel BiLSTM-Transformer networks for forecasting chaotic dynamics
von: Ma, Junwen, et al.
Veröffentlicht: (2025)
von: Ma, Junwen, et al.
Veröffentlicht: (2025)
SE-Enhanced ViT and BiLSTM-Based Intrusion Detection for Secure IIoT and IoMT Environments
von: Gueriani, Afrah, et al.
Veröffentlicht: (2026)
von: Gueriani, Afrah, et al.
Veröffentlicht: (2026)
Alzheimer's Disease Prediction Using EffNetViTLoRA and BiLSTM with Multimodal Longitudinal MRI Data
von: Khatooni, Mahdieh Behjat, et al.
Veröffentlicht: (2025)
von: Khatooni, Mahdieh Behjat, et al.
Veröffentlicht: (2025)
ViBERTgrid BiLSTM-CRF: Multimodal Key Information Extraction from Unstructured Financial Documents
von: Pala, Furkan, et al.
Veröffentlicht: (2024)
von: Pala, Furkan, et al.
Veröffentlicht: (2024)
Evaluating BiLSTM and CNN+GRU Approaches for Human Activity Recognition Using WiFi CSI Data
von: Wakili, Almustapha A., et al.
Veröffentlicht: (2025)
von: Wakili, Almustapha A., et al.
Veröffentlicht: (2025)
Error Patterns in Historical OCR: A Comparative Analysis of TrOCR and a Vision-Language Model
von: Vesalainen, Ari, et al.
Veröffentlicht: (2026)
von: Vesalainen, Ari, et al.
Veröffentlicht: (2026)
A CNN-BiLSTM Model with Attention Mechanism for Earthquake Prediction
von: Kavianpour, Parisa, et al.
Veröffentlicht: (2021)
von: Kavianpour, Parisa, et al.
Veröffentlicht: (2021)
Semantic and Contextual Modeling for Malicious Comment Detection with BERT-BiLSTM
von: Fang, Zhou, et al.
Veröffentlicht: (2025)
von: Fang, Zhou, et al.
Veröffentlicht: (2025)
Explore BiLSTM-CRF-Based Models for Open Relation Extraction
von: Ni, Tao, et al.
Veröffentlicht: (2021)
von: Ni, Tao, et al.
Veröffentlicht: (2021)
TFT-ACB-XML: Decision-Level Integration of Customized Temporal Fusion Transformer and Attention-BiLSTM with XGBoost Meta-Learner for BTC Price Forecasting
von: Din, Raiz Ud, et al.
Veröffentlicht: (2026)
von: Din, Raiz Ud, et al.
Veröffentlicht: (2026)
Enhancing Interpretability of AR-SSVEP-Based Motor Intention Recognition via CNN-BiLSTM and SHAP Analysis on EEG Data
von: Yang, Lin, et al.
Veröffentlicht: (2025)
von: Yang, Lin, et al.
Veröffentlicht: (2025)
Application of Attention Mechanism with Bidirectional Long Short-Term Memory (BiLSTM) and CNN for Human Conflict Detection using Computer Vision
von: Farias, Erick da Silva, et al.
Veröffentlicht: (2025)
von: Farias, Erick da Silva, et al.
Veröffentlicht: (2025)
BET‐BiLSTM Model: A Robust Solution for Automated Requirements Classification
von: Jalil Abbas, et al.
Veröffentlicht: (2025)
von: Jalil Abbas, et al.
Veröffentlicht: (2025)
Evaluating the Sensitivity of BiLSTM Forecasting Models to Sequence Length and Input Noise
von: Albelali, Salma, et al.
Veröffentlicht: (2025)
von: Albelali, Salma, et al.
Veröffentlicht: (2025)
Adaptive Model Integration for Stock Price Forecasting With CNN‐ATTN‐BiLSTM
von: Yi Xiao, et al.
Veröffentlicht: (2026)
von: Yi Xiao, et al.
Veröffentlicht: (2026)
FastTab: A Fast Table Recognizer with a Tiny Recursive Module and 1D Transformers
von: Hamdi, Laziz, et al.
Veröffentlicht: (2026)
von: Hamdi, Laziz, et al.
Veröffentlicht: (2026)
Sheet Music Transformer: End-To-End Optical Music Recognition Beyond Monophonic Transcription
von: Ríos-Vila, Antonio, et al.
Veröffentlicht: (2024)
von: Ríos-Vila, Antonio, et al.
Veröffentlicht: (2024)
Benchmarking LightGBM and BiLSTM for Sentiment Analysis on Indonesian E-Commerce Reviews
von: Marpaung, Lidia Natasyah, et al.
Veröffentlicht: (2026)
von: Marpaung, Lidia Natasyah, et al.
Veröffentlicht: (2026)
CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy
von: Yang, Zhibo, et al.
Veröffentlicht: (2024)
von: Yang, Zhibo, et al.
Veröffentlicht: (2024)
See Without Decoding: Motion-Vector-Based Tracking in Compressed Video
von: Duché, Axel, et al.
Veröffentlicht: (2026)
von: Duché, Axel, et al.
Veröffentlicht: (2026)
DenTab: A Dataset for Table Recognition and Visual QA on Real-World Dental Estimates
von: Hamdi, Laziz, et al.
Veröffentlicht: (2026)
von: Hamdi, Laziz, et al.
Veröffentlicht: (2026)
Short‐Term Wind Power Forecasting Based on an Learning OOA and Transformer‐ BiLSTM
von: Yang Xue, et al.
Veröffentlicht: (2026)
von: Yang Xue, et al.
Veröffentlicht: (2026)
A Hybrid CNN–BiLSTM–GRU Model for Electric Energy Consumption Forecasting
von: Luciano Roberto da Silva Leal, et al.
Veröffentlicht: (2026)
von: Luciano Roberto da Silva Leal, et al.
Veröffentlicht: (2026)
RoBERTa-BiLSTM: A Context-Aware Hybrid Model for Sentiment Analysis
von: Rahman, Md. Mostafizer, et al.
Veröffentlicht: (2024)
von: Rahman, Md. Mostafizer, et al.
Veröffentlicht: (2024)
An Adaptive CSI Feedback Model Based on BiLSTM for Massive MIMO-OFDM Systems
von: Shen, Hongrui, et al.
Veröffentlicht: (2024)
von: Shen, Hongrui, et al.
Veröffentlicht: (2024)
Efficient Machine Translation with a BiLSTM-Attention Approach
von: Wu, Yuxu, et al.
Veröffentlicht: (2024)
von: Wu, Yuxu, et al.
Veröffentlicht: (2024)
Improving OCR for Historical Texts of Multiple Languages
von: Westerdijk, Hylke, et al.
Veröffentlicht: (2025)
von: Westerdijk, Hylke, et al.
Veröffentlicht: (2025)
Fine-Grained Emotion Detection on GoEmotions: Experimental Comparison of Classical Machine Learning, BiLSTM, and Transformer Models
von: Harutyunyan, Ani, et al.
Veröffentlicht: (2026)
von: Harutyunyan, Ani, et al.
Veröffentlicht: (2026)
Fetch-A-Set: A Large-Scale OCR-Free Benchmark for Historical Document Retrieval
von: Molina, Adrià, et al.
Veröffentlicht: (2024)
von: Molina, Adrià, et al.
Veröffentlicht: (2024)
KazakhOCR: A Synthetic Benchmark for Evaluating Multimodal Models in Low-Resource Kazakh Script OCR
von: Gagnier, Henry, et al.
Veröffentlicht: (2026)
von: Gagnier, Henry, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Few-shot Writer Adaptation via Multimodal In-Context Learning
von: Simon, Tom, et al.
Veröffentlicht: (2026) -
Classifying the Unknown: In-Context Learning for Open-Vocabulary Text and Symbol Recognition
von: Simon, Tom, et al.
Veröffentlicht: (2025) -
DANIEL: A fast Document Attention Network for Information Extraction and Labelling of handwritten documents
von: Constum, Thomas, et al.
Veröffentlicht: (2024) -
End-to-end information extraction in handwritten documents: Understanding Paris marriage records from 1880 to 1940
von: Constum, Thomas, et al.
Veröffentlicht: (2024) -
PILOT: A Promptable Interleaved Layout-aware OCR Transformer
von: Hamdi, Laziz, et al.
Veröffentlicht: (2025)