DIVA-DAF: A Deep Learning Framework for Historical Document Image Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vögtlin, Lars, Scius-Bertrand, Anna, Maergner, Paul, Fischer, Andreas, Ingold, Rolf |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-Shot Prompting and Few-Shot Fine-Tuning: Revisiting Document Image Classification Using Large Language Models
von: Scius-Bertrand, Anna, et al.
Veröffentlicht: (2024)
von: Scius-Bertrand, Anna, et al.
Veröffentlicht: (2024)
Contrastive Learning for Character Detection in Ancient Greek Papyri
von: Nakka, Vedasri, et al.
Veröffentlicht: (2024)
von: Nakka, Vedasri, et al.
Veröffentlicht: (2024)
CTC Transcription Alignment of the Bullinger Letters: Automatic Improvement of Annotation Quality
von: Peer, Marco, et al.
Veröffentlicht: (2025)
von: Peer, Marco, et al.
Veröffentlicht: (2025)
BullingerDB: A Dataset for Handwritten Text Recognition and Writer Retrieval
von: Peer, Marco, et al.
Veröffentlicht: (2026)
von: Peer, Marco, et al.
Veröffentlicht: (2026)
From Pen Strokes to Sleep States: Detecting Low-Recovery Days Using Sigma-Lognormal Handwriting Features
von: Tanaka, Chisa, et al.
Veröffentlicht: (2026)
von: Tanaka, Chisa, et al.
Veröffentlicht: (2026)
Prediction of Grade, Gender, and Academic Performance of Children and Teenagers from Handwriting Using the Sigma-Lognormal Model
von: Iste, Adrian, et al.
Veröffentlicht: (2026)
von: Iste, Adrian, et al.
Veröffentlicht: (2026)
Rule-Based Reinforcement Learning for Document Image Classification with Vision Language Models
von: Jungo, Michael, et al.
Veröffentlicht: (2025)
von: Jungo, Michael, et al.
Veröffentlicht: (2025)
DIVA: Harnessing the Representation Divergence in Unified Multimodal Models for Mutual Reinforcement
von: Lu, Renjie, et al.
Veröffentlicht: (2026)
von: Lu, Renjie, et al.
Veröffentlicht: (2026)
A GAN-Enhanced Deep Learning Framework for Rooftop Detection from Historical Aerial Imagery
von: Chen, Pengyu, et al.
Veröffentlicht: (2025)
von: Chen, Pengyu, et al.
Veröffentlicht: (2025)
DAF-Net: A Dual-Branch Feature Decomposition Fusion Network with Domain Adaptive for Infrared and Visible Image Fusion
von: Xu, Jian, et al.
Veröffentlicht: (2024)
von: Xu, Jian, et al.
Veröffentlicht: (2024)
Annotation Cost-Efficient Active Learning for Deep Metric Learning Driven Remote Sensing Image Retrieval
von: Hoxha, Genc, et al.
Veröffentlicht: (2024)
von: Hoxha, Genc, et al.
Veröffentlicht: (2024)
Deep Learning Framework for Early Detection of Pancreatic Cancer Using Multi-Modal Medical Imaging Analysis
von: Slobodzian, Dennis, et al.
Veröffentlicht: (2025)
von: Slobodzian, Dennis, et al.
Veröffentlicht: (2025)
Accurate Fine-grained Layout Analysis for the Historical Tibetan Document Based on the Instance Segmentation
von: Zhao, Penghai, et al.
Veröffentlicht: (2021)
von: Zhao, Penghai, et al.
Veröffentlicht: (2021)
Synthetic Data Augmentation for Table Detection: Re-evaluating TableNet's Performance with Automatically Generated Document Images
von: Sahukara, Krishna, et al.
Veröffentlicht: (2025)
von: Sahukara, Krishna, et al.
Veröffentlicht: (2025)
Handwriting Recognition in Historical Documents with Multimodal LLM
von: Li, Lucian
Veröffentlicht: (2024)
von: Li, Lucian
Veröffentlicht: (2024)
Predicting the Original Appearance of Damaged Historical Documents
von: Yang, Zhenhua, et al.
Veröffentlicht: (2024)
von: Yang, Zhenhua, et al.
Veröffentlicht: (2024)
A Fair Evaluation of Various Deep Learning-Based Document Image Binarization Approaches
von: Sukesh, Richin, et al.
Veröffentlicht: (2024)
von: Sukesh, Richin, et al.
Veröffentlicht: (2024)
Visual Bias and Interpretability in Deep Learning for Dermatological Image Analysis
von: Taufik, Enam Ahmed, et al.
Veröffentlicht: (2025)
von: Taufik, Enam Ahmed, et al.
Veröffentlicht: (2025)
Coarse-to-Fine Non-rigid Multi-modal Image Registration for Historical Panel Paintings based on Crack Structures
von: Sindel, Aline, et al.
Veröffentlicht: (2026)
von: Sindel, Aline, et al.
Veröffentlicht: (2026)
Deep Learning Approaches for Seizure Video Analysis: A Review
von: Ahmedt-Aristizabal, David, et al.
Veröffentlicht: (2023)
von: Ahmedt-Aristizabal, David, et al.
Veröffentlicht: (2023)
Npix2Cpix: A GAN-Based Image-to-Image Translation Network With Retrieval- Classification Integration for Watermark Retrieval From Historical Document Images
von: Saha, Utsab, et al.
Veröffentlicht: (2024)
von: Saha, Utsab, et al.
Veröffentlicht: (2024)
PQ-DAF: Pose-driven Quality-controlled Data Augmentation for Data-scarce Driver Distraction Detection
von: Sun, Haibin, et al.
Veröffentlicht: (2025)
von: Sun, Haibin, et al.
Veröffentlicht: (2025)
A Framework for Critical Evaluation of Text-to-Image Models: Integrating Art Historical Analysis, Artistic Exploration, and Critical Prompt Engineering
von: Foka, Amalia
Veröffentlicht: (2024)
von: Foka, Amalia
Veröffentlicht: (2024)
PyPotteryLens: An Open-Source Deep Learning Framework for Automated Digitisation of Archaeological Pottery Documentation
von: Cardarelli, Lorenzo
Veröffentlicht: (2024)
von: Cardarelli, Lorenzo
Veröffentlicht: (2024)
Beyond the Pipeline: Analyzing Key Factors in End-to-End Deep Learning for Historical Writer Identification
von: Rasyidi, Hanif, et al.
Veröffentlicht: (2025)
von: Rasyidi, Hanif, et al.
Veröffentlicht: (2025)
CUTS: A Deep Learning and Topological Framework for Multigranular Unsupervised Medical Image Segmentation
von: Liu, Chen, et al.
Veröffentlicht: (2022)
von: Liu, Chen, et al.
Veröffentlicht: (2022)
KonfAI: A Modular and Fully Configurable Framework for Deep Learning in Medical Imaging
von: Boussot, Valentin, et al.
Veröffentlicht: (2025)
von: Boussot, Valentin, et al.
Veröffentlicht: (2025)
A Deep-Learning Framework for Land-Sliding Classification from Remote Sensing Image
von: Tang, Hieu, et al.
Veröffentlicht: (2025)
von: Tang, Hieu, et al.
Veröffentlicht: (2025)
SegHist: A General Segmentation-based Framework for Chinese Historical Document Text Line Detection
von: Hu, Xingjian, et al.
Veröffentlicht: (2024)
von: Hu, Xingjian, et al.
Veröffentlicht: (2024)
DIVA-VQA: Detecting Inter-frame Variations in UGC Video Quality
von: Wang, Xinyi, et al.
Veröffentlicht: (2025)
von: Wang, Xinyi, et al.
Veröffentlicht: (2025)
Deep Companion Learning: Enhancing Generalization Through Historical Consistency
von: Zhu, Ruizhao, et al.
Veröffentlicht: (2024)
von: Zhu, Ruizhao, et al.
Veröffentlicht: (2024)
Covariance Descriptors Meet General Vision Encoders: Riemannian Deep Learning for Medical Image Classification
von: Mayr, Josef, et al.
Veröffentlicht: (2025)
von: Mayr, Josef, et al.
Veröffentlicht: (2025)
MedSR-Vision: Deep Learning Framework for Multi-Domain Medical Image Super-Resolution
von: Gurappa, Subhash, et al.
Veröffentlicht: (2026)
von: Gurappa, Subhash, et al.
Veröffentlicht: (2026)
ClusterMark: Towards Robust Watermarking for Autoregressive Image Generators with Visual Token Clustering
von: Lukovnikov, Denis, et al.
Veröffentlicht: (2025)
von: Lukovnikov, Denis, et al.
Veröffentlicht: (2025)
Structuring Quantitative Image Analysis with Object Prominence
von: Arnold, Christian, et al.
Veröffentlicht: (2024)
von: Arnold, Christian, et al.
Veröffentlicht: (2024)
TextBite: A Historical Czech Document Dataset for Logical Page Segmentation
von: Kostelník, Martin, et al.
Veröffentlicht: (2025)
von: Kostelník, Martin, et al.
Veröffentlicht: (2025)
Integrating Vision and Location with Transformers: A Multimodal Deep Learning Framework for Medical Wound Analysis
von: Mousa, Ramin, et al.
Veröffentlicht: (2025)
von: Mousa, Ramin, et al.
Veröffentlicht: (2025)
Addressing Fairness Issues in Deep Learning-Based Medical Image Analysis: A Systematic Review
von: Xu, Zikang, et al.
Veröffentlicht: (2022)
von: Xu, Zikang, et al.
Veröffentlicht: (2022)
acia-workflows: Automated Single-cell Imaging Analysis for Scalable and Deep Learning-based Live-cell Imaging Analysis Workflows
von: Seiffarth, Johannes, et al.
Veröffentlicht: (2025)
von: Seiffarth, Johannes, et al.
Veröffentlicht: (2025)
Parking Analytics Framework using Deep Learning
von: Benjdira, Bilel, et al.
Veröffentlicht: (2022)
von: Benjdira, Bilel, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Zero-Shot Prompting and Few-Shot Fine-Tuning: Revisiting Document Image Classification Using Large Language Models
von: Scius-Bertrand, Anna, et al.
Veröffentlicht: (2024) -
Contrastive Learning for Character Detection in Ancient Greek Papyri
von: Nakka, Vedasri, et al.
Veröffentlicht: (2024) -
CTC Transcription Alignment of the Bullinger Letters: Automatic Improvement of Annotation Quality
von: Peer, Marco, et al.
Veröffentlicht: (2025) -
BullingerDB: A Dataset for Handwritten Text Recognition and Writer Retrieval
von: Peer, Marco, et al.
Veröffentlicht: (2026) -
From Pen Strokes to Sleep States: Detecting Low-Recovery Days Using Sigma-Lognormal Handwriting Features
von: Tanaka, Chisa, et al.
Veröffentlicht: (2026)