DocXPand-25k: a large and diverse benchmark dataset for identity documents analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lerouge, Julien, Betmont, Guillaume, Bres, Thomas, Stepankevich, Evgeny, Bergès, Alexis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A large-scale image-text dataset benchmark for farmland segmentation
von: Tao, Chao, et al.
Veröffentlicht: (2025)
von: Tao, Chao, et al.
Veröffentlicht: (2025)
A benchmark multimodal oro-dental dataset for large vision-language models
von: Lv, Haoxin, et al.
Veröffentlicht: (2025)
von: Lv, Haoxin, et al.
Veröffentlicht: (2025)
MatPredict: a dataset and benchmark for learning material properties of diverse indoor objects
von: Chen, Yuzhen, et al.
Veröffentlicht: (2025)
von: Chen, Yuzhen, et al.
Veröffentlicht: (2025)
Comics Datasets Framework: Mix of Comics datasets for detection benchmarking
von: Vivoli, Emanuele, et al.
Veröffentlicht: (2024)
von: Vivoli, Emanuele, et al.
Veröffentlicht: (2024)
A large-scale multicenter breast cancer DCE-MRI benchmark dataset with expert segmentations
von: Garrucho, Lidia, et al.
Veröffentlicht: (2024)
von: Garrucho, Lidia, et al.
Veröffentlicht: (2024)
FungiTastic: A multi-modal dataset and benchmark for image categorization
von: Picek, Lukas, et al.
Veröffentlicht: (2024)
von: Picek, Lukas, et al.
Veröffentlicht: (2024)
NordFKB: a fine-grained benchmark dataset for geospatial AI in Norway
von: Jyhne, Sander Riisøen, et al.
Veröffentlicht: (2025)
von: Jyhne, Sander Riisøen, et al.
Veröffentlicht: (2025)
MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark
von: Osmulski, Radek, et al.
Veröffentlicht: (2025)
von: Osmulski, Radek, et al.
Veröffentlicht: (2025)
Fine-Grained Customized Fashion Design with Image-into-Prompt benchmark and dataset from LMM
von: Li, Hui, et al.
Veröffentlicht: (2025)
von: Li, Hui, et al.
Veröffentlicht: (2025)
FantasyID: A dataset for detecting digital manipulations of ID-documents
von: Korshunov, Pavel, et al.
Veröffentlicht: (2025)
von: Korshunov, Pavel, et al.
Veröffentlicht: (2025)
DeepSea MOT: A benchmark dataset for multi-object tracking on deep-sea video
von: Barnard, Kevin, et al.
Veröffentlicht: (2025)
von: Barnard, Kevin, et al.
Veröffentlicht: (2025)
KidSat: satellite imagery to map childhood poverty dataset and benchmark
von: Sharma, Makkunda, et al.
Veröffentlicht: (2024)
von: Sharma, Makkunda, et al.
Veröffentlicht: (2024)
A benchmark dataset for deep learning-based airplane detection: HRPlanes
von: Bakirman, Tolga, et al.
Veröffentlicht: (2022)
von: Bakirman, Tolga, et al.
Veröffentlicht: (2022)
CogDoc: Towards Unified thinking in Documents
von: Xu, Qixin, et al.
Veröffentlicht: (2025)
von: Xu, Qixin, et al.
Veröffentlicht: (2025)
AniDoc: Animation Creation Made Easier
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
A Sentinel-2 multi-year, multi-country benchmark dataset for crop classification and segmentation with deep learning
von: Sykas, Dimitrios, et al.
Veröffentlicht: (2022)
von: Sykas, Dimitrios, et al.
Veröffentlicht: (2022)
PlanarTrack: A high-quality and challenging benchmark for large-scale planar object tracking
von: Jiao, Yifan, et al.
Veröffentlicht: (2025)
von: Jiao, Yifan, et al.
Veröffentlicht: (2025)
600k-ks-ocr: a large-scale synthetic dataset for optical character recognition in kashmiri script
von: Malik, Haq Nawaz
Veröffentlicht: (2026)
von: Malik, Haq Nawaz
Veröffentlicht: (2026)
WildlifeReID-10k: Wildlife re-identification dataset with 10k individual animals
von: Adam, Lukáš, et al.
Veröffentlicht: (2024)
von: Adam, Lukáš, et al.
Veröffentlicht: (2024)
Open-ended VQA benchmarking of Vision-Language models by exploiting Classification datasets and their semantic hierarchy
von: Ging, Simon, et al.
Veröffentlicht: (2024)
von: Ging, Simon, et al.
Veröffentlicht: (2024)
A benchmark for video-based laparoscopic skill analysis and assessment
von: Funke, Isabel, et al.
Veröffentlicht: (2026)
von: Funke, Isabel, et al.
Veröffentlicht: (2026)
Re-assembling the past: The RePAIR dataset and benchmark for real world 2D and 3D puzzle solving
von: Tsesmelis, Theodore, et al.
Veröffentlicht: (2024)
von: Tsesmelis, Theodore, et al.
Veröffentlicht: (2024)
Supervised makeup transfer with a curated dataset: Decoupling identity and makeup features for enhanced transformation
von: Pan, Qihe, et al.
Veröffentlicht: (2026)
von: Pan, Qihe, et al.
Veröffentlicht: (2026)
DocSLM: A Small Vision-Language Model for Long Multimodal Document Understanding
von: Hannan, Tanveer, et al.
Veröffentlicht: (2025)
von: Hannan, Tanveer, et al.
Veröffentlicht: (2025)
Scale-invariant brain morphometry: application to sulcal depth
von: Dieudonné, Maxime, et al.
Veröffentlicht: (2025)
von: Dieudonné, Maxime, et al.
Veröffentlicht: (2025)
CardioSyntax: end-to-end SYNTAX score prediction -- dataset, benchmark and method
von: Ponomarchuk, Alexander, et al.
Veröffentlicht: (2024)
von: Ponomarchuk, Alexander, et al.
Veröffentlicht: (2024)
Lumbar spine segmentation in MR images: a dataset and a public benchmark
von: van der Graaf, Jasper W., et al.
Veröffentlicht: (2023)
von: van der Graaf, Jasper W., et al.
Veröffentlicht: (2023)
DocDeshadower: Frequency-Aware Transformer for Document Shadow Removal
von: Zhou, Ziyang, et al.
Veröffentlicht: (2023)
von: Zhou, Ziyang, et al.
Veröffentlicht: (2023)
DocRevive: A Unified Pipeline for Document Text Restoration
von: Purkayastha, Kunal, et al.
Veröffentlicht: (2026)
von: Purkayastha, Kunal, et al.
Veröffentlicht: (2026)
A large-scale dataset for end-to-end table recognition in the wild
von: Yang, Fan, et al.
Veröffentlicht: (2023)
von: Yang, Fan, et al.
Veröffentlicht: (2023)
Symbrain: A large-scale dataset of MRI images for neonatal brain symmetry analysis
von: Gucciardi, Arnaud, et al.
Veröffentlicht: (2024)
von: Gucciardi, Arnaud, et al.
Veröffentlicht: (2024)
A comprehensive multimodal dataset and benchmark for ulcerative colitis scoring in endoscopy
von: Ghatwary, Noha, et al.
Veröffentlicht: (2026)
von: Ghatwary, Noha, et al.
Veröffentlicht: (2026)
M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
von: Cho, Jaemin, et al.
Veröffentlicht: (2024)
von: Cho, Jaemin, et al.
Veröffentlicht: (2024)
MIRAGE: Multimodal foundation model and benchmark for comprehensive retinal OCT image analysis
von: Morano, José, et al.
Veröffentlicht: (2025)
von: Morano, José, et al.
Veröffentlicht: (2025)
DocTron-Formula: Generalized Formula Recognition in Complex and Structured Scenarios
von: Zhong, Yufeng, et al.
Veröffentlicht: (2025)
von: Zhong, Yufeng, et al.
Veröffentlicht: (2025)
CowScreeningDB: A public benchmark dataset for lameness detection in dairy cows
von: Ismail, Shahid, et al.
Veröffentlicht: (2024)
von: Ismail, Shahid, et al.
Veröffentlicht: (2024)
Synthetic dataset of ID and Travel Document
von: Boned, Carlos, et al.
Veröffentlicht: (2024)
von: Boned, Carlos, et al.
Veröffentlicht: (2024)
FISBe: A real-world benchmark dataset for instance segmentation of long-range thin filamentous structures
von: Mais, Lisa, et al.
Veröffentlicht: (2024)
von: Mais, Lisa, et al.
Veröffentlicht: (2024)
An evaluation of Deep Learning based stereo dense matching dataset shift from aerial images and a large scale stereo dataset
von: Wu, Teng, et al.
Veröffentlicht: (2024)
von: Wu, Teng, et al.
Veröffentlicht: (2024)
Uni-DocDiff: A Unified Document Restoration Model Based on Diffusion
von: Zhao, Fangmin, et al.
Veröffentlicht: (2025)
von: Zhao, Fangmin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A large-scale image-text dataset benchmark for farmland segmentation
von: Tao, Chao, et al.
Veröffentlicht: (2025) -
A benchmark multimodal oro-dental dataset for large vision-language models
von: Lv, Haoxin, et al.
Veröffentlicht: (2025) -
MatPredict: a dataset and benchmark for learning material properties of diverse indoor objects
von: Chen, Yuzhen, et al.
Veröffentlicht: (2025) -
Comics Datasets Framework: Mix of Comics datasets for detection benchmarking
von: Vivoli, Emanuele, et al.
Veröffentlicht: (2024) -
A large-scale multicenter breast cancer DCE-MRI benchmark dataset with expert segmentations
von: Garrucho, Lidia, et al.
Veröffentlicht: (2024)