DenTab: A Dataset for Table Recognition and Visual QA on Real-World Dental Estimates
Fuente:
arXiv
Saved in:
| Main Authors: | Hamdi, Laziz, Tamasna, Amine, Paquet, Thierry |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FastTab: A Fast Table Recognizer with a Tiny Recursive Module and 1D Transformers
by: Hamdi, Laziz, et al.
Published: (2026)
by: Hamdi, Laziz, et al.
Published: (2026)
TableSeq: Unified Generation of Structure, Content, and Layout
by: Hamdi, Laziz, et al.
Published: (2026)
by: Hamdi, Laziz, et al.
Published: (2026)
PILOT: A Promptable Interleaved Layout-aware OCR Transformer
by: Hamdi, Laziz, et al.
Published: (2025)
by: Hamdi, Laziz, et al.
Published: (2025)
MMIS: Multimodal Dataset for Interior Scene Visual Generation and Recognition
by: Kassab, Hozaifa, et al.
Published: (2024)
by: Kassab, Hozaifa, et al.
Published: (2024)
Classifying the Unknown: In-Context Learning for Open-Vocabulary Text and Symbol Recognition
by: Simon, Tom, et al.
Published: (2025)
by: Simon, Tom, et al.
Published: (2025)
End-to-End Full-Page Optical Music Recognition for Pianoform Sheet Music
by: Ríos-Vila, Antonio, et al.
Published: (2024)
by: Ríos-Vila, Antonio, et al.
Published: (2024)
Sheet Music Transformer: End-To-End Optical Music Recognition Beyond Monophonic Transcription
by: Ríos-Vila, Antonio, et al.
Published: (2024)
by: Ríos-Vila, Antonio, et al.
Published: (2024)
DenVisCoM: Dense Vision Correspondence Mamba for Efficient and Real-time Optical Flow and Stereo Estimation
by: Anand, Tushar, et al.
Published: (2026)
by: Anand, Tushar, et al.
Published: (2026)
RoadscapesQA: A Multitask, Multimodal Dataset for Visual Question Answering on Indian Roads
by: Iyer, Vijayasri, et al.
Published: (2026)
by: Iyer, Vijayasri, et al.
Published: (2026)
VeriSciQA: An Auto-Verified Dataset for Scientific Visual Question Answering
by: Li, Yuyi, et al.
Published: (2025)
by: Li, Yuyi, et al.
Published: (2025)
DetReIDX: A Stress-Test Dataset for Real-World UAV-Based Person Recognition
by: Hambarde, Kailash A., et al.
Published: (2025)
by: Hambarde, Kailash A., et al.
Published: (2025)
Entity6K: A Large Open-Domain Evaluation Dataset for Real-World Entity Recognition
by: Qiu, Jielin, et al.
Published: (2024)
by: Qiu, Jielin, et al.
Published: (2024)
TabSniper: Towards Accurate Table Detection & Structure Recognition for Bank Statements
by: Trivedi, Abhishek, et al.
Published: (2024)
by: Trivedi, Abhishek, et al.
Published: (2024)
STRIDE-QA: Visual Question Answering Dataset for Spatiotemporal Reasoning in Urban Driving Scenes
by: Ishihara, Keishi, et al.
Published: (2025)
by: Ishihara, Keishi, et al.
Published: (2025)
TabPedia: Towards Comprehensive Visual Table Understanding with Concept Synergy
by: Zhao, Weichao, et al.
Published: (2024)
by: Zhao, Weichao, et al.
Published: (2024)
UniTabNet: Bridging Vision and Language Models for Enhanced Table Structure Recognition
by: Zhang, Zhenrong, et al.
Published: (2024)
by: Zhang, Zhenrong, et al.
Published: (2024)
Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images
by: Lompo, Boammani Aser, et al.
Published: (2025)
by: Lompo, Boammani Aser, et al.
Published: (2025)
GCF: Graph Convolutional Networks for Facial Expression Recognition
by: Kassab, Hozaifa, et al.
Published: (2024)
by: Kassab, Hozaifa, et al.
Published: (2024)
ICTPolarReal: A Polarized Reflection and Material Dataset of Real World Objects
by: Yang, Jing, et al.
Published: (2026)
by: Yang, Jing, et al.
Published: (2026)
VisualTrans: A Benchmark for Real-World Visual Transformation Reasoning
by: Ji, Yuheng, et al.
Published: (2025)
by: Ji, Yuheng, et al.
Published: (2025)
Data-Augmented Multimodal Feature Fusion for Multiclass Visual Recognition of Oral Cancer Lesions
by: Naoum, Joy, et al.
Published: (2025)
by: Naoum, Joy, et al.
Published: (2025)
Dynamically Modulating Visual Place Recognition Sequence Length For Minimum Acceptable Performance Scenarios
by: Malone, Connor, et al.
Published: (2024)
by: Malone, Connor, et al.
Published: (2024)
Helvipad: A Real-World Dataset for Omnidirectional Stereo Depth Estimation
by: Zayene, Mehdi, et al.
Published: (2024)
by: Zayene, Mehdi, et al.
Published: (2024)
End-to-end information extraction in handwritten documents: Understanding Paris marriage records from 1880 to 1940
by: Constum, Thomas, et al.
Published: (2024)
by: Constum, Thomas, et al.
Published: (2024)
Dual-Imbalance Continual Learning for Real-World Food Recognition
by: Zhang, Xiaoyan, et al.
Published: (2026)
by: Zhang, Xiaoyan, et al.
Published: (2026)
Real-World Transferable Adversarial Attack on Face-Recognition Systems
by: Kaznacheev, Andrey, et al.
Published: (2025)
by: Kaznacheev, Andrey, et al.
Published: (2025)
Synthetic-to-Real Domain Adaptation for Action Recognition: A Dataset and Baseline Performances
by: Reddy, Arun V., et al.
Published: (2023)
by: Reddy, Arun V., et al.
Published: (2023)
ISLES'24 -- A Real-World Longitudinal Multimodal Stroke Dataset
by: Riedel, Evamaria Olga, et al.
Published: (2024)
by: Riedel, Evamaria Olga, et al.
Published: (2024)
WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning
by: Zhang, Yuanhan, et al.
Published: (2024)
by: Zhang, Yuanhan, et al.
Published: (2024)
On the Estimation of Image-matching Uncertainty in Visual Place Recognition
by: Zaffar, Mubariz, et al.
Published: (2024)
by: Zaffar, Mubariz, et al.
Published: (2024)
ActionAtlas: A VideoQA Benchmark for Domain-specialized Action Recognition
by: Salehi, Mohammadreza, et al.
Published: (2024)
by: Salehi, Mohammadreza, et al.
Published: (2024)
UniMERNet: A Universal Network for Real-World Mathematical Expression Recognition
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
Massively Annotated Datasets for Assessment of Synthetic and Real Data in Face Recognition
by: Neto, Pedro C., et al.
Published: (2024)
by: Neto, Pedro C., et al.
Published: (2024)
WorldPose: A World Cup Dataset for Global 3D Human Pose Estimation
by: Jiang, Tianjian, et al.
Published: (2025)
by: Jiang, Tianjian, et al.
Published: (2025)
Dens3R: A Foundation Model for 3D Geometry Prediction
by: Fang, Xianze, et al.
Published: (2025)
by: Fang, Xianze, et al.
Published: (2025)
Real-IAD: A Real-World Multi-View Dataset for Benchmarking Versatile Industrial Anomaly Detection
by: Wang, Chengjie, et al.
Published: (2024)
by: Wang, Chengjie, et al.
Published: (2024)
UDC-VIT: A Real-World Video Dataset for Under-Display Cameras
by: Ahn, Kyusu, et al.
Published: (2025)
by: Ahn, Kyusu, et al.
Published: (2025)
DVD: A Comprehensive Dataset for Advancing Violence Detection in Real-World Scenarios
by: Kollias, Dimitrios, et al.
Published: (2025)
by: Kollias, Dimitrios, et al.
Published: (2025)
A New Dataset and Framework for Real-World Blurred Images Super-Resolution
by: Qin, Rui, et al.
Published: (2024)
by: Qin, Rui, et al.
Published: (2024)
WildPPG: A Real-World PPG Dataset of Long Continuous Recordings
by: Meier, Manuel, et al.
Published: (2024)
by: Meier, Manuel, et al.
Published: (2024)
Similar Items
-
FastTab: A Fast Table Recognizer with a Tiny Recursive Module and 1D Transformers
by: Hamdi, Laziz, et al.
Published: (2026) -
TableSeq: Unified Generation of Structure, Content, and Layout
by: Hamdi, Laziz, et al.
Published: (2026) -
PILOT: A Promptable Interleaved Layout-aware OCR Transformer
by: Hamdi, Laziz, et al.
Published: (2025) -
MMIS: Multimodal Dataset for Interior Scene Visual Generation and Recognition
by: Kassab, Hozaifa, et al.
Published: (2024) -
Classifying the Unknown: In-Context Learning for Open-Vocabulary Text and Symbol Recognition
by: Simon, Tom, et al.
Published: (2025)