DocXplain: A Novel Model-Agnostic Explainability Method for Document Image Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Saifullah, Saifullah, Agne, Stefan, Dengel, Andreas, Ahmed, Sheraz |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DocVCE: Diffusion-based Visual Counterfactual Explanations for Document Image Classification
by: Saifullah, Saifullah, et al.
Published: (2025)
by: Saifullah, Saifullah, et al.
Published: (2025)
Latent Diffusion for Guided Document Table Generation
by: Hamdani, Syed Jawwad Haider, et al.
Published: (2024)
by: Hamdani, Syed Jawwad Haider, et al.
Published: (2024)
WordVIS: A Color Worth A Thousand Words
by: Khan, Umar, et al.
Published: (2024)
by: Khan, Umar, et al.
Published: (2024)
Privacy Meets Explainability: A Comprehensive Impact Benchmark
by: Saifullah, Saifullah, et al.
Published: (2022)
by: Saifullah, Saifullah, et al.
Published: (2022)
DP-DocLDM: Differentially Private Document Image Generation using Latent Diffusion Models
by: Saifullah, Saifullah, et al.
Published: (2025)
by: Saifullah, Saifullah, et al.
Published: (2025)
StylusAI: Stylistic Adaptation for Robust German Handwritten Text Generation
by: Riaz, Nauman, et al.
Published: (2024)
by: Riaz, Nauman, et al.
Published: (2024)
LD-ViCE: Latent Diffusion Model for Video Counterfactual Explanations
by: Varshney, Payal, et al.
Published: (2025)
by: Varshney, Payal, et al.
Published: (2025)
Discovering Concept Directions from Diffusion-based Counterfactuals via Latent Clustering
by: Varshney, Payal, et al.
Published: (2025)
by: Varshney, Payal, et al.
Published: (2025)
MedXplain-VQA: Multi-Component Explainable Medical Visual Question Answering
by: Nguyen, Hai-Dang, et al.
Published: (2025)
by: Nguyen, Hai-Dang, et al.
Published: (2025)
ForestVO: Enhancing Visual Odometry in Forest Environments through ForestGlue
by: Pritchard, Thomas, et al.
Published: (2025)
by: Pritchard, Thomas, et al.
Published: (2025)
Comparative Analysis of Image Enhancement Techniques for Brain Tumor Segmentation: Contrast, Histogram, and Hybrid Approaches
by: Saifullah, Shoffan, et al.
Published: (2024)
by: Saifullah, Shoffan, et al.
Published: (2024)
DocThinker: Explainable Multimodal Large Language Models with Rule-based Reinforcement Learning for Document Understanding
by: Yu, Wenwen, et al.
Published: (2025)
by: Yu, Wenwen, et al.
Published: (2025)
GlobalDoc: A Cross-Modal Vision-Language Framework for Real-World Document Image Retrieval and Classification
by: Bakkali, Souhail, et al.
Published: (2023)
by: Bakkali, Souhail, et al.
Published: (2023)
Automated Diabetic Screening via Anterior Segment Ocular Imaging: A Deep Learning and Explainable AI Approach
by: Maqsood, Hasaan, et al.
Published: (2026)
by: Maqsood, Hasaan, et al.
Published: (2026)
DocRes: A Generalist Model Toward Unifying Document Image Restoration Tasks
by: Zhang, Jiaxin, et al.
Published: (2024)
by: Zhang, Jiaxin, et al.
Published: (2024)
PSO-UNet: Particle Swarm-Optimized U-Net Framework for Precise Multimodal Brain Tumor Segmentation
by: Saifullah, Shoffan, et al.
Published: (2025)
by: Saifullah, Shoffan, et al.
Published: (2025)
Multi-Label Scene Classification in Remote Sensing Benefits from Image Super-Resolution
by: Mudraje, Ashitha, et al.
Published: (2025)
by: Mudraje, Ashitha, et al.
Published: (2025)
GazeXplain: Learning to Predict Natural Language Explanations of Visual Scanpaths
by: Chen, Xianyu, et al.
Published: (2024)
by: Chen, Xianyu, et al.
Published: (2024)
Q-Doc: Benchmarking Document Image Quality Assessment Capabilities in Multi-modal Large Language Models
by: Huang, Jiaxi, et al.
Published: (2025)
by: Huang, Jiaxi, et al.
Published: (2025)
CogDoc: Towards Unified thinking in Documents
by: Xu, Qixin, et al.
Published: (2025)
by: Xu, Qixin, et al.
Published: (2025)
Uni-DocDiff: A Unified Document Restoration Model Based on Diffusion
by: Zhao, Fangmin, et al.
Published: (2025)
by: Zhao, Fangmin, et al.
Published: (2025)
Rule-Based Reinforcement Learning for Document Image Classification with Vision Language Models
by: Jungo, Michael, et al.
Published: (2025)
by: Jungo, Michael, et al.
Published: (2025)
DeQA-Doc: Adapting DeQA-Score to Document Image Quality Assessment
by: Gao, Junjie, et al.
Published: (2025)
by: Gao, Junjie, et al.
Published: (2025)
Spherical Dense Text-to-Image Synthesis
by: Winter, Timon, et al.
Published: (2025)
by: Winter, Timon, et al.
Published: (2025)
Learning Interpretable Queries for Explainable Image Classification with Information Pursuit
by: Kolek, Stefan, et al.
Published: (2023)
by: Kolek, Stefan, et al.
Published: (2023)
In the Search for Optimal Multi-view Learning Models for Crop Classification with Global Remote Sensing Data
by: Mena, Francisco, et al.
Published: (2024)
by: Mena, Francisco, et al.
Published: (2024)
CountXplain: Interpretable Cell Counting with Prototype-Based Density Map Estimation
by: Mohammed, Abdurahman Ali, et al.
Published: (2025)
by: Mohammed, Abdurahman Ali, et al.
Published: (2025)
EMF: Event Meta Formers for Event-based Real-time Traffic Object Detection
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
DocRevive: A Unified Pipeline for Document Text Restoration
by: Purkayastha, Kunal, et al.
Published: (2026)
by: Purkayastha, Kunal, et al.
Published: (2026)
ObjBlur: A Curriculum Learning Approach With Progressive Object-Level Blurring for Improved Layout-to-Image Generation
by: Frolov, Stanislav, et al.
Published: (2024)
by: Frolov, Stanislav, et al.
Published: (2024)
DocSLM: A Small Vision-Language Model for Long Multimodal Document Understanding
by: Hannan, Tanveer, et al.
Published: (2025)
by: Hannan, Tanveer, et al.
Published: (2025)
MeDocVL: A Visual Language Model for Medical Document Understanding and Parsing
by: Wang, Wenjie, et al.
Published: (2026)
by: Wang, Wenjie, et al.
Published: (2026)
Class-Agnostic Region-of-Interest Matching in Document Images
by: Zhang, Demin, et al.
Published: (2025)
by: Zhang, Demin, et al.
Published: (2025)
DocShaDiffusion: Diffusion Model in Latent Space for Document Image Shadow Removal
by: Liu, Wenjie, et al.
Published: (2025)
by: Liu, Wenjie, et al.
Published: (2025)
DocAtlas: Multilingual Document Understanding Across 80+ Languages
by: Heakl, Ahmed, et al.
Published: (2026)
by: Heakl, Ahmed, et al.
Published: (2026)
Fine-Grained Scene Image Classification with Modality-Agnostic Adapter
by: Wang, Yiqun, et al.
Published: (2024)
by: Wang, Yiqun, et al.
Published: (2024)
SynthDoc: Bilingual Documents Synthesis for Visual Document Understanding
by: Ding, Chuanghao, et al.
Published: (2024)
by: Ding, Chuanghao, et al.
Published: (2024)
DocKylin: A Large Multimodal Model for Visual Document Understanding with Efficient Visual Slimming
by: Zhang, Jiaxin, et al.
Published: (2024)
by: Zhang, Jiaxin, et al.
Published: (2024)
DocDeshadower: Frequency-Aware Transformer for Document Shadow Removal
by: Zhou, Ziyang, et al.
Published: (2023)
by: Zhou, Ziyang, et al.
Published: (2023)
Modular Graph Extraction for Handwritten Circuit Diagram Images
by: Bayer, Johannes, et al.
Published: (2024)
by: Bayer, Johannes, et al.
Published: (2024)
Similar Items
-
DocVCE: Diffusion-based Visual Counterfactual Explanations for Document Image Classification
by: Saifullah, Saifullah, et al.
Published: (2025) -
Latent Diffusion for Guided Document Table Generation
by: Hamdani, Syed Jawwad Haider, et al.
Published: (2024) -
WordVIS: A Color Worth A Thousand Words
by: Khan, Umar, et al.
Published: (2024) -
Privacy Meets Explainability: A Comprehensive Impact Benchmark
by: Saifullah, Saifullah, et al.
Published: (2022) -
DP-DocLDM: Differentially Private Document Image Generation using Latent Diffusion Models
by: Saifullah, Saifullah, et al.
Published: (2025)