Privacy-Aware Document Visual Question Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Tito, Rubèn, Nguyen, Khanh, Tobaben, Marlon, Kerkouche, Raouf, Souibgui, Mohamed Ali, Jung, Kangsoo, Jälkö, Joonas, D'Andecy, Vincent Poulain, Joseph, Aurelie, Kang, Lei, Valveny, Ernest, Honkela, Antti, Fritz, Mario, Karatzas, Dimosthenis |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DocVXQA: Context-Aware Visual Explanations for Document Question Answering
by: Souibgui, Mohamed Ali, et al.
Published: (2025)
by: Souibgui, Mohamed Ali, et al.
Published: (2025)
Multi-Page Document Visual Question Answering using Self-Attention Scoring Mechanism
by: Kang, Lei, et al.
Published: (2024)
by: Kang, Lei, et al.
Published: (2024)
NeurIPS 2023 Competition: Privacy Preserving Federated Learning Document VQA
by: Tobaben, Marlon, et al.
Published: (2024)
by: Tobaben, Marlon, et al.
Published: (2024)
Hyperparameters in Score-Based Membership Inference Attacks
by: Pradhan, Gauri, et al.
Published: (2025)
by: Pradhan, Gauri, et al.
Published: (2025)
Impact of Dataset Properties on Membership Inference Vulnerability of Deep Transfer Learning
by: Tobaben, Marlon, et al.
Published: (2024)
by: Tobaben, Marlon, et al.
Published: (2024)
Efficient and Scalable Implementation of Differentially Private Deep Learning without Shortcuts
by: Beltran, Sebastian Rodriguez, et al.
Published: (2024)
by: Beltran, Sebastian Rodriguez, et al.
Published: (2024)
DocMIA: Document-Level Membership Inference Attacks against DocVQA Models
by: Nguyen, Khanh, et al.
Published: (2025)
by: Nguyen, Khanh, et al.
Published: (2025)
Federated Document Visual Question Answering: A Pilot Study
by: Nguyen, Khanh, et al.
Published: (2024)
by: Nguyen, Khanh, et al.
Published: (2024)
Learning Quantifiable Visual Explanations Without Ground-Truth
by: Singh, Amritpal, et al.
Published: (2026)
by: Singh, Amritpal, et al.
Published: (2026)
Subsampling is not Magic: Why Large Batch Sizes Work for Differentially Private Stochastic Optimisation
by: Räisä, Ossi, et al.
Published: (2024)
by: Räisä, Ossi, et al.
Published: (2024)
Noise-Aware Differentially Private Variational Inference
by: Alrawajfeh, Talal, et al.
Published: (2024)
by: Alrawajfeh, Talal, et al.
Published: (2024)
Beyond Membership: Limitations of Add/Remove Adjacency in Differential Privacy
by: Pradhan, Gauri, et al.
Published: (2025)
by: Pradhan, Gauri, et al.
Published: (2025)
Machine Unlearning for Document Classification
by: Kang, Lei, et al.
Published: (2024)
by: Kang, Lei, et al.
Published: (2024)
Empirical Comparison of Membership Inference Attacks in Deep Transfer Learning
by: Bai, Yuxuan, et al.
Published: (2025)
by: Bai, Yuxuan, et al.
Published: (2025)
ComicsPAP: understanding comic strips by picking the correct panel
by: Vivoli, Emanuele, et al.
Published: (2025)
by: Vivoli, Emanuele, et al.
Published: (2025)
On Reliability of Efficient Membership Inference Vulnerability Evaluation
by: Jälkö, Joonas, et al.
Published: (2026)
by: Jälkö, Joonas, et al.
Published: (2026)
Privacy Leakage via Output Label Space and Differentially Private Continual Learning
by: Tobaben, Marlon, et al.
Published: (2024)
by: Tobaben, Marlon, et al.
Published: (2024)
Image-text matching for large-scale book collections
by: Llabrés, Artemis, et al.
Published: (2024)
by: Llabrés, Artemis, et al.
Published: (2024)
Retrieval Augmented Verification for Zero-Shot Detection of Multimodal Disinformation
by: Dey, Arka Ujjal, et al.
Published: (2024)
by: Dey, Arka Ujjal, et al.
Published: (2024)
Preserving Privacy Without Compromising Accuracy: Machine Unlearning for Handwritten Text Recognition
by: Kang, Lei, et al.
Published: (2025)
by: Kang, Lei, et al.
Published: (2025)
Multimodal Transformer for Comics Text-Cloze
by: Vivoli, Emanuele, et al.
Published: (2024)
by: Vivoli, Emanuele, et al.
Published: (2024)
GRIF-DM: Generation of Rich Impression Fonts using Diffusion Models
by: Kang, Lei, et al.
Published: (2024)
by: Kang, Lei, et al.
Published: (2024)
Reading in the Dark: Low-light Scene Text Recognition
by: Fu, Xuanshuo, et al.
Published: (2026)
by: Fu, Xuanshuo, et al.
Published: (2026)
xLSTM-ECG: Multi-label ECG Classification via Feature Fusion with xLSTM
by: Kang, Lei, et al.
Published: (2025)
by: Kang, Lei, et al.
Published: (2025)
RAPTOR: Refined Approach for Product Table Object Recognition
by: Thomas, Eliott, et al.
Published: (2025)
by: Thomas, Eliott, et al.
Published: (2025)
QUEST: Quality-aware Semi-supervised Table Extraction for Business Documents
by: Thomas, Eliott, et al.
Published: (2025)
by: Thomas, Eliott, et al.
Published: (2025)
A Benchmark for Symbolic Reasoning from Pixel Sequences: Grid-Level Visual Completion and Correction
by: Kang, Lei, et al.
Published: (2025)
by: Kang, Lei, et al.
Published: (2025)
Spatially Grounded Explanations in Vision Language Models for Document Visual Question Answering
by: Lagos, Maximiliano Hormazábal, et al.
Published: (2025)
by: Lagos, Maximiliano Hormazábal, et al.
Published: (2025)
Counterfeit Answers: Adversarial Forgery against OCR-Free Document Visual Question Answering
by: Pintore, Marco, et al.
Published: (2025)
by: Pintore, Marco, et al.
Published: (2025)
LLM-Driven Medical Document Analysis: Enhancing Trustworthy Pathology and Differential Diagnosis
by: Kang, Lei, et al.
Published: (2025)
by: Kang, Lei, et al.
Published: (2025)
Accuracy-First Rényi Differential Privacy and Post-Processing Immunity
by: Räisä, Ossi, et al.
Published: (2025)
by: Räisä, Ossi, et al.
Published: (2025)
Pascal, Catalan, Motzkin triangles and tensor product multiplicities
by: d'Andecy, L. Poulain
Published: (2026)
by: d'Andecy, L. Poulain
Published: (2026)
AVIR: Adaptive Visual In-Document Retrieval for Efficient Multi-Page Document Question Answering
by: Li, Zongmin, et al.
Published: (2026)
by: Li, Zongmin, et al.
Published: (2026)
Noise-Aware Differentially Private Regression via Meta-Learning
by: Räisä, Ossi, et al.
Published: (2024)
by: Räisä, Ossi, et al.
Published: (2024)
$f$-Differential Privacy Filters: Validity and Approximate Solutions
by: Tran, Long, et al.
Published: (2026)
by: Tran, Long, et al.
Published: (2026)
One missing piece in Vision and Language: A Survey on Comics Understanding
by: Vivoli, Emanuele, et al.
Published: (2024)
by: Vivoli, Emanuele, et al.
Published: (2024)
A Fast Hierarchical Method for Multi-script and Arbitrary Oriented Scene Text Extraction
by: Gomez, Lluis, et al.
Published: (2014)
by: Gomez, Lluis, et al.
Published: (2014)
Framization of Schur--Weyl duality and Yokonuma--Hecke type algebras
by: Lacabanne, Abel, et al.
Published: (2023)
by: Lacabanne, Abel, et al.
Published: (2023)
Kazhdan-Lusztig bases of parabolic Hecke algebras and applications to Schur-Weyl duality
by: Guilhot, Jeremie, et al.
Published: (2026)
by: Guilhot, Jeremie, et al.
Published: (2026)
Little $q$-Jacobi polynomials and symmetry breaking operators for $U_q(sl_2)$
by: Labriet, Quentin, et al.
Published: (2025)
by: Labriet, Quentin, et al.
Published: (2025)
Similar Items
-
DocVXQA: Context-Aware Visual Explanations for Document Question Answering
by: Souibgui, Mohamed Ali, et al.
Published: (2025) -
Multi-Page Document Visual Question Answering using Self-Attention Scoring Mechanism
by: Kang, Lei, et al.
Published: (2024) -
NeurIPS 2023 Competition: Privacy Preserving Federated Learning Document VQA
by: Tobaben, Marlon, et al.
Published: (2024) -
Hyperparameters in Score-Based Membership Inference Attacks
by: Pradhan, Gauri, et al.
Published: (2025) -
Impact of Dataset Properties on Membership Inference Vulnerability of Deep Transfer Learning
by: Tobaben, Marlon, et al.
Published: (2024)