DocShield: Towards AI Document Safety via Evidence-Grounded Agentic Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Fanwei, Miao, Changtao, Huang, Jing, Tan, Zhiya, Gong, Shutao, Yu, Xiaoming, Wang, Yang, Yao, Weibin, Zhou, Joey Tianyi, Li, Jianshu, Yan, Yin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FlexDoc: Parameterized Sampling for Diverse Multilingual Synthetic Documents for Training Document Understanding Models
by: Dua, Karan, et al.
Published: (2025)
by: Dua, Karan, et al.
Published: (2025)
LaV-CoT: Language-Aware Visual CoT with Multi-Aspect Reward Optimization for Real-World Multilingual VQA
by: Huang, Jing, et al.
Published: (2025)
by: Huang, Jing, et al.
Published: (2025)
LogicLens: Visual-Logical Co-Reasoning for Text-Centric Forgery Analysis
by: Zeng, Fanwei, et al.
Published: (2025)
by: Zeng, Fanwei, et al.
Published: (2025)
SH17: A Dataset for Human Safety and Personal Protective Equipment Detection in Manufacturing Industry
by: Ahmad, Hafiz Mughees, et al.
Published: (2024)
by: Ahmad, Hafiz Mughees, et al.
Published: (2024)
Grounding Synthetic Data Generation With Vision and Language Models
by: Çağlar, Ümit Mert, et al.
Published: (2026)
by: Çağlar, Ümit Mert, et al.
Published: (2026)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
by: Tourani, Ali, et al.
Published: (2023)
by: Tourani, Ali, et al.
Published: (2023)
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
by: Radwan, Ahmed, et al.
Published: (2024)
by: Radwan, Ahmed, et al.
Published: (2024)
DIsoN: Decentralized Isolation Networks for Out-of-Distribution Detection in Medical Imaging
by: Wagner, Felix, et al.
Published: (2025)
by: Wagner, Felix, et al.
Published: (2025)
Decoupled Sensitivity-Consistency Learning for Weakly Supervised Video Anomaly Detection
by: Zheng, Hantao, et al.
Published: (2026)
by: Zheng, Hantao, et al.
Published: (2026)
Can LLM Agents Respond to Disasters? Benchmarking Heterogeneous Geospatial Reasoning in Emergency Operations
by: Wang, Junjue, et al.
Published: (2026)
by: Wang, Junjue, et al.
Published: (2026)
Data Augmentation in Earth Observation: A Diffusion Model Approach
by: Sousa, Tiago, et al.
Published: (2024)
by: Sousa, Tiago, et al.
Published: (2024)
Chat-Driven Text Generation and Interaction for Person Retrieval
by: Xie, Zequn, et al.
Published: (2025)
by: Xie, Zequn, et al.
Published: (2025)
Advanced Long-term Earth System Forecasting
by: Wu, Hao, et al.
Published: (2025)
by: Wu, Hao, et al.
Published: (2025)
IAMAP: Unlocking Deep Learning in QGIS for non-coders and limited computing resources
by: Tresson, Paul, et al.
Published: (2025)
by: Tresson, Paul, et al.
Published: (2025)
Optimizing Multi-Scale Representations to Detect Effect Heterogeneity Using Earth Observation and Computer Vision: Applications to Two Anti-Poverty RCTs
by: Zhu, Fucheng Warren, et al.
Published: (2024)
by: Zhu, Fucheng Warren, et al.
Published: (2024)
Bridge Diffusion Model: Bridge Chinese Text-to-Image Diffusion Model with English Communities
by: Liu, Shanyuan, et al.
Published: (2023)
by: Liu, Shanyuan, et al.
Published: (2023)
RipVIS: Rip Currents Video Instance Segmentation Benchmark for Beach Monitoring and Safety
by: Dumitriu, Andrei, et al.
Published: (2025)
by: Dumitriu, Andrei, et al.
Published: (2025)
Low-Cost Tree Crown Dieback Estimation Using Deep Learning-Based Segmentation
by: Allen, M. J., et al.
Published: (2024)
by: Allen, M. J., et al.
Published: (2024)
PTB-XL-Image-17K: A Large-Scale Synthetic ECG Image Dataset with Comprehensive Ground Truth for Deep Learning-Based Digitization
by: Mehdi, Naqcho Ali, et al.
Published: (2026)
by: Mehdi, Naqcho Ali, et al.
Published: (2026)
Estimating optical vegetation indices and biophysical variables for temperate forests with Sentinel-1 SAR data using machine learning techniques: A case study for Czechia
by: Paluba, Daniel, et al.
Published: (2023)
by: Paluba, Daniel, et al.
Published: (2023)
Rapid Adaptation of Earth Observation Foundation Models for Segmentation
by: Selvam, Karthick Panner, et al.
Published: (2024)
by: Selvam, Karthick Panner, et al.
Published: (2024)
GPT4o-Receipt: A Dataset and Human Study for AI-Generated Document Forensics
by: Zhang, Yan, et al.
Published: (2026)
by: Zhang, Yan, et al.
Published: (2026)
CANSURF: An ASV-View Can Dataset and Benchmark for Detection and Tracking of Surface-Level Debris
by: Aljundi, Zaid, et al.
Published: (2026)
by: Aljundi, Zaid, et al.
Published: (2026)
Pairwise Spatiotemporal Partial Trajectory Matching for Co-movement Analysis
by: Cardei, Maria, et al.
Published: (2024)
by: Cardei, Maria, et al.
Published: (2024)
SelvaMask: Segmenting Trees in Tropical Forests and Beyond
by: Duguay, Simon-Olivier, et al.
Published: (2026)
by: Duguay, Simon-Olivier, et al.
Published: (2026)
MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation
by: Bartkowiak, Patryk, et al.
Published: (2026)
by: Bartkowiak, Patryk, et al.
Published: (2026)
Event-based Solutions for Human-centered Applications: A Comprehensive Review
by: Adra, Mira, et al.
Published: (2025)
by: Adra, Mira, et al.
Published: (2025)
Optimizing the image correction pipeline for pedestrian detection in the thermal-infrared domain
by: Karam, Christophe, et al.
Published: (2024)
by: Karam, Christophe, et al.
Published: (2024)
Data Augmentation with Diffusion Models for Colon Polyp Localization on the Low Data Regime: How much real data is enough?
by: Tormos, Adrian, et al.
Published: (2024)
by: Tormos, Adrian, et al.
Published: (2024)
KidsNanny: A Two-Stage Multimodal Content Moderation Pipeline Integrating Visual Classification, Object Detection, OCR, and Contextual Reasoning for Child Safety
by: Panchal, Viraj, et al.
Published: (2026)
by: Panchal, Viraj, et al.
Published: (2026)
DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models
by: Zhou, Yue, et al.
Published: (2026)
by: Zhou, Yue, et al.
Published: (2026)
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
by: Lu, Wanglong, et al.
Published: (2025)
by: Lu, Wanglong, et al.
Published: (2025)
MdaIF: Robust One-Stop Multi-Degradation-Aware Image Fusion with Language-Driven Semantics
by: Li, Jing, et al.
Published: (2025)
by: Li, Jing, et al.
Published: (2025)
Optimal Blackjack Strategy Recommender: A Comprehensive Study on Computer Vision Integration for Enhanced Gameplay
by: Gupta, Krishnanshu, et al.
Published: (2024)
by: Gupta, Krishnanshu, et al.
Published: (2024)
A Light Perspective for 3D Object Detection
by: Pederiva, Marcelo Eduardo, et al.
Published: (2025)
by: Pederiva, Marcelo Eduardo, et al.
Published: (2025)
Neural Fields for 3D Tracking of Anatomy and Surgical Instruments in Monocular Laparoscopic Video Clips
by: Gerats, Beerend G. A., et al.
Published: (2024)
by: Gerats, Beerend G. A., et al.
Published: (2024)
Combining Absolute and Semi-Generalized Relative Poses for Visual Localization
by: Panek, Vojtech, et al.
Published: (2024)
by: Panek, Vojtech, et al.
Published: (2024)
A Guide to Structureless Visual Localization
by: Panek, Vojtech, et al.
Published: (2025)
by: Panek, Vojtech, et al.
Published: (2025)
Ego-Motion Aware Target Prediction Module for Robust Multi-Object Tracking
by: Mahdian, Navid, et al.
Published: (2024)
by: Mahdian, Navid, et al.
Published: (2024)
Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models
by: Seo, Huichan, et al.
Published: (2025)
by: Seo, Huichan, et al.
Published: (2025)
Similar Items
-
FlexDoc: Parameterized Sampling for Diverse Multilingual Synthetic Documents for Training Document Understanding Models
by: Dua, Karan, et al.
Published: (2025) -
LaV-CoT: Language-Aware Visual CoT with Multi-Aspect Reward Optimization for Real-World Multilingual VQA
by: Huang, Jing, et al.
Published: (2025) -
LogicLens: Visual-Logical Co-Reasoning for Text-Centric Forgery Analysis
by: Zeng, Fanwei, et al.
Published: (2025) -
SH17: A Dataset for Human Safety and Personal Protective Equipment Detection in Manufacturing Industry
by: Ahmad, Hafiz Mughees, et al.
Published: (2024) -
Grounding Synthetic Data Generation With Vision and Language Models
by: Çağlar, Ümit Mert, et al.
Published: (2026)