Saved in:
| Main Authors: | Rashid, Maisha Binte, Rivas, Pablo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2502.16361 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MetaCloak-JPEG: JPEG-Robust Adversarial Perturbation for Preventing Unauthorized DreamBooth-Based Deepfake Generation
by: Fardin, Tanjim Rahaman, et al.
Published: (2026)
by: Fardin, Tanjim Rahaman, et al.
Published: (2026)
Vision Token Masking Alone Cannot Prevent PHI Leakage in Medical Document OCR: A Systematic Evaluation
by: Young, Richard J.
Published: (2025)
by: Young, Richard J.
Published: (2025)
Fixed-Threshold Evaluation of a Hybrid CNN-ViT for AI-Generated Image Detection Across Photos and Art
by: Khan, Md Ashik, et al.
Published: (2025)
by: Khan, Md Ashik, et al.
Published: (2025)
Detecting Inpainted Video with Frequency Domain Insights
by: Tang, Quanhui, et al.
Published: (2024)
by: Tang, Quanhui, et al.
Published: (2024)
Optimizing the image correction pipeline for pedestrian detection in the thermal-infrared domain
by: Karam, Christophe, et al.
Published: (2024)
by: Karam, Christophe, et al.
Published: (2024)
A Real-Time Diminished Reality Approach to Privacy in MR Collaboration
by: Fane, Christian
Published: (2025)
by: Fane, Christian
Published: (2025)
Threats and Opportunities in AI-generated Images for Armed Forces
by: Meier, Raphael
Published: (2025)
by: Meier, Raphael
Published: (2025)
The Architecture of Trust: A Framework for AI-Augmented Real Estate Valuation in the Era of Structured Data
by: Teikari, Petteri, et al.
Published: (2025)
by: Teikari, Petteri, et al.
Published: (2025)
SelvaMask: Segmenting Trees in Tropical Forests and Beyond
by: Duguay, Simon-Olivier, et al.
Published: (2026)
by: Duguay, Simon-Olivier, et al.
Published: (2026)
KidsNanny: A Two-Stage Multimodal Content Moderation Pipeline Integrating Visual Classification, Object Detection, OCR, and Contextual Reasoning for Child Safety
by: Panchal, Viraj, et al.
Published: (2026)
by: Panchal, Viraj, et al.
Published: (2026)
Positive Style Accumulation: A Style Screening and Continuous Utilization Framework for Federated DG-ReID
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
A Light Perspective for 3D Object Detection
by: Pederiva, Marcelo Eduardo, et al.
Published: (2025)
by: Pederiva, Marcelo Eduardo, et al.
Published: (2025)
Combining Absolute and Semi-Generalized Relative Poses for Visual Localization
by: Panek, Vojtech, et al.
Published: (2024)
by: Panek, Vojtech, et al.
Published: (2024)
A Guide to Structureless Visual Localization
by: Panek, Vojtech, et al.
Published: (2025)
by: Panek, Vojtech, et al.
Published: (2025)
Reference Dataset and Benchmark for Reconstructing Laser Parameters from On-axis Video in Powder Bed Fusion of Bulk Stainless Steel
by: Blanc, Cyril, et al.
Published: (2024)
by: Blanc, Cyril, et al.
Published: (2024)
Facial Attribute Based Text Guided Face Anonymization
by: Muştu, Mustafa İzzet, et al.
Published: (2025)
by: Muştu, Mustafa İzzet, et al.
Published: (2025)
Privacy-Preserving Structureless Visual Localization via Image Obfuscation
by: Panek, Vojtech, et al.
Published: (2026)
by: Panek, Vojtech, et al.
Published: (2026)
HATL: Hierarchical Adaptive-Transfer Learning Framework for Sign Language Machine Translation
by: Shahin, Nada, et al.
Published: (2026)
by: Shahin, Nada, et al.
Published: (2026)
Hybrid Knowledge Transfer through Attention and Logit Distillation for On-Device Vision Systems in Agricultural IoT
by: Mugisha, Stanley, et al.
Published: (2025)
by: Mugisha, Stanley, et al.
Published: (2025)
SH17: A Dataset for Human Safety and Personal Protective Equipment Detection in Manufacturing Industry
by: Ahmad, Hafiz Mughees, et al.
Published: (2024)
by: Ahmad, Hafiz Mughees, et al.
Published: (2024)
Parameter-efficient fine-tuning (PEFT) of Vision Foundation Models for Atypical Mitotic Figure Classification
by: Ramchandani, Lavish, et al.
Published: (2025)
by: Ramchandani, Lavish, et al.
Published: (2025)
MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation
by: Bartkowiak, Patryk, et al.
Published: (2026)
by: Bartkowiak, Patryk, et al.
Published: (2026)
Ego-Motion Aware Target Prediction Module for Robust Multi-Object Tracking
by: Mahdian, Navid, et al.
Published: (2024)
by: Mahdian, Navid, et al.
Published: (2024)
NumeriKontrol: Adding Numeric Control to Diffusion Transformers for Instruction-based Image Editing
by: Xu, Zhenyu, et al.
Published: (2025)
by: Xu, Zhenyu, et al.
Published: (2025)
Accelerating Post-Tornado Disaster Assessment Using Advanced Deep Learning Models
by: Umeike, Robinson, et al.
Published: (2024)
by: Umeike, Robinson, et al.
Published: (2024)
Bridge Diffusion Model: Bridge Chinese Text-to-Image Diffusion Model with English Communities
by: Liu, Shanyuan, et al.
Published: (2023)
by: Liu, Shanyuan, et al.
Published: (2023)
Revealing an Unattractivity Bias in Mental Reconstruction of Occluded Faces using Generative Image Models
by: Riedmann, Frederik, et al.
Published: (2024)
by: Riedmann, Frederik, et al.
Published: (2024)
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
by: Käs, Stephanie, et al.
Published: (2025)
by: Käs, Stephanie, et al.
Published: (2025)
Rethinking VLMs for Image Forgery Detection and Localization
by: Guo, Shaofeng, et al.
Published: (2026)
by: Guo, Shaofeng, et al.
Published: (2026)
Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models
by: Seo, Huichan, et al.
Published: (2025)
by: Seo, Huichan, et al.
Published: (2025)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
Evolving to the Aesthetics of a Vision-Language Model
by: Krol, Stephen James, et al.
Published: (2026)
by: Krol, Stephen James, et al.
Published: (2026)
AVControl: Efficient Framework for Training Audio-Visual Controls
by: Ben-Yosef, Matan, et al.
Published: (2026)
by: Ben-Yosef, Matan, et al.
Published: (2026)
A Comprehensive Review of Fish Feeding Behavior Analysis in Aquaculture: Tasks, Techniques, and Applications
by: Zhang, Shulong, et al.
Published: (2025)
by: Zhang, Shulong, et al.
Published: (2025)
Towards Localizing Structural Elements: Merging Geometrical Detection with Semantic Verification in RGB-D Data
by: Tourani, Ali, et al.
Published: (2024)
by: Tourani, Ali, et al.
Published: (2024)
Vision transformers in domain adaptation and domain generalization: a study of robustness
by: Alijani, Shadi, et al.
Published: (2024)
by: Alijani, Shadi, et al.
Published: (2024)
A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors
by: Doan, Gia-Bao, et al.
Published: (2026)
by: Doan, Gia-Bao, et al.
Published: (2026)
GPT4o-Receipt: A Dataset and Human Study for AI-Generated Document Forensics
by: Zhang, Yan, et al.
Published: (2026)
by: Zhang, Yan, et al.
Published: (2026)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
by: Tourani, Ali, et al.
Published: (2023)
by: Tourani, Ali, et al.
Published: (2023)
Intrinsic Image Fusion for Multi-View 3D Material Reconstruction
by: Kocsis, Peter, et al.
Published: (2025)
by: Kocsis, Peter, et al.
Published: (2025)
Similar Items
-
MetaCloak-JPEG: JPEG-Robust Adversarial Perturbation for Preventing Unauthorized DreamBooth-Based Deepfake Generation
by: Fardin, Tanjim Rahaman, et al.
Published: (2026) -
Vision Token Masking Alone Cannot Prevent PHI Leakage in Medical Document OCR: A Systematic Evaluation
by: Young, Richard J.
Published: (2025) -
Fixed-Threshold Evaluation of a Hybrid CNN-ViT for AI-Generated Image Detection Across Photos and Art
by: Khan, Md Ashik, et al.
Published: (2025) -
Detecting Inpainted Video with Frequency Domain Insights
by: Tang, Quanhui, et al.
Published: (2024) -
Optimizing the image correction pipeline for pedestrian detection in the thermal-infrared domain
by: Karam, Christophe, et al.
Published: (2024)