VAAS: Vision-Attention Anomaly Scoring for Image Manipulation Detection in Digital Forensics
Fuente:
arXiv
Guardado en:
| Autores principales: | Bamigbade, Opeyemi, Scanlon, Mark, Sheppard, John |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Computer Vision for Multimedia Geolocation in Human Trafficking Investigation: A Systematic Literature Review
por: Bamigbade, Opeyemi, et al.
Publicado: (2024)
por: Bamigbade, Opeyemi, et al.
Publicado: (2024)
Hiding Local Manipulations on SAR Images: a Counter-Forensic Attack
por: Mandelli, Sara, et al.
Publicado: (2024)
por: Mandelli, Sara, et al.
Publicado: (2024)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
por: Zhang, Zhenxing, et al.
Publicado: (2024)
por: Zhang, Zhenxing, et al.
Publicado: (2024)
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection
por: Cheung, Tsun-Hin, et al.
Publicado: (2024)
por: Cheung, Tsun-Hin, et al.
Publicado: (2024)
Segmentation-Based Attention Entropy: Detecting and Mitigating Object Hallucinations in Large Vision-Language Models
por: Song, Jiale, et al.
Publicado: (2026)
por: Song, Jiale, et al.
Publicado: (2026)
Backbone is All You Need: Assessing Vulnerabilities of Frozen Foundation Models in Synthetic Image Forensics
por: Musso, Chiara, et al.
Publicado: (2026)
por: Musso, Chiara, et al.
Publicado: (2026)
Harmfully Manipulated Images Matter in Multimodal Misinformation Detection
por: Wang, Bing, et al.
Publicado: (2024)
por: Wang, Bing, et al.
Publicado: (2024)
Do LLMs Understand Visual Anomalies? Uncovering LLM's Capabilities in Zero-shot Anomaly Detection
por: Zhu, Jiaqi, et al.
Publicado: (2024)
por: Zhu, Jiaqi, et al.
Publicado: (2024)
LookupForensics: A Large-Scale Multi-Task Dataset for Multi-Phase Image-Based Fact Verification
por: Cui, Shuhan, et al.
Publicado: (2024)
por: Cui, Shuhan, et al.
Publicado: (2024)
Efficient Vision Language Model Fine-tuning for Text-based Person Anomaly Search
por: He, Jiayi, et al.
Publicado: (2025)
por: He, Jiayi, et al.
Publicado: (2025)
QPT V2: Masked Image Modeling Advances Visual Scoring
por: Xie, Qizhi, et al.
Publicado: (2024)
por: Xie, Qizhi, et al.
Publicado: (2024)
GMFVAD: Using Grained Multi-modal Feature to Improve Video Anomaly Detection
por: Dai, Guangyu, et al.
Publicado: (2025)
por: Dai, Guangyu, et al.
Publicado: (2025)
Robust Modality-incomplete Anomaly Detection: A Modality-instructive Framework with Benchmark
por: Miao, Bingchen, et al.
Publicado: (2024)
por: Miao, Bingchen, et al.
Publicado: (2024)
Generalized Video Anomaly Event Detection: Systematic Taxonomy and Comparison of Deep Models
por: Liu, Yang, et al.
Publicado: (2023)
por: Liu, Yang, et al.
Publicado: (2023)
MSCT: Differential Cross-Modal Attention for Deepfake Detection
por: Wei, Fangda, et al.
Publicado: (2026)
por: Wei, Fangda, et al.
Publicado: (2026)
EntroAD: Structural Entropy-Guided Prompt Adaptation for Zero-Shot Anomaly Detection
por: Zhao, Xinyu, et al.
Publicado: (2026)
por: Zhao, Xinyu, et al.
Publicado: (2026)
Embedded Heterogeneous Attention Transformer for Cross-lingual Image Captioning
por: Song, Zijie, et al.
Publicado: (2023)
por: Song, Zijie, et al.
Publicado: (2023)
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
por: Xie, Liping, et al.
Publicado: (2025)
por: Xie, Liping, et al.
Publicado: (2025)
Group-based Distinctive Image Captioning with Memory Difference Encoding and Attention
por: Wang, Jiuniu, et al.
Publicado: (2025)
por: Wang, Jiuniu, et al.
Publicado: (2025)
Mitigating Image Captioning Hallucinations in Vision-Language Models
por: Zhao, Fei, et al.
Publicado: (2025)
por: Zhao, Fei, et al.
Publicado: (2025)
Where Does Vision Meet Language? Understanding and Refining Visual Fusion in MLLMs via Contrastive Attention
por: Song, Shezheng, et al.
Publicado: (2026)
por: Song, Shezheng, et al.
Publicado: (2026)
Beyond Walking: A Large-Scale Image-Text Benchmark for Text-based Person Anomaly Search
por: Yang, Shuyu, et al.
Publicado: (2024)
por: Yang, Shuyu, et al.
Publicado: (2024)
Off-the-shelf Vision Models Benefit Image Manipulation Localization
por: Zhang, Zhengxuan, et al.
Publicado: (2026)
por: Zhang, Zhengxuan, et al.
Publicado: (2026)
Fine-grained Image Retrieval via Dual-Vision Adaptation
por: Jiang, Xin, et al.
Publicado: (2025)
por: Jiang, Xin, et al.
Publicado: (2025)
TALE: Training-free Cross-domain Image Composition via Adaptive Latent Manipulation and Energy-guided Optimization
por: Pham, Kien T., et al.
Publicado: (2024)
por: Pham, Kien T., et al.
Publicado: (2024)
Cross-modal Proxy Evolving for OOD Detection with Vision-Language Models
por: Tang, Hao, et al.
Publicado: (2026)
por: Tang, Hao, et al.
Publicado: (2026)
A Perspective on Deep Vision Performance with Standard Image and Video Codecs
por: Reich, Christoph, et al.
Publicado: (2024)
por: Reich, Christoph, et al.
Publicado: (2024)
ESIQA: Perceptual Quality Assessment of Vision-Pro-based Egocentric Spatial Images
por: Zhu, Xilei, et al.
Publicado: (2024)
por: Zhu, Xilei, et al.
Publicado: (2024)
AttentionBender: Manipulating Cross-Attention in Video Diffusion Transformers as a Creative Probe
por: Cole, Adam, et al.
Publicado: (2026)
por: Cole, Adam, et al.
Publicado: (2026)
Anomaly Detection and Localization for Speech Deepfakes via Feature Pyramid Matching
por: Coletta, Emma, et al.
Publicado: (2025)
por: Coletta, Emma, et al.
Publicado: (2025)
Generalized Face Forgery Detection via Adaptive Learning for Pre-trained Vision Transformer
por: Luo, Anwei, et al.
Publicado: (2023)
por: Luo, Anwei, et al.
Publicado: (2023)
Vision-Language Models Learn Super Images for Efficient Partially Relevant Video Retrieval
por: Nishimura, Taichi, et al.
Publicado: (2023)
por: Nishimura, Taichi, et al.
Publicado: (2023)
Media Forensics and Deepfake Systematic Survey
por: CH, Nadeem Jabbar, et al.
Publicado: (2024)
por: CH, Nadeem Jabbar, et al.
Publicado: (2024)
PDA: Text-Augmented Defense Framework for Robust Vision-Language Models against Adversarial Image Attacks
por: Xu, Jingning, et al.
Publicado: (2026)
por: Xu, Jingning, et al.
Publicado: (2026)
Towards Real-World Adverse Weather Image Restoration: Enhancing Clearness and Semantics with Vision-Language Models
por: Xu, Jiaqi, et al.
Publicado: (2024)
por: Xu, Jiaqi, et al.
Publicado: (2024)
Enhancing Interactive Image Retrieval With Query Rewriting Using Large Language Models and Vision Language Models
por: Zhu, Hongyi, et al.
Publicado: (2024)
por: Zhu, Hongyi, et al.
Publicado: (2024)
Depth and Image Fusion for Road Obstacle Detection Using Stereo Camera
por: Perezyabov, Oleg, et al.
Publicado: (2025)
por: Perezyabov, Oleg, et al.
Publicado: (2025)
Copy-Move Forgery Detection and Question Answering for Remote Sensing Image
por: Zhang, Ze, et al.
Publicado: (2024)
por: Zhang, Ze, et al.
Publicado: (2024)
Learning Efficient Unsupervised Satellite Image-based Building Damage Detection
por: Zhang, Yiyun, et al.
Publicado: (2023)
por: Zhang, Yiyun, et al.
Publicado: (2023)
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration
por: Fazli, Mehrdad, et al.
Publicado: (2025)
por: Fazli, Mehrdad, et al.
Publicado: (2025)
Ejemplares similares
-
Computer Vision for Multimedia Geolocation in Human Trafficking Investigation: A Systematic Literature Review
por: Bamigbade, Opeyemi, et al.
Publicado: (2024) -
Hiding Local Manipulations on SAR Images: a Counter-Forensic Attack
por: Mandelli, Sara, et al.
Publicado: (2024) -
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
por: Zhang, Zhenxing, et al.
Publicado: (2024) -
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection
por: Cheung, Tsun-Hin, et al.
Publicado: (2024) -
Segmentation-Based Attention Entropy: Detecting and Mitigating Object Hallucinations in Large Vision-Language Models
por: Song, Jiale, et al.
Publicado: (2026)