Exploiting Modality-Specific Features For Multi-Modal Manipulation Detection And Grounding
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Jiazhen, Liu, Bin, Miao, Changtao, Zhao, Zhiwei, Zhuang, Wanyi, Chu, Qi, Yu, Nenghai |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-spectral Class Center Network for Face Manipulation Detection and Localization
by: Miao, Changtao, et al.
Published: (2023)
by: Miao, Changtao, et al.
Published: (2023)
Mixture-of-Noises Enhanced Forgery-Aware Predictor for Multi-Face Manipulation Detection and Localization
by: Miao, Changtao, et al.
Published: (2024)
by: Miao, Changtao, et al.
Published: (2024)
Training-Free In-Context Forensic Chain for Image Manipulation Detection and Localization
by: Chen, Rui, et al.
Published: (2025)
by: Chen, Rui, et al.
Published: (2025)
LAKAN: Landmark-assisted Adaptive Kolmogorov-Arnold Network for Face Forgery Detection
by: Jiang, Jiayao, et al.
Published: (2025)
by: Jiang, Jiayao, et al.
Published: (2025)
SAPL: Semantic-Agnostic Prompt Learning in CLIP for Weakly Supervised Image Manipulation Localization
by: Wang, Xinghao, et al.
Published: (2026)
by: Wang, Xinghao, et al.
Published: (2026)
Unleashing the Potential of Consistency Learning for Detecting and Grounding Multi-Modal Media Manipulation
by: Li, Yiheng, et al.
Published: (2025)
by: Li, Yiheng, et al.
Published: (2025)
Context-Aware Weakly Supervised Image Manipulation Localization with SAM Refinement
by: Wang, Xinghao, et al.
Published: (2025)
by: Wang, Xinghao, et al.
Published: (2025)
Rethinking Multi-Modal Object Detection from the Perspective of Mono-Modality Feature Learning
by: Zhao, Tianyi, et al.
Published: (2025)
by: Zhao, Tianyi, et al.
Published: (2025)
Advancing Aesthetic Image Generation via Composition Transfer
by: Zou, Kai, et al.
Published: (2026)
by: Zou, Kai, et al.
Published: (2026)
GuardTrace-VL: Detecting Unsafe Multimodel Reasoning via Iterative Safety Supervision
by: Xiang, Yuxiao, et al.
Published: (2025)
by: Xiang, Yuxiao, et al.
Published: (2025)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
by: Zhang, Zhenxing, et al.
Published: (2024)
by: Zhang, Zhenxing, et al.
Published: (2024)
MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection
by: Wang, Zitian, et al.
Published: (2024)
by: Wang, Zitian, et al.
Published: (2024)
TCI-Former: Thermal Conduction-Inspired Transformer for Infrared Small Target Detection
by: Chen, Tianxiang, et al.
Published: (2024)
by: Chen, Tianxiang, et al.
Published: (2024)
Multi-modal Learning with Missing Modality via Shared-Specific Feature Modelling
by: Wang, Hu, et al.
Published: (2023)
by: Wang, Hu, et al.
Published: (2023)
OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation
by: Huang, Qidong, et al.
Published: (2023)
by: Huang, Qidong, et al.
Published: (2023)
Multi-Modal Face Anti-Spoofing via Cross-Modal Feature Transitions
by: Chong, Jun-Xiong, et al.
Published: (2025)
by: Chong, Jun-Xiong, et al.
Published: (2025)
Robust Modality-incomplete Anomaly Detection: A Modality-instructive Framework with Benchmark
by: Miao, Bingchen, et al.
Published: (2024)
by: Miao, Bingchen, et al.
Published: (2024)
Counterfactual Intervention Feature Transfer for Visible-Infrared Person Re-identification
by: Li, Xulin, et al.
Published: (2022)
by: Li, Xulin, et al.
Published: (2022)
Benchmarking Unified Face Attack Detection via Hierarchical Prompt Tuning
by: Liu, Ajian, et al.
Published: (2025)
by: Liu, Ajian, et al.
Published: (2025)
UniMMAD: Unified Multi-Modal and Multi-Class Anomaly Detection via MoE-Driven Feature Decompression
by: Zhao, Yuan, et al.
Published: (2025)
by: Zhao, Yuan, et al.
Published: (2025)
ModaVerse: Efficiently Transforming Modalities with LLMs
by: Wang, Xinyu, et al.
Published: (2024)
by: Wang, Xinyu, et al.
Published: (2024)
TUNI: Real-time RGB-T Semantic Segmentation with Unified Multi-Modal Feature Extraction and Cross-Modal Feature Fusion
by: Guo, Xiaodong, et al.
Published: (2025)
by: Guo, Xiaodong, et al.
Published: (2025)
Robust Pedestrian Detection with Uncertain Modality
by: Bie, Qian, et al.
Published: (2026)
by: Bie, Qian, et al.
Published: (2026)
Modality-Specific Enhancement and Complementary Fusion for Semi-Supervised Multi-Modal Brain Tumor Segmentation
by: Chung, Tien-Dat, et al.
Published: (2025)
by: Chung, Tien-Dat, et al.
Published: (2025)
MSCT: Differential Cross-Modal Attention for Deepfake Detection
by: Wei, Fangda, et al.
Published: (2026)
by: Wei, Fangda, et al.
Published: (2026)
Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration Rate
by: Huang, Qidong, et al.
Published: (2024)
by: Huang, Qidong, et al.
Published: (2024)
The Courtroom Trial of Pixels: Robust Image Manipulation Localization via Adversarial Evidence and Reinforcement Learning Judgment
by: Li, Songlin, et al.
Published: (2026)
by: Li, Songlin, et al.
Published: (2026)
Modality-Agnostic Prompt Learning for Multi-Modal Camouflaged Object Detection
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Feature Fusion and Knowledge-Distilled Multi-Modal Multi-Target Detection
by: Do, Ngoc Tuyen, et al.
Published: (2025)
by: Do, Ngoc Tuyen, et al.
Published: (2025)
Decoupling Feature Representations of Ego and Other Modalities for Incomplete Multi-modal Brain Tumor Segmentation
by: Yang, Kaixiang, et al.
Published: (2024)
by: Yang, Kaixiang, et al.
Published: (2024)
MLVTG: Mamba-Based Feature Alignment and LLM-Driven Purification for Multi-Modal Video Temporal Grounding
by: Zhu, Zhiyi, et al.
Published: (2025)
by: Zhu, Zhiyi, et al.
Published: (2025)
MultiMAE for Brain MRIs: Robustness to Missing Inputs Using Multi-Modal Masked Autoencoder
by: Erdur, Ayhan Can, et al.
Published: (2025)
by: Erdur, Ayhan Can, et al.
Published: (2025)
MiM-ISTD: Mamba-in-Mamba for Efficient Infrared Small Target Detection
by: Chen, Tianxiang, et al.
Published: (2024)
by: Chen, Tianxiang, et al.
Published: (2024)
Modality-Aware Feature Matching: A Comprehensive Review of Single- and Cross-Modality Techniques
by: Liu, Weide, et al.
Published: (2025)
by: Liu, Weide, et al.
Published: (2025)
DEYOLO: Dual-Feature-Enhancement YOLO for Cross-Modality Object Detection
by: Chen, Yishuo, et al.
Published: (2024)
by: Chen, Yishuo, et al.
Published: (2024)
MIFNet: Learning Modality-Invariant Features for Generalizable Multimodal Image Matching
by: Liu, Yepeng, et al.
Published: (2025)
by: Liu, Yepeng, et al.
Published: (2025)
Modality-Specific Hierarchical Enhancement for RGB-D Camouflaged Object Detection
by: Niu, Yuzhen, et al.
Published: (2026)
by: Niu, Yuzhen, et al.
Published: (2026)
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
CMF-IoU: Multi-Stage Cross-Modal Fusion 3D Object Detection with IoU Joint Prediction
by: Ning, Zhiwei, et al.
Published: (2025)
by: Ning, Zhiwei, et al.
Published: (2025)
M2I2HA: Multi-modal Object Detection Based on Intra- and Inter-Modal Hypergraph Attention
by: Yang, Xiaofan, et al.
Published: (2026)
by: Yang, Xiaofan, et al.
Published: (2026)
Similar Items
-
Multi-spectral Class Center Network for Face Manipulation Detection and Localization
by: Miao, Changtao, et al.
Published: (2023) -
Mixture-of-Noises Enhanced Forgery-Aware Predictor for Multi-Face Manipulation Detection and Localization
by: Miao, Changtao, et al.
Published: (2024) -
Training-Free In-Context Forensic Chain for Image Manipulation Detection and Localization
by: Chen, Rui, et al.
Published: (2025) -
LAKAN: Landmark-assisted Adaptive Kolmogorov-Arnold Network for Face Forgery Detection
by: Jiang, Jiayao, et al.
Published: (2025) -
SAPL: Semantic-Agnostic Prompt Learning in CLIP for Weakly Supervised Image Manipulation Localization
by: Wang, Xinghao, et al.
Published: (2026)