Multi-axis Analysis of Image Manipulation Localization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nichols, Keanu, Appapogu, Divya, Biamby, Giscard, Bashkirova, Dina, Rohrbach, Anna, Plummer, Bryan A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FuTCR: Future-Targeted Contrast and Repulsion for Continual Panoptic Segmentation
von: Ikechukwu, Nicholas, et al.
Veröffentlicht: (2026)
von: Ikechukwu, Nicholas, et al.
Veröffentlicht: (2026)
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
von: Qraitem, Maan, et al.
Veröffentlicht: (2023)
von: Qraitem, Maan, et al.
Veröffentlicht: (2023)
LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
von: Niu, Dantong, et al.
Veröffentlicht: (2024)
von: Niu, Dantong, et al.
Veröffentlicht: (2024)
RECAST: Reparameterized, Compact weight Adaptation for Sequential Tasks
von: Tasnim, Nazia, et al.
Veröffentlicht: (2024)
von: Tasnim, Nazia, et al.
Veröffentlicht: (2024)
ChA-MAEViT: Unifying Channel-Aware Masked Autoencoders and Multi-Channel Vision Transformers for Improved Cross-Channel Learning
von: Pham, Chau, et al.
Veröffentlicht: (2025)
von: Pham, Chau, et al.
Veröffentlicht: (2025)
Tuning Just Enough: Lightweight Backdoor Attacks on Multi-Encoder Diffusion Models
von: Chen, Ziyuan, et al.
Veröffentlicht: (2026)
von: Chen, Ziyuan, et al.
Veröffentlicht: (2026)
Noise-Aware Generalization: Robustness to In-Domain Noise and Out-of-Domain Generalization
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2026)
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2026)
Spurious-Aware Prototype Refinement for Reliable Out-of-Distribution Detection
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2025)
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2025)
Seeing Isn't Orienting: A Cognitively Grounded Benchmark Reveals Systematic Orientation Failures in MLLMs Supplementary
von: Tasnim, Nazia, et al.
Veröffentlicht: (2026)
von: Tasnim, Nazia, et al.
Veröffentlicht: (2026)
Seeing Isn't Orienting: A Cognitively Grounded Benchmark Reveals Systematic Orientation Failures in MLLMs
von: Tasnim, Nazia, et al.
Veröffentlicht: (2025)
von: Tasnim, Nazia, et al.
Veröffentlicht: (2025)
OP-LoRA: The Blessing of Dimensionality
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
Is Large-Scale Pretraining the Secret to Good Domain Generalization?
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
ERM++: An Improved Baseline for Domain Generalization
von: Teterwak, Piotr, et al.
Veröffentlicht: (2023)
von: Teterwak, Piotr, et al.
Veröffentlicht: (2023)
Topology Aware Neural Interpolation of Scalar Fields
von: Kissi, Mohamed, et al.
Veröffentlicht: (2025)
von: Kissi, Mohamed, et al.
Veröffentlicht: (2025)
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
von: Qraitem, Maan, et al.
Veröffentlicht: (2024)
von: Qraitem, Maan, et al.
Veröffentlicht: (2024)
Efficient Pre-training for Localized Instruction Generation of Videos
von: Batra, Anil, et al.
Veröffentlicht: (2023)
von: Batra, Anil, et al.
Veröffentlicht: (2023)
Omni-IML: Towards Unified Image Manipulation Localization
von: Qu, Chenfan, et al.
Veröffentlicht: (2024)
von: Qu, Chenfan, et al.
Veröffentlicht: (2024)
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
von: Liu, Aoming, et al.
Veröffentlicht: (2025)
von: Liu, Aoming, et al.
Veröffentlicht: (2025)
Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark
von: Wu, Tsung-Han, et al.
Veröffentlicht: (2024)
von: Wu, Tsung-Han, et al.
Veröffentlicht: (2024)
Beyond Memorization: Selective Learning for Copyright-Safe Diffusion Model Training
von: Kothandaraman, Divya, et al.
Veröffentlicht: (2025)
von: Kothandaraman, Divya, et al.
Veröffentlicht: (2025)
BEEM: Boosting Performance of Early Exit DNNs using Multi-Exit Classifiers as Experts
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
Sharp-It: A Multi-view to Multi-view Diffusion Model for 3D Synthesis and Manipulation
von: Edelstein, Yiftach, et al.
Veröffentlicht: (2024)
von: Edelstein, Yiftach, et al.
Veröffentlicht: (2024)
Local Policies Enable Zero-shot Long-horizon Manipulation
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
V$^2$Dial: Unification of Video and Visual Dialog via Multimodal Experts
von: Abdessaied, Adnen, et al.
Veröffentlicht: (2025)
von: Abdessaied, Adnen, et al.
Veröffentlicht: (2025)
ForgerySleuth: Empowering Multimodal Large Language Models for Image Manipulation Detection
von: Sun, Zhihao, et al.
Veröffentlicht: (2024)
von: Sun, Zhihao, et al.
Veröffentlicht: (2024)
FREE: Fast and Robust Vision Language Models with Early Exits
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
TwoHead-SwinFPN: A Unified DL Architecture for Synthetic Manipulation, Detection and Localization in Identity Documents
von: Naseeb, Chan, et al.
Veröffentlicht: (2026)
von: Naseeb, Chan, et al.
Veröffentlicht: (2026)
Scaling Image Geo-Localization to Continent Level
von: Lindenberger, Philipp, et al.
Veröffentlicht: (2025)
von: Lindenberger, Philipp, et al.
Veröffentlicht: (2025)
Contrastive Localized Language-Image Pre-Training
von: Chen, Hong-You, et al.
Veröffentlicht: (2024)
von: Chen, Hong-You, et al.
Veröffentlicht: (2024)
FedDistill: Global Model Distillation for Local Model De-Biasing in Non-IID Federated Learning
von: Song, Changlin, et al.
Veröffentlicht: (2024)
von: Song, Changlin, et al.
Veröffentlicht: (2024)
LNL+K: Enhancing Learning with Noisy Labels Through Noise Source Knowledge Integration
von: Wang, Siqi, et al.
Veröffentlicht: (2023)
von: Wang, Siqi, et al.
Veröffentlicht: (2023)
Shape-Guided Diffusion with Inside-Outside Attention
von: Park, Dong Huk, et al.
Veröffentlicht: (2022)
von: Park, Dong Huk, et al.
Veröffentlicht: (2022)
Enabling Local Editing in Diffusion Models by Joint and Individual Component Analysis
von: Kouzelis, Theodoros, et al.
Veröffentlicht: (2024)
von: Kouzelis, Theodoros, et al.
Veröffentlicht: (2024)
DEFAME: Dynamic Evidence-based FAct-checking with Multimodal Experts
von: Braun, Tobias, et al.
Veröffentlicht: (2024)
von: Braun, Tobias, et al.
Veröffentlicht: (2024)
Test-time augmentation improves efficiency in conformal prediction
von: Shanmugam, Divya, et al.
Veröffentlicht: (2025)
von: Shanmugam, Divya, et al.
Veröffentlicht: (2025)
LoFi: Neural Local Fields for Scalable Image Reconstruction
von: Khorashadizadeh, AmirEhsan, et al.
Veröffentlicht: (2024)
von: Khorashadizadeh, AmirEhsan, et al.
Veröffentlicht: (2024)
Localizing Anomalies via Multiscale Score Matching Analysis
von: Mahmood, Ahsan, et al.
Veröffentlicht: (2024)
von: Mahmood, Ahsan, et al.
Veröffentlicht: (2024)
HamVision: Hamiltonian Dynamics as Inductive Bias for Medical Image Analysis
von: Mabrok, Mohamed A
Veröffentlicht: (2026)
von: Mabrok, Mohamed A
Veröffentlicht: (2026)
The Narrow Gate: Localized Image-Text Communication in Native Multimodal Models
von: Serra, Alessandro Pietro, et al.
Veröffentlicht: (2024)
von: Serra, Alessandro Pietro, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FuTCR: Future-Targeted Contrast and Repulsion for Continual Panoptic Segmentation
von: Ikechukwu, Nicholas, et al.
Veröffentlicht: (2026) -
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
von: Qraitem, Maan, et al.
Veröffentlicht: (2023) -
LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
von: Niu, Dantong, et al.
Veröffentlicht: (2024) -
RECAST: Reparameterized, Compact weight Adaptation for Sequential Tasks
von: Tasnim, Nazia, et al.
Veröffentlicht: (2024) -
ChA-MAEViT: Unifying Channel-Aware Masked Autoencoders and Multi-Channel Vision Transformers for Improved Cross-Channel Learning
von: Pham, Chau, et al.
Veröffentlicht: (2025)