Multi-axis Analysis of Image Manipulation Localization
Fuente:
arXiv
Salvato in:
| Autori principali: | Nichols, Keanu, Appapogu, Divya, Biamby, Giscard, Bashkirova, Dina, Rohrbach, Anna, Plummer, Bryan A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FuTCR: Future-Targeted Contrast and Repulsion for Continual Panoptic Segmentation
di: Ikechukwu, Nicholas, et al.
Pubblicazione: (2026)
di: Ikechukwu, Nicholas, et al.
Pubblicazione: (2026)
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
di: Qraitem, Maan, et al.
Pubblicazione: (2023)
di: Qraitem, Maan, et al.
Pubblicazione: (2023)
LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
di: Niu, Dantong, et al.
Pubblicazione: (2024)
di: Niu, Dantong, et al.
Pubblicazione: (2024)
RECAST: Reparameterized, Compact weight Adaptation for Sequential Tasks
di: Tasnim, Nazia, et al.
Pubblicazione: (2024)
di: Tasnim, Nazia, et al.
Pubblicazione: (2024)
ChA-MAEViT: Unifying Channel-Aware Masked Autoencoders and Multi-Channel Vision Transformers for Improved Cross-Channel Learning
di: Pham, Chau, et al.
Pubblicazione: (2025)
di: Pham, Chau, et al.
Pubblicazione: (2025)
Tuning Just Enough: Lightweight Backdoor Attacks on Multi-Encoder Diffusion Models
di: Chen, Ziyuan, et al.
Pubblicazione: (2026)
di: Chen, Ziyuan, et al.
Pubblicazione: (2026)
Noise-Aware Generalization: Robustness to In-Domain Noise and Out-of-Domain Generalization
di: Wang, Siqi, et al.
Pubblicazione: (2025)
di: Wang, Siqi, et al.
Pubblicazione: (2025)
HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models
di: Zohrabi, Reihaneh, et al.
Pubblicazione: (2026)
di: Zohrabi, Reihaneh, et al.
Pubblicazione: (2026)
Spurious-Aware Prototype Refinement for Reliable Out-of-Distribution Detection
di: Zohrabi, Reihaneh, et al.
Pubblicazione: (2025)
di: Zohrabi, Reihaneh, et al.
Pubblicazione: (2025)
Seeing Isn't Orienting: A Cognitively Grounded Benchmark Reveals Systematic Orientation Failures in MLLMs Supplementary
di: Tasnim, Nazia, et al.
Pubblicazione: (2026)
di: Tasnim, Nazia, et al.
Pubblicazione: (2026)
Seeing Isn't Orienting: A Cognitively Grounded Benchmark Reveals Systematic Orientation Failures in MLLMs
di: Tasnim, Nazia, et al.
Pubblicazione: (2025)
di: Tasnim, Nazia, et al.
Pubblicazione: (2025)
OP-LoRA: The Blessing of Dimensionality
di: Teterwak, Piotr, et al.
Pubblicazione: (2024)
di: Teterwak, Piotr, et al.
Pubblicazione: (2024)
Is Large-Scale Pretraining the Secret to Good Domain Generalization?
di: Teterwak, Piotr, et al.
Pubblicazione: (2024)
di: Teterwak, Piotr, et al.
Pubblicazione: (2024)
ERM++: An Improved Baseline for Domain Generalization
di: Teterwak, Piotr, et al.
Pubblicazione: (2023)
di: Teterwak, Piotr, et al.
Pubblicazione: (2023)
Topology Aware Neural Interpolation of Scalar Fields
di: Kissi, Mohamed, et al.
Pubblicazione: (2025)
di: Kissi, Mohamed, et al.
Pubblicazione: (2025)
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
di: Qraitem, Maan, et al.
Pubblicazione: (2024)
di: Qraitem, Maan, et al.
Pubblicazione: (2024)
Efficient Pre-training for Localized Instruction Generation of Videos
di: Batra, Anil, et al.
Pubblicazione: (2023)
di: Batra, Anil, et al.
Pubblicazione: (2023)
Omni-IML: Towards Unified Image Manipulation Localization
di: Qu, Chenfan, et al.
Pubblicazione: (2024)
di: Qu, Chenfan, et al.
Pubblicazione: (2024)
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
di: Liu, Aoming, et al.
Pubblicazione: (2025)
di: Liu, Aoming, et al.
Pubblicazione: (2025)
Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark
di: Wu, Tsung-Han, et al.
Pubblicazione: (2024)
di: Wu, Tsung-Han, et al.
Pubblicazione: (2024)
Beyond Memorization: Selective Learning for Copyright-Safe Diffusion Model Training
di: Kothandaraman, Divya, et al.
Pubblicazione: (2025)
di: Kothandaraman, Divya, et al.
Pubblicazione: (2025)
BEEM: Boosting Performance of Early Exit DNNs using Multi-Exit Classifiers as Experts
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
Sharp-It: A Multi-view to Multi-view Diffusion Model for 3D Synthesis and Manipulation
di: Edelstein, Yiftach, et al.
Pubblicazione: (2024)
di: Edelstein, Yiftach, et al.
Pubblicazione: (2024)
Local Policies Enable Zero-shot Long-horizon Manipulation
di: Dalal, Murtaza, et al.
Pubblicazione: (2024)
di: Dalal, Murtaza, et al.
Pubblicazione: (2024)
V$^2$Dial: Unification of Video and Visual Dialog via Multimodal Experts
di: Abdessaied, Adnen, et al.
Pubblicazione: (2025)
di: Abdessaied, Adnen, et al.
Pubblicazione: (2025)
ForgerySleuth: Empowering Multimodal Large Language Models for Image Manipulation Detection
di: Sun, Zhihao, et al.
Pubblicazione: (2024)
di: Sun, Zhihao, et al.
Pubblicazione: (2024)
FREE: Fast and Robust Vision Language Models with Early Exits
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
TwoHead-SwinFPN: A Unified DL Architecture for Synthetic Manipulation, Detection and Localization in Identity Documents
di: Naseeb, Chan, et al.
Pubblicazione: (2026)
di: Naseeb, Chan, et al.
Pubblicazione: (2026)
Scaling Image Geo-Localization to Continent Level
di: Lindenberger, Philipp, et al.
Pubblicazione: (2025)
di: Lindenberger, Philipp, et al.
Pubblicazione: (2025)
Contrastive Localized Language-Image Pre-Training
di: Chen, Hong-You, et al.
Pubblicazione: (2024)
di: Chen, Hong-You, et al.
Pubblicazione: (2024)
FedDistill: Global Model Distillation for Local Model De-Biasing in Non-IID Federated Learning
di: Song, Changlin, et al.
Pubblicazione: (2024)
di: Song, Changlin, et al.
Pubblicazione: (2024)
LNL+K: Enhancing Learning with Noisy Labels Through Noise Source Knowledge Integration
di: Wang, Siqi, et al.
Pubblicazione: (2023)
di: Wang, Siqi, et al.
Pubblicazione: (2023)
Shape-Guided Diffusion with Inside-Outside Attention
di: Park, Dong Huk, et al.
Pubblicazione: (2022)
di: Park, Dong Huk, et al.
Pubblicazione: (2022)
Enabling Local Editing in Diffusion Models by Joint and Individual Component Analysis
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2024)
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2024)
DEFAME: Dynamic Evidence-based FAct-checking with Multimodal Experts
di: Braun, Tobias, et al.
Pubblicazione: (2024)
di: Braun, Tobias, et al.
Pubblicazione: (2024)
Test-time augmentation improves efficiency in conformal prediction
di: Shanmugam, Divya, et al.
Pubblicazione: (2025)
di: Shanmugam, Divya, et al.
Pubblicazione: (2025)
LoFi: Neural Local Fields for Scalable Image Reconstruction
di: Khorashadizadeh, AmirEhsan, et al.
Pubblicazione: (2024)
di: Khorashadizadeh, AmirEhsan, et al.
Pubblicazione: (2024)
Localizing Anomalies via Multiscale Score Matching Analysis
di: Mahmood, Ahsan, et al.
Pubblicazione: (2024)
di: Mahmood, Ahsan, et al.
Pubblicazione: (2024)
HamVision: Hamiltonian Dynamics as Inductive Bias for Medical Image Analysis
di: Mabrok, Mohamed A
Pubblicazione: (2026)
di: Mabrok, Mohamed A
Pubblicazione: (2026)
The Narrow Gate: Localized Image-Text Communication in Native Multimodal Models
di: Serra, Alessandro Pietro, et al.
Pubblicazione: (2024)
di: Serra, Alessandro Pietro, et al.
Pubblicazione: (2024)
Documenti analoghi
-
FuTCR: Future-Targeted Contrast and Repulsion for Continual Panoptic Segmentation
di: Ikechukwu, Nicholas, et al.
Pubblicazione: (2026) -
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
di: Qraitem, Maan, et al.
Pubblicazione: (2023) -
LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
di: Niu, Dantong, et al.
Pubblicazione: (2024) -
RECAST: Reparameterized, Compact weight Adaptation for Sequential Tasks
di: Tasnim, Nazia, et al.
Pubblicazione: (2024) -
ChA-MAEViT: Unifying Channel-Aware Masked Autoencoders and Multi-Channel Vision Transformers for Improved Cross-Channel Learning
di: Pham, Chau, et al.
Pubblicazione: (2025)