Learning Universal Features for Generalizable Image Forgery Localization
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhao, Hengrun, Zhuge, Yunzhi, Wang, Yifan, Wang, Lijun, Lu, Huchuan, Zeng, Yu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AVS-Mamba: Exploring Temporal and Multi-modal Mamba for Audio-Visual Segmentation
di: Gong, Sitong, et al.
Pubblicazione: (2025)
di: Gong, Sitong, et al.
Pubblicazione: (2025)
Complementary and Contrastive Learning for Audio-Visual Segmentation
di: Gong, Sitong, et al.
Pubblicazione: (2025)
di: Gong, Sitong, et al.
Pubblicazione: (2025)
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation
di: Zhuge, Yunzhi, et al.
Pubblicazione: (2025)
di: Zhuge, Yunzhi, et al.
Pubblicazione: (2025)
Boosting Continual Learning of Vision-Language Models via Mixture-of-Experts Adapters
di: Yu, Jiazuo, et al.
Pubblicazione: (2024)
di: Yu, Jiazuo, et al.
Pubblicazione: (2024)
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding
di: Xiong, Haomiao, et al.
Pubblicazione: (2025)
di: Xiong, Haomiao, et al.
Pubblicazione: (2025)
Parameter Aware Mamba Model for Multi-task Dense Prediction
di: Yu, Xinzhuo, et al.
Pubblicazione: (2025)
di: Yu, Xinzhuo, et al.
Pubblicazione: (2025)
FineRS: Fine-grained Reasoning and Segmentation of Small Objects with Reinforcement Learning
di: Zhang, Lu, et al.
Pubblicazione: (2025)
di: Zhang, Lu, et al.
Pubblicazione: (2025)
Reinforcing Video Reasoning Segmentation to Think Before It Segments
di: Gong, Sitong, et al.
Pubblicazione: (2025)
di: Gong, Sitong, et al.
Pubblicazione: (2025)
Towards Cross-Platform Generalization: Domain Adaptive 3D Detection with Augmentation and Pseudo-Labeling
di: Feng, Xiyan, et al.
Pubblicazione: (2026)
di: Feng, Xiyan, et al.
Pubblicazione: (2026)
Learning Expressive And Generalizable Motion Features For Face Forgery Detection
di: Zhang, Jingyi, et al.
Pubblicazione: (2024)
di: Zhang, Jingyi, et al.
Pubblicazione: (2024)
Revisiting Salient Object Detection from an Observer-Centric Perspective
di: Zhang, Fuxi, et al.
Pubblicazione: (2026)
di: Zhang, Fuxi, et al.
Pubblicazione: (2026)
StableIdentity: Inserting Anybody into Anywhere at First Sight
di: Wang, Qinghe, et al.
Pubblicazione: (2024)
di: Wang, Qinghe, et al.
Pubblicazione: (2024)
The Devil is in Temporal Token: High Quality Video Reasoning Segmentation
di: Gong, Sitong, et al.
Pubblicazione: (2025)
di: Gong, Sitong, et al.
Pubblicazione: (2025)
Bootstraping Clustering of Gaussians for View-consistent 3D Scene Understanding
di: Zhang, Wenbo, et al.
Pubblicazione: (2024)
di: Zhang, Wenbo, et al.
Pubblicazione: (2024)
SHERL: Synthesizing High Accuracy and Efficient Memory for Resource-Limited Transfer Learning
di: Diao, Haiwen, et al.
Pubblicazione: (2024)
di: Diao, Haiwen, et al.
Pubblicazione: (2024)
DreamMix: Decoupling Object Attributes for Enhanced Editability in Customized Image Inpainting
di: Yang, Yicheng, et al.
Pubblicazione: (2024)
di: Yang, Yicheng, et al.
Pubblicazione: (2024)
EDGER: EDge-Guided with HEatmap Refinement for Generalizable Image Forgery Localization
di: Le-Phan, Minh-Khoa, et al.
Pubblicazione: (2026)
di: Le-Phan, Minh-Khoa, et al.
Pubblicazione: (2026)
Image Copy-Move Forgery Detection and Localization Scheme: How to Avoid Missed Detection and False Alarm
di: Jiang, Li, et al.
Pubblicazione: (2024)
di: Jiang, Li, et al.
Pubblicazione: (2024)
BEV-IO: Enhancing Bird's-Eye-View 3D Detection with Instance Occupancy
di: Zhang, Zaibin, et al.
Pubblicazione: (2023)
di: Zhang, Zaibin, et al.
Pubblicazione: (2023)
Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge
di: Xiong, Haomiao, et al.
Pubblicazione: (2025)
di: Xiong, Haomiao, et al.
Pubblicazione: (2025)
Towards Open-Vocabulary Remote Sensing Image Semantic Segmentation
di: Ye, Chengyang, et al.
Pubblicazione: (2024)
di: Ye, Chengyang, et al.
Pubblicazione: (2024)
ForgeLens: Data-Efficient Forgery Focus for Generalizable Forgery Image Detection
di: Chen, Yingjian, et al.
Pubblicazione: (2024)
di: Chen, Yingjian, et al.
Pubblicazione: (2024)
MFVLR: Multi-domain Fine-grained Vision-Language Reconstruction for Generalizable Diffusion Face Forgery Detection and Localization
di: Zhang, Yaning, et al.
Pubblicazione: (2026)
di: Zhang, Yaning, et al.
Pubblicazione: (2026)
From Forecasting to Planning: Policy World Model for Collaborative State-Action Prediction
di: Zhao, Zhida, et al.
Pubblicazione: (2025)
di: Zhao, Zhida, et al.
Pubblicazione: (2025)
VFXMaster: Unlocking Dynamic Visual Effect Generation via In-Context Learning
di: Li, Baolu, et al.
Pubblicazione: (2025)
di: Li, Baolu, et al.
Pubblicazione: (2025)
VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?
di: Liu, Qing'an, et al.
Pubblicazione: (2026)
di: Liu, Qing'an, et al.
Pubblicazione: (2026)
AR-MOT: Autoregressive Multi-object Tracking
di: Jia, Lianjie, et al.
Pubblicazione: (2026)
di: Jia, Lianjie, et al.
Pubblicazione: (2026)
Toward Generalizable Forgery Detection and Reasoning
di: Gao, Yueying, et al.
Pubblicazione: (2025)
di: Gao, Yueying, et al.
Pubblicazione: (2025)
Other Tokens Matter: Exploring Global and Local Features of Vision Transformers for Object Re-Identification
di: Wang, Yingquan, et al.
Pubblicazione: (2024)
di: Wang, Yingquan, et al.
Pubblicazione: (2024)
ReAlign: Generalizable Image Forgery Detection via Reasoning-Aligned Representation
di: Huang, Qing, et al.
Pubblicazione: (2026)
di: Huang, Qing, et al.
Pubblicazione: (2026)
Image Forgery Localization with State Space Models
di: Lou, Zijie, et al.
Pubblicazione: (2024)
di: Lou, Zijie, et al.
Pubblicazione: (2024)
Low-rank Orthogonal Subspace Intervention for Generalizable Face Forgery Detection
di: Wang, Chi, et al.
Pubblicazione: (2026)
di: Wang, Chi, et al.
Pubblicazione: (2026)
Loupe: A Generalizable and Adaptive Framework for Image Forgery Detection
di: Jiang, Yuchu, et al.
Pubblicazione: (2025)
di: Jiang, Yuchu, et al.
Pubblicazione: (2025)
ForgeryVCR: Visual-Centric Reasoning via Efficient Forensic Tools in MLLMs for Image Forgery Detection and Localization
di: Wang, Youqi, et al.
Pubblicazione: (2026)
di: Wang, Youqi, et al.
Pubblicazione: (2026)
AD-H: Language-guided Autonomous Driving with Hierarchical Agents
di: Zhang, Zaibin, et al.
Pubblicazione: (2024)
di: Zhang, Zaibin, et al.
Pubblicazione: (2024)
Generalizable Face Forgery Detection via Separable Prompt Learning
di: Yang, Enrui, et al.
Pubblicazione: (2026)
di: Yang, Enrui, et al.
Pubblicazione: (2026)
CBDiff:Conditional Bernoulli Diffusion Models for Image Forgery Localization
di: Lei, Zhou, et al.
Pubblicazione: (2025)
di: Lei, Zhou, et al.
Pubblicazione: (2025)
Mono2Stereo: A Benchmark and Empirical Study for Stereo Conversion
di: Yu, Songsong, et al.
Pubblicazione: (2025)
di: Yu, Songsong, et al.
Pubblicazione: (2025)
Learning to Discover Forgery Cues for Face Forgery Detection
di: Tian, Jiahe, et al.
Pubblicazione: (2024)
di: Tian, Jiahe, et al.
Pubblicazione: (2024)
Word-Anchored Temporal Forgery Localization
di: Wang, Tianyi, et al.
Pubblicazione: (2026)
di: Wang, Tianyi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
AVS-Mamba: Exploring Temporal and Multi-modal Mamba for Audio-Visual Segmentation
di: Gong, Sitong, et al.
Pubblicazione: (2025) -
Complementary and Contrastive Learning for Audio-Visual Segmentation
di: Gong, Sitong, et al.
Pubblicazione: (2025) -
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation
di: Zhuge, Yunzhi, et al.
Pubblicazione: (2025) -
Boosting Continual Learning of Vision-Language Models via Mixture-of-Experts Adapters
di: Yu, Jiazuo, et al.
Pubblicazione: (2024) -
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding
di: Xiong, Haomiao, et al.
Pubblicazione: (2025)