DeformTrace: A Deformable State Space Model with Relay Tokens for Temporal Forgery Localization
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhu, Xiaodong, Wang, Suting, Zheng, Yuanming, Yang, Junqi, Liao, Yangxu, Yang, Yuhong, Tu, Weiping, Wang, Zhongyuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GEM-TFL: Bridging Weak and Full Supervision for Forgery Localization through EM-Guided Decomposition and Temporal Refinement
di: Zhu, Xiaodong, et al.
Pubblicazione: (2026)
di: Zhu, Xiaodong, et al.
Pubblicazione: (2026)
Improving Speech Enhancement by Integrating Inter-Channel and Band Features with Dual-branch Conformer
di: Li, Jizhen, et al.
Pubblicazione: (2024)
di: Li, Jizhen, et al.
Pubblicazione: (2024)
DDNet: A Dual-Stream Graph Learning and Disentanglement Framework for Temporal Forgery Localization
di: Zhao, Boyang, et al.
Pubblicazione: (2026)
di: Zhao, Boyang, et al.
Pubblicazione: (2026)
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
di: Xu, Junhao, et al.
Pubblicazione: (2025)
di: Xu, Junhao, et al.
Pubblicazione: (2025)
Context-aware TFL: A Universal Context-aware Contrastive Learning Framework for Temporal Forgery Localization
di: Yin, Qilin, et al.
Pubblicazione: (2025)
di: Yin, Qilin, et al.
Pubblicazione: (2025)
Weakly-supervised Audio Temporal Forgery Localization via Progressive Audio-language Co-learning Network
di: Wu, Junyan, et al.
Pubblicazione: (2025)
di: Wu, Junyan, et al.
Pubblicazione: (2025)
Magic Tokens: Select Diverse Tokens for Multi-modal Object Re-Identification
di: Zhang, Pingping, et al.
Pubblicazione: (2024)
di: Zhang, Pingping, et al.
Pubblicazione: (2024)
Identity-Driven Multimedia Forgery Detection via Reference Assistance
di: Xu, Junhao, et al.
Pubblicazione: (2024)
di: Xu, Junhao, et al.
Pubblicazione: (2024)
Coarse-to-Fine Proposal Refinement Framework for Audio Temporal Forgery Detection and Localization
di: Wu, Junyan, et al.
Pubblicazione: (2024)
di: Wu, Junyan, et al.
Pubblicazione: (2024)
Optimization-Free Universal Watermark Forgery with Regenerative Diffusion Models
di: Zhu, Chaoyi, et al.
Pubblicazione: (2025)
di: Zhu, Chaoyi, et al.
Pubblicazione: (2025)
PRINTER:Deformation-Aware Adversarial Learning for Virtual IHC Staining with In Situ Fidelity
di: Yuan, Yizhe, et al.
Pubblicazione: (2025)
di: Yuan, Yizhe, et al.
Pubblicazione: (2025)
Stepwise Schema-Guided Prompting Framework with Parameter Efficient Instruction Tuning for Multimedia Event Extraction
di: Yuan, Xiang, et al.
Pubblicazione: (2025)
di: Yuan, Xiang, et al.
Pubblicazione: (2025)
IDEA: Inverted Text with Cooperative Deformable Aggregation for Multi-modal Object Re-Identification
di: Wang, Yuhao, et al.
Pubblicazione: (2025)
di: Wang, Yuhao, et al.
Pubblicazione: (2025)
Deformable Audio Transformer for Audio Event Detection
di: Zhu, Wentao
Pubblicazione: (2023)
di: Zhu, Wentao
Pubblicazione: (2023)
Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries
di: Kim, Minyoung, et al.
Pubblicazione: (2025)
di: Kim, Minyoung, et al.
Pubblicazione: (2025)
Challenging Dataset and Multi-modal Gated Mixture of Experts Model for Remote Sensing Copy-Move Forgery Understanding
di: Zhang, Ze, et al.
Pubblicazione: (2025)
di: Zhang, Ze, et al.
Pubblicazione: (2025)
Towards Flexible Evaluation for Generative Visual Question Answering
di: Ji, Huishan, et al.
Pubblicazione: (2024)
di: Ji, Huishan, et al.
Pubblicazione: (2024)
PopSim: Social Network Simulation for Social Media Popularity Prediction
di: Liu, Yijun, et al.
Pubblicazione: (2025)
di: Liu, Yijun, et al.
Pubblicazione: (2025)
3DMambaIPF: A State Space Model for Iterative Point Cloud Filtering via Differentiable Rendering
di: Zhou, Qingyuan, et al.
Pubblicazione: (2024)
di: Zhou, Qingyuan, et al.
Pubblicazione: (2024)
Towards Universal Modal Tracking with Online Dense Temporal Token Learning
di: Zheng, Yaozong, et al.
Pubblicazione: (2025)
di: Zheng, Yaozong, et al.
Pubblicazione: (2025)
Enhancing Cross-Prompt Transferability in Vision-Language Models through Contextual Injection of Target Tokens
di: Yang, Xikang, et al.
Pubblicazione: (2024)
di: Yang, Xikang, et al.
Pubblicazione: (2024)
Listening Deepfake Detection: A New Perspective Beyond Speaking-Centric Forgery Analysis
di: Liu, Miao, et al.
Pubblicazione: (2026)
di: Liu, Miao, et al.
Pubblicazione: (2026)
Band-Attention Modulated RetNet for Face Forgery Detection
di: Zhang, Zhida, et al.
Pubblicazione: (2024)
di: Zhang, Zhida, et al.
Pubblicazione: (2024)
Other Tokens Matter: Exploring Global and Local Features of Vision Transformers for Object Re-Identification
di: Wang, Yingquan, et al.
Pubblicazione: (2024)
di: Wang, Yingquan, et al.
Pubblicazione: (2024)
SpikEmo: Enhancing Emotion Recognition With Spiking Temporal Dynamics in Conversations
di: Yu, Xiaomin, et al.
Pubblicazione: (2024)
di: Yu, Xiaomin, et al.
Pubblicazione: (2024)
Compression Metadata-assisted RoI Extraction and Adaptive Inference for Efficient Video Analytics
di: Wang, Chengzhi, et al.
Pubblicazione: (2025)
di: Wang, Chengzhi, et al.
Pubblicazione: (2025)
Synchronized Video Storytelling: Generating Video Narrations with Structured Storyline
di: Yang, Dingyi, et al.
Pubblicazione: (2024)
di: Yang, Dingyi, et al.
Pubblicazione: (2024)
Deepfake Detection: A Comprehensive Survey from the Reliability Perspective
di: Wang, Tianyi, et al.
Pubblicazione: (2022)
di: Wang, Tianyi, et al.
Pubblicazione: (2022)
Period-conscious Time-series Reconstruction under Local Differential Privacy
di: Wang, Yaxuan, et al.
Pubblicazione: (2026)
di: Wang, Yaxuan, et al.
Pubblicazione: (2026)
RA-BLIP: Multimodal Adaptive Retrieval-Augmented Bootstrapping Language-Image Pre-training
di: Ding, Muhe, et al.
Pubblicazione: (2024)
di: Ding, Muhe, et al.
Pubblicazione: (2024)
Video Echoed in Music: Semantic, Temporal, and Rhythmic Alignment for Video-to-Music Generation
di: Tong, Xinyi, et al.
Pubblicazione: (2025)
di: Tong, Xinyi, et al.
Pubblicazione: (2025)
Enhancing Self-Supervised Talking Head Forgery Detection via a Training-Free Dual-System Framework
di: Liu, Ke, et al.
Pubblicazione: (2026)
di: Liu, Ke, et al.
Pubblicazione: (2026)
Copy-Move Forgery Detection and Question Answering for Remote Sensing Image
di: Zhang, Ze, et al.
Pubblicazione: (2024)
di: Zhang, Ze, et al.
Pubblicazione: (2024)
Beyond Walking: A Large-Scale Image-Text Benchmark for Text-based Person Anomaly Search
di: Yang, Shuyu, et al.
Pubblicazione: (2024)
di: Yang, Shuyu, et al.
Pubblicazione: (2024)
CLEAR: Null-Space Projection for Cross-Modal De-Redundancy in Multimodal Recommendation
di: Zhan, Hao, et al.
Pubblicazione: (2026)
di: Zhan, Hao, et al.
Pubblicazione: (2026)
Look, Compare and Draw: Differential Query Transformer for Automatic Oil Painting
di: Liu, Lingyu, et al.
Pubblicazione: (2026)
di: Liu, Lingyu, et al.
Pubblicazione: (2026)
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing
di: Chen, Yaru, et al.
Pubblicazione: (2025)
di: Chen, Yaru, et al.
Pubblicazione: (2025)
OOD-GraphLLM: Graph Large Language Model for Out-of-Distribution Generalized Drug Synergy Prediction
di: Wang, Xin, et al.
Pubblicazione: (2026)
di: Wang, Xin, et al.
Pubblicazione: (2026)
Hybrid Local-Global Context Learning for Neural Video Compression
di: Zhai, Yongqi, et al.
Pubblicazione: (2024)
di: Zhai, Yongqi, et al.
Pubblicazione: (2024)
Muse: A Multimodal Conversational Recommendation Dataset with Scenario-Grounded User Profiles
di: Wang, Zihan, et al.
Pubblicazione: (2024)
di: Wang, Zihan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
GEM-TFL: Bridging Weak and Full Supervision for Forgery Localization through EM-Guided Decomposition and Temporal Refinement
di: Zhu, Xiaodong, et al.
Pubblicazione: (2026) -
Improving Speech Enhancement by Integrating Inter-Channel and Band Features with Dual-branch Conformer
di: Li, Jizhen, et al.
Pubblicazione: (2024) -
DDNet: A Dual-Stream Graph Learning and Disentanglement Framework for Temporal Forgery Localization
di: Zhao, Boyang, et al.
Pubblicazione: (2026) -
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
di: Xu, Junhao, et al.
Pubblicazione: (2025) -
Context-aware TFL: A Universal Context-aware Contrastive Learning Framework for Temporal Forgery Localization
di: Yin, Qilin, et al.
Pubblicazione: (2025)