HydraPrompt: An Adaptive and Asymmetric Framework of Vision-Language Models for Synthetic Image Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Shi, Senyuan, Tan, Hao, Tan, Zichang, Feng, Shuhan, Liu, Ajian, Escalera, Sergio, Wan, Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Veritas: Generalizable Deepfake Detection via Pattern-Aware Reasoning
di: Tan, Hao, et al.
Pubblicazione: (2025)
di: Tan, Hao, et al.
Pubblicazione: (2025)
Recover and Match: Open-Vocabulary Multi-Label Recognition through Knowledge-Constrained Optimal Transport
di: Tan, Hao, et al.
Pubblicazione: (2025)
di: Tan, Hao, et al.
Pubblicazione: (2025)
PVLR: Prompt-driven Visual-Linguistic Representation Learning for Multi-Label Image Recognition
di: Tan, Hao, et al.
Pubblicazione: (2024)
di: Tan, Hao, et al.
Pubblicazione: (2024)
VideoVeritas: AI-Generated Video Detection via Perception Pretext Reinforcement Learning
di: Tan, Hao, et al.
Pubblicazione: (2026)
di: Tan, Hao, et al.
Pubblicazione: (2026)
SSPA: Split-and-Synthesize Prompting with Gated Alignments for Multi-Label Image Recognition
di: Tan, Hao, et al.
Pubblicazione: (2024)
di: Tan, Hao, et al.
Pubblicazione: (2024)
Towards Generalizable AI-Generated Image Detection via Image-Adaptive Prompt Learning
di: Li, Yiheng, et al.
Pubblicazione: (2025)
di: Li, Yiheng, et al.
Pubblicazione: (2025)
CFPL-FAS: Class Free Prompt Learning for Generalizable Face Anti-spoofing
di: Liu, Ajian, et al.
Pubblicazione: (2024)
di: Liu, Ajian, et al.
Pubblicazione: (2024)
Mixture-of-Attack-Experts with Class Regularization for Unified Physical-Digital Face Attack Detection
di: Chen, Shunxin, et al.
Pubblicazione: (2025)
di: Chen, Shunxin, et al.
Pubblicazione: (2025)
RVLF: A Reinforcing Vision-Language Framework for Gloss-Free Sign Language Translation
di: Rao, Zhi, et al.
Pubblicazione: (2025)
di: Rao, Zhi, et al.
Pubblicazione: (2025)
Training-Free Unsupervised Prompt for Vision-Language Models
di: Long, Sifan, et al.
Pubblicazione: (2024)
di: Long, Sifan, et al.
Pubblicazione: (2024)
Unified Physical-Digital Attack Detection Challenge
di: Yuan, Haocheng, et al.
Pubblicazione: (2024)
di: Yuan, Haocheng, et al.
Pubblicazione: (2024)
Unified Physical-Digital Face Attack Detection
di: Fang, Hao, et al.
Pubblicazione: (2024)
di: Fang, Hao, et al.
Pubblicazione: (2024)
NCL++: Nested Collaborative Learning for Long-Tailed Visual Recognition
di: Tan, Zichang, et al.
Pubblicazione: (2023)
di: Tan, Zichang, et al.
Pubblicazione: (2023)
A Transformer Model for Boundary Detection in Continuous Sign Language
di: Rastgoo, Razieh, et al.
Pubblicazione: (2024)
di: Rastgoo, Razieh, et al.
Pubblicazione: (2024)
AGC: Adaptive Geodesic Correction for Adversarial Robustness on Vision-Language Models
di: Li, Zhiwei, et al.
Pubblicazione: (2026)
di: Li, Zhiwei, et al.
Pubblicazione: (2026)
Reduce the Artifacts Bias for More Generalizable AI-Generated Image Detection
di: Li, Yiheng, et al.
Pubblicazione: (2026)
di: Li, Yiheng, et al.
Pubblicazione: (2026)
CTForensics: A Comprehensive Dataset and Method for AI-Generated CT Image Detection
di: Li, Yiheng, et al.
Pubblicazione: (2026)
di: Li, Yiheng, et al.
Pubblicazione: (2026)
Interactive Post-Training for Vision-Language-Action Models
di: Tan, Shuhan, et al.
Pubblicazione: (2025)
di: Tan, Shuhan, et al.
Pubblicazione: (2025)
FA^{3}-CLIP: Frequency-Aware Cues Fusion and Attack-Agnostic Prompt Learning for Unified Face Attack Detection
di: Li, Yongze, et al.
Pubblicazione: (2025)
di: Li, Yongze, et al.
Pubblicazione: (2025)
MLLM-Enhanced Face Forgery Detection: A Vision-Language Fusion Solution
di: Peng, Siran, et al.
Pubblicazione: (2025)
di: Peng, Siran, et al.
Pubblicazione: (2025)
PrismVAU: Prompt-Refined Inference System for Multimodal Video Anomaly Understanding
di: Erregue, Iñaki, et al.
Pubblicazione: (2026)
di: Erregue, Iñaki, et al.
Pubblicazione: (2026)
L-SWAG: Layer-Sample Wise Activation with Gradients information for Zero-Shot NAS on Vision Transformers
di: Casarin, Sofia, et al.
Pubblicazione: (2025)
di: Casarin, Sofia, et al.
Pubblicazione: (2025)
Human-Free Automated Prompting for Vision-Language Anomaly Detection: Prompt Optimization with Meta-guiding Prompt Scheme
di: Chen, Pi-Wei, et al.
Pubblicazione: (2024)
di: Chen, Pi-Wei, et al.
Pubblicazione: (2024)
Multi-Turn Adaptive Prompting Attack on Large Vision-Language Models
di: Choi, In Chong, et al.
Pubblicazione: (2026)
di: Choi, In Chong, et al.
Pubblicazione: (2026)
La-SoftMoE CLIP for Unified Physical-Digital Face Attack Detection
di: Zou, Hang, et al.
Pubblicazione: (2024)
di: Zou, Hang, et al.
Pubblicazione: (2024)
Cross-Image Contrastive Decoding: Precise, Lossless Suppression of Language Priors in Large Vision-Language Models
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
Benchmarking Unified Face Attack Detection via Hierarchical Prompt Tuning
di: Liu, Ajian, et al.
Pubblicazione: (2025)
di: Liu, Ajian, et al.
Pubblicazione: (2025)
BEVSpread: Spread Voxel Pooling for Bird's-Eye-View Representation in Vision-based Roadside 3D Object Detection
di: Wang, Wenjie, et al.
Pubblicazione: (2024)
di: Wang, Wenjie, et al.
Pubblicazione: (2024)
Quantized Prompt for Efficient Generalization of Vision-Language Models
di: Hao, Tianxiang, et al.
Pubblicazione: (2024)
di: Hao, Tianxiang, et al.
Pubblicazione: (2024)
Improving Synthetic Image Detection Towards Generalization: An Image Transformation Perspective
di: Li, Ouxiang, et al.
Pubblicazione: (2024)
di: Li, Ouxiang, et al.
Pubblicazione: (2024)
LightningDrag: Lightning Fast and Accurate Drag-based Image Editing Emerging from Videos
di: Shi, Yujun, et al.
Pubblicazione: (2024)
di: Shi, Yujun, et al.
Pubblicazione: (2024)
ControlFusion: A Controllable Image Fusion Framework with Language-Vision Degradation Prompts
di: Tang, Linfeng, et al.
Pubblicazione: (2025)
di: Tang, Linfeng, et al.
Pubblicazione: (2025)
Unleashing the Potential of Consistency Learning for Detecting and Grounding Multi-Modal Media Manipulation
di: Li, Yiheng, et al.
Pubblicazione: (2025)
di: Li, Yiheng, et al.
Pubblicazione: (2025)
FA: Forced Prompt Learning of Vision-Language Models for Out-of-Distribution Detection
di: Lu, Xinhua, et al.
Pubblicazione: (2025)
di: Lu, Xinhua, et al.
Pubblicazione: (2025)
Asymmetric Visual Semantic Embedding Framework for Efficient Vision-Language Alignment
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
White-Balance First, Adjust Later: Cross-Camera Color Constancy via Vision-Language Evaluation
di: Li, Shuwei, et al.
Pubblicazione: (2026)
di: Li, Shuwei, et al.
Pubblicazione: (2026)
Data Adaptive Traceback for Vision-Language Foundation Models in Image Classification
di: Peng, Wenshuo, et al.
Pubblicazione: (2024)
di: Peng, Wenshuo, et al.
Pubblicazione: (2024)
Hydra: Computer Vision for Data Quality Monitoring
di: Britton, Thomas, et al.
Pubblicazione: (2024)
di: Britton, Thomas, et al.
Pubblicazione: (2024)
Enhancing Zero-Shot Vision Models by Label-Free Prompt Distribution Learning and Bias Correcting
di: Zhu, Xingyu, et al.
Pubblicazione: (2024)
di: Zhu, Xingyu, et al.
Pubblicazione: (2024)
Spoofing-aware Prompt Learning for Unified Physical-Digital Facial Attack Detection
di: Guo, Jiabao, et al.
Pubblicazione: (2025)
di: Guo, Jiabao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Veritas: Generalizable Deepfake Detection via Pattern-Aware Reasoning
di: Tan, Hao, et al.
Pubblicazione: (2025) -
Recover and Match: Open-Vocabulary Multi-Label Recognition through Knowledge-Constrained Optimal Transport
di: Tan, Hao, et al.
Pubblicazione: (2025) -
PVLR: Prompt-driven Visual-Linguistic Representation Learning for Multi-Label Image Recognition
di: Tan, Hao, et al.
Pubblicazione: (2024) -
VideoVeritas: AI-Generated Video Detection via Perception Pretext Reinforcement Learning
di: Tan, Hao, et al.
Pubblicazione: (2026) -
SSPA: Split-and-Synthesize Prompting with Gated Alignments for Multi-Label Image Recognition
di: Tan, Hao, et al.
Pubblicazione: (2024)