Neural Antidote: Class-Wise Prompt Tuning for Purifying Backdoors in CLIP
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kong, Jiawei, Fang, Hao, Guo, Sihang, Qing, Chenxi, Gao, Kuofeng, Chen, Bin, Xia, Shu-Tao, Xu, Ke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIP
von: Bai, Jiawang, et al.
Veröffentlicht: (2023)
von: Bai, Jiawang, et al.
Veröffentlicht: (2023)
Revisiting Backdoor Attacks on LLMs: A Stealthy and Practical Poisoning Framework via Harmless Inputs
von: Kong, Jiawei, et al.
Veröffentlicht: (2025)
von: Kong, Jiawei, et al.
Veröffentlicht: (2025)
Retrievals Can Be Detrimental: Unveiling the Backdoor Vulnerability of Retrieval-Augmented Diffusion Models
von: Fang, Hao, et al.
Veröffentlicht: (2025)
von: Fang, Hao, et al.
Veröffentlicht: (2025)
Grounding Language with Vision: A Conditional Mutual Information Calibrated Decoding Strategy for Reducing Hallucinations in LVLMs
von: Fang, Hao, et al.
Veröffentlicht: (2025)
von: Fang, Hao, et al.
Veröffentlicht: (2025)
CLIP-Guided Generative Networks for Transferable Targeted Adversarial Attacks
von: Fang, Hao, et al.
Veröffentlicht: (2024)
von: Fang, Hao, et al.
Veröffentlicht: (2024)
Seeing Through the Chain: Mitigate Hallucination in Multimodal Reasoning Models via CoT Compression and Contrastive Preference Optimization
von: Fang, Hao, et al.
Veröffentlicht: (2026)
von: Fang, Hao, et al.
Veröffentlicht: (2026)
Towards Distillation-Resistant Large Language Models: An Information-Theoretic Perspective
von: Fang, Hao, et al.
Veröffentlicht: (2026)
von: Fang, Hao, et al.
Veröffentlicht: (2026)
Not All Prompts Are Secure: A Switchable Backdoor Attack Against Pre-trained Vision Transformers
von: Yang, Sheng, et al.
Veröffentlicht: (2024)
von: Yang, Sheng, et al.
Veröffentlicht: (2024)
Your Language Model Can Secretly Write Like Humans: Contrastive Paraphrase Attacks on LLM-Generated Text Detectors
von: Fang, Hao, et al.
Veröffentlicht: (2025)
von: Fang, Hao, et al.
Veröffentlicht: (2025)
Explicit Uncertainty Modeling for Active CLIP Adaptation with Dual Prompt Tuning
von: Wang, Qian-Wei, et al.
Veröffentlicht: (2026)
von: Wang, Qian-Wei, et al.
Veröffentlicht: (2026)
Benchmarking Open-ended Audio Dialogue Understanding for Large Audio-Language Models
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)
Class-Conditional Neural Polarizer: A Lightweight and Effective Backdoor Defense by Purifying Poisoned Features
von: Zhu, Mingli, et al.
Veröffentlicht: (2025)
von: Zhu, Mingli, et al.
Veröffentlicht: (2025)
CLIP-Inspector: Model-Level Backdoor Detection for Prompt-Tuned CLIP via OOD Trigger Inversion
von: Jindal, Akshit, et al.
Veröffentlicht: (2026)
von: Jindal, Akshit, et al.
Veröffentlicht: (2026)
Adversarial Backdoor Defense in CLIP
von: Kuang, Junhao, et al.
Veröffentlicht: (2024)
von: Kuang, Junhao, et al.
Veröffentlicht: (2024)
CodePurify: Defend Backdoor Attacks on Neural Code Models via Entropy-based Purification
von: Mu, Fangwen, et al.
Veröffentlicht: (2024)
von: Mu, Fangwen, et al.
Veröffentlicht: (2024)
JPRO: Automated Multimodal Jailbreaking via Multi-Agent Collaboration Framework
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
ToxicTextCLIP: Text-Based Poisoning and Backdoor Attacks on CLIP Pre-training
von: Yao, Xin, et al.
Veröffentlicht: (2025)
von: Yao, Xin, et al.
Veröffentlicht: (2025)
Looking Back and Forth: Cross-Image Attention Calibration and Attentive Preference Learning for Multi-Image Hallucination Mitigation
von: Yang, Xiaochen, et al.
Veröffentlicht: (2026)
von: Yang, Xiaochen, et al.
Veröffentlicht: (2026)
Lethe: Purifying Backdoored Large Language Models with Knowledge Dilution
von: Chen, Chen, et al.
Veröffentlicht: (2025)
von: Chen, Chen, et al.
Veröffentlicht: (2025)
Prompt Tuning for CLIP on the Pretrained Manifold
von: Yang, Xi, et al.
Veröffentlicht: (2026)
von: Yang, Xi, et al.
Veröffentlicht: (2026)
Adversarial Robustness for Visual Grounding of Multimodal Large Language Models
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)
Video Watermarking: Safeguarding Your Video from (Unauthorized) Annotations by Video-based LLMs
von: Li, Jinmin, et al.
Veröffentlicht: (2024)
von: Li, Jinmin, et al.
Veröffentlicht: (2024)
One Perturbation is Enough: On Generating Universal Adversarial Perturbations against Vision-Language Pre-training Models
von: Fang, Hao, et al.
Veröffentlicht: (2024)
von: Fang, Hao, et al.
Veröffentlicht: (2024)
Mistletoe: Stealthy Acceleration-Collapse Attacks on Speculative Decoding
von: Sun, Shuoyang, et al.
Veröffentlicht: (2026)
von: Sun, Shuoyang, et al.
Veröffentlicht: (2026)
C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts
von: Qing, Chenxi, et al.
Veröffentlicht: (2026)
von: Qing, Chenxi, et al.
Veröffentlicht: (2026)
Triple Antidot Molecules
von: Mizuno, Naomi, et al.
Veröffentlicht: (2026)
von: Mizuno, Naomi, et al.
Veröffentlicht: (2026)
Enhancing Gradient Inversion Attacks in Federated Learning via Hierarchical Feature Optimization
von: Fang, Hao, et al.
Veröffentlicht: (2026)
von: Fang, Hao, et al.
Veröffentlicht: (2026)
GI-NAS: Boosting Gradient Inversion Attacks Through Adaptive Neural Architecture Search
von: Yu, Wenbo, et al.
Veröffentlicht: (2024)
von: Yu, Wenbo, et al.
Veröffentlicht: (2024)
Backdooring CLIP through Concept Confusion
von: Hu, Lijie, et al.
Veröffentlicht: (2025)
von: Hu, Lijie, et al.
Veröffentlicht: (2025)
"No Matter What You Do": Purifying GNN Models via Backdoor Unlearning
von: Zhang, Jiale, et al.
Veröffentlicht: (2024)
von: Zhang, Jiale, et al.
Veröffentlicht: (2024)
Stealthy Dual-Trigger Backdoors: Attacking Prompt Tuning in LM-Empowered Graph Foundation Models
von: Xue, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Xue, Xiaoyu, et al.
Veröffentlicht: (2025)
Attention to the Burstiness in Visual Prompt Tuning!
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
Denial-of-Service Poisoning Attacks against Large Language Models
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)
Reasoning Matters: Mitigate Hallucination in Multimodal Large Reasoning Models via Reasoning-Conditioned Preference Optimization
von: Kong, Jiawei, et al.
Veröffentlicht: (2026)
von: Kong, Jiawei, et al.
Veröffentlicht: (2026)
Privacy Leakage on DNNs: A Survey of Model Inversion Attacks and Defenses
von: Fang, Hao, et al.
Veröffentlicht: (2024)
von: Fang, Hao, et al.
Veröffentlicht: (2024)
Enhancing CLIP with CLIP: Exploring Pseudolabeling for Limited-Label Prompt Tuning
von: Menghini, Cristina, et al.
Veröffentlicht: (2023)
von: Menghini, Cristina, et al.
Veröffentlicht: (2023)
Protecting Your Video Content: Disrupting Automated Video-based LLM Annotations
von: Liu, Haitong, et al.
Veröffentlicht: (2025)
von: Liu, Haitong, et al.
Veröffentlicht: (2025)
Antidote or poison: The relationship between “lying flat” tendency and mental health
von: Huanhua Lu, et al.
Veröffentlicht: (2024)
von: Huanhua Lu, et al.
Veröffentlicht: (2024)
Learning Generalizable Prompt for CLIP with Class Similarity Knowledge
von: Jung, Sehun, et al.
Veröffentlicht: (2025)
von: Jung, Sehun, et al.
Veröffentlicht: (2025)
Augmented Neural Fine-Tuning for Efficient Backdoor Purification
von: Karim, Nazmul, et al.
Veröffentlicht: (2024)
von: Karim, Nazmul, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
BadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIP
von: Bai, Jiawang, et al.
Veröffentlicht: (2023) -
Revisiting Backdoor Attacks on LLMs: A Stealthy and Practical Poisoning Framework via Harmless Inputs
von: Kong, Jiawei, et al.
Veröffentlicht: (2025) -
Retrievals Can Be Detrimental: Unveiling the Backdoor Vulnerability of Retrieval-Augmented Diffusion Models
von: Fang, Hao, et al.
Veröffentlicht: (2025) -
Grounding Language with Vision: A Conditional Mutual Information Calibrated Decoding Strategy for Reducing Hallucinations in LVLMs
von: Fang, Hao, et al.
Veröffentlicht: (2025) -
CLIP-Guided Generative Networks for Transferable Targeted Adversarial Attacks
von: Fang, Hao, et al.
Veröffentlicht: (2024)