Lie Detector: Unified Backdoor Detection via Cross-Examination Framework
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Xuan, Liang, Siyuan, Liao, Dongping, Fang, Han, Liu, Aishan, Cao, Xiaochun, Lu, Yu-liang, Chang, Ee-Chien, Gao, Xitong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BadCLIP: Dual-Embedding Guided Backdoor Attack on Multimodal Contrastive Learning
por: Liang, Siyuan, et al.
Publicado: (2023)
por: Liang, Siyuan, et al.
Publicado: (2023)
Poison Once, Control Anywhere: Clean-Text Visual Backdoors in VLM-based Mobile Agents
por: Wang, Xuan, et al.
Publicado: (2025)
por: Wang, Xuan, et al.
Publicado: (2025)
VL-Trojan: Multimodal Instruction Backdoor Attacks against Autoregressive Visual Language Models
por: Liang, Jiawei, et al.
Publicado: (2024)
por: Liang, Jiawei, et al.
Publicado: (2024)
TrapFlow: Controllable Website Fingerprinting Defense via Dynamic Backdoor Learning
por: Liang, Siyuan, et al.
Publicado: (2024)
por: Liang, Siyuan, et al.
Publicado: (2024)
Object Detectors in the Open Environment: Challenges, Solutions, and Outlook
por: Liang, Siyuan, et al.
Publicado: (2024)
por: Liang, Siyuan, et al.
Publicado: (2024)
Unlearning Backdoor Threats: Enhancing Backdoor Defense in Multimodal Contrastive Learning via Local Token Unlearning
por: Liang, Siyuan, et al.
Publicado: (2024)
por: Liang, Siyuan, et al.
Publicado: (2024)
SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents
por: Liang, Siyuan, et al.
Publicado: (2025)
por: Liang, Siyuan, et al.
Publicado: (2025)
Poisoned Forgery Face: Towards Backdoor Attacks on Face Forgery Detection
por: Liang, Jiawei, et al.
Publicado: (2024)
por: Liang, Jiawei, et al.
Publicado: (2024)
Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs
por: Li, Xiaoxia, et al.
Publicado: (2024)
por: Li, Xiaoxia, et al.
Publicado: (2024)
Revisiting Backdoor Attacks against Large Vision-Language Models from Domain Shift
por: Liang, Siyuan, et al.
Publicado: (2024)
por: Liang, Siyuan, et al.
Publicado: (2024)
Exploring Inconsistent Knowledge Distillation for Object Detection with Data Augmentation
por: Liang, Jiawei, et al.
Publicado: (2022)
por: Liang, Jiawei, et al.
Publicado: (2022)
Towards Robust Physical-world Backdoor Attacks on Lane Detection
por: Zhang, Xinwei, et al.
Publicado: (2024)
por: Zhang, Xinwei, et al.
Publicado: (2024)
Adversarial Backdoor Defense in CLIP
por: Kuang, Junhao, et al.
Publicado: (2024)
por: Kuang, Junhao, et al.
Publicado: (2024)
ICLShield: Exploring and Mitigating In-Context Learning Backdoor Attacks
por: Ren, Zhiyao, et al.
Publicado: (2025)
por: Ren, Zhiyao, et al.
Publicado: (2025)
ME: Trigger Element Combination Backdoor Attack on Copyright Infringement
por: Yang, Feiyu, et al.
Publicado: (2025)
por: Yang, Feiyu, et al.
Publicado: (2025)
FLIP: Towards Comprehensive and Reliable Evaluation of Federated Prompt Learning
por: Liao, Dongping, et al.
Publicado: (2025)
por: Liao, Dongping, et al.
Publicado: (2025)
Robust Anti-Backdoor Instruction Tuning in LVLMs
por: Xun, Yuan, et al.
Publicado: (2025)
por: Xun, Yuan, et al.
Publicado: (2025)
Domain Bridge: Generative model-based domain forensic for black-box models
por: Zhang, Jiyi, et al.
Publicado: (2024)
por: Zhang, Jiyi, et al.
Publicado: (2024)
Efficient Backdoor Defense in Multimodal Contrastive Learning: A Token-Level Unlearning Method for Mitigating Threats
por: Liu, Kuanrong, et al.
Publicado: (2024)
por: Liu, Kuanrong, et al.
Publicado: (2024)
RoboView-Bias: Benchmarking Visual Bias in Embodied Agents for Robotic Manipulation
por: Liu, Enguang, et al.
Publicado: (2025)
por: Liu, Enguang, et al.
Publicado: (2025)
CleanerCLIP: Fine-grained Counterfactual Semantic Augmentation for Backdoor Defense in Contrastive Learning
por: Xun, Yuan, et al.
Publicado: (2024)
por: Xun, Yuan, et al.
Publicado: (2024)
BadCLIP++: Stealthy and Persistent Backdoors in Multimodal Contrastive Learning
por: Liang, Siyuan, et al.
Publicado: (2026)
por: Liang, Siyuan, et al.
Publicado: (2026)
Does Few-shot Learning Suffer from Backdoor Attacks?
por: Liu, Xinwei, et al.
Publicado: (2023)
por: Liu, Xinwei, et al.
Publicado: (2023)
Proof-of-Authorship for Diffusion-based AI Generated Content
por: Lee, De Zhang, et al.
Publicado: (2026)
por: Lee, De Zhang, et al.
Publicado: (2026)
T2VShield: Model-Agnostic Jailbreak Defense for Text-to-Video Models
por: Liang, Siyuan, et al.
Publicado: (2025)
por: Liang, Siyuan, et al.
Publicado: (2025)
ELBA-Bench: An Efficient Learning Backdoor Attacks Benchmark for Large Language Models
por: Liu, Xuxu, et al.
Publicado: (2025)
por: Liu, Xuxu, et al.
Publicado: (2025)
Removal Attack and Defense on AI-generated Content Latent-based Watermarking
por: Lee, De Zhang, et al.
Publicado: (2025)
por: Lee, De Zhang, et al.
Publicado: (2025)
WFCAT: Augmenting Website Fingerprinting with Channel-wise Attention on Timing Features
por: Gong, Jiajun, et al.
Publicado: (2024)
por: Gong, Jiajun, et al.
Publicado: (2024)
Mixture of Weight-shared Heterogeneous Group Attention Experts for Dynamic Token-wise KV Optimization
por: Song, Guanghui, et al.
Publicado: (2025)
por: Song, Guanghui, et al.
Publicado: (2025)
R-PGA: Robust Physical Adversarial Camouflage Generation via Relightable 3D Gaussian Splatting
por: Lou, Tianrui, et al.
Publicado: (2026)
por: Lou, Tianrui, et al.
Publicado: (2026)
ResetEdit: Precise Text-guided Editing of Generated Image via Resettable Starting Latent
por: Wang, Hanyi, et al.
Publicado: (2026)
por: Wang, Hanyi, et al.
Publicado: (2026)
ResGuard: Enhancing Robustness Against Known Original Attacks in Deep Watermarking
por: Wang, Hanyi, et al.
Publicado: (2026)
por: Wang, Hanyi, et al.
Publicado: (2026)
SRD: Reinforcement-Learned Semantic Perturbation for Backdoor Defense in VLMs
por: Xu, Shuhan, et al.
Publicado: (2025)
por: Xu, Shuhan, et al.
Publicado: (2025)
Towards Robust Object Detection: Identifying and Removing Backdoors via Module Inconsistency Analysis
por: Zhang, Xianda, et al.
Publicado: (2024)
por: Zhang, Xianda, et al.
Publicado: (2024)
BridgeNet: A Unified Multimodal Framework for Bridging 2D and 3D Industrial Anomaly Detection
por: Xiang, An, et al.
Publicado: (2025)
por: Xiang, An, et al.
Publicado: (2025)
SnapGuard: Lightweight Prompt Injection Detection for Screenshot-Based Web Agents
por: Du, Mengyao, et al.
Publicado: (2026)
por: Du, Mengyao, et al.
Publicado: (2026)
BDefects4NN: A Backdoor Defect Database for Controlled Localization Studies in Neural Networks
por: Xiao, Yisong, et al.
Publicado: (2024)
por: Xiao, Yisong, et al.
Publicado: (2024)
Text Adversarial Attacks with Dynamic Outputs
por: Wang, Wenqiang, et al.
Publicado: (2025)
por: Wang, Wenqiang, et al.
Publicado: (2025)
CopyrightShield: Enhancing Diffusion Model Security against Copyright Infringement Attacks
por: Guo, Zhixiang, et al.
Publicado: (2024)
por: Guo, Zhixiang, et al.
Publicado: (2024)
Compromising Embodied Agents with Contextual Backdoor Attacks
por: Liu, Aishan, et al.
Publicado: (2024)
por: Liu, Aishan, et al.
Publicado: (2024)
Ejemplares similares
-
BadCLIP: Dual-Embedding Guided Backdoor Attack on Multimodal Contrastive Learning
por: Liang, Siyuan, et al.
Publicado: (2023) -
Poison Once, Control Anywhere: Clean-Text Visual Backdoors in VLM-based Mobile Agents
por: Wang, Xuan, et al.
Publicado: (2025) -
VL-Trojan: Multimodal Instruction Backdoor Attacks against Autoregressive Visual Language Models
por: Liang, Jiawei, et al.
Publicado: (2024) -
TrapFlow: Controllable Website Fingerprinting Defense via Dynamic Backdoor Learning
por: Liang, Siyuan, et al.
Publicado: (2024) -
Object Detectors in the Open Environment: Challenges, Solutions, and Outlook
por: Liang, Siyuan, et al.
Publicado: (2024)