Saved in:
| Main Authors: | He, Bowei, Yin, Lihao, Zhen, Hui-Ling, Zhang, Jianping, Hong, Lanqing, Yuan, Mingxuan, Ma, Chen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2502.06892 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PASER: Post-Training Data Selection for Efficient Pruned Large Language Model Recovery
by: He, Bowei, et al.
Published: (2025)
by: He, Bowei, et al.
Published: (2025)
Attention-Aware GNN-based Input Defense against Multi-Turn LLM Jailbreak
by: Huang, Zixuan, et al.
Published: (2025)
by: Huang, Zixuan, et al.
Published: (2025)
PECAN: A Deterministic Certified Defense Against Backdoor Attacks
by: Zhang, Yuhao, et al.
Published: (2023)
by: Zhang, Yuhao, et al.
Published: (2023)
Distributed Backdoor Attacks on Federated Graph Learning and Certified Defenses
by: Yang, Yuxin, et al.
Published: (2024)
by: Yang, Yuxin, et al.
Published: (2024)
Multi-Faceted Attack: Exposing Cross-Model Vulnerabilities in Defense-Equipped Vision-Language Models
by: Yang, Yijun, et al.
Published: (2025)
by: Yang, Yijun, et al.
Published: (2025)
Preserving LLM Capabilities through Calibration Data Curation: From Analysis to Optimization
by: He, Bowei, et al.
Published: (2025)
by: He, Bowei, et al.
Published: (2025)
Robust and Transferable Backdoor Attacks Against Deep Image Compression With Selective Frequency Prior
by: Yu, Yi, et al.
Published: (2024)
by: Yu, Yi, et al.
Published: (2024)
Cert-SSBD: Certified Backdoor Defense with Sample-Specific Smoothing Noises
by: Qiao, Ting, et al.
Published: (2025)
by: Qiao, Ting, et al.
Published: (2025)
TrojVLM: Backdoor Attack Against Vision Language Models
by: Lyu, Weimin, et al.
Published: (2024)
by: Lyu, Weimin, et al.
Published: (2024)
Certifiable Black-Box Attacks with Randomized Adversarial Examples: Breaking Defenses with Provable Confidence
by: Hong, Hanbin, et al.
Published: (2023)
by: Hong, Hanbin, et al.
Published: (2023)
What Matters For Safety Alignment?
by: Li, Xing, et al.
Published: (2026)
by: Li, Xing, et al.
Published: (2026)
Defense Against Syntactic Textual Backdoor Attacks with Token Substitution
by: Li, Xinglin, et al.
Published: (2024)
by: Li, Xinglin, et al.
Published: (2024)
Robust Defense Strategies for Multimodal Contrastive Learning: Efficient Fine-tuning Against Backdoor Attacks
by: Hossain, Md. Iqbal, et al.
Published: (2025)
by: Hossain, Md. Iqbal, et al.
Published: (2025)
Certified Causal Defense with Generalizable Robustness
by: Qiao, Yiran, et al.
Published: (2024)
by: Qiao, Yiran, et al.
Published: (2024)
TED-LaST: Towards Robust Backdoor Defense Against Adaptive Attacks
by: Mo, Xiaoxing, et al.
Published: (2025)
by: Mo, Xiaoxing, et al.
Published: (2025)
Towards Strong Certified Defense with Universal Asymmetric Randomization
by: Hong, Hanbin, et al.
Published: (2025)
by: Hong, Hanbin, et al.
Published: (2025)
Variance-Based Defense Against Blended Backdoor Attacks
by: Aseervatham, Sujeevan, et al.
Published: (2025)
by: Aseervatham, Sujeevan, et al.
Published: (2025)
BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models
by: Li, Yige, et al.
Published: (2024)
by: Li, Yige, et al.
Published: (2024)
A Whole-Process Certifiably Robust Aggregation Method Against Backdoor Attacks in Federated Learning
by: Zhou, Anqi, et al.
Published: (2024)
by: Zhou, Anqi, et al.
Published: (2024)
DiLA: Enhancing LLM Tool Learning with Differential Logic Layer
by: Zhang, Yu, et al.
Published: (2024)
by: Zhang, Yu, et al.
Published: (2024)
Jailbreak Attacks and Defenses Against Large Language Models: A Survey
by: Yi, Sibo, et al.
Published: (2024)
by: Yi, Sibo, et al.
Published: (2024)
Authority Backdoor: A Certifiable Backdoor Mechanism for Authoring DNNs
by: Yang, Han, et al.
Published: (2025)
by: Yang, Han, et al.
Published: (2025)
PEFTGuard: Detecting Backdoor Attacks Against Parameter-Efficient Fine-Tuning
by: Sun, Zhen, et al.
Published: (2024)
by: Sun, Zhen, et al.
Published: (2024)
E-SAGE: Explainability-based Defense Against Backdoor Attacks on Graph Neural Networks
by: Yuan, Dingqiang, et al.
Published: (2024)
by: Yuan, Dingqiang, et al.
Published: (2024)
GANcrop: A Contrastive Defense Against Backdoor Attacks in Federated Learning
by: Gan, Xiaoyun, et al.
Published: (2024)
by: Gan, Xiaoyun, et al.
Published: (2024)
Laplace-Bridged Randomized Smoothing for Fast Certified Robustness
by: Lin, Miao, et al.
Published: (2026)
by: Lin, Miao, et al.
Published: (2026)
Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning
by: Liu, Shijie, et al.
Published: (2025)
by: Liu, Shijie, et al.
Published: (2025)
General Lipschitz: Certified Robustness Against Resolvable Semantic Transformations via Transformation-Dependent Randomized Smoothing
by: Korzh, Dmitrii, et al.
Published: (2023)
by: Korzh, Dmitrii, et al.
Published: (2023)
Certified Robustness for Deep Equilibrium Models via Serialized Random Smoothing
by: Gao, Weizhi, et al.
Published: (2024)
by: Gao, Weizhi, et al.
Published: (2024)
Composite Backdoor Attacks Against Large Language Models
by: Huang, Hai, et al.
Published: (2023)
by: Huang, Hai, et al.
Published: (2023)
Adversarially Guided Stateful Defense Against Backdoor Attacks in Federated Deep Learning
by: Ali, Hassan, et al.
Published: (2024)
by: Ali, Hassan, et al.
Published: (2024)
MOSS: Efficient and Accurate FP8 LLM Training with Microscaling and Automatic Scaling
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
Efficient Based on Improved Random Forest Defense System Against Application‐Layer DDoS Attacks
by: Junjiang He, et al.
Published: (2024)
by: Junjiang He, et al.
Published: (2024)
HoneypotNet: Backdoor Attacks Against Model Extraction
by: Wang, Yixu, et al.
Published: (2025)
by: Wang, Yixu, et al.
Published: (2025)
Fuzzing: Randomness? Reasoning! Efficient Directed Fuzzing via Large Language Models
by: Feng, Xiaotao, et al.
Published: (2025)
by: Feng, Xiaotao, et al.
Published: (2025)
BadMerging: Backdoor Attacks Against Model Merging
by: Zhang, Jinghuai, et al.
Published: (2024)
by: Zhang, Jinghuai, et al.
Published: (2024)
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
by: Wei, Shaokui, et al.
Published: (2024)
by: Wei, Shaokui, et al.
Published: (2024)
SPLITZ: Certifiable Robustness via Split Lipschitz Randomized Smoothing
by: Zhong, Meiyu, et al.
Published: (2024)
by: Zhong, Meiyu, et al.
Published: (2024)
Certifiably-Robust Federated Adversarial Learning via Randomized Smoothing
by: Chen, Cheng, et al.
Published: (2021)
by: Chen, Cheng, et al.
Published: (2021)
Certified Adversarial Robustness via Partition-based Randomized Smoothing
by: Goli, Hossein, et al.
Published: (2024)
by: Goli, Hossein, et al.
Published: (2024)
Similar Items
-
PASER: Post-Training Data Selection for Efficient Pruned Large Language Model Recovery
by: He, Bowei, et al.
Published: (2025) -
Attention-Aware GNN-based Input Defense against Multi-Turn LLM Jailbreak
by: Huang, Zixuan, et al.
Published: (2025) -
PECAN: A Deterministic Certified Defense Against Backdoor Attacks
by: Zhang, Yuhao, et al.
Published: (2023) -
Distributed Backdoor Attacks on Federated Graph Learning and Certified Defenses
by: Yang, Yuxin, et al.
Published: (2024) -
Multi-Faceted Attack: Exposing Cross-Model Vulnerabilities in Defense-Equipped Vision-Language Models
by: Yang, Yijun, et al.
Published: (2025)