Breaking the Stealth-Potency Trade-off in Clean-Image Backdoors with Generative Trigger Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Binyan, Yang, Fan, Tang, Di, Dai, Xilin, Zhang, Kehuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLIP-Guided Backdoor Defense through Entropy-Based Poisoned Dataset Separation
von: Xu, Binyan, et al.
Veröffentlicht: (2025)
von: Xu, Binyan, et al.
Veröffentlicht: (2025)
One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP
von: Xu, Binyan, et al.
Veröffentlicht: (2025)
von: Xu, Binyan, et al.
Veröffentlicht: (2025)
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
von: Dang, Kieu, et al.
Veröffentlicht: (2025)
von: Dang, Kieu, et al.
Veröffentlicht: (2025)
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
von: Hill, Brennen, et al.
Veröffentlicht: (2025)
von: Hill, Brennen, et al.
Veröffentlicht: (2025)
BreakFun: Jailbreaking LLMs via Schema Exploitation
von: Oskooei, Amirkia Rafiei, et al.
Veröffentlicht: (2025)
von: Oskooei, Amirkia Rafiei, et al.
Veröffentlicht: (2025)
"Abuse Risks are Often Inherent to Product Features": Exploring AI Vendors' Bug Bounty and Responsible Disclosure Policies
von: Piao, Yangheran, et al.
Veröffentlicht: (2025)
von: Piao, Yangheran, et al.
Veröffentlicht: (2025)
A Dual-Path Generative Framework for Zero-Day Fraud Detection in Banking Systems
von: Ismail, Nasim Abdirahman, et al.
Veröffentlicht: (2026)
von: Ismail, Nasim Abdirahman, et al.
Veröffentlicht: (2026)
Explainable Attention-Based LSTM Framework for Early Detection of AI-Assisted Ransomware via File System Behavioral Analysis
von: Nayak, Prabhudarshi, et al.
Veröffentlicht: (2026)
von: Nayak, Prabhudarshi, et al.
Veröffentlicht: (2026)
PACT: Reducing Alert Fatigue in Low-Prevalence SOC Streams with Triggered Active Learning
von: Ndichu, Samuel, et al.
Veröffentlicht: (2026)
von: Ndichu, Samuel, et al.
Veröffentlicht: (2026)
Latent Sculpting for Zero-Shot Generalization: A Manifold Learning Approach to Out-of-Distribution Anomaly Detection
von: Chhetri, Rajeeb Thapa, et al.
Veröffentlicht: (2025)
von: Chhetri, Rajeeb Thapa, et al.
Veröffentlicht: (2025)
Robust DDoS-Attack Classification with 3D CNNs Against Adversarial Methods
von: Bragg, Landon, et al.
Veröffentlicht: (2025)
von: Bragg, Landon, et al.
Veröffentlicht: (2025)
The Age of Sensorial Zero Trust: Why We Can No Longer Trust Our Senses
von: Xavier, Fabio Correa
Veröffentlicht: (2025)
von: Xavier, Fabio Correa
Veröffentlicht: (2025)
A V2X-based Privacy Preserving Federated Measuring and Learning System
von: Alekszejenkó, Levente, et al.
Veröffentlicht: (2024)
von: Alekszejenkó, Levente, et al.
Veröffentlicht: (2024)
Safety, Security, and Cognitive Risks in World Models
von: Parmar, Manoj
Veröffentlicht: (2026)
von: Parmar, Manoj
Veröffentlicht: (2026)
Machine Unlearning for Class Removal through SISA-based Deep Neural Network Architectures
von: Mahi, Ishrak Hamim, et al.
Veröffentlicht: (2026)
von: Mahi, Ishrak Hamim, et al.
Veröffentlicht: (2026)
DeTrigger: A Gradient-Centric Approach to Backdoor Attack Mitigation in Federated Learning
von: Lee, Kichang, et al.
Veröffentlicht: (2024)
von: Lee, Kichang, et al.
Veröffentlicht: (2024)
The Inhibitor: ReLU and Addition-Based Attention for Efficient Transformers under Fully Homomorphic Encryption on the Torus
von: Brännvall, Rickard, et al.
Veröffentlicht: (2023)
von: Brännvall, Rickard, et al.
Veröffentlicht: (2023)
DRIFT: Drift-Resilient Invariant-Feature Transformer for DGA Detection
von: Lee, Chaeyoung, et al.
Veröffentlicht: (2026)
von: Lee, Chaeyoung, et al.
Veröffentlicht: (2026)
Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3
von: Nitarach, Natapong
Veröffentlicht: (2026)
von: Nitarach, Natapong
Veröffentlicht: (2026)
Step-Aware Residual-Guided Diffusion for EEG Spatial Super-Resolution
von: Liu, Hongjun, et al.
Veröffentlicht: (2025)
von: Liu, Hongjun, et al.
Veröffentlicht: (2025)
Portable Agent Memory: A Protocol for Cryptographically-Verified Memory Transfer Across Heterogeneous AI Agents
von: Ravindran, Santhosh Kumar
Veröffentlicht: (2026)
von: Ravindran, Santhosh Kumar
Veröffentlicht: (2026)
Illuminating the Black Box: Real-Time Monitoring of Backdoor Unlearning in CNNs via Explainable AI
von: Hoang, Tien Dat
Veröffentlicht: (2025)
von: Hoang, Tien Dat
Veröffentlicht: (2025)
From Features to Graphs: Exploring Graph Structures and Pairwise Interactions via GNNs
von: Yamchote, Phaphontee, et al.
Veröffentlicht: (2025)
von: Yamchote, Phaphontee, et al.
Veröffentlicht: (2025)
NeuronSpark: A Spiking Neural Network Language Model with Selective State Space Dynamics
von: Tang, Zhengzheng
Veröffentlicht: (2026)
von: Tang, Zhengzheng
Veröffentlicht: (2026)
Semantically Guided Adversarial Testing of Vision Models Using Language Models
von: Filus, Katarzyna, et al.
Veröffentlicht: (2025)
von: Filus, Katarzyna, et al.
Veröffentlicht: (2025)
An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning
von: Saini, Saurabh, et al.
Veröffentlicht: (2026)
von: Saini, Saurabh, et al.
Veröffentlicht: (2026)
Operationalizing Cybersecurity Governance for Mitigation Planning with Attack-Path Modeling and Reinforcement Learning
von: Huff, Philip, et al.
Veröffentlicht: (2026)
von: Huff, Philip, et al.
Veröffentlicht: (2026)
Verifiable Fine-Tuning for LLMs: Zero-Knowledge Training Proofs Bound to Data Provenance and Policy
von: Akgul, Hasan, et al.
Veröffentlicht: (2025)
von: Akgul, Hasan, et al.
Veröffentlicht: (2025)
Upside Down Reinforcement Learning with Policy Generators
von: Di Ventura, Jacopo, et al.
Veröffentlicht: (2025)
von: Di Ventura, Jacopo, et al.
Veröffentlicht: (2025)
Breaking Boundaries: Balancing Performance and Robustness in Deep Wireless Traffic Forecasting
von: Ilbert, Romain, et al.
Veröffentlicht: (2023)
von: Ilbert, Romain, et al.
Veröffentlicht: (2023)
Optimizing Inference in Transformer-Based Models: A Multi-Method Benchmark
von: Ho, Siu Hang, et al.
Veröffentlicht: (2025)
von: Ho, Siu Hang, et al.
Veröffentlicht: (2025)
Deep Hedging Under Non-Convexity: Limitations and a Case for AlphaZero
von: Maggiolo, Matteo, et al.
Veröffentlicht: (2025)
von: Maggiolo, Matteo, et al.
Veröffentlicht: (2025)
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
Stage-wise Dynamics of Classifier-Free Guidance in Diffusion Models
von: Jin, Cheng, et al.
Veröffentlicht: (2025)
von: Jin, Cheng, et al.
Veröffentlicht: (2025)
Semantic Reward Collapse and the Preservation of Epistemic Integrity in Adaptive AI Systems
von: Parris, William
Veröffentlicht: (2026)
von: Parris, William
Veröffentlicht: (2026)
Drift-Resilient TabPFN: In-Context Learning Temporal Distribution Shifts on Tabular Data
von: Helli, Kai, et al.
Veröffentlicht: (2024)
von: Helli, Kai, et al.
Veröffentlicht: (2024)
FeNeC: Enhancing Continual Learning via Feature Clustering with Neighbor- or Logit-Based Classification
von: Książek, Kamil, et al.
Veröffentlicht: (2025)
von: Książek, Kamil, et al.
Veröffentlicht: (2025)
X-Factor: Quality Is a Dataset-Intrinsic Property
von: Couch, Josiah, et al.
Veröffentlicht: (2025)
von: Couch, Josiah, et al.
Veröffentlicht: (2025)
A Constraint-Preserving Neural Network Approach for Solving Mean-Field Games Equilibrium
von: Liu, Jinwei, et al.
Veröffentlicht: (2025)
von: Liu, Jinwei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CLIP-Guided Backdoor Defense through Entropy-Based Poisoned Dataset Separation
von: Xu, Binyan, et al.
Veröffentlicht: (2025) -
One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP
von: Xu, Binyan, et al.
Veröffentlicht: (2025) -
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025) -
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
von: Dang, Kieu, et al.
Veröffentlicht: (2025) -
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
von: Hill, Brennen, et al.
Veröffentlicht: (2025)