Amplifying Machine Learning Attacks Through Strategic Compositions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yugeng, Li, Zheng, Huang, Hai, Backes, Michael, Zhang, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Watermarking LLM-Generated Datasets in Downstream Tasks
von: Liu, Yugeng, et al.
Veröffentlicht: (2025)
von: Liu, Yugeng, et al.
Veröffentlicht: (2025)
$\texttt{ModSCAN}$: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
von: Jiang, Yukun, et al.
Veröffentlicht: (2024)
von: Jiang, Yukun, et al.
Veröffentlicht: (2024)
Robustness Over Time: Understanding Adversarial Examples' Effectiveness on Longitudinal Versions of Large Language Models
von: Liu, Yugeng, et al.
Veröffentlicht: (2023)
von: Liu, Yugeng, et al.
Veröffentlicht: (2023)
JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
von: Chu, Junjie, et al.
Veröffentlicht: (2024)
von: Chu, Junjie, et al.
Veröffentlicht: (2024)
Composite Backdoor Attacks Against Large Language Models
von: Huang, Hai, et al.
Veröffentlicht: (2023)
von: Huang, Hai, et al.
Veröffentlicht: (2023)
Membership Inference Attacks Against In-Context Learning
von: Wen, Rui, et al.
Veröffentlicht: (2024)
von: Wen, Rui, et al.
Veröffentlicht: (2024)
Understanding Data Importance in Machine Learning Attacks: Does Valuable Data Pose Greater Harm?
von: Wen, Rui, et al.
Veröffentlicht: (2024)
von: Wen, Rui, et al.
Veröffentlicht: (2024)
SoK: Data Reconstruction Attacks Against Machine Learning Models: Definition, Metrics, and Benchmark
von: Wen, Rui, et al.
Veröffentlicht: (2025)
von: Wen, Rui, et al.
Veröffentlicht: (2025)
Transferable Availability Poisoning Attacks
von: Liu, Yiyong, et al.
Veröffentlicht: (2023)
von: Liu, Yiyong, et al.
Veröffentlicht: (2023)
Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
Vera Verto: Multimodal Hijacking Attack
von: Zhang, Minxing, et al.
Veröffentlicht: (2024)
von: Zhang, Minxing, et al.
Veröffentlicht: (2024)
Excessive Reasoning Attack on Reasoning LLMs
von: Si, Wai Man, et al.
Veröffentlicht: (2025)
von: Si, Wai Man, et al.
Veröffentlicht: (2025)
Voice Jailbreak Attacks Against GPT-4o
von: Shen, Xinyue, et al.
Veröffentlicht: (2024)
von: Shen, Xinyue, et al.
Veröffentlicht: (2024)
Prompt Stealing Attacks Against Text-to-Image Generation Models
von: Shen, Xinyue, et al.
Veröffentlicht: (2023)
von: Shen, Xinyue, et al.
Veröffentlicht: (2023)
AttackPilot: Autonomous Inference Attacks Against ML Services With LLM-Based Agents
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
Invisibility Cloak: Disappearance under Human Pose Estimation via Backdoor Attacks
von: Zhang, Minxing, et al.
Veröffentlicht: (2024)
von: Zhang, Minxing, et al.
Veröffentlicht: (2024)
Trust Me, Import This: Dependency Steering Attacks via Malicious Agent Skills
von: Liu, Yiyong, et al.
Veröffentlicht: (2026)
von: Liu, Yiyong, et al.
Veröffentlicht: (2026)
Efficient Data-Free Model Stealing with Label Diversity
von: Liu, Yiyong, et al.
Veröffentlicht: (2024)
von: Liu, Yiyong, et al.
Veröffentlicht: (2024)
Adjacent Words, Divergent Intents: Jailbreaking Large Language Models via Task Concurrency
von: Jiang, Yukun, et al.
Veröffentlicht: (2025)
von: Jiang, Yukun, et al.
Veröffentlicht: (2025)
Secure Composition of Robust and Optimising Compilers
von: Kruse, Matthis, et al.
Veröffentlicht: (2023)
von: Kruse, Matthis, et al.
Veröffentlicht: (2023)
Instruction Backdoor Attacks Against Customized LLMs
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
SOS! Soft Prompt Attack Against Open-Source Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2024)
von: Yang, Ziqing, et al.
Veröffentlicht: (2024)
MGTBench: Benchmarking Machine-Generated Text Detection
von: He, Xinlei, et al.
Veröffentlicht: (2023)
von: He, Xinlei, et al.
Veröffentlicht: (2023)
ICLGuard: Controlling In-Context Learning Behavior for Applicability Authorization
von: Si, Wai Man, et al.
Veröffentlicht: (2024)
von: Si, Wai Man, et al.
Veröffentlicht: (2024)
BadBone: Backdoor Attacks Against Backbone Models in Visual Prompt Learning
von: Yang, Ziqing, et al.
Veröffentlicht: (2026)
von: Yang, Ziqing, et al.
Veröffentlicht: (2026)
CAMH: Advancing Model Hijacking Attack in Machine Learning
von: He, Xing, et al.
Veröffentlicht: (2024)
von: He, Xing, et al.
Veröffentlicht: (2024)
Sparse Models, Sparse Safety: Unsafe Routes in Mixture-of-Experts LLMs
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
Peering Behind the Shield: Guardrail Identification in Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
Label Leakage Attacks in Machine Unlearning: A Parameter and Inversion-Based Approach
von: Zheng, Weidong, et al.
Veröffentlicht: (2026)
von: Zheng, Weidong, et al.
Veröffentlicht: (2026)
Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities
von: Qu, Yiting, et al.
Veröffentlicht: (2025)
von: Qu, Yiting, et al.
Veröffentlicht: (2025)
Link Stealing Attacks Against Inductive Graph Neural Networks
von: Wu, Yixin, et al.
Veröffentlicht: (2024)
von: Wu, Yixin, et al.
Veröffentlicht: (2024)
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification
von: Zhang, Boyang, et al.
Veröffentlicht: (2024)
von: Zhang, Boyang, et al.
Veröffentlicht: (2024)
On the Proactive Generation of Unsafe Images From Text-To-Image Models Using Benign Prompts
von: Wu, Yixin, et al.
Veröffentlicht: (2023)
von: Wu, Yixin, et al.
Veröffentlicht: (2023)
How Secure is Forgetting? Linking Machine Unlearning to Machine Learning Attacks
von: P., Muhammed Shafi K., et al.
Veröffentlicht: (2025)
von: P., Muhammed Shafi K., et al.
Veröffentlicht: (2025)
Detecting Quishing Attacks with Machine Learning Techniques Through QR Code Analysis
von: Trad, Fouad, et al.
Veröffentlicht: (2025)
von: Trad, Fouad, et al.
Veröffentlicht: (2025)
A Comprehensive Review of Adversarial Attacks on Machine Learning
von: Ahmed, Syed Quiser, et al.
Veröffentlicht: (2024)
von: Ahmed, Syed Quiser, et al.
Veröffentlicht: (2024)
Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate
von: Qi, Senmao, et al.
Veröffentlicht: (2025)
von: Qi, Senmao, et al.
Veröffentlicht: (2025)
Machine Learning Power Side-Channel Attack on SNOW-V
von: Deepak, et al.
Veröffentlicht: (2025)
von: Deepak, et al.
Veröffentlicht: (2025)
RLCracker: Evaluating the Worst-Case Vulnerability of LLM Watermarks with Adaptive RL Attacks
von: Huang, Hanbo, et al.
Veröffentlicht: (2025)
von: Huang, Hanbo, et al.
Veröffentlicht: (2025)
Enhancing Privacy of Spatiotemporal Federated Learning against Gradient Inversion Attacks
von: Zheng, Lele, et al.
Veröffentlicht: (2024)
von: Zheng, Lele, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Watermarking LLM-Generated Datasets in Downstream Tasks
von: Liu, Yugeng, et al.
Veröffentlicht: (2025) -
$\texttt{ModSCAN}$: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
von: Jiang, Yukun, et al.
Veröffentlicht: (2024) -
Robustness Over Time: Understanding Adversarial Examples' Effectiveness on Longitudinal Versions of Large Language Models
von: Liu, Yugeng, et al.
Veröffentlicht: (2023) -
JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
von: Chu, Junjie, et al.
Veröffentlicht: (2024) -
Composite Backdoor Attacks Against Large Language Models
von: Huang, Hai, et al.
Veröffentlicht: (2023)