Attacks Meet Interpretability (AmI) Evaluation and Findings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Qian, Ye, Ziping, Mehnaz, Shagufta |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GNNBleed: Inference Attacks to Unveil Private Edges in Graphs with Realistic Access to GNN Models
von: Song, Zeyu, et al.
Veröffentlicht: (2023)
von: Song, Zeyu, et al.
Veröffentlicht: (2023)
Baiting AI: Deceptive Adversary Against AI-Protected Industrial Infrastructures
von: Pasikhani, Aryan, et al.
Veröffentlicht: (2026)
von: Pasikhani, Aryan, et al.
Veröffentlicht: (2026)
Enhancing Adversarial Example Detection Through Model Explanation
von: Ma, Qian, et al.
Veröffentlicht: (2025)
von: Ma, Qian, et al.
Veröffentlicht: (2025)
I Stolenly Swear That I Am Up to (No) Good: Design and Evaluation of Model Stealing Attacks
von: Oliynyk, Daryna, et al.
Veröffentlicht: (2025)
von: Oliynyk, Daryna, et al.
Veröffentlicht: (2025)
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage
von: Rashid, Md Rafi Ur, et al.
Veröffentlicht: (2024)
von: Rashid, Md Rafi Ur, et al.
Veröffentlicht: (2024)
Gradient-Free Privacy Leakage in Federated Language Models through Selective Weight Tampering
von: Rashid, Md Rafi Ur, et al.
Veröffentlicht: (2023)
von: Rashid, Md Rafi Ur, et al.
Veröffentlicht: (2023)
LogicEval: A Systematic Framework for Evaluating Automated Repair Techniques for Logical Vulnerabilities in Real-World Software
von: Rashid, Syed Md Mukit, et al.
Veröffentlicht: (2026)
von: Rashid, Syed Md Mukit, et al.
Veröffentlicht: (2026)
Finding Software Supply Chain Attack Paths with Logical Attack Graphs
von: Soeiro, Luıs, et al.
Veröffentlicht: (2025)
von: Soeiro, Luıs, et al.
Veröffentlicht: (2025)
Am I Infected? Lessons from Operating a Large-Scale IoT Security Diagnostic Service
von: Sasaki, Takayuki, et al.
Veröffentlicht: (2025)
von: Sasaki, Takayuki, et al.
Veröffentlicht: (2025)
Cross-Service Token: Finding Attacks in 5G Core Networks
von: Chen, Anqi, et al.
Veröffentlicht: (2025)
von: Chen, Anqi, et al.
Veröffentlicht: (2025)
Comprehensive Evaluation of Cloaking Backdoor Attacks on Object Detector in Real-World
von: Ma, Hua, et al.
Veröffentlicht: (2025)
von: Ma, Hua, et al.
Veröffentlicht: (2025)
Rethinking Fraud Safety Evaluation: Multi-Round Attacks Reveal Safety-Utility Tradeoffs in Graph-Context LLM Defenders
von: Jiang, Laura, et al.
Veröffentlicht: (2026)
von: Jiang, Laura, et al.
Veröffentlicht: (2026)
Breaking the Prompt Wall (I): A Real-World Case Study of Attacking ChatGPT via Lightweight Prompt Injection
von: Chang, Xiangyu, et al.
Veröffentlicht: (2025)
von: Chang, Xiangyu, et al.
Veröffentlicht: (2025)
Bones of Contention: Exploring Query-Efficient Attacks against Skeleton Recognition Systems
von: Cao, Yuxin, et al.
Veröffentlicht: (2025)
von: Cao, Yuxin, et al.
Veröffentlicht: (2025)
Interpreting Adversarial Attacks and Defences using Architectures with Enhanced Interpretability
von: Rao, Akshay G, et al.
Veröffentlicht: (2025)
von: Rao, Akshay G, et al.
Veröffentlicht: (2025)
Model Inversion Attacks Meet Cryptographic Fuzzy Extractors
von: Prabhakar, Mallika, et al.
Veröffentlicht: (2025)
von: Prabhakar, Mallika, et al.
Veröffentlicht: (2025)
Defensive Prompt Patch: A Robust and Interpretable Defense of LLMs against Jailbreak Attacks
von: Xiong, Chen, et al.
Veröffentlicht: (2024)
von: Xiong, Chen, et al.
Veröffentlicht: (2024)
Quantization Aware Attack: Enhancing Transferable Adversarial Attacks by Model Quantization
von: Yang, Yulong, et al.
Veröffentlicht: (2023)
von: Yang, Yulong, et al.
Veröffentlicht: (2023)
Quantum Meet-in-the-Middle Attacks on Key-Length Extension Constructions
von: Liang, Min, et al.
Veröffentlicht: (2025)
von: Liang, Min, et al.
Veröffentlicht: (2025)
"Training robust watermarking model may hurt authentication!'' Exploring and Mitigating the Identity Leakage in Robust Watermarking
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026)
Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
Nip in the Bud: Forecasting and Interpreting Post-exploitation Attacks in Real-time through Cyber Threat Intelligence Reports
von: Zhu, Tiantian, et al.
Veröffentlicht: (2024)
von: Zhu, Tiantian, et al.
Veröffentlicht: (2024)
Infighting in the Dark: Multi-Label Backdoor Attack in Federated Learning
von: Li, Ye, et al.
Veröffentlicht: (2024)
von: Li, Ye, et al.
Veröffentlicht: (2024)
Cybersecurity AI: Evaluating Agentic Cybersecurity in Attack/Defense CTFs
von: Balassone, Francesco, et al.
Veröffentlicht: (2025)
von: Balassone, Francesco, et al.
Veröffentlicht: (2025)
Attack Effect Model based Malicious Behavior Detection
von: Wang, Limin, et al.
Veröffentlicht: (2025)
von: Wang, Limin, et al.
Veröffentlicht: (2025)
AmBox: Device-to-Blockchain Ambient Sensing for Food Traceability
von: Fernandes, João Miguel Guerreiro, et al.
Veröffentlicht: (2026)
von: Fernandes, João Miguel Guerreiro, et al.
Veröffentlicht: (2026)
Examining I2P Resilience: Effect of Centrality-based Attack
von: Akanbi, Kemi, et al.
Veröffentlicht: (2025)
von: Akanbi, Kemi, et al.
Veröffentlicht: (2025)
ExtendAttack: Attacking Servers of LRMs via Extending Reasoning
von: Zhu, Zhenhao, et al.
Veröffentlicht: (2025)
von: Zhu, Zhenhao, et al.
Veröffentlicht: (2025)
Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models
von: Ma, Jiachen, et al.
Veröffentlicht: (2024)
von: Ma, Jiachen, et al.
Veröffentlicht: (2024)
Evaluating the Defense Potential of Machine Unlearning against Membership Inference Attacks
von: Tsiolakis, Theodoros, et al.
Veröffentlicht: (2025)
von: Tsiolakis, Theodoros, et al.
Veröffentlicht: (2025)
Data Exfiltration by Compression Attack: Definition and Evaluation on Medical Image Data
von: Li, Huiyu, et al.
Veröffentlicht: (2025)
von: Li, Huiyu, et al.
Veröffentlicht: (2025)
Library-Attack: Reverse Engineering Approach for Evaluating Hardware IP Protection
von: Dasgupta, Aritra, et al.
Veröffentlicht: (2025)
von: Dasgupta, Aritra, et al.
Veröffentlicht: (2025)
CIBER: A Comprehensive Benchmark for Security Evaluation of Code Interpreter Agents
von: Ba, Lei, et al.
Veröffentlicht: (2026)
von: Ba, Lei, et al.
Veröffentlicht: (2026)
Defending Against Neural Network Model Inversion Attacks via Data Poisoning
von: Zhou, Shuai, et al.
Veröffentlicht: (2024)
von: Zhou, Shuai, et al.
Veröffentlicht: (2024)
WebGPU-SPY: Finding Fingerprints in the Sandbox through GPU Cache Attacks
von: Ferguson, Ethan, et al.
Veröffentlicht: (2024)
von: Ferguson, Ethan, et al.
Veröffentlicht: (2024)
Shrinking the Kernel Attack Surface Through Static and Dynamic Syscall Limitation
von: Zhan, Dongyang, et al.
Veröffentlicht: (2025)
von: Zhan, Dongyang, et al.
Veröffentlicht: (2025)
Membership Inference Attacks and Defenses in Federated Learning: A Survey
von: Bai, Li, et al.
Veröffentlicht: (2024)
von: Bai, Li, et al.
Veröffentlicht: (2024)
A Sample-Level Evaluation and Generative Framework for Model Inversion Attacks
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
RLCracker: Evaluating the Worst-Case Vulnerability of LLM Watermarks with Adaptive RL Attacks
von: Huang, Hanbo, et al.
Veröffentlicht: (2025)
von: Huang, Hanbo, et al.
Veröffentlicht: (2025)
Finding Missing Input Validation in TEEs via LLM-Assisted Symbolic Execution
von: Ma, Chengyan, et al.
Veröffentlicht: (2026)
von: Ma, Chengyan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
GNNBleed: Inference Attacks to Unveil Private Edges in Graphs with Realistic Access to GNN Models
von: Song, Zeyu, et al.
Veröffentlicht: (2023) -
Baiting AI: Deceptive Adversary Against AI-Protected Industrial Infrastructures
von: Pasikhani, Aryan, et al.
Veröffentlicht: (2026) -
Enhancing Adversarial Example Detection Through Model Explanation
von: Ma, Qian, et al.
Veröffentlicht: (2025) -
I Stolenly Swear That I Am Up to (No) Good: Design and Evaluation of Model Stealing Attacks
von: Oliynyk, Daryna, et al.
Veröffentlicht: (2025) -
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage
von: Rashid, Md Rafi Ur, et al.
Veröffentlicht: (2024)