Attacks and Defenses Against LLM Fingerprinting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kurian, Kevin, Holland, Ethan, Oesch, Sean |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Living Off the LLM: How LLMs Will Change Adversary Tactics
von: Oesch, Sean, et al.
Veröffentlicht: (2025)
von: Oesch, Sean, et al.
Veröffentlicht: (2025)
Optimal Defenses Against Gradient Reconstruction Attacks
von: Chen, Yuxiao, et al.
Veröffentlicht: (2024)
von: Chen, Yuxiao, et al.
Veröffentlicht: (2024)
LeakSealer: A Semisupervised Defense for LLMs Against Prompt Injection and Leakage Attacks
von: Panebianco, Francesco, et al.
Veröffentlicht: (2025)
von: Panebianco, Francesco, et al.
Veröffentlicht: (2025)
Can Adversarial Code Comments Fool AI Security Reviewers -- Large-Scale Empirical Study of Comment-Based Attacks and Defenses Against LLM Code Analysis
von: Thornton, Scott
Veröffentlicht: (2026)
von: Thornton, Scott
Veröffentlicht: (2026)
Are Robust LLM Fingerprints Adversarially Robust?
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)
Jailbreak Attacks and Defenses Against Large Language Models: A Survey
von: Yi, Sibo, et al.
Veröffentlicht: (2024)
von: Yi, Sibo, et al.
Veröffentlicht: (2024)
FLARE: A Wireless Side-Channel Fingerprinting Attack on Federated Learning
von: Shuvo, Md Nahid Hasan, et al.
Veröffentlicht: (2025)
von: Shuvo, Md Nahid Hasan, et al.
Veröffentlicht: (2025)
A Causal Perspective for Enhancing Jailbreak Attack and Defense
von: Pan, Licheng, et al.
Veröffentlicht: (2026)
von: Pan, Licheng, et al.
Veröffentlicht: (2026)
Stealthy Poisoning Attacks Bypass Defenses in Regression Settings
von: Carnerero-Cano, Javier, et al.
Veröffentlicht: (2026)
von: Carnerero-Cano, Javier, et al.
Veröffentlicht: (2026)
Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacks
von: Halloran, John T., et al.
Veröffentlicht: (2026)
von: Halloran, John T., et al.
Veröffentlicht: (2026)
SoK: Benchmarking Poisoning Attacks and Defenses in Federated Learning
von: Zhang, Heyi, et al.
Veröffentlicht: (2025)
von: Zhang, Heyi, et al.
Veröffentlicht: (2025)
Attacks and Defenses for Generative Diffusion Models: A Comprehensive Survey
von: Truong, Vu Tuan, et al.
Veröffentlicht: (2024)
von: Truong, Vu Tuan, et al.
Veröffentlicht: (2024)
LLM Security: Vulnerabilities, Attacks, Defenses, and Countermeasures
von: Aguilera-Martínez, Francisco, et al.
Veröffentlicht: (2025)
von: Aguilera-Martínez, Francisco, et al.
Veröffentlicht: (2025)
Beyond a Single Perspective: Towards a Realistic Evaluation of Website Fingerprinting Attacks
von: Deng, Xinhao, et al.
Veröffentlicht: (2025)
von: Deng, Xinhao, et al.
Veröffentlicht: (2025)
Backdoor Vectors: a Task Arithmetic View on Backdoor Attacks and Defenses
von: Pawlak, Stanisław, et al.
Veröffentlicht: (2025)
von: Pawlak, Stanisław, et al.
Veröffentlicht: (2025)
A Survey on Model Extraction Attacks and Defenses for Large Language Models
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
A Survey of Model Extraction Attacks and Defenses in Distributed Computing Environments
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
Semantic Chameleon: Corpus-Dependent Poisoning Attacks and Defenses in RAG Systems
von: Thornton, Scott
Veröffentlicht: (2026)
von: Thornton, Scott
Veröffentlicht: (2026)
DeepStage: Learning Autonomous Defense Policies Against Multi-Stage APT Campaigns
von: Phan, Trung V., et al.
Veröffentlicht: (2026)
von: Phan, Trung V., et al.
Veröffentlicht: (2026)
Robust and Reliable Early-Stage Website Fingerprinting Attacks via Spatial-Temporal Distribution Analysis
von: Deng, Xinhao, et al.
Veröffentlicht: (2024)
von: Deng, Xinhao, et al.
Veröffentlicht: (2024)
Attacking LLMs and AI Agents: Advertisement Embedding Attacks Against Large Language Models
von: Guo, Qiming, et al.
Veröffentlicht: (2025)
von: Guo, Qiming, et al.
Veröffentlicht: (2025)
Defending the Edge: Representative-Attention Defense against Backdoor Attacks in Federated Learning
von: Obioma, Chibueze Peace, et al.
Veröffentlicht: (2025)
von: Obioma, Chibueze Peace, et al.
Veröffentlicht: (2025)
A Systematic Survey of Model Extraction Attacks and Defenses: State-of-the-Art and Perspectives
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey
von: Huang, Tiansheng, et al.
Veröffentlicht: (2024)
von: Huang, Tiansheng, et al.
Veröffentlicht: (2024)
The Autonomy Tax: Defense Training Breaks LLM Agents
von: Li, Shawn, et al.
Veröffentlicht: (2026)
von: Li, Shawn, et al.
Veröffentlicht: (2026)
One Step to the Side: Why Defenses Against Malicious Finetuning Fail Under Adaptive Adversaries
von: Zloczower, Itay, et al.
Veröffentlicht: (2026)
von: Zloczower, Itay, et al.
Veröffentlicht: (2026)
Enhancing Security in Deep Reinforcement Learning: A Comprehensive Survey on Adversarial Attacks and Defenses
von: Yichao, Wu, et al.
Veröffentlicht: (2025)
von: Yichao, Wu, et al.
Veröffentlicht: (2025)
Center-Based Relaxed Learning Against Membership Inference Attacks
von: Fang, Xingli, et al.
Veröffentlicht: (2024)
von: Fang, Xingli, et al.
Veröffentlicht: (2024)
Improved Membership Inference Attacks Against Language Classification Models
von: Shachor, Shlomit, et al.
Veröffentlicht: (2023)
von: Shachor, Shlomit, et al.
Veröffentlicht: (2023)
Seed Hijacking of LLM Sampling and Quantum Random Number Defense
von: You, Ziyang, et al.
Veröffentlicht: (2026)
von: You, Ziyang, et al.
Veröffentlicht: (2026)
A White-Box Adversarial Attack Against a Digital Twin
von: Patterson, Wilson, et al.
Veröffentlicht: (2022)
von: Patterson, Wilson, et al.
Veröffentlicht: (2022)
Enhancing IoT Security Against DDoS Attacks through Federated Learning
von: Shirvani, Ghazaleh, et al.
Veröffentlicht: (2024)
von: Shirvani, Ghazaleh, et al.
Veröffentlicht: (2024)
Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language Models
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM
von: Cao, Bochuan, et al.
Veröffentlicht: (2023)
von: Cao, Bochuan, et al.
Veröffentlicht: (2023)
Can Differentially Private Fine-tuning LLMs Protect Against Privacy Attacks?
von: Du, Hao, et al.
Veröffentlicht: (2025)
von: Du, Hao, et al.
Veröffentlicht: (2025)
Training RL Agents for Multi-Objective Network Defense Tasks
von: Molina-Markham, Andres, et al.
Veröffentlicht: (2025)
von: Molina-Markham, Andres, et al.
Veröffentlicht: (2025)
A Systematic Literature Review on LLM Defenses Against Prompt Injection and Jailbreaking: Expanding NIST Taxonomy
von: Correia, Pedro H. Barcha, et al.
Veröffentlicht: (2026)
von: Correia, Pedro H. Barcha, et al.
Veröffentlicht: (2026)
DCMI: A Differential Calibration Membership Inference Attack Against Retrieval-Augmented Generation
von: Gao, Xinyu, et al.
Veröffentlicht: (2025)
von: Gao, Xinyu, et al.
Veröffentlicht: (2025)
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
von: Liu, Yupei, et al.
Veröffentlicht: (2023)
von: Liu, Yupei, et al.
Veröffentlicht: (2023)
Prompt2Fingerprint: Plug-and-Play LLM Fingerprinting via Text-to-Weight Generation
von: Chen, Sixu, et al.
Veröffentlicht: (2026)
von: Chen, Sixu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Living Off the LLM: How LLMs Will Change Adversary Tactics
von: Oesch, Sean, et al.
Veröffentlicht: (2025) -
Optimal Defenses Against Gradient Reconstruction Attacks
von: Chen, Yuxiao, et al.
Veröffentlicht: (2024) -
LeakSealer: A Semisupervised Defense for LLMs Against Prompt Injection and Leakage Attacks
von: Panebianco, Francesco, et al.
Veröffentlicht: (2025) -
Can Adversarial Code Comments Fool AI Security Reviewers -- Large-Scale Empirical Study of Comment-Based Attacks and Defenses Against LLM Code Analysis
von: Thornton, Scott
Veröffentlicht: (2026) -
Are Robust LLM Fingerprints Adversarially Robust?
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)