Intrinsic Fingerprint of LLMs: Continue Training is NOT All You Need to Steal A Model!
Fuente:
arXiv
Salvato in:
| Autori principali: | Yoon, Do-hyeon, Chun, Minsoo, Allen, Thomas, Müller, Hans, Wang, Min, Sharma, Rajesh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
You Can't Steal Nothing: Mitigating Prompt Leakages in LLMs via System Vectors
di: Cao, Bochuan, et al.
Pubblicazione: (2025)
di: Cao, Bochuan, et al.
Pubblicazione: (2025)
GraphSteal: Structural Knowledge Stealing from Graph RAG via Traversal Reconstruction
di: Gu, Jinze, et al.
Pubblicazione: (2026)
di: Gu, Jinze, et al.
Pubblicazione: (2026)
Prompt Stealing Attacks Against Large Language Models
di: Sha, Zeyang, et al.
Pubblicazione: (2024)
di: Sha, Zeyang, et al.
Pubblicazione: (2024)
Fingerprinting LLMs via Prompt Injection
di: Hu, Yuepeng, et al.
Pubblicazione: (2025)
di: Hu, Yuepeng, et al.
Pubblicazione: (2025)
Reformulation is All You Need: Addressing Malicious Text Features in DNNs
di: Jiang, Yi, et al.
Pubblicazione: (2025)
di: Jiang, Yi, et al.
Pubblicazione: (2025)
SecEncoder: Logs are All You Need in Security
di: Bulut, Muhammed Fatih, et al.
Pubblicazione: (2024)
di: Bulut, Muhammed Fatih, et al.
Pubblicazione: (2024)
Prompt Pirates Need a Map: Stealing Seeds helps Stealing Prompts
di: Mächtle, Felix, et al.
Pubblicazione: (2025)
di: Mächtle, Felix, et al.
Pubblicazione: (2025)
Teach LLMs to Phish: Stealing Private Information from Language Models
di: Panda, Ashwinee, et al.
Pubblicazione: (2024)
di: Panda, Ashwinee, et al.
Pubblicazione: (2024)
Security Steerability is All You Need
di: Hazan, Itay, et al.
Pubblicazione: (2025)
di: Hazan, Itay, et al.
Pubblicazione: (2025)
PRSA: Prompt Stealing Attacks against Real-World Prompt Services
di: Yang, Yong, et al.
Pubblicazione: (2024)
di: Yang, Yong, et al.
Pubblicazione: (2024)
KinGuard: Hierarchical Kinship-Aware Fingerprinting to Defend Against Large Language Model Stealing
di: Xu, Zhenhua, et al.
Pubblicazione: (2026)
di: Xu, Zhenhua, et al.
Pubblicazione: (2026)
Attention is All You Need to Defend Against Indirect Prompt Injection Attacks in LLMs
di: Zhong, Yinan, et al.
Pubblicazione: (2025)
di: Zhong, Yinan, et al.
Pubblicazione: (2025)
One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs
di: Li, Linbao, et al.
Pubblicazione: (2025)
di: Li, Linbao, et al.
Pubblicazione: (2025)
Implicit Identity Technologies for LLMs: Fingerprinting and Watermarking across Datasets, Models, and Generated Content
di: Liu, Bing, et al.
Pubblicazione: (2026)
di: Liu, Bing, et al.
Pubblicazione: (2026)
Layerwise Convergence Fingerprints for Runtime Misbehavior Detection in Large Language Models
di: Min, Nay Myat, et al.
Pubblicazione: (2026)
di: Min, Nay Myat, et al.
Pubblicazione: (2026)
All You Need is "Leet": Evading Hate-speech Detection AI
di: Kahu, Sampanna Yashwant, et al.
Pubblicazione: (2025)
di: Kahu, Sampanna Yashwant, et al.
Pubblicazione: (2025)
Dynamic Graph-based Fingerprinting of In-browser Cryptomining
di: Sermchaiwong, Tanapoom, et al.
Pubblicazione: (2025)
di: Sermchaiwong, Tanapoom, et al.
Pubblicazione: (2025)
Indirect Prompt Injections: Are Firewalls All You Need, or Stronger Benchmarks?
di: Bhagwatkar, Rishika, et al.
Pubblicazione: (2025)
di: Bhagwatkar, Rishika, et al.
Pubblicazione: (2025)
Stealing Training Data from Large Language Models in Decentralized Training through Activation Inversion Attack
di: Dai, Chenxi, et al.
Pubblicazione: (2025)
di: Dai, Chenxi, et al.
Pubblicazione: (2025)
Merger-as-a-Stealer: Stealing Targeted PII from Aligned LLMs with Model Merging
di: Lu, Lin, et al.
Pubblicazione: (2025)
di: Lu, Lin, et al.
Pubblicazione: (2025)
Fingerprint Vector: Enabling Scalable and Efficient Model Fingerprint Transfer via Vector Addition
di: Xu, Zhenhua, et al.
Pubblicazione: (2024)
di: Xu, Zhenhua, et al.
Pubblicazione: (2024)
Stealing Part of a Production Language Model
di: Carlini, Nicholas, et al.
Pubblicazione: (2024)
di: Carlini, Nicholas, et al.
Pubblicazione: (2024)
DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection
di: Yan, Yuliang, et al.
Pubblicazione: (2025)
di: Yan, Yuliang, et al.
Pubblicazione: (2025)
BarkBeetle: Stealing Decision Tree Models with Fault Injection
di: Wang, Qifan, et al.
Pubblicazione: (2025)
di: Wang, Qifan, et al.
Pubblicazione: (2025)
SeedPrints: Fingerprints Can Even Tell Which Seed Your Large Language Model Was Trained From
di: Tong, Yao, et al.
Pubblicazione: (2025)
di: Tong, Yao, et al.
Pubblicazione: (2025)
Continual Pretraining on Encrypted Synthetic Data for Privacy-Preserving LLMs
di: Liu, Honghao, et al.
Pubblicazione: (2026)
di: Liu, Honghao, et al.
Pubblicazione: (2026)
All Your Knowledge Belongs to Us: Stealing Knowledge Graphs via Reasoning APIs
di: Xi, Zhaohan
Pubblicazione: (2025)
di: Xi, Zhaohan
Pubblicazione: (2025)
Ingest-And-Ground: Dispelling Hallucinations from Continually-Pretrained LLMs with RAG
di: Fang, Chenhao, et al.
Pubblicazione: (2024)
di: Fang, Chenhao, et al.
Pubblicazione: (2024)
Stealing User Prompts from Mixture of Experts
di: Yona, Itay, et al.
Pubblicazione: (2024)
di: Yona, Itay, et al.
Pubblicazione: (2024)
Transpose Attack: Stealing Datasets with Bidirectional Training
di: Amit, Guy, et al.
Pubblicazione: (2023)
di: Amit, Guy, et al.
Pubblicazione: (2023)
REEF: Representation Encoding Fingerprints for Large Language Models
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
FNF: Functional Network Fingerprint for Large Language Models
di: Liu, Yiheng, et al.
Pubblicazione: (2026)
di: Liu, Yiheng, et al.
Pubblicazione: (2026)
TRUCE: Private Benchmarking to Prevent Contamination and Improve Comparative Evaluation of LLMs
di: Rajore, Tanmay, et al.
Pubblicazione: (2024)
di: Rajore, Tanmay, et al.
Pubblicazione: (2024)
Efficient Data-Free Model Stealing with Label Diversity
di: Liu, Yiyong, et al.
Pubblicazione: (2024)
di: Liu, Yiyong, et al.
Pubblicazione: (2024)
The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training
di: Zhang, Rui, et al.
Pubblicazione: (2026)
di: Zhang, Rui, et al.
Pubblicazione: (2026)
You Can Run But You Can't Hide: Runtime Protection Against Malicious Package Updates For Node.js
di: Ohm, Marc, et al.
Pubblicazione: (2023)
di: Ohm, Marc, et al.
Pubblicazione: (2023)
Stealing Training Graphs from Graph Neural Networks
di: Lin, Minhua, et al.
Pubblicazione: (2024)
di: Lin, Minhua, et al.
Pubblicazione: (2024)
SoK: Large Language Model Copyright Auditing via Fingerprinting
di: Shao, Shuo, et al.
Pubblicazione: (2025)
di: Shao, Shuo, et al.
Pubblicazione: (2025)
Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in LLMs
di: Sekar, Anirudh, et al.
Pubblicazione: (2026)
di: Sekar, Anirudh, et al.
Pubblicazione: (2026)
All You Need Is A Fuzzing Brain: An LLM-Powered System for Automated Vulnerability Detection and Patching
di: Sheng, Ze, et al.
Pubblicazione: (2025)
di: Sheng, Ze, et al.
Pubblicazione: (2025)
Documenti analoghi
-
You Can't Steal Nothing: Mitigating Prompt Leakages in LLMs via System Vectors
di: Cao, Bochuan, et al.
Pubblicazione: (2025) -
GraphSteal: Structural Knowledge Stealing from Graph RAG via Traversal Reconstruction
di: Gu, Jinze, et al.
Pubblicazione: (2026) -
Prompt Stealing Attacks Against Large Language Models
di: Sha, Zeyang, et al.
Pubblicazione: (2024) -
Fingerprinting LLMs via Prompt Injection
di: Hu, Yuepeng, et al.
Pubblicazione: (2025) -
Reformulation is All You Need: Addressing Malicious Text Features in DNNs
di: Jiang, Yi, et al.
Pubblicazione: (2025)