DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yan, Yuliang, Tang, Haochun, Yan, Shuo, Dai, Enyan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimization
von: Tang, Haochun, et al.
Veröffentlicht: (2026)
von: Tang, Haochun, et al.
Veröffentlicht: (2026)
Have You Merged My Model? On The Robustness of Large Language Model IP Protection Methods Against Model Merging
von: Cong, Tianshuo, et al.
Veröffentlicht: (2024)
von: Cong, Tianshuo, et al.
Veröffentlicht: (2024)
SoK: Large Language Model Copyright Auditing via Fingerprinting
von: Shao, Shuo, et al.
Veröffentlicht: (2025)
von: Shao, Shuo, et al.
Veröffentlicht: (2025)
Protecting Your LLMs with Information Bottleneck
von: Liu, Zichuan, et al.
Veröffentlicht: (2024)
von: Liu, Zichuan, et al.
Veröffentlicht: (2024)
Reading Between the Lines: Towards Reliable Black-box LLM Fingerprinting via Zeroth-order Gradient Estimation
von: Shao, Shuo, et al.
Veröffentlicht: (2025)
von: Shao, Shuo, et al.
Veröffentlicht: (2025)
Feint and Attack: Attention-Based Strategies for Jailbreaking and Protecting LLMs
von: Pu, Rui, et al.
Veröffentlicht: (2024)
von: Pu, Rui, et al.
Veröffentlicht: (2024)
IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
von: Wang, Liwen, et al.
Veröffentlicht: (2025)
von: Wang, Liwen, et al.
Veröffentlicht: (2025)
Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs
von: Yan, Dong, et al.
Veröffentlicht: (2026)
von: Yan, Dong, et al.
Veröffentlicht: (2026)
HARMONIC: Harnessing LLMs for Tabular Data Synthesis and Privacy Protection
von: Wang, Yuxin, et al.
Veröffentlicht: (2024)
von: Wang, Yuxin, et al.
Veröffentlicht: (2024)
REEF: Representation Encoding Fingerprints for Large Language Models
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
FNF: Functional Network Fingerprint for Large Language Models
von: Liu, Yiheng, et al.
Veröffentlicht: (2026)
von: Liu, Yiheng, et al.
Veröffentlicht: (2026)
Dagger Behind Smile: Fool LLMs with a Happy Ending Story
von: Song, Xurui, et al.
Veröffentlicht: (2025)
von: Song, Xurui, et al.
Veröffentlicht: (2025)
SAMark: A Self-Anchored Text Watermarking with Paragraph-Level Paraphrase Robustness
von: Huo, Jiahao, et al.
Veröffentlicht: (2026)
von: Huo, Jiahao, et al.
Veröffentlicht: (2026)
Layerwise Convergence Fingerprints for Runtime Misbehavior Detection in Large Language Models
von: Min, Nay Myat, et al.
Veröffentlicht: (2026)
von: Min, Nay Myat, et al.
Veröffentlicht: (2026)
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters
von: Yang, Yan, et al.
Veröffentlicht: (2024)
von: Yang, Yan, et al.
Veröffentlicht: (2024)
Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia
von: Shen, Guangyu, et al.
Veröffentlicht: (2024)
von: Shen, Guangyu, et al.
Veröffentlicht: (2024)
Topology Matters: Measuring Memory Leakage in Multi-Agent LLMs
von: Liu, Jinbo, et al.
Veröffentlicht: (2025)
von: Liu, Jinbo, et al.
Veröffentlicht: (2025)
Waterfall: Framework for Robust and Scalable Text Watermarking and Provenance for LLMs
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)
SeedPrints: Fingerprints Can Even Tell Which Seed Your Large Language Model Was Trained From
von: Tong, Yao, et al.
Veröffentlicht: (2025)
von: Tong, Yao, et al.
Veröffentlicht: (2025)
DMFI: A Dual-Modality Log Analysis Framework for Insider Threat Detection with LoRA-Tuned Language Models
von: Kong, Kaichuan, et al.
Veröffentlicht: (2025)
von: Kong, Kaichuan, et al.
Veröffentlicht: (2025)
Cognitive Control Architecture (CCA): A Lifecycle Supervision Framework for Robustly Aligned AI Agents
von: Liang, Zhibo, et al.
Veröffentlicht: (2025)
von: Liang, Zhibo, et al.
Veröffentlicht: (2025)
LLMs Have Rhythm: Fingerprinting Large Language Models Using Inter-Token Times and Network Traffic Analysis
von: Alhazbi, Saeif, et al.
Veröffentlicht: (2025)
von: Alhazbi, Saeif, et al.
Veröffentlicht: (2025)
Prompt2Fingerprint: Plug-and-Play LLM Fingerprinting via Text-to-Weight Generation
von: Chen, Sixu, et al.
Veröffentlicht: (2026)
von: Chen, Sixu, et al.
Veröffentlicht: (2026)
AdaPPA: Adaptive Position Pre-Fill Jailbreak Attack Approach Targeting LLMs
von: Lv, Lijia, et al.
Veröffentlicht: (2024)
von: Lv, Lijia, et al.
Veröffentlicht: (2024)
Instructional Fingerprinting of Large Language Models
von: Xu, Jiashu, et al.
Veröffentlicht: (2024)
von: Xu, Jiashu, et al.
Veröffentlicht: (2024)
Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs
von: Kim, Jinhwa, et al.
Veröffentlicht: (2025)
von: Kim, Jinhwa, et al.
Veröffentlicht: (2025)
Pruning for Protection: Increasing Jailbreak Resistance in Aligned LLMs Without Fine-Tuning
von: Hasan, Adib, et al.
Veröffentlicht: (2024)
von: Hasan, Adib, et al.
Veröffentlicht: (2024)
Protecting Users From Themselves: Safeguarding Contextual Privacy in Interactions with Conversational Agents
von: Ngong, Ivoline, et al.
Veröffentlicht: (2025)
von: Ngong, Ivoline, et al.
Veröffentlicht: (2025)
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs
von: Liu, Songyang, et al.
Veröffentlicht: (2025)
von: Liu, Songyang, et al.
Veröffentlicht: (2025)
DNF: Dual-Layer Nested Fingerprinting for Large Language Model Intellectual Property Protection
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026)
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026)
Knowledge-to-Jailbreak: Investigating Knowledge-driven Jailbreaking Attacks for Large Language Models
von: Tu, Shangqing, et al.
Veröffentlicht: (2024)
von: Tu, Shangqing, et al.
Veröffentlicht: (2024)
Attention Slipping: A Mechanistic Understanding of Jailbreak Attacks and Defenses in LLMs
von: Hu, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Hu, Xiaomeng, et al.
Veröffentlicht: (2025)
ShieldLearner: A New Paradigm for Jailbreak Attack Defense in LLMs
von: Ni, Ziyi, et al.
Veröffentlicht: (2025)
von: Ni, Ziyi, et al.
Veröffentlicht: (2025)
Defend LLMs Through Self-Consciousness
von: Huang, Boshi, et al.
Veröffentlicht: (2025)
von: Huang, Boshi, et al.
Veröffentlicht: (2025)
SELF: A Robust Singular Value and Eigenvalue Approach for LLM Fingerprinting
von: Zhang, Hanxiu, et al.
Veröffentlicht: (2025)
von: Zhang, Hanxiu, et al.
Veröffentlicht: (2025)
AgentAlign: Navigating Safety Alignment in the Shift from Informative to Agentic Large Language Models
von: Zhang, Jinchuan, et al.
Veröffentlicht: (2025)
von: Zhang, Jinchuan, et al.
Veröffentlicht: (2025)
Distract Large Language Models for Automatic Jailbreak Attack
von: Xiao, Zeguan, et al.
Veröffentlicht: (2024)
von: Xiao, Zeguan, et al.
Veröffentlicht: (2024)
From Vulnerabilities to Remediation: A Systematic Literature Review of LLMs in Code Security
von: Basic, Enna, et al.
Veröffentlicht: (2024)
von: Basic, Enna, et al.
Veröffentlicht: (2024)
Supporting Artifact Evaluation with LLMs: A Study with Published Security Research Papers
von: Heye, David, et al.
Veröffentlicht: (2026)
von: Heye, David, et al.
Veröffentlicht: (2026)
Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs
von: Xu, Zhao, et al.
Veröffentlicht: (2024)
von: Xu, Zhao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimization
von: Tang, Haochun, et al.
Veröffentlicht: (2026) -
Have You Merged My Model? On The Robustness of Large Language Model IP Protection Methods Against Model Merging
von: Cong, Tianshuo, et al.
Veröffentlicht: (2024) -
SoK: Large Language Model Copyright Auditing via Fingerprinting
von: Shao, Shuo, et al.
Veröffentlicht: (2025) -
Protecting Your LLMs with Information Bottleneck
von: Liu, Zichuan, et al.
Veröffentlicht: (2024) -
Reading Between the Lines: Towards Reliable Black-box LLM Fingerprinting via Zeroth-order Gradient Estimation
von: Shao, Shuo, et al.
Veröffentlicht: (2025)