MergePrint: Merge-Resistant Fingerprints for Robust Black-box Ownership Verification of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yamabe, Shojiro, Waseda, Futa, Takahashi, Tsubasa, Wataoka, Koki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Understanding Sensitivity of Differential Attention through the Lens of Adversarial Robustness
von: Takahashi, Tsubasa, et al.
Veröffentlicht: (2025)
von: Takahashi, Tsubasa, et al.
Veröffentlicht: (2025)
Text-Printed Image: Bridging the Image-Text Modality Gap for Text-centric Training of Large Vision-Language Models
von: Yamabe, Shojiro, et al.
Veröffentlicht: (2025)
von: Yamabe, Shojiro, et al.
Veröffentlicht: (2025)
FIT-Print: Towards False-claim-resistant Model Ownership Verification via Targeted Fingerprint
von: Shao, Shuo, et al.
Veröffentlicht: (2025)
von: Shao, Shuo, et al.
Veröffentlicht: (2025)
BadMerging: Backdoor Attacks Against Model Merging
von: Zhang, Jinghuai, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghuai, et al.
Veröffentlicht: (2024)
Targeted Data Poisoning for Black-Box Audio Datasets Ownership Verification
von: Bouaziz, Wassim, et al.
Veröffentlicht: (2025)
von: Bouaziz, Wassim, et al.
Veröffentlicht: (2025)
Read or Ignore? A Unified Benchmark for Typographic-Attack Robustness and Text Recognition in Vision-Language Models
von: Waseda, Futa, et al.
Veröffentlicht: (2025)
von: Waseda, Futa, et al.
Veröffentlicht: (2025)
Scalable Fingerprinting of Large Language Models
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)
Differentially Private Model Merging
von: Yin, Qichuan, et al.
Veröffentlicht: (2026)
von: Yin, Qichuan, et al.
Veröffentlicht: (2026)
Fingerprinting Inference Systems of Large Language Models
von: Wimbauer, Anna, et al.
Veröffentlicht: (2026)
von: Wimbauer, Anna, et al.
Veröffentlicht: (2026)
RouteMark: A Fingerprint for Intellectual Property Attribution in Routing-based Model Merging
von: He, Xin, et al.
Veröffentlicht: (2025)
von: He, Xin, et al.
Veröffentlicht: (2025)
Mitigating the Backdoor Effect for Multi-Task Model Merging via Safety-Aware Subspace
von: Yang, Jinluan, et al.
Veröffentlicht: (2024)
von: Yang, Jinluan, et al.
Veröffentlicht: (2024)
Data Taggants: Dataset Ownership Verification via Harmless Targeted Data Poisoning
von: Bouaziz, Wassim, et al.
Veröffentlicht: (2024)
von: Bouaziz, Wassim, et al.
Veröffentlicht: (2024)
AttnDiff: Attention-based Differential Fingerprinting for Large Language Models
von: Zhang, Haobo, et al.
Veröffentlicht: (2026)
von: Zhang, Haobo, et al.
Veröffentlicht: (2026)
Merge Hijacking: Backdoor Attacks to Model Merging of Large Language Models
von: Yuan, Zenghui, et al.
Veröffentlicht: (2025)
von: Yuan, Zenghui, et al.
Veröffentlicht: (2025)
FedTracker: Furnishing Ownership Verification and Traceability for Federated Learning Model
von: Shao, Shuo, et al.
Veröffentlicht: (2022)
von: Shao, Shuo, et al.
Veröffentlicht: (2022)
Traceable Black-box Watermarks for Federated Learning
von: Xu, Jiahao, et al.
Veröffentlicht: (2025)
von: Xu, Jiahao, et al.
Veröffentlicht: (2025)
ProFLingo: A Fingerprinting-based Intellectual Property Protection Scheme for Large Language Models
von: Jin, Heng, et al.
Veröffentlicht: (2024)
von: Jin, Heng, et al.
Veröffentlicht: (2024)
LoBAM: LoRA-Based Backdoor Attack on Model Merging
von: Yin, Ming, et al.
Veröffentlicht: (2024)
von: Yin, Ming, et al.
Veröffentlicht: (2024)
Evading Black-box Classifiers Without Breaking Eggs
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2023)
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2023)
Black-box Optimization of LLM Outputs by Asking for Directions
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
Disrupting Model Merging: A Parameter-Level Defense Without Sacrificing Accuracy
von: Junhao, Wei, et al.
Veröffentlicht: (2025)
von: Junhao, Wei, et al.
Veröffentlicht: (2025)
Black-box Adversarial Transferability: An Empirical Study in Cybersecurity Perspective
von: Roshan, Khushnaseeb, et al.
Veröffentlicht: (2024)
von: Roshan, Khushnaseeb, et al.
Veröffentlicht: (2024)
Dynamic Black-box Backdoor Attacks on IoT Sensory Data
von: Chathoth, Ajesh Koyatan, et al.
Veröffentlicht: (2025)
von: Chathoth, Ajesh Koyatan, et al.
Veröffentlicht: (2025)
The Challenge of Identifying the Origin of Black-Box Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
SEA: Shareable and Explainable Attribution for Query-based Black-box Attacks
von: Gao, Yue, et al.
Veröffentlicht: (2023)
von: Gao, Yue, et al.
Veröffentlicht: (2023)
LoRAGuard: An Effective Black-box Watermarking Approach for LoRAs
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025)
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025)
Instructional Fingerprinting of Large Language Models
von: Xu, Jiashu, et al.
Veröffentlicht: (2024)
von: Xu, Jiashu, et al.
Veröffentlicht: (2024)
Are Robust LLM Fingerprints Adversarially Robust?
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)
Online Poisoning Attack Against Reinforcement Learning under Black-box Environments
von: Li, Jianhui, et al.
Veröffentlicht: (2024)
von: Li, Jianhui, et al.
Veröffentlicht: (2024)
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
von: Chao, Patrick, et al.
Veröffentlicht: (2024)
von: Chao, Patrick, et al.
Veröffentlicht: (2024)
EvadeDroid: A Practical Evasion Attack on Machine Learning for Black-box Android Malware Detection
von: Bostani, Hamid, et al.
Veröffentlicht: (2021)
von: Bostani, Hamid, et al.
Veröffentlicht: (2021)
Queries, Representation & Detection: The Next 100 Model Fingerprinting Schemes
von: Godinot, Augustin, et al.
Veröffentlicht: (2024)
von: Godinot, Augustin, et al.
Veröffentlicht: (2024)
Practicable Black-box Evasion Attacks on Link Prediction in Dynamic Graphs -- A Graph Sequential Embedding Method
von: Li, Jiate, et al.
Veröffentlicht: (2024)
von: Li, Jiate, et al.
Veröffentlicht: (2024)
Multi-granular Adversarial Attacks against Black-box Neural Ranking Models
von: Liu, Yu-An, et al.
Veröffentlicht: (2024)
von: Liu, Yu-An, et al.
Veröffentlicht: (2024)
Black-Box Detection of Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
DeepCore: Simple Fingerprint Construction for Differentiating Homologous and Piracy Models
von: Sun, Haifeng, et al.
Veröffentlicht: (2024)
von: Sun, Haifeng, et al.
Veröffentlicht: (2024)
iSeal: Encrypted Fingerprinting for Reliable LLM Ownership Verification
von: Xiong, Zixun, et al.
Veröffentlicht: (2025)
von: Xiong, Zixun, et al.
Veröffentlicht: (2025)
DP-TRAE: A Dual-Phase Merging Transferable Reversible Adversarial Example for Image Privacy Protection
von: Du, Xia, et al.
Veröffentlicht: (2025)
von: Du, Xia, et al.
Veröffentlicht: (2025)
Have You Merged My Model? On The Robustness of Large Language Model IP Protection Methods Against Model Merging
von: Cong, Tianshuo, et al.
Veröffentlicht: (2024)
von: Cong, Tianshuo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Understanding Sensitivity of Differential Attention through the Lens of Adversarial Robustness
von: Takahashi, Tsubasa, et al.
Veröffentlicht: (2025) -
Text-Printed Image: Bridging the Image-Text Modality Gap for Text-centric Training of Large Vision-Language Models
von: Yamabe, Shojiro, et al.
Veröffentlicht: (2025) -
FIT-Print: Towards False-claim-resistant Model Ownership Verification via Targeted Fingerprint
von: Shao, Shuo, et al.
Veröffentlicht: (2025) -
BadMerging: Backdoor Attacks Against Model Merging
von: Zhang, Jinghuai, et al.
Veröffentlicht: (2024) -
Targeted Data Poisoning for Black-Box Audio Datasets Ownership Verification
von: Bouaziz, Wassim, et al.
Veröffentlicht: (2025)