ModelLock: Locking Your Model With a Spell
Fuente:
arXiv
Salvato in:
| Autori principali: | Gao, Yifeng, Sun, Yuhua, Ma, Xingjun, Wu, Zuxuan, Jiang, Yu-Gang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Identity Lock: Locking API Fine-tuned LLMs With Identity-based Wake Words
di: Su, Hongyu, et al.
Pubblicazione: (2025)
di: Su, Hongyu, et al.
Pubblicazione: (2025)
Locking Machine Learning Models into Hardware
di: Clifford, Eleanor, et al.
Pubblicazione: (2024)
di: Clifford, Eleanor, et al.
Pubblicazione: (2024)
FedEGG: Federated Learning with Explicit Global Guidance
di: Zhai, Kun, et al.
Pubblicazione: (2024)
di: Zhai, Kun, et al.
Pubblicazione: (2024)
Stress-Testing Capability Elicitation With Password-Locked Models
di: Greenblatt, Ryan, et al.
Pubblicazione: (2024)
di: Greenblatt, Ryan, et al.
Pubblicazione: (2024)
Locket: Robust Feature-Locking Technique for Language Models
di: He, Lipeng, et al.
Pubblicazione: (2025)
di: He, Lipeng, et al.
Pubblicazione: (2025)
Extracting Training Data from Unconditional Diffusion Models
di: Chen, Yunhao, et al.
Pubblicazione: (2024)
di: Chen, Yunhao, et al.
Pubblicazione: (2024)
Steering the Verifiability of Multimodal AI Hallucinations
di: Pang, Jianhong, et al.
Pubblicazione: (2026)
di: Pang, Jianhong, et al.
Pubblicazione: (2026)
AIM: Additional Image Guided Generation of Transferable Adversarial Attacks
di: Li, Teng, et al.
Pubblicazione: (2025)
di: Li, Teng, et al.
Pubblicazione: (2025)
Thinking with Deltas: Incentivizing Reinforcement Learning via Differential Visual Reasoning Policy
di: Gao, Shujian, et al.
Pubblicazione: (2026)
di: Gao, Shujian, et al.
Pubblicazione: (2026)
ForgerySleuth: Empowering Multimodal Large Language Models for Image Manipulation Detection
di: Sun, Zhihao, et al.
Pubblicazione: (2024)
di: Sun, Zhihao, et al.
Pubblicazione: (2024)
DarkLLM: Learning Language-Driven Adversarial Attacks with Large Language Models
di: Sun, Ye, et al.
Pubblicazione: (2026)
di: Sun, Ye, et al.
Pubblicazione: (2026)
EnJa: Ensemble Jailbreak on Large Language Models
di: Zhang, Jiahao, et al.
Pubblicazione: (2024)
di: Zhang, Jiahao, et al.
Pubblicazione: (2024)
Spelling Bee Embeddings for Language Modeling
di: Rabe, Markus N., et al.
Pubblicazione: (2026)
di: Rabe, Markus N., et al.
Pubblicazione: (2026)
DistilLock: Safeguarding LLMs from Unauthorized Knowledge Distillation on the Edge
di: Mohanty, Asmita, et al.
Pubblicazione: (2025)
di: Mohanty, Asmita, et al.
Pubblicazione: (2025)
Adaptive Retention & Correction: Test-Time Training for Continual Learning
di: Chen, Haoran, et al.
Pubblicazione: (2024)
di: Chen, Haoran, et al.
Pubblicazione: (2024)
LLA: Enhancing Security and Privacy for Generative Models with Logic-Locked Accelerators
di: Li, You, et al.
Pubblicazione: (2025)
di: Li, You, et al.
Pubblicazione: (2025)
Shackled Dancing: A Bit-Locked Diffusion Algorithm for Lossless and Controllable Image Steganography
di: Zhang, Tianshuo, et al.
Pubblicazione: (2025)
di: Zhang, Tianshuo, et al.
Pubblicazione: (2025)
AID: Adapting Image2Video Diffusion Models for Instruction-guided Video Prediction
di: Xing, Zhen, et al.
Pubblicazione: (2024)
di: Xing, Zhen, et al.
Pubblicazione: (2024)
Shortcuts Everywhere and Nowhere: Exploring Multi-Trigger Backdoor Attacks
di: Li, Yige, et al.
Pubblicazione: (2024)
di: Li, Yige, et al.
Pubblicazione: (2024)
UnSeg: One Universal Unlearnable Example Generator is Enough against All Image Segmentation
di: Sun, Ye, et al.
Pubblicazione: (2024)
di: Sun, Ye, et al.
Pubblicazione: (2024)
Deep-Lock: Secure Authorization for Deep Neural Networks
di: Alam, Manaar, et al.
Pubblicazione: (2020)
di: Alam, Manaar, et al.
Pubblicazione: (2020)
An Imitative Reinforcement Learning Framework for Pursuit-Lock-Launch Missions
di: Li, Siyuan, et al.
Pubblicazione: (2024)
di: Li, Siyuan, et al.
Pubblicazione: (2024)
BlueSuffix: Reinforced Blue Teaming for Vision-Language Models Against Jailbreak Attacks
di: Zhao, Yunhan, et al.
Pubblicazione: (2024)
di: Zhao, Yunhan, et al.
Pubblicazione: (2024)
Locking Pretrained Weights via Deep Low-Rank Residual Distillation
di: Sakamoto, Keitaro, et al.
Pubblicazione: (2026)
di: Sakamoto, Keitaro, et al.
Pubblicazione: (2026)
The Lock-in Hypothesis: Stagnation by Algorithm
di: Qiu, Tianyi Alex, et al.
Pubblicazione: (2025)
di: Qiu, Tianyi Alex, et al.
Pubblicazione: (2025)
Breaking Model Lock-in: Cost-Efficient Zero-Shot LLM Routing via a Universal Latent Space
di: Yan, Cheng, et al.
Pubblicazione: (2026)
di: Yan, Cheng, et al.
Pubblicazione: (2026)
Causal-Transformer with Adaptive Mutation-Locking for Early Prediction of Acute Kidney Injury
di: Nie, Weizhi, et al.
Pubblicazione: (2026)
di: Nie, Weizhi, et al.
Pubblicazione: (2026)
TiSpell: A Semi-Masked Methodology for Tibetan Spelling Correction covering Multi-Level Error with Data Augmentation
di: Liu, Yutong, et al.
Pubblicazione: (2025)
di: Liu, Yutong, et al.
Pubblicazione: (2025)
WildDeepfake: A Challenging Real-World Dataset for Deepfake Detection
di: Zi, Bojia, et al.
Pubblicazione: (2021)
di: Zi, Bojia, et al.
Pubblicazione: (2021)
On the Limits of Latent Reuse in Diffusion Models
di: Yu, Yifeng, et al.
Pubblicazione: (2026)
di: Yu, Yifeng, et al.
Pubblicazione: (2026)
A Survey on Video Diffusion Models
di: Xing, Zhen, et al.
Pubblicazione: (2023)
di: Xing, Zhen, et al.
Pubblicazione: (2023)
Downstream Transfer Attack: Adversarial Attacks on Downstream Models with Pre-trained Vision Transformers
di: Zheng, Weijie, et al.
Pubblicazione: (2024)
di: Zheng, Weijie, et al.
Pubblicazione: (2024)
RedTopic: Toward Topic-Diverse Red Teaming of Large Language Models
di: Ding, Jiale, et al.
Pubblicazione: (2025)
di: Ding, Jiale, et al.
Pubblicazione: (2025)
Phase-Locked SNR Band Selection for Weak Mineral Signal Detection in Hyperspectral Imagery
di: Yang, Judy X
Pubblicazione: (2025)
di: Yang, Judy X
Pubblicazione: (2025)
Feature Repulsion and Spectral Lock-in: An Empirical Study of Two-Layer Network Grokking
di: Xu, Yongzhong
Pubblicazione: (2026)
di: Xu, Yongzhong
Pubblicazione: (2026)
Robust Load Prediction of Power Network Clusters Based on Cloud-Model-Improved Transformer
di: Jiang, Cheng, et al.
Pubblicazione: (2024)
di: Jiang, Cheng, et al.
Pubblicazione: (2024)
Advancing Wasserstein Convergence Analysis of Score-Based Models: Insights from Discretization and Second-Order Acceleration
di: Yu, Yifeng, et al.
Pubblicazione: (2025)
di: Yu, Yifeng, et al.
Pubblicazione: (2025)
From Order to Distribution: A Spectral Characterization of Forgetting in Continual Learning
di: Xu, Zonghuan, et al.
Pubblicazione: (2026)
di: Xu, Zonghuan, et al.
Pubblicazione: (2026)
Gradient Leakage Defense with Key-Lock Module for Federated Learning
di: Ren, Hanchi, et al.
Pubblicazione: (2023)
di: Ren, Hanchi, et al.
Pubblicazione: (2023)
Sign Lock-In: Randomly Initialized Weight Signs Persist and Bottleneck Sub-Bit Model Compression
di: Sakai, Akira, et al.
Pubblicazione: (2026)
di: Sakai, Akira, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Identity Lock: Locking API Fine-tuned LLMs With Identity-based Wake Words
di: Su, Hongyu, et al.
Pubblicazione: (2025) -
Locking Machine Learning Models into Hardware
di: Clifford, Eleanor, et al.
Pubblicazione: (2024) -
FedEGG: Federated Learning with Explicit Global Guidance
di: Zhai, Kun, et al.
Pubblicazione: (2024) -
Stress-Testing Capability Elicitation With Password-Locked Models
di: Greenblatt, Ryan, et al.
Pubblicazione: (2024) -
Locket: Robust Feature-Locking Technique for Language Models
di: He, Lipeng, et al.
Pubblicazione: (2025)