Locket: Robust Feature-Locking Technique for Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | He, Lipeng, Duddu, Vasisht, Asokan, N. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Combining Machine Learning Defenses without Conflicts
por: Duddu, Vasisht, et al.
Publicado: (2024)
por: Duddu, Vasisht, et al.
Publicado: (2024)
SoK: Unintended Interactions among Machine Learning Defenses and Risks
por: Duddu, Vasisht, et al.
Publicado: (2023)
por: Duddu, Vasisht, et al.
Publicado: (2023)
Attesting Distributional Properties of Training Data for Machine Learning
por: Duddu, Vasisht, et al.
Publicado: (2023)
por: Duddu, Vasisht, et al.
Publicado: (2023)
Privacy Bias in Language Models: A Contextual Integrity-based Auditing Metric
por: Shvartzshnaider, Yan, et al.
Publicado: (2024)
por: Shvartzshnaider, Yan, et al.
Publicado: (2024)
On the Alignment of Group Fairness with Attribute Privacy
por: Aalmoes, Jan, et al.
Publicado: (2022)
por: Aalmoes, Jan, et al.
Publicado: (2022)
Espresso: Robust Concept Filtering in Text-to-Image Models
por: Das, Anudeep, et al.
Publicado: (2024)
por: Das, Anudeep, et al.
Publicado: (2024)
Laminator: Verifiable ML Property Cards using Hardware-assisted Attestations
por: Duddu, Vasisht, et al.
Publicado: (2024)
por: Duddu, Vasisht, et al.
Publicado: (2024)
PATCH: Mitigating PII Leakage in Language Models with Privacy-Aware Targeted Circuit PatcHing
por: Hughes, Anthony, et al.
Publicado: (2025)
por: Hughes, Anthony, et al.
Publicado: (2025)
PAL*M: Property Attestation for Large Generative Models
por: Chantasantitam, Prach, et al.
Publicado: (2026)
por: Chantasantitam, Prach, et al.
Publicado: (2026)
Extracting Training Data from Diffusion Language Models via Infilling
por: Wang, Yihan, et al.
Publicado: (2026)
por: Wang, Yihan, et al.
Publicado: (2026)
Amulet: a Python Library for Assessing Interactions Among ML Defenses and Risks
por: Waheed, Asim, et al.
Publicado: (2025)
por: Waheed, Asim, et al.
Publicado: (2025)
Deep-Lock: Secure Authorization for Deep Neural Networks
por: Alam, Manaar, et al.
Publicado: (2020)
por: Alam, Manaar, et al.
Publicado: (2020)
LLA: Enhancing Security and Privacy for Generative Models with Logic-Locked Accelerators
por: Li, You, et al.
Publicado: (2025)
por: Li, You, et al.
Publicado: (2025)
Locking Machine Learning Models into Hardware
por: Clifford, Eleanor, et al.
Publicado: (2024)
por: Clifford, Eleanor, et al.
Publicado: (2024)
DistilLock: Safeguarding LLMs from Unauthorized Knowledge Distillation on the Edge
por: Mohanty, Asmita, et al.
Publicado: (2025)
por: Mohanty, Asmita, et al.
Publicado: (2025)
JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
por: Chao, Patrick, et al.
Publicado: (2024)
por: Chao, Patrick, et al.
Publicado: (2024)
FLARE: Feature-based Lightweight Aggregation for Robust Evaluation of IoT Intrusion Detection
por: Boswell, Bradley, et al.
Publicado: (2025)
por: Boswell, Bradley, et al.
Publicado: (2025)
Robust Feature Inference: A Test-time Defense Strategy using Spectral Projections
por: Singh, Anurag, et al.
Publicado: (2023)
por: Singh, Anurag, et al.
Publicado: (2023)
Capability and Robustness Cannot Both Be Free: An Information-Theoretic Bound for Vision-Language-Action Models
por: Tai, Jianwei
Publicado: (2026)
por: Tai, Jianwei
Publicado: (2026)
Differentially Private Random Feature Model
por: Liao, Chunyang, et al.
Publicado: (2024)
por: Liao, Chunyang, et al.
Publicado: (2024)
Adaptive and Robust Data Poisoning Detection and Sanitization in Wearable IoT Systems using Large Language Models
por: Mithsara, W. K. M, et al.
Publicado: (2025)
por: Mithsara, W. K. M, et al.
Publicado: (2025)
MergePrint: Merge-Resistant Fingerprints for Robust Black-box Ownership Verification of Large Language Models
por: Yamabe, Shojiro, et al.
Publicado: (2024)
por: Yamabe, Shojiro, et al.
Publicado: (2024)
Knowledge-Driven Multi-Turn Jailbreaking on Large Language Models
por: Li, Songze, et al.
Publicado: (2026)
por: Li, Songze, et al.
Publicado: (2026)
Level Up with ML Vulnerability Identification: Leveraging Domain Constraints in Feature Space for Robust Android Malware Detection
por: Bostani, Hamid, et al.
Publicado: (2022)
por: Bostani, Hamid, et al.
Publicado: (2022)
Large Language Models in Cybersecurity: Applications, Vulnerabilities, and Defense Techniques
por: Jaffal, Niveen O., et al.
Publicado: (2025)
por: Jaffal, Niveen O., et al.
Publicado: (2025)
Improved Techniques for Optimization-Based Jailbreaking on Large Language Models
por: Jia, Xiaojun, et al.
Publicado: (2024)
por: Jia, Xiaojun, et al.
Publicado: (2024)
Analysis of Privacy Leakage in Federated Large Language Models
por: Vu, Minh N., et al.
Publicado: (2024)
por: Vu, Minh N., et al.
Publicado: (2024)
Towards Scalable and Robust Model Versioning
por: Ding, Wenxin, et al.
Publicado: (2024)
por: Ding, Wenxin, et al.
Publicado: (2024)
Multimodal Techniques for Malware Classification
por: Jiang, Jonathan, et al.
Publicado: (2025)
por: Jiang, Jonathan, et al.
Publicado: (2025)
Breaking Distortion-free Watermarks in Large Language Models
por: Reynolds, Shayleen, et al.
Publicado: (2025)
por: Reynolds, Shayleen, et al.
Publicado: (2025)
Improving the Trade-off Between Watermark Strength and Speculative Sampling Efficiency for Language Models
por: He, Weiqing, et al.
Publicado: (2026)
por: He, Weiqing, et al.
Publicado: (2026)
SoK: Privacy-aware LLM in Healthcare: Threat Model, Privacy Techniques, Challenges and Recommendations
por: Tahera, Mohoshin Ara, et al.
Publicado: (2026)
por: Tahera, Mohoshin Ara, et al.
Publicado: (2026)
RedChronos: A Large Language Model-Based Log Analysis System for Insider Threat Detection in Enterprises
por: Li, Chenyu, et al.
Publicado: (2025)
por: Li, Chenyu, et al.
Publicado: (2025)
Integrating Feature Attention and Temporal Modeling for Collaborative Financial Risk Assessment
por: Yao, Yue, et al.
Publicado: (2025)
por: Yao, Yue, et al.
Publicado: (2025)
FIMBA: Evaluating the Robustness of AI in Genomics via Feature Importance Adversarial Attacks
por: Skovorodnikov, Heorhii, et al.
Publicado: (2024)
por: Skovorodnikov, Heorhii, et al.
Publicado: (2024)
MM-FusionNet: Context-Aware Dynamic Fusion for Multi-modal Fake News Detection with Large Vision-Language Models
por: He, Junhao, et al.
Publicado: (2025)
por: He, Junhao, et al.
Publicado: (2025)
Randomization Techniques to Mitigate the Risk of Copyright Infringement
por: Chen, Wei-Ning, et al.
Publicado: (2024)
por: Chen, Wei-Ning, et al.
Publicado: (2024)
Robust Distortion-free Watermarks for Language Models
por: Kuditipudi, Rohith, et al.
Publicado: (2023)
por: Kuditipudi, Rohith, et al.
Publicado: (2023)
FDINet: Protecting against DNN Model Extraction via Feature Distortion Index
por: Yao, Hongwei, et al.
Publicado: (2023)
por: Yao, Hongwei, et al.
Publicado: (2023)
PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
por: Zhu, Kaijie, et al.
Publicado: (2023)
por: Zhu, Kaijie, et al.
Publicado: (2023)
Ejemplares similares
-
Combining Machine Learning Defenses without Conflicts
por: Duddu, Vasisht, et al.
Publicado: (2024) -
SoK: Unintended Interactions among Machine Learning Defenses and Risks
por: Duddu, Vasisht, et al.
Publicado: (2023) -
Attesting Distributional Properties of Training Data for Machine Learning
por: Duddu, Vasisht, et al.
Publicado: (2023) -
Privacy Bias in Language Models: A Contextual Integrity-based Auditing Metric
por: Shvartzshnaider, Yan, et al.
Publicado: (2024) -
On the Alignment of Group Fairness with Attribute Privacy
por: Aalmoes, Jan, et al.
Publicado: (2022)