Towards Provable (In)Secure Model Weight Release Schemes
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Xin, Tang, Bintao, Wang, Yuhao, Ji, Zimo, Zhang, Terry Jingchen, Jiang, Wenyuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Measuring the Permission Gate: A Stress-Test Evaluation of Claude Code's Auto Mode
by: Ji, Zimo, et al.
Published: (2026)
by: Ji, Zimo, et al.
Published: (2026)
Provably Secure Agent Guardrail
by: Wu, Benlong, et al.
Published: (2026)
by: Wu, Benlong, et al.
Published: (2026)
Provably Secure Retrieval-Augmented Generation
by: Zhou, Pengcheng, et al.
Published: (2025)
by: Zhou, Pengcheng, et al.
Published: (2025)
Training with Differential Privacy: A Gradient-Preserving Noise Reduction Approach with Provable Security
by: Wang, Haodi, et al.
Published: (2024)
by: Wang, Haodi, et al.
Published: (2024)
CryptoScope: Utilizing Large Language Models for Automated Cryptographic Logic Vulnerability Detection
by: Li, Zhihao, et al.
Published: (2025)
by: Li, Zhihao, et al.
Published: (2025)
Taylor Unswift: Secured Weight Release for Large Language Models via Taylor Expansion
by: Wang, Guanchu, et al.
Published: (2024)
by: Wang, Guanchu, et al.
Published: (2024)
Backdoor Sentinel: Detecting and Detoxifying Backdoors in Diffusion Models via Temporal Noise Consistency
by: Wang, Bingzheng, et al.
Published: (2026)
by: Wang, Bingzheng, et al.
Published: (2026)
SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents
by: Ouyang, Yipeng, et al.
Published: (2026)
by: Ouyang, Yipeng, et al.
Published: (2026)
Towards Understanding and Enhancing Security of Proof-of-Training for DNN Model Ownership Verification
by: Chang, Yijia, et al.
Published: (2024)
by: Chang, Yijia, et al.
Published: (2024)
Towards Secure and Explainable Smart Contract Generation with Security-Aware Group Relative Policy Optimization
by: Yu, Lei, et al.
Published: (2025)
by: Yu, Lei, et al.
Published: (2025)
Differentially Private and Communication Efficient Large Language Model Split Inference via Stochastic Quantization and Soft Prompt
by: Gu, Yujie, et al.
Published: (2026)
by: Gu, Yujie, et al.
Published: (2026)
INTEGRALBENCH: Benchmarking LLMs with Definite Integral Problems
by: Tang, Bintao, et al.
Published: (2025)
by: Tang, Bintao, et al.
Published: (2025)
RobWE: Robust Watermark Embedding for Personalized Federated Learning Model Ownership Protection
by: Xu, Yang, et al.
Published: (2024)
by: Xu, Yang, et al.
Published: (2024)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
by: Zhu, Kaijie, et al.
Published: (2025)
by: Zhu, Kaijie, et al.
Published: (2025)
Secured Communication Schemes for UAVs in 5G: CRYSTALS-Kyber and IDS
by: Sharma, Taneya, et al.
Published: (2025)
by: Sharma, Taneya, et al.
Published: (2025)
Toward a Unified Security Framework for AI Agents: Trust, Risk, and Liability
by: Mo, Jiayun, et al.
Published: (2025)
by: Mo, Jiayun, et al.
Published: (2025)
FedMPQ: Secure and Communication-Efficient Federated Learning with Multi-codebook Product Quantization
by: Yang, Xu, et al.
Published: (2024)
by: Yang, Xu, et al.
Published: (2024)
A Mixture of Linear Corrections Generates Secure Code
by: Yu, Weichen, et al.
Published: (2025)
by: Yu, Weichen, et al.
Published: (2025)
Towards Secure Logging: Characterizing and Benchmarking Logging Code Security Issues with LLMs
by: Yuan, He Yang, et al.
Published: (2026)
by: Yuan, He Yang, et al.
Published: (2026)
Towards Secure and Private AI: A Framework for Decentralized Inference
by: Zhang, Hongyang, et al.
Published: (2024)
by: Zhang, Hongyang, et al.
Published: (2024)
CryptoTensors: A Light-Weight Large Language Model File Format for Highly-Secure Model Distribution
by: Zhu, Huifeng, et al.
Published: (2025)
by: Zhu, Huifeng, et al.
Published: (2025)
Towards more Practical Threat Models in Artificial Intelligence Security
by: Grosse, Kathrin, et al.
Published: (2023)
by: Grosse, Kathrin, et al.
Published: (2023)
Optimized Ensemble Model Towards Secured Industrial IoT Devices
by: Injadat, MohammadNoor
Published: (2024)
by: Injadat, MohammadNoor
Published: (2024)
On the Security Risks of ML-based Malware Detection Systems: A Survey
by: He, Ping, et al.
Published: (2025)
by: He, Ping, et al.
Published: (2025)
Towards Effective Complementary Security Analysis using Large Language Models
by: Wagner, Jonas, et al.
Published: (2025)
by: Wagner, Jonas, et al.
Published: (2025)
A Survey: Towards Privacy and Security in Mobile Large Language Models
by: Xu, Honghui, et al.
Published: (2025)
by: Xu, Honghui, et al.
Published: (2025)
Towards Small Language Models for Security Query Generation in SOC Workflows
by: Muzammil, Saleha, et al.
Published: (2025)
by: Muzammil, Saleha, et al.
Published: (2025)
Towards Secure Retrieval-Augmented Generation: A Comprehensive Review of Threats, Defenses and Benchmarks
by: Mu, Yanming, et al.
Published: (2026)
by: Mu, Yanming, et al.
Published: (2026)
Towards Secure Agent Skills: Architecture, Threat Taxonomy, and Security Analysis
by: Li, Zhiyuan, et al.
Published: (2026)
by: Li, Zhiyuan, et al.
Published: (2026)
Automating Security Audit Using Large Language Model based Agent: An Exploration Experiment
by: Chin, Jia Hui, et al.
Published: (2025)
by: Chin, Jia Hui, et al.
Published: (2025)
Taxonomy, Evaluation and Exploitation of IPI-Centric LLM Agent Defense Frameworks
by: Ji, Zimo, et al.
Published: (2025)
by: Ji, Zimo, et al.
Published: (2025)
ClawLess: A Security Model of AI Agents
by: Lu, Hongyi, et al.
Published: (2026)
by: Lu, Hongyi, et al.
Published: (2026)
Secure Confidential Business Information When Sharing Machine Learning Models
by: Yang, Yunfan, et al.
Published: (2025)
by: Yang, Yunfan, et al.
Published: (2025)
SeCodePLT: A Unified Platform for Evaluating the Security of Code GenAI
by: Nie, Yuzhou, et al.
Published: (2024)
by: Nie, Yuzhou, et al.
Published: (2024)
Federated Learning-Based Data Collaboration Method for Enhancing Edge Cloud AI System Security Using Large Language Models
by: Luo, Huaiying, et al.
Published: (2025)
by: Luo, Huaiying, et al.
Published: (2025)
Federated Large Language Models: Feasibility, Robustness, Security and Future Directions
by: Jiang, Wenhao, et al.
Published: (2025)
by: Jiang, Wenhao, et al.
Published: (2025)
Secure Code Generation via Online Reinforcement Learning with Vulnerability Reward Model
by: Wu, Tianyi, et al.
Published: (2026)
by: Wu, Tianyi, et al.
Published: (2026)
The Security Threat of Compressed Projectors in Large Vision-Language Models
by: Zhang, Yudong, et al.
Published: (2025)
by: Zhang, Yudong, et al.
Published: (2025)
Invariant-based Robust Weights Watermark for Large Language Models
by: Guo, Qingxiao, et al.
Published: (2025)
by: Guo, Qingxiao, et al.
Published: (2025)
Large Language Models for Cyber Security: A Systematic Literature Review
by: Xu, Hanxiang, et al.
Published: (2024)
by: Xu, Hanxiang, et al.
Published: (2024)
Similar Items
-
Measuring the Permission Gate: A Stress-Test Evaluation of Claude Code's Auto Mode
by: Ji, Zimo, et al.
Published: (2026) -
Provably Secure Agent Guardrail
by: Wu, Benlong, et al.
Published: (2026) -
Provably Secure Retrieval-Augmented Generation
by: Zhou, Pengcheng, et al.
Published: (2025) -
Training with Differential Privacy: A Gradient-Preserving Noise Reduction Approach with Provable Security
by: Wang, Haodi, et al.
Published: (2024) -
CryptoScope: Utilizing Large Language Models for Automated Cryptographic Logic Vulnerability Detection
by: Li, Zhihao, et al.
Published: (2025)