Aggressive Compression Enables LLM Weight Theft
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Brown, Davis, Rivera, Juan-Pablo, Hendrycks, Dan, Mazeika, Mantas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Best Practices for Biorisk Evaluations on Open-Weight Bio-Foundation Models
von: Wei, Boyi, et al.
Veröffentlicht: (2025)
von: Wei, Boyi, et al.
Veröffentlicht: (2025)
A Novel DDPM-based Ensemble Approach for Energy Theft Detection in Smart Grids
von: Yuan, Xun, et al.
Veröffentlicht: (2023)
von: Yuan, Xun, et al.
Veröffentlicht: (2023)
Fewer Weights, More Problems: A Practical Attack on LLM Pruning
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
MarkTune: Improving the Quality-Detectability Trade-off in Open-Weight LLM Watermarking
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
Quantized Delta Weight Is Safety Keeper
von: Liu, Yule, et al.
Veröffentlicht: (2024)
von: Liu, Yule, et al.
Veröffentlicht: (2024)
WARP: Weight Teleportation for Attack-Resilient Unlearning Protocols
von: Maheri, Mohammad M, et al.
Veröffentlicht: (2025)
von: Maheri, Mohammad M, et al.
Veröffentlicht: (2025)
Learnability and Privacy Vulnerability are Entangled in a Few Critical Weights
von: Fang, Xingli, et al.
Veröffentlicht: (2026)
von: Fang, Xingli, et al.
Veröffentlicht: (2026)
Trusted Weights, Treacherous Optimizations? Optimization-Triggered Backdoor Attacks on LLMs
von: Wang, Yifei, et al.
Veröffentlicht: (2026)
von: Wang, Yifei, et al.
Veröffentlicht: (2026)
Challenges in Enabling Private Data Valuation
von: Fu, Yiwei, et al.
Veröffentlicht: (2026)
von: Fu, Yiwei, et al.
Veröffentlicht: (2026)
Local Data Quantity-Aware Weighted Averaging for Federated Learning with Dishonest Clients
von: Wu, Leming, et al.
Veröffentlicht: (2025)
von: Wu, Leming, et al.
Veröffentlicht: (2025)
Inducing Uncertainty on Open-Weight Models for Test-Time Privacy in Image Recognition
von: Ashiq, Muhammad H., et al.
Veröffentlicht: (2025)
von: Ashiq, Muhammad H., et al.
Veröffentlicht: (2025)
Too Good to be True? Turn Any Model Differentially Private With DP-Weights
von: Zagardo, David
Veröffentlicht: (2024)
von: Zagardo, David
Veröffentlicht: (2024)
Exploiting LLM Quantization
von: Egashira, Kazuki, et al.
Veröffentlicht: (2024)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2024)
Blockchain-Enabled Explainable AI for Trusted Healthcare Systems
von: Mohsin, Md Talha
Veröffentlicht: (2025)
von: Mohsin, Md Talha
Veröffentlicht: (2025)
Robust Federated Learning with Confidence-Weighted Filtering and GAN-Based Completion under Noisy and Incomplete Data
von: Gokcen, Alpaslan, et al.
Veröffentlicht: (2025)
von: Gokcen, Alpaslan, et al.
Veröffentlicht: (2025)
Prompt2Fingerprint: Plug-and-Play LLM Fingerprinting via Text-to-Weight Generation
von: Chen, Sixu, et al.
Veröffentlicht: (2026)
von: Chen, Sixu, et al.
Veröffentlicht: (2026)
Optimistic Verifiable Training by Controlling Hardware Nondeterminism
von: Srivastava, Megha, et al.
Veröffentlicht: (2024)
von: Srivastava, Megha, et al.
Veröffentlicht: (2024)
A Unified Framework for LLM Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
Attacks and Defenses Against LLM Fingerprinting
von: Kurian, Kevin, et al.
Veröffentlicht: (2025)
von: Kurian, Kevin, et al.
Veröffentlicht: (2025)
Are Robust LLM Fingerprints Adversarially Robust?
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)
von: Nasery, Anshul, et al.
Veröffentlicht: (2025)
How Not to Detect Prompt Injections with an LLM
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
MPC-Minimized Secure LLM Inference
von: Rathee, Deevashwer, et al.
Veröffentlicht: (2024)
von: Rathee, Deevashwer, et al.
Veröffentlicht: (2024)
Multi-Continental Healthcare Modelling Using Blockchain-Enabled Federated Learning
von: Sun, Rui, et al.
Veröffentlicht: (2024)
von: Sun, Rui, et al.
Veröffentlicht: (2024)
PRO: Enabling Precise and Robust Text Watermark for Open-Source LLMs
von: Xue, Jiaqi, et al.
Veröffentlicht: (2025)
von: Xue, Jiaqi, et al.
Veröffentlicht: (2025)
LLM Benchmark Datasets Should Be Contamination-Resistant
von: Al-Lawati, Ali, et al.
Veröffentlicht: (2026)
von: Al-Lawati, Ali, et al.
Veröffentlicht: (2026)
Reliable Weak-to-Strong Monitoring of LLM Agents
von: Kale, Neil, et al.
Veröffentlicht: (2025)
von: Kale, Neil, et al.
Veröffentlicht: (2025)
Mimicking the Familiar: Dynamic Command Generation for Information Theft Attacks in LLM Tool-Learning System
von: Jiang, Ziyou, et al.
Veröffentlicht: (2025)
von: Jiang, Ziyou, et al.
Veröffentlicht: (2025)
Evaluating False Alarm and Missing Attacks in CAN IDS
von: Hossain, Nirab, et al.
Veröffentlicht: (2026)
von: Hossain, Nirab, et al.
Veröffentlicht: (2026)
Enhancing Reliability in LLM-Based Secure Code Generation
von: Kharma, Mohammed F., et al.
Veröffentlicht: (2026)
von: Kharma, Mohammed F., et al.
Veröffentlicht: (2026)
The Autonomy Tax: Defense Training Breaks LLM Agents
von: Li, Shawn, et al.
Veröffentlicht: (2026)
von: Li, Shawn, et al.
Veröffentlicht: (2026)
Syntax- and Compilation-Preserving Evasion of LLM Vulnerability Detectors
von: Sun, Luze, et al.
Veröffentlicht: (2026)
von: Sun, Luze, et al.
Veröffentlicht: (2026)
GuardReasoner: Towards Reasoning-based LLM Safeguards
von: Liu, Yue, et al.
Veröffentlicht: (2025)
von: Liu, Yue, et al.
Veröffentlicht: (2025)
How Catastrophic is Your LLM? Certifying Risk in Conversation
von: Wang, Chengxiao, et al.
Veröffentlicht: (2025)
von: Wang, Chengxiao, et al.
Veröffentlicht: (2025)
SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents
von: Kutasov, Jonathan, et al.
Veröffentlicht: (2025)
von: Kutasov, Jonathan, et al.
Veröffentlicht: (2025)
Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
KnowGraph: Knowledge-Enabled Anomaly Detection via Logical Reasoning on Graph Data
von: Zhou, Andy, et al.
Veröffentlicht: (2024)
von: Zhou, Andy, et al.
Veröffentlicht: (2024)
HybridGuard: Enhancing Minority-Class Intrusion Detection in Dew-Enabled Edge-of-Things Networks
von: Kara, Binayak, et al.
Veröffentlicht: (2025)
von: Kara, Binayak, et al.
Veröffentlicht: (2025)
Seed Hijacking of LLM Sampling and Quantum Random Number Defense
von: You, Ziyang, et al.
Veröffentlicht: (2026)
von: You, Ziyang, et al.
Veröffentlicht: (2026)
HARP: Measuring Harm Amplification in Multi-Agent LLM Systems
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2026)
von: Rahman, Md Hafizur, et al.
Veröffentlicht: (2026)
Learning to Watermark LLM-generated Text via Reinforcement Learning
von: Xu, Xiaojun, et al.
Veröffentlicht: (2024)
von: Xu, Xiaojun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Best Practices for Biorisk Evaluations on Open-Weight Bio-Foundation Models
von: Wei, Boyi, et al.
Veröffentlicht: (2025) -
A Novel DDPM-based Ensemble Approach for Energy Theft Detection in Smart Grids
von: Yuan, Xun, et al.
Veröffentlicht: (2023) -
Fewer Weights, More Problems: A Practical Attack on LLM Pruning
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025) -
MarkTune: Improving the Quality-Detectability Trade-off in Open-Weight LLM Watermarking
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025) -
Quantized Delta Weight Is Safety Keeper
von: Liu, Yule, et al.
Veröffentlicht: (2024)