Verifying LLM Inference to Detect Model Weight Exfiltration
Fuente:
arXiv
Saved in:
| Main Authors: | Rinberg, Roy, Karvonen, Adam, Hoover, Alexander, Reuter, Daniel, Warr, Keri |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiFR: Inference Verification Despite Nondeterminism
by: Karvonen, Adam, et al.
Published: (2025)
by: Karvonen, Adam, et al.
Published: (2025)
Improving DNS Exfiltration Detection via Transformer Pretraining
by: Tomić, Miloš, et al.
Published: (2026)
by: Tomić, Miloš, et al.
Published: (2026)
Beyond Laplace and Gaussian: Exploring the Generalized Gaussian Mechanism for Private Machine Learning
by: Rinberg, Roy, et al.
Published: (2025)
by: Rinberg, Roy, et al.
Published: (2025)
Have it your way: Individualized Privacy Assignment for DP-SGD
by: Boenisch, Franziska, et al.
Published: (2023)
by: Boenisch, Franziska, et al.
Published: (2023)
OrgForge-IT: A Verifiable Synthetic Benchmark for LLM-Based Insider Threat Detection
by: Flynt, Jeffrey
Published: (2026)
by: Flynt, Jeffrey
Published: (2026)
Privacy-Preserving Verifiable Neural Network Inference Service
by: Riasi, Arman, et al.
Published: (2024)
by: Riasi, Arman, et al.
Published: (2024)
Privacy-Preserving Mechanisms Enable Cheap Verifiable Inference of LLMs
by: Pal, Arka, et al.
Published: (2026)
by: Pal, Arka, et al.
Published: (2026)
MarkTune: Improving the Quality-Detectability Trade-off in Open-Weight LLM Watermarking
by: Zhao, Yizhou, et al.
Published: (2025)
by: Zhao, Yizhou, et al.
Published: (2025)
Efficient and Verifiable Privacy-Preserving Convolutional Computation for CNN Inference with Untrusted Clouds
by: Lu, Jinyu, et al.
Published: (2025)
by: Lu, Jinyu, et al.
Published: (2025)
VeriLLM: A Lightweight Framework for Publicly Verifiable Decentralized Inference
by: Wang, Ke, et al.
Published: (2025)
by: Wang, Ke, et al.
Published: (2025)
Generate-then-Verify: Reconstructing Data from Limited Published Statistics
by: Liu, Terrance, et al.
Published: (2025)
by: Liu, Terrance, et al.
Published: (2025)
MemHunter: Automated and Verifiable Memorization Detection at Dataset-scale in LLMs
by: Wu, Zhenpeng, et al.
Published: (2024)
by: Wu, Zhenpeng, et al.
Published: (2024)
NANOZK: Layerwise Zero-Knowledge Proofs for Verifiable Large Language Model Inference
by: Wang, Zhaohui Geoffrey
Published: (2026)
by: Wang, Zhaohui Geoffrey
Published: (2026)
Cascade: Token-Sharded Private LLM Inference
by: Thomas, Rahul, et al.
Published: (2025)
by: Thomas, Rahul, et al.
Published: (2025)
Order of Magnitude Speedups for LLM Membership Inference
by: Zhang, Rongting, et al.
Published: (2024)
by: Zhang, Rongting, et al.
Published: (2024)
Verifiable Unlearning on Edge
by: Maheri, Mohammad M, et al.
Published: (2025)
by: Maheri, Mohammad M, et al.
Published: (2025)
The Sample Complexity of Membership Inference and Privacy Auditing
by: Haghifam, Mahdi, et al.
Published: (2025)
by: Haghifam, Mahdi, et al.
Published: (2025)
AERO: Entropy-Guided Framework for Private LLM Inference
by: Jha, Nandan Kumar, et al.
Published: (2024)
by: Jha, Nandan Kumar, et al.
Published: (2024)
Open-Weight LLM Fine-Tuning Defenses are Susceptible to Simple Attacks
by: Kuo, Kevin, et al.
Published: (2026)
by: Kuo, Kevin, et al.
Published: (2026)
LLM Dataset Inference: Did you train on my dataset?
by: Maini, Pratyush, et al.
Published: (2024)
by: Maini, Pratyush, et al.
Published: (2024)
Filter-then-Verify: A Multiphase GNN and ModernBERT Framework for Social Engineering Detection in Email Networks
by: Khadka, Barsat, et al.
Published: (2026)
by: Khadka, Barsat, et al.
Published: (2026)
Command-line Obfuscation Detection using Small Language Models
by: Outrata, Vojtech, et al.
Published: (2024)
by: Outrata, Vojtech, et al.
Published: (2024)
TruncFormer: Private LLM Inference Using Only Truncations
by: Yubeaton, Patrick, et al.
Published: (2024)
by: Yubeaton, Patrick, et al.
Published: (2024)
PVMark: Enabling Public Verifiability for LLM Watermarking Schemes
by: Duan, Haohua, et al.
Published: (2025)
by: Duan, Haohua, et al.
Published: (2025)
ZKBoost: Zero-Knowledge Verifiable Training for XGBoost
by: Melissaris, Nikolas, et al.
Published: (2026)
by: Melissaris, Nikolas, et al.
Published: (2026)
SLIP: Securing LLMs IP Using Weights Decomposition
by: Refael, Yehonathan, et al.
Published: (2024)
by: Refael, Yehonathan, et al.
Published: (2024)
Malicious Internet Entity Detection Using Local Graph Inference
by: Mandlik, Simon, et al.
Published: (2024)
by: Mandlik, Simon, et al.
Published: (2024)
Expressive Losses for Verified Robustness via Convex Combinations
by: De Palma, Alessandro, et al.
Published: (2023)
by: De Palma, Alessandro, et al.
Published: (2023)
Fine-tuning Large Language Models for DGA and DNS Exfiltration Detection
by: Sayed, Md Abu, et al.
Published: (2024)
by: Sayed, Md Abu, et al.
Published: (2024)
Automated Membership Inference Attacks: Discovering MIA Signal Computations using LLM Agents
by: Tran, Toan, et al.
Published: (2026)
by: Tran, Toan, et al.
Published: (2026)
Weights Shuffling for Improving DPSGD in Transformer-based Models
by: Yang, Jungang, et al.
Published: (2024)
by: Yang, Jungang, et al.
Published: (2024)
LLM-Generated Samples for Android Malware Detection
by: Rollinson, Nik, et al.
Published: (2025)
by: Rollinson, Nik, et al.
Published: (2025)
Token-Efficient Change Detection in LLM APIs
by: Chauvin, Timothée, et al.
Published: (2026)
by: Chauvin, Timothée, et al.
Published: (2026)
Auditing Privacy Mechanisms via Label Inference Attacks
by: Busa-Fekete, Róbert István, et al.
Published: (2024)
by: Busa-Fekete, Róbert István, et al.
Published: (2024)
Membership Inference Attacks on Sequence Models
by: Rossi, Lorenzo, et al.
Published: (2025)
by: Rossi, Lorenzo, et al.
Published: (2025)
SVIP: Towards Verifiable Inference of Open-source Large Language Models
by: Sun, Yifan, et al.
Published: (2024)
by: Sun, Yifan, et al.
Published: (2024)
What Really is a Member? Discrediting Membership Inference via Poisoning
by: Mangaokar, Neal, et al.
Published: (2025)
by: Mangaokar, Neal, et al.
Published: (2025)
Privacy-Preserving Inference for Quantized BERT Models
by: Lu, Tianpei, et al.
Published: (2025)
by: Lu, Tianpei, et al.
Published: (2025)
Fingerprinting Inference Systems of Large Language Models
by: Wimbauer, Anna, et al.
Published: (2026)
by: Wimbauer, Anna, et al.
Published: (2026)
Privacy-Preserving Hierarchical Model-Distributed Inference
by: Dehkordi, Fatemeh Jafarian, et al.
Published: (2024)
by: Dehkordi, Fatemeh Jafarian, et al.
Published: (2024)
Similar Items
-
DiFR: Inference Verification Despite Nondeterminism
by: Karvonen, Adam, et al.
Published: (2025) -
Improving DNS Exfiltration Detection via Transformer Pretraining
by: Tomić, Miloš, et al.
Published: (2026) -
Beyond Laplace and Gaussian: Exploring the Generalized Gaussian Mechanism for Private Machine Learning
by: Rinberg, Roy, et al.
Published: (2025) -
Have it your way: Individualized Privacy Assignment for DP-SGD
by: Boenisch, Franziska, et al.
Published: (2023) -
OrgForge-IT: A Verifiable Synthetic Benchmark for LLM-Based Insider Threat Detection
by: Flynt, Jeffrey
Published: (2026)