Private Transformer Inference in MLaaS: A Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yang, Zhou, Xinyu, Wang, Yitong, Qian, Liangxin, Zhao, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Survey on Private Transformer Inference
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
ERASER: Machine Unlearning in MLaaS via an Inference Serving-Aware Approach
by: Hu, Yuke, et al.
Published: (2023)
by: Hu, Yuke, et al.
Published: (2023)
BDFirewall: Towards Effective and Expeditiously Black-Box Backdoor Defense in MLaaS
by: Li, Ye, et al.
Published: (2025)
by: Li, Ye, et al.
Published: (2025)
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
Towards Secure and Private AI: A Framework for Decentralized Inference
by: Zhang, Hongyang, et al.
Published: (2024)
by: Zhang, Hongyang, et al.
Published: (2024)
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference
by: Lin, Chenqi, et al.
Published: (2024)
by: Lin, Chenqi, et al.
Published: (2024)
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
Comet: Accelerating Private Inference for Large Language Model by Predicting Activation Sparsity
by: Yan, Guang, et al.
Published: (2025)
by: Yan, Guang, et al.
Published: (2025)
Differentially Private Worst-group Risk Minimization
by: Zhou, Xinyu, et al.
Published: (2024)
by: Zhou, Xinyu, et al.
Published: (2024)
Nimbus: Secure and Efficient Two-Party Inference for Transformers
by: Li, Zhengyi, et al.
Published: (2024)
by: Li, Zhengyi, et al.
Published: (2024)
PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
UFO: Unlocking Ultra-Efficient Quantized Private Inference with Protocol and Algorithm Co-Optimization
by: Zeng, Wenxuan, et al.
Published: (2026)
by: Zeng, Wenxuan, et al.
Published: (2026)
Comet: A Communication-efficient and Performant Approximation for Private Transformer Inference
by: Xu, Xiangrui, et al.
Published: (2024)
by: Xu, Xiangrui, et al.
Published: (2024)
On the (In-)Security of the Shuffling Defense in the Transformer Secure Inference
by: Li, Zhengyi, et al.
Published: (2026)
by: Li, Zhengyi, et al.
Published: (2026)
Locally Differentially Private In-Context Learning
by: Zheng, Chunyan, et al.
Published: (2024)
by: Zheng, Chunyan, et al.
Published: (2024)
Optimized Layerwise Approximation for Efficient Private Inference on Fully Homomorphic Encryption
by: Lee, Junghyun, et al.
Published: (2023)
by: Lee, Junghyun, et al.
Published: (2023)
Private Seeds, Public LLMs: Realistic and Privacy-Preserving Synthetic Data Generation
by: Ma, Qian, et al.
Published: (2026)
by: Ma, Qian, et al.
Published: (2026)
From Easy to Hard: Building a Shortcut for Differentially Private Image Synthesis
by: Li, Kecen, et al.
Published: (2025)
by: Li, Kecen, et al.
Published: (2025)
PIR-RAG: A System for Private Information Retrieval in Retrieval-Augmented Generation
by: Wang, Baiqiang, et al.
Published: (2025)
by: Wang, Baiqiang, et al.
Published: (2025)
DPImageBench: A Unified Benchmark for Differentially Private Image Synthesis
by: Gong, Chen, et al.
Published: (2025)
by: Gong, Chen, et al.
Published: (2025)
FHAIM: Fully Homomorphic AIM For Private Synthetic Data Generation
by: Kumar, Mayank, et al.
Published: (2026)
by: Kumar, Mayank, et al.
Published: (2026)
Exposing Hidden Interfaces: LLM-Guided Type Inference for Reverse Engineering macOS Private Frameworks
by: Kharlamova, Arina, et al.
Published: (2026)
by: Kharlamova, Arina, et al.
Published: (2026)
A First Look At Efficient And Secure On-Device LLM Inference Against KV Leakage
by: Yang, Huan, et al.
Published: (2024)
by: Yang, Huan, et al.
Published: (2024)
Generating Is Believing: Membership Inference Attacks against Retrieval-Augmented Generation
by: Li, Yuying, et al.
Published: (2024)
by: Li, Yuying, et al.
Published: (2024)
Differentially Private and Communication Efficient Large Language Model Split Inference via Stochastic Quantization and Soft Prompt
by: Gu, Yujie, et al.
Published: (2026)
by: Gu, Yujie, et al.
Published: (2026)
Revisiting Training-Inference Trigger Intensity in Backdoor Attacks
by: Lin, Chenhao, et al.
Published: (2025)
by: Lin, Chenhao, et al.
Published: (2025)
A Survey for Deep Reinforcement Learning Based Network Intrusion Detection
by: Yang, Wanrong, et al.
Published: (2024)
by: Yang, Wanrong, et al.
Published: (2024)
SecureRouter: Encrypted Routing for Efficient Secure Inference
by: Zhang, Yukuan, et al.
Published: (2026)
by: Zhang, Yukuan, et al.
Published: (2026)
A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations
by: Zhou, Yihe, et al.
Published: (2025)
by: Zhou, Yihe, et al.
Published: (2025)
PathSeeker: Exploring LLM Security Vulnerabilities with a Reinforcement Learning-Based Jailbreak Approach
by: Lin, Zhihao, et al.
Published: (2024)
by: Lin, Zhihao, et al.
Published: (2024)
Linearizing Models for Efficient yet Robust Private Inference
by: Sarkar, Sreetama, et al.
Published: (2024)
by: Sarkar, Sreetama, et al.
Published: (2024)
Differentially Private Federated Low Rank Adaptation Beyond Fixed-Matrix
by: Wen, Ming, et al.
Published: (2025)
by: Wen, Ming, et al.
Published: (2025)
A New Linear Scaling Rule for Private Adaptive Hyperparameter Optimization
by: Panda, Ashwinee, et al.
Published: (2022)
by: Panda, Ashwinee, et al.
Published: (2022)
A Survey on Data Security in Large Language Models
by: Chen, Kang, et al.
Published: (2025)
by: Chen, Kang, et al.
Published: (2025)
Deep Learning-based Intrusion Detection Systems: A Survey
by: Xu, Zhiwei, et al.
Published: (2025)
by: Xu, Zhiwei, et al.
Published: (2025)
You Told Me to Do It: Measuring Instructional Text-induced Private Data Leakage in LLM Agents
by: Kao, Ching-Yu, et al.
Published: (2026)
by: Kao, Ching-Yu, et al.
Published: (2026)
Dual-Priv Pruning : Efficient Differential Private Fine-Tuning in Multimodal Large Language Models
by: Wei, Qianshan, et al.
Published: (2025)
by: Wei, Qianshan, et al.
Published: (2025)
VaultGemma: A Differentially Private Gemma Model
by: Sinha, Amer, et al.
Published: (2025)
by: Sinha, Amer, et al.
Published: (2025)
Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
by: Gao, Fengyu, et al.
Published: (2024)
by: Gao, Fengyu, et al.
Published: (2024)
Data Poisoning in Deep Learning: A Survey
by: Zhao, Pinlong, et al.
Published: (2025)
by: Zhao, Pinlong, et al.
Published: (2025)
Similar Items
-
A Survey on Private Transformer Inference
by: Li, Yang, et al.
Published: (2024) -
ERASER: Machine Unlearning in MLaaS via an Inference Serving-Aware Approach
by: Hu, Yuke, et al.
Published: (2023) -
BDFirewall: Towards Effective and Expeditiously Black-Box Backdoor Defense in MLaaS
by: Li, Ye, et al.
Published: (2025) -
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
by: Xu, Tianshi, et al.
Published: (2024) -
Towards Secure and Private AI: A Framework for Decentralized Inference
by: Zhang, Hongyang, et al.
Published: (2024)