A Survey on Private Transformer Inference
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Yang, Zhou, Xinyu, Wang, Yitong, Qian, Liangxin, Zhao, Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Private Transformer Inference in MLaaS: A Survey
di: Li, Yang, et al.
Pubblicazione: (2025)
di: Li, Yang, et al.
Pubblicazione: (2025)
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
di: Xu, Tianshi, et al.
Pubblicazione: (2024)
di: Xu, Tianshi, et al.
Pubblicazione: (2024)
Towards Secure and Private AI: A Framework for Decentralized Inference
di: Zhang, Hongyang, et al.
Pubblicazione: (2024)
di: Zhang, Hongyang, et al.
Pubblicazione: (2024)
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference
di: Lin, Chenqi, et al.
Pubblicazione: (2024)
di: Lin, Chenqi, et al.
Pubblicazione: (2024)
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
di: Xu, Tianshi, et al.
Pubblicazione: (2024)
di: Xu, Tianshi, et al.
Pubblicazione: (2024)
Comet: Accelerating Private Inference for Large Language Model by Predicting Activation Sparsity
di: Yan, Guang, et al.
Pubblicazione: (2025)
di: Yan, Guang, et al.
Pubblicazione: (2025)
Differentially Private Worst-group Risk Minimization
di: Zhou, Xinyu, et al.
Pubblicazione: (2024)
di: Zhou, Xinyu, et al.
Pubblicazione: (2024)
Nimbus: Secure and Efficient Two-Party Inference for Transformers
di: Li, Zhengyi, et al.
Pubblicazione: (2024)
di: Li, Zhengyi, et al.
Pubblicazione: (2024)
PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization
di: Xu, Tianshi, et al.
Pubblicazione: (2024)
di: Xu, Tianshi, et al.
Pubblicazione: (2024)
UFO: Unlocking Ultra-Efficient Quantized Private Inference with Protocol and Algorithm Co-Optimization
di: Zeng, Wenxuan, et al.
Pubblicazione: (2026)
di: Zeng, Wenxuan, et al.
Pubblicazione: (2026)
Comet: A Communication-efficient and Performant Approximation for Private Transformer Inference
di: Xu, Xiangrui, et al.
Pubblicazione: (2024)
di: Xu, Xiangrui, et al.
Pubblicazione: (2024)
On the (In-)Security of the Shuffling Defense in the Transformer Secure Inference
di: Li, Zhengyi, et al.
Pubblicazione: (2026)
di: Li, Zhengyi, et al.
Pubblicazione: (2026)
Locally Differentially Private In-Context Learning
di: Zheng, Chunyan, et al.
Pubblicazione: (2024)
di: Zheng, Chunyan, et al.
Pubblicazione: (2024)
Optimized Layerwise Approximation for Efficient Private Inference on Fully Homomorphic Encryption
di: Lee, Junghyun, et al.
Pubblicazione: (2023)
di: Lee, Junghyun, et al.
Pubblicazione: (2023)
Private Seeds, Public LLMs: Realistic and Privacy-Preserving Synthetic Data Generation
di: Ma, Qian, et al.
Pubblicazione: (2026)
di: Ma, Qian, et al.
Pubblicazione: (2026)
From Easy to Hard: Building a Shortcut for Differentially Private Image Synthesis
di: Li, Kecen, et al.
Pubblicazione: (2025)
di: Li, Kecen, et al.
Pubblicazione: (2025)
PIR-RAG: A System for Private Information Retrieval in Retrieval-Augmented Generation
di: Wang, Baiqiang, et al.
Pubblicazione: (2025)
di: Wang, Baiqiang, et al.
Pubblicazione: (2025)
DPImageBench: A Unified Benchmark for Differentially Private Image Synthesis
di: Gong, Chen, et al.
Pubblicazione: (2025)
di: Gong, Chen, et al.
Pubblicazione: (2025)
FHAIM: Fully Homomorphic AIM For Private Synthetic Data Generation
di: Kumar, Mayank, et al.
Pubblicazione: (2026)
di: Kumar, Mayank, et al.
Pubblicazione: (2026)
A First Look At Efficient And Secure On-Device LLM Inference Against KV Leakage
di: Yang, Huan, et al.
Pubblicazione: (2024)
di: Yang, Huan, et al.
Pubblicazione: (2024)
Exposing Hidden Interfaces: LLM-Guided Type Inference for Reverse Engineering macOS Private Frameworks
di: Kharlamova, Arina, et al.
Pubblicazione: (2026)
di: Kharlamova, Arina, et al.
Pubblicazione: (2026)
Generating Is Believing: Membership Inference Attacks against Retrieval-Augmented Generation
di: Li, Yuying, et al.
Pubblicazione: (2024)
di: Li, Yuying, et al.
Pubblicazione: (2024)
Differentially Private and Communication Efficient Large Language Model Split Inference via Stochastic Quantization and Soft Prompt
di: Gu, Yujie, et al.
Pubblicazione: (2026)
di: Gu, Yujie, et al.
Pubblicazione: (2026)
Revisiting Training-Inference Trigger Intensity in Backdoor Attacks
di: Lin, Chenhao, et al.
Pubblicazione: (2025)
di: Lin, Chenhao, et al.
Pubblicazione: (2025)
A Survey for Deep Reinforcement Learning Based Network Intrusion Detection
di: Yang, Wanrong, et al.
Pubblicazione: (2024)
di: Yang, Wanrong, et al.
Pubblicazione: (2024)
SecureRouter: Encrypted Routing for Efficient Secure Inference
di: Zhang, Yukuan, et al.
Pubblicazione: (2026)
di: Zhang, Yukuan, et al.
Pubblicazione: (2026)
A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations
di: Zhou, Yihe, et al.
Pubblicazione: (2025)
di: Zhou, Yihe, et al.
Pubblicazione: (2025)
PathSeeker: Exploring LLM Security Vulnerabilities with a Reinforcement Learning-Based Jailbreak Approach
di: Lin, Zhihao, et al.
Pubblicazione: (2024)
di: Lin, Zhihao, et al.
Pubblicazione: (2024)
Linearizing Models for Efficient yet Robust Private Inference
di: Sarkar, Sreetama, et al.
Pubblicazione: (2024)
di: Sarkar, Sreetama, et al.
Pubblicazione: (2024)
Differentially Private Federated Low Rank Adaptation Beyond Fixed-Matrix
di: Wen, Ming, et al.
Pubblicazione: (2025)
di: Wen, Ming, et al.
Pubblicazione: (2025)
A New Linear Scaling Rule for Private Adaptive Hyperparameter Optimization
di: Panda, Ashwinee, et al.
Pubblicazione: (2022)
di: Panda, Ashwinee, et al.
Pubblicazione: (2022)
A Survey on Data Security in Large Language Models
di: Chen, Kang, et al.
Pubblicazione: (2025)
di: Chen, Kang, et al.
Pubblicazione: (2025)
Deep Learning-based Intrusion Detection Systems: A Survey
di: Xu, Zhiwei, et al.
Pubblicazione: (2025)
di: Xu, Zhiwei, et al.
Pubblicazione: (2025)
You Told Me to Do It: Measuring Instructional Text-induced Private Data Leakage in LLM Agents
di: Kao, Ching-Yu, et al.
Pubblicazione: (2026)
di: Kao, Ching-Yu, et al.
Pubblicazione: (2026)
Dual-Priv Pruning : Efficient Differential Private Fine-Tuning in Multimodal Large Language Models
di: Wei, Qianshan, et al.
Pubblicazione: (2025)
di: Wei, Qianshan, et al.
Pubblicazione: (2025)
VaultGemma: A Differentially Private Gemma Model
di: Sinha, Amer, et al.
Pubblicazione: (2025)
di: Sinha, Amer, et al.
Pubblicazione: (2025)
Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
di: Gao, Fengyu, et al.
Pubblicazione: (2024)
di: Gao, Fengyu, et al.
Pubblicazione: (2024)
Data Poisoning in Deep Learning: A Survey
di: Zhao, Pinlong, et al.
Pubblicazione: (2025)
di: Zhao, Pinlong, et al.
Pubblicazione: (2025)
Fluent: Round-efficient Secure Aggregation for Private Federated Learning
di: Li, Xincheng, et al.
Pubblicazione: (2024)
di: Li, Xincheng, et al.
Pubblicazione: (2024)
Opal: Private Memory for Personal AI
di: Kaviani, Darya, et al.
Pubblicazione: (2026)
di: Kaviani, Darya, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Private Transformer Inference in MLaaS: A Survey
di: Li, Yang, et al.
Pubblicazione: (2025) -
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
di: Xu, Tianshi, et al.
Pubblicazione: (2024) -
Towards Secure and Private AI: A Framework for Decentralized Inference
di: Zhang, Hongyang, et al.
Pubblicazione: (2024) -
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference
di: Lin, Chenqi, et al.
Pubblicazione: (2024) -
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
di: Xu, Tianshi, et al.
Pubblicazione: (2024)