Accelerating Private Large Transformers Inference through Fine-grained Collaborative Computation
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Yuntian, Tang, Zhanyong, Lu, Tianpei, Zhang, Bingsheng, Shi, Zhiying, Wang, Zheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Privacy-Preserving Inference for Quantized BERT Models
by: Lu, Tianpei, et al.
Published: (2025)
by: Lu, Tianpei, et al.
Published: (2025)
The Communication-Friendly Privacy-Preserving Machine Learning against Malicious Adversaries
by: Lu, Tianpei, et al.
Published: (2024)
by: Lu, Tianpei, et al.
Published: (2024)
SuperEar: Eavesdropping on Mobile Voice Calls via Stealthy Acoustic Metamaterials
by: Ning, Zhiyuan, et al.
Published: (2025)
by: Ning, Zhiyuan, et al.
Published: (2025)
Comet: Accelerating Private Inference for Large Language Model by Predicting Activation Sparsity
by: Yan, Guang, et al.
Published: (2025)
by: Yan, Guang, et al.
Published: (2025)
Breaking the Layer Barrier: Remodeling Private Transformer Inference with Hybrid CKKS and MPC
by: Xu, Tianshi, et al.
Published: (2025)
by: Xu, Tianshi, et al.
Published: (2025)
PermLLM: Private Inference of Large Language Models within 3 Seconds under WAN
by: Zheng, Fei, et al.
Published: (2024)
by: Zheng, Fei, et al.
Published: (2024)
CipherPrune: Efficient and Scalable Private Transformer Inference
by: Zhang, Yancheng, et al.
Published: (2025)
by: Zhang, Yancheng, et al.
Published: (2025)
CryptPEFT: Efficient and Private Neural Network Inference via Parameter-Efficient Fine-Tuning
by: Xia, Saisai, et al.
Published: (2025)
by: Xia, Saisai, et al.
Published: (2025)
Differentially Private Subspace Fine-Tuning for Large Language Models
by: Zheng, Lele, et al.
Published: (2026)
by: Zheng, Lele, et al.
Published: (2026)
CipherFormer: Efficient Transformer Private Inference with Low Round Complexity
by: Wang, Weize, et al.
Published: (2024)
by: Wang, Weize, et al.
Published: (2024)
A Survey on Private Transformer Inference
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
A Portable and Stealthy Inaudible Voice Attack Based on Acoustic Metamaterials
by: Ning, Zhiyuan, et al.
Published: (2025)
by: Ning, Zhiyuan, et al.
Published: (2025)
zkVC: Fast Zero-Knowledge Proof for Private and Verifiable Computing
by: Zhang, Yancheng, et al.
Published: (2025)
by: Zhang, Yancheng, et al.
Published: (2025)
Differentially Private Parameter-Efficient Fine-tuning for Large ASR Models
by: Liu, Hongbin, et al.
Published: (2024)
by: Liu, Hongbin, et al.
Published: (2024)
PSA: Private Set Alignment for Secure and Collaborative Analytics on Large-Scale Data
by: Wang, Jiabo, et al.
Published: (2024)
by: Wang, Jiabo, et al.
Published: (2024)
Private Transformer Inference in MLaaS: A Survey
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
Practical Secure Inference Algorithm for Fine-tuned Large Language Model Based on Fully Homomorphic Encryption
by: Ruoyan, Zhang, et al.
Published: (2025)
by: Ruoyan, Zhang, et al.
Published: (2025)
Network and Compiler Optimizations for Efficient Linear Algebra Kernels in Private Transformer Inference
by: Garimella, Karthik, et al.
Published: (2025)
by: Garimella, Karthik, et al.
Published: (2025)
Moderator: Moderating Text-to-Image Diffusion Models through Fine-grained Context-based Policies
by: Wang, Peiran, et al.
Published: (2024)
by: Wang, Peiran, et al.
Published: (2024)
Kangaroo: A Private and Amortized Inference Framework over WAN for Large-Scale Decision Tree Evaluation
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
Private Collaborative Edge Inference via Over-the-Air Computation
by: Yilmaz, Selim F., et al.
Published: (2024)
by: Yilmaz, Selim F., et al.
Published: (2024)
Prompt Inversion Attack against Collaborative Inference of Large Language Models
by: Qu, Wenjie, et al.
Published: (2025)
by: Qu, Wenjie, et al.
Published: (2025)
Inferentially-Private Private Information
by: Wang, Shuaiqi, et al.
Published: (2024)
by: Wang, Shuaiqi, et al.
Published: (2024)
Fine-grained Data Access Control for Collaborative Process Execution on Blockchain
by: Marangone, Edoardo, et al.
Published: (2022)
by: Marangone, Edoardo, et al.
Published: (2022)
DP-SelFT: Differentially Private Selective Fine-Tuning for Large Language Models
by: Sha, Haichao, et al.
Published: (2026)
by: Sha, Haichao, et al.
Published: (2026)
Almost-Free Queue Jumping for Prior Inputs in Private Neural Inference
by: Zhang, Qiao, et al.
Published: (2026)
by: Zhang, Qiao, et al.
Published: (2026)
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
Fine-grained Manipulation Attacks to Local Differential Privacy Protocols for Data Streams
by: Li, Xinyu, et al.
Published: (2025)
by: Li, Xinyu, et al.
Published: (2025)
ProvAgent: Threat Detection Based on Identity-Behavior Binding and Multi-Agent Collaborative Attack Investigation
by: Yan, Wenhao, et al.
Published: (2026)
by: Yan, Wenhao, et al.
Published: (2026)
Styx: Collaborative and Private Data Processing With TEE-Enforced Sticky Policy
by: Zhao, Shixuan, et al.
Published: (2026)
by: Zhao, Shixuan, et al.
Published: (2026)
DP-SAPF: Saliency-Aware Parameter Fine-tuning of Public Models for Differentially Private Image Synthesis
by: Gong, Chen, et al.
Published: (2026)
by: Gong, Chen, et al.
Published: (2026)
CachePrune: Privacy-Aware and Fine-Grained KV Cache Sharing for Efficient LLM Inference
by: Wu, Guanlong, et al.
Published: (2026)
by: Wu, Guanlong, et al.
Published: (2026)
Private and Collaborative Kaplan-Meier Estimators
by: Rahimian, Shadi, et al.
Published: (2023)
by: Rahimian, Shadi, et al.
Published: (2023)
Dash: Accelerating Distributed Private Convolutional Neural Network Inference with Arithmetic Garbled Circuits
by: Sander, Jonas, et al.
Published: (2023)
by: Sander, Jonas, et al.
Published: (2023)
Private Fine-tuning of Large Language Models with Zeroth-order Optimization
by: Tang, Xinyu, et al.
Published: (2024)
by: Tang, Xinyu, et al.
Published: (2024)
MPCache: MPC-Friendly KV Cache Eviction for Efficient Private LLM Inference
by: Zeng, Wenxuan, et al.
Published: (2025)
by: Zeng, Wenxuan, et al.
Published: (2025)
Efficient and High-Accuracy Private CNN Inference with Helper-Assisted Malicious Security
by: Wang, Kaiwen, et al.
Published: (2025)
by: Wang, Kaiwen, et al.
Published: (2025)
Inner-product Functional Encryption with Fine-grained Revocation for Flexible EHR Sharing
by: Han, Yue, et al.
Published: (2025)
by: Han, Yue, et al.
Published: (2025)
Breaking Euston: Recovering Private Inputs from Secure Inference by Exploiting Subspace Leakage
by: Zhao, Jiaqi, et al.
Published: (2026)
by: Zhao, Jiaqi, et al.
Published: (2026)
RobPI: Robust Private Inference against Malicious Client
by: Xue, Jiaqi, et al.
Published: (2026)
by: Xue, Jiaqi, et al.
Published: (2026)
Similar Items
-
Privacy-Preserving Inference for Quantized BERT Models
by: Lu, Tianpei, et al.
Published: (2025) -
The Communication-Friendly Privacy-Preserving Machine Learning against Malicious Adversaries
by: Lu, Tianpei, et al.
Published: (2024) -
SuperEar: Eavesdropping on Mobile Voice Calls via Stealthy Acoustic Metamaterials
by: Ning, Zhiyuan, et al.
Published: (2025) -
Comet: Accelerating Private Inference for Large Language Model by Predicting Activation Sparsity
by: Yan, Guang, et al.
Published: (2025) -
Breaking the Layer Barrier: Remodeling Private Transformer Inference with Hybrid CKKS and MPC
by: Xu, Tianshi, et al.
Published: (2025)