Breaking the Layer Barrier: Remodeling Private Transformer Inference with Hybrid CKKS and MPC
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Tianshi, Lu, Wen-jie, Yu, Jiangrui, Yi, Chen, Lin, Chenqi, Wang, Runsheng, Li, Meng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference
by: Lin, Chenqi, et al.
Published: (2024)
by: Lin, Chenqi, et al.
Published: (2024)
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
EQO: Exploring Ultra-Efficient Private Inference with Winograd-Based Protocol and Quantization Co-Optimization
by: Zeng, Wenxuan, et al.
Published: (2024)
by: Zeng, Wenxuan, et al.
Published: (2024)
PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
MPCache: MPC-Friendly KV Cache Eviction for Efficient Private LLM Inference
by: Zeng, Wenxuan, et al.
Published: (2025)
by: Zeng, Wenxuan, et al.
Published: (2025)
UFO: Unlocking Ultra-Efficient Quantized Private Inference with Protocol and Algorithm Co-Optimization
by: Zeng, Wenxuan, et al.
Published: (2026)
by: Zeng, Wenxuan, et al.
Published: (2026)
Efficient Mod Approximation and Its Applications to CKKS Ciphertexts
by: Zhou, Yufei
Published: (2025)
by: Zhou, Yufei
Published: (2025)
Efficient Ranking, Order Statistics, and Sorting under CKKS
by: Mazzone, Federico, et al.
Published: (2024)
by: Mazzone, Federico, et al.
Published: (2024)
CipherFormer: Efficient Transformer Private Inference with Low Round Complexity
by: Wang, Weize, et al.
Published: (2024)
by: Wang, Weize, et al.
Published: (2024)
FHE-Agent: Automating CKKS Configuration for Practical Encrypted Inference via an LLM-Guided Agentic Framework
by: Xu, Nuo, et al.
Published: (2025)
by: Xu, Nuo, et al.
Published: (2025)
CryptoMoE: Privacy-Preserving and Scalable Mixture of Experts Inference via Balanced Expert Routing
by: Zhou, Yifan, et al.
Published: (2025)
by: Zhou, Yifan, et al.
Published: (2025)
Accelerating Private Large Transformers Inference through Fine-grained Collaborative Computation
by: Chen, Yuntian, et al.
Published: (2024)
by: Chen, Yuntian, et al.
Published: (2024)
Triple-Hoisted Baby-Step Giant-Step Linear Transformation over CKKS Homomorphic Encryption and Hardware Accelerator
by: Akherati, Sajjad, et al.
Published: (2026)
by: Akherati, Sajjad, et al.
Published: (2026)
PRIVMARK: Private Large Language Models Watermarking with MPC
by: Fargues, Thomas, et al.
Published: (2025)
by: Fargues, Thomas, et al.
Published: (2025)
Nimbus: Secure and Efficient Two-Party Inference for Transformers
by: Li, Zhengyi, et al.
Published: (2024)
by: Li, Zhengyi, et al.
Published: (2024)
EinHops: Einsum Notation for Expressive Homomorphic Operations on RNS-CKKS Tensors
by: Garimella, Karthik, et al.
Published: (2025)
by: Garimella, Karthik, et al.
Published: (2025)
Characterizing the Sensitivity to Individual Bit Flips in Client-Side Operations of the CKKS Scheme
by: Mazzanti, Matias, et al.
Published: (2025)
by: Mazzanti, Matias, et al.
Published: (2025)
Resource Estimation of CGGI and CKKS scheme workloads on FracTLcore Computing Fabric
by: Ovichinnikov, Denis, et al.
Published: (2025)
by: Ovichinnikov, Denis, et al.
Published: (2025)
Breaking Euston: Recovering Private Inputs from Secure Inference by Exploiting Subspace Leakage
by: Zhao, Jiaqi, et al.
Published: (2026)
by: Zhao, Jiaqi, et al.
Published: (2026)
Taiyi: A high-performance CKKS accelerator for Practical Fully Homomorphic Encryption
by: Fan, Shengyu, et al.
Published: (2024)
by: Fan, Shengyu, et al.
Published: (2024)
FIDESlib: A Fully-Fledged Open-Source FHE Library for Efficient CKKS on GPUs
by: Agulló-Domingo, Carlos, et al.
Published: (2025)
by: Agulló-Domingo, Carlos, et al.
Published: (2025)
Ditto: Quantization-aware Secure Inference of Transformers upon MPC
by: Wu, Haoqi, et al.
Published: (2024)
by: Wu, Haoqi, et al.
Published: (2024)
Scaling up Privacy-Preserving ML: A CKKS Implementation of Llama-2-7B
by: Park, Jaiyoung, et al.
Published: (2026)
by: Park, Jaiyoung, et al.
Published: (2026)
Differentially Private Diffusion Models
by: Dockhorn, Tim, et al.
Published: (2022)
by: Dockhorn, Tim, et al.
Published: (2022)
A Survey on Private Transformer Inference
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
Breaking Free: Efficient Multi-Party Private Set Union Without Non-Collusion Assumptions
by: Dong, Minglang, et al.
Published: (2024)
by: Dong, Minglang, et al.
Published: (2024)
PUMA: Secure Inference of LLaMA-7B in Five Minutes
by: Dong, Ye, et al.
Published: (2023)
by: Dong, Ye, et al.
Published: (2023)
Network and Compiler Optimizations for Efficient Linear Algebra Kernels in Private Transformer Inference
by: Garimella, Karthik, et al.
Published: (2025)
by: Garimella, Karthik, et al.
Published: (2025)
Large-Scale MPC: Scaling Private Iris Code Uniqueness Checks to Millions of Users
by: Bloemen, Remco, et al.
Published: (2024)
by: Bloemen, Remco, et al.
Published: (2024)
An Attack to Break Permutation-Based Private Third-Party Inference Schemes for LLMs
by: Thomas, Rahul, et al.
Published: (2025)
by: Thomas, Rahul, et al.
Published: (2025)
Private Transformer Inference in MLaaS: A Survey
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
Robust and Verifiable MPC with Applications to Linear Machine Learning Inference
by: Wang, Tzu-Shen, et al.
Published: (2025)
by: Wang, Tzu-Shen, et al.
Published: (2025)
LRD-MPC: Efficient MPC Inference through Low-rank Decomposition
by: Tang, Tingting, et al.
Published: (2026)
by: Tang, Tingting, et al.
Published: (2026)
CipherPrune: Efficient and Scalable Private Transformer Inference
by: Zhang, Yancheng, et al.
Published: (2025)
by: Zhang, Yancheng, et al.
Published: (2025)
Breaking 5G on The Lower Layer
by: Shanto, Subangkar Karmaker, et al.
Published: (2026)
by: Shanto, Subangkar Karmaker, et al.
Published: (2026)
Practical and Private Hybrid ML Inference with Fully Homomorphic Encryption
by: Biswas, Sayan, et al.
Published: (2025)
by: Biswas, Sayan, et al.
Published: (2025)
Flash: A Hybrid Private Inference Protocol for Deep CNNs with High Accuracy and Low Latency on CPU
by: Roh, Hyeri, et al.
Published: (2024)
by: Roh, Hyeri, et al.
Published: (2024)
Almost-Free Queue Jumping for Prior Inputs in Private Neural Inference
by: Zhang, Qiao, et al.
Published: (2026)
by: Zhang, Qiao, et al.
Published: (2026)
SelectFormer: Private and Practical Data Selection for Transformers
by: Ouyang, Xu, et al.
Published: (2023)
by: Ouyang, Xu, et al.
Published: (2023)
Similar Items
-
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference
by: Lin, Chenqi, et al.
Published: (2024) -
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
by: Xu, Tianshi, et al.
Published: (2024) -
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
by: Xu, Tianshi, et al.
Published: (2024) -
EQO: Exploring Ultra-Efficient Private Inference with Winograd-Based Protocol and Quantization Co-Optimization
by: Zeng, Wenxuan, et al.
Published: (2024) -
PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization
by: Xu, Tianshi, et al.
Published: (2024)