EQO: Exploring Ultra-Efficient Private Inference with Winograd-Based Protocol and Quantization Co-Optimization
Fuente:
arXiv
Guardado en:
| Autores principales: | Zeng, Wenxuan, Xu, Tianshi, Li, Meng, Wang, Runsheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization
por: Xu, Tianshi, et al.
Publicado: (2024)
por: Xu, Tianshi, et al.
Publicado: (2024)
UFO: Unlocking Ultra-Efficient Quantized Private Inference with Protocol and Algorithm Co-Optimization
por: Zeng, Wenxuan, et al.
Publicado: (2026)
por: Zeng, Wenxuan, et al.
Publicado: (2026)
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
por: Xu, Tianshi, et al.
Publicado: (2024)
por: Xu, Tianshi, et al.
Publicado: (2024)
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
por: Xu, Tianshi, et al.
Publicado: (2024)
por: Xu, Tianshi, et al.
Publicado: (2024)
MPCache: MPC-Friendly KV Cache Eviction for Efficient Private LLM Inference
por: Zeng, Wenxuan, et al.
Publicado: (2025)
por: Zeng, Wenxuan, et al.
Publicado: (2025)
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference
por: Lin, Chenqi, et al.
Publicado: (2024)
por: Lin, Chenqi, et al.
Publicado: (2024)
Breaking the Layer Barrier: Remodeling Private Transformer Inference with Hybrid CKKS and MPC
por: Xu, Tianshi, et al.
Publicado: (2025)
por: Xu, Tianshi, et al.
Publicado: (2025)
Towards Efficient Privacy-Preserving Machine Learning: A Systematic Review from Protocol, Model, and System Perspectives
por: Zeng, Wenxuan, et al.
Publicado: (2025)
por: Zeng, Wenxuan, et al.
Publicado: (2025)
CryptoMoE: Privacy-Preserving and Scalable Mixture of Experts Inference via Balanced Expert Routing
por: Zhou, Yifan, et al.
Publicado: (2025)
por: Zhou, Yifan, et al.
Publicado: (2025)
CryptPEFT: Efficient and Private Neural Network Inference via Parameter-Efficient Fine-Tuning
por: Xia, Saisai, et al.
Publicado: (2025)
por: Xia, Saisai, et al.
Publicado: (2025)
Vectorised Hashing Based on Bernstein-Rabin-Winograd Polynomials over Prime Order Fields
por: Nath, Kaushik, et al.
Publicado: (2025)
por: Nath, Kaushik, et al.
Publicado: (2025)
Network and Compiler Optimizations for Efficient Linear Algebra Kernels in Private Transformer Inference
por: Garimella, Karthik, et al.
Publicado: (2025)
por: Garimella, Karthik, et al.
Publicado: (2025)
Differentially Private and Communication Efficient Large Language Model Split Inference via Stochastic Quantization and Soft Prompt
por: Gu, Yujie, et al.
Publicado: (2026)
por: Gu, Yujie, et al.
Publicado: (2026)
CipherFormer: Efficient Transformer Private Inference with Low Round Complexity
por: Wang, Weize, et al.
Publicado: (2024)
por: Wang, Weize, et al.
Publicado: (2024)
Publicly Verifiable Private Information Retrieval Protocols Based on Function Secret Sharing
por: Zhu, Lin, et al.
Publicado: (2025)
por: Zhu, Lin, et al.
Publicado: (2025)
Differentially Private Diffusion Models
por: Dockhorn, Tim, et al.
Publicado: (2022)
por: Dockhorn, Tim, et al.
Publicado: (2022)
Efficient and High-Accuracy Private CNN Inference with Helper-Assisted Malicious Security
por: Wang, Kaiwen, et al.
Publicado: (2025)
por: Wang, Kaiwen, et al.
Publicado: (2025)
Hyena: Optimizing Homomorphically Encrypted Convolution for Private CNN Inference
por: Roh, Hyeri, et al.
Publicado: (2023)
por: Roh, Hyeri, et al.
Publicado: (2023)
Flash: A Hybrid Private Inference Protocol for Deep CNNs with High Accuracy and Low Latency on CPU
por: Roh, Hyeri, et al.
Publicado: (2024)
por: Roh, Hyeri, et al.
Publicado: (2024)
Optimized Layerwise Approximation for Efficient Private Inference on Fully Homomorphic Encryption
por: Lee, Junghyun, et al.
Publicado: (2023)
por: Lee, Junghyun, et al.
Publicado: (2023)
DeepReShape: Redesigning Neural Networks for Efficient Private Inference
por: Jha, Nandan Kumar, et al.
Publicado: (2023)
por: Jha, Nandan Kumar, et al.
Publicado: (2023)
Lightweight Protocols for Distributed Private Quantile Estimation
por: Aamand, Anders, et al.
Publicado: (2025)
por: Aamand, Anders, et al.
Publicado: (2025)
Multi-Party Private Set Intersection: A Circuit-Based Protocol with Jaccard Similarity for Secure and Efficient Anomaly Detection in Network Traffic
por: Su, Jiuheng, et al.
Publicado: (2024)
por: Su, Jiuheng, et al.
Publicado: (2024)
Almost-Free Queue Jumping for Prior Inputs in Private Neural Inference
por: Zhang, Qiao, et al.
Publicado: (2026)
por: Zhang, Qiao, et al.
Publicado: (2026)
CipherPrune: Efficient and Scalable Private Transformer Inference
por: Zhang, Yancheng, et al.
Publicado: (2025)
por: Zhang, Yancheng, et al.
Publicado: (2025)
HE-LRM: Efficient Private Embedding Lookups for Neural Inference Using Fully Homomorphic Encryption
por: Garimella, Karthik, et al.
Publicado: (2025)
por: Garimella, Karthik, et al.
Publicado: (2025)
A Survey on Private Transformer Inference
por: Li, Yang, et al.
Publicado: (2024)
por: Li, Yang, et al.
Publicado: (2024)
An Improved Quantum Private Set Intersection Protocol Based on Hadamard Gates
por: Liu, Wenjie, et al.
Publicado: (2023)
por: Liu, Wenjie, et al.
Publicado: (2023)
Towards Reliable and Generalizable Differentially Private Machine Learning (Extended Version)
por: Bao, Wenxuan, et al.
Publicado: (2025)
por: Bao, Wenxuan, et al.
Publicado: (2025)
Fast and Private Inference of Deep Neural Networks by Co-designing Activation Functions
por: Diaa, Abdulrahman, et al.
Publicado: (2023)
por: Diaa, Abdulrahman, et al.
Publicado: (2023)
Computationally Differentially Private Inner Product Protocols Imply Oblivious Transfer
por: Haitner, Iftach, et al.
Publicado: (2025)
por: Haitner, Iftach, et al.
Publicado: (2025)
Revealing the True Cost of Locally Differentially Private Protocols: An Auditing Perspective
por: Arcolezi, Héber H., et al.
Publicado: (2023)
por: Arcolezi, Héber H., et al.
Publicado: (2023)
Efficient Fuzzy Private Set Intersection from Secret-shared OPRF
por: Yang, Xinpeng, et al.
Publicado: (2026)
por: Yang, Xinpeng, et al.
Publicado: (2026)
Exploring the Robustness and Transferability of Patch-Based Adversarial Attacks in Quantized Neural Networks
por: Guesmi, Amira, et al.
Publicado: (2024)
por: Guesmi, Amira, et al.
Publicado: (2024)
Automatic State Machine Inference for Binary Protocol Reverse Engineering
por: Yang, Junhai, et al.
Publicado: (2024)
por: Yang, Junhai, et al.
Publicado: (2024)
Private Transformer Inference in MLaaS: A Survey
por: Li, Yang, et al.
Publicado: (2025)
por: Li, Yang, et al.
Publicado: (2025)
Data Poisoning Attacks to Locally Differentially Private Frequent Itemset Mining Protocols
por: Tong, Wei, et al.
Publicado: (2024)
por: Tong, Wei, et al.
Publicado: (2024)
AdaPI: Facilitating DNN Model Adaptivity for Efficient Private Inference in Edge Computing
por: Zhou, Tong, et al.
Publicado: (2024)
por: Zhou, Tong, et al.
Publicado: (2024)
Secure Transformer Inference Protocol
por: Yuan, Mu, et al.
Publicado: (2023)
por: Yuan, Mu, et al.
Publicado: (2023)
Kangaroo: A Private and Amortized Inference Framework over WAN for Large-Scale Decision Tree Evaluation
por: Xu, Wei, et al.
Publicado: (2025)
por: Xu, Wei, et al.
Publicado: (2025)
Ejemplares similares
-
PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization
por: Xu, Tianshi, et al.
Publicado: (2024) -
UFO: Unlocking Ultra-Efficient Quantized Private Inference with Protocol and Algorithm Co-Optimization
por: Zeng, Wenxuan, et al.
Publicado: (2026) -
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
por: Xu, Tianshi, et al.
Publicado: (2024) -
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
por: Xu, Tianshi, et al.
Publicado: (2024) -
MPCache: MPC-Friendly KV Cache Eviction for Efficient Private LLM Inference
por: Zeng, Wenxuan, et al.
Publicado: (2025)