PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Tianshi, Zhong, Shuzhang, Zeng, Wenxuan, Wang, Runsheng, Li, Meng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
by: Xu, Tianshi, et al.
Published: (2024)
by: Xu, Tianshi, et al.
Published: (2024)
EQO: Exploring Ultra-Efficient Private Inference with Winograd-Based Protocol and Quantization Co-Optimization
by: Zeng, Wenxuan, et al.
Published: (2024)
by: Zeng, Wenxuan, et al.
Published: (2024)
UFO: Unlocking Ultra-Efficient Quantized Private Inference with Protocol and Algorithm Co-Optimization
by: Zeng, Wenxuan, et al.
Published: (2026)
by: Zeng, Wenxuan, et al.
Published: (2026)
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference
by: Lin, Chenqi, et al.
Published: (2024)
by: Lin, Chenqi, et al.
Published: (2024)
Towards Efficient Privacy-Preserving Machine Learning: A Systematic Review from Protocol, Model, and System Perspectives
by: Zeng, Wenxuan, et al.
Published: (2025)
by: Zeng, Wenxuan, et al.
Published: (2025)
Dual-Priv Pruning : Efficient Differential Private Fine-Tuning in Multimodal Large Language Models
by: Wei, Qianshan, et al.
Published: (2025)
by: Wei, Qianshan, et al.
Published: (2025)
Differentially Private and Communication Efficient Large Language Model Split Inference via Stochastic Quantization and Soft Prompt
by: Gu, Yujie, et al.
Published: (2026)
by: Gu, Yujie, et al.
Published: (2026)
PrivSpike: Employing Homomorphic Encryption for Private Inference of Deep Spiking Neural Networks
by: Njungle, Nges Brian, et al.
Published: (2025)
by: Njungle, Nges Brian, et al.
Published: (2025)
Optimized Layerwise Approximation for Efficient Private Inference on Fully Homomorphic Encryption
by: Lee, Junghyun, et al.
Published: (2023)
by: Lee, Junghyun, et al.
Published: (2023)
MPCache: MPC-Friendly KV Cache Eviction for Efficient Private LLM Inference
by: Zeng, Wenxuan, et al.
Published: (2025)
by: Zeng, Wenxuan, et al.
Published: (2025)
Breaking the Layer Barrier: Remodeling Private Transformer Inference with Hybrid CKKS and MPC
by: Xu, Tianshi, et al.
Published: (2025)
by: Xu, Tianshi, et al.
Published: (2025)
PrivScope: Task-scoped Disclosure Control for Hybrid Agentic Systems
by: Seeam, Shafizur Rahman, et al.
Published: (2026)
by: Seeam, Shafizur Rahman, et al.
Published: (2026)
A Survey on Private Transformer Inference
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
FedMPQ: Secure and Communication-Efficient Federated Learning with Multi-codebook Product Quantization
by: Yang, Xu, et al.
Published: (2024)
by: Yang, Xu, et al.
Published: (2024)
Private Transformer Inference in MLaaS: A Survey
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
PrivSTRUCT: Untangling Data Purpose Compliance of Privacy Policies in Google Play Store
by: Silva, Bhanuka, et al.
Published: (2026)
by: Silva, Bhanuka, et al.
Published: (2026)
PrivLLMSwarm: Privacy-Preserving LLM-Driven UAV Swarms for Secure IoT Surveillance
by: Ayana, Jifar Wakuma, et al.
Published: (2025)
by: Ayana, Jifar Wakuma, et al.
Published: (2025)
Comet: A Communication-efficient and Performant Approximation for Private Transformer Inference
by: Xu, Xiangrui, et al.
Published: (2024)
by: Xu, Xiangrui, et al.
Published: (2024)
Not My Agent, Not My Boundary? Elicitation of Personal Privacy Boundaries in AI-Delegated Information Sharing
by: Guo, Bingcan, et al.
Published: (2025)
by: Guo, Bingcan, et al.
Published: (2025)
Agentic LLMs as Powerful Deanonymizers: Re-identification of Participants in the Anthropic Interviewer Dataset
by: Li, Tianshi
Published: (2026)
by: Li, Tianshi
Published: (2026)
Comet: Accelerating Private Inference for Large Language Model by Predicting Activation Sparsity
by: Yan, Guang, et al.
Published: (2025)
by: Yan, Guang, et al.
Published: (2025)
PrivComp-KG : Leveraging Knowledge Graph and Large Language Models for Privacy Policy Compliance Verification
by: Garza, Leon, et al.
Published: (2024)
by: Garza, Leon, et al.
Published: (2024)
Linearizing Models for Efficient yet Robust Private Inference
by: Sarkar, Sreetama, et al.
Published: (2024)
by: Sarkar, Sreetama, et al.
Published: (2024)
Towards Secure and Private AI: A Framework for Decentralized Inference
by: Zhang, Hongyang, et al.
Published: (2024)
by: Zhang, Hongyang, et al.
Published: (2024)
PrivTune: Efficient and Privacy-Preserving Fine-Tuning of Large Language Models via Device-Cloud Collaboration
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
Ents: An Efficient Three-party Training Framework for Decision Trees by Communication Optimization
by: Lin, Guopeng, et al.
Published: (2024)
by: Lin, Guopeng, et al.
Published: (2024)
SecMoE: Communication-Efficient Secure MoE Inference via Select-Then-Compute
by: Shen, Bowen, et al.
Published: (2026)
by: Shen, Bowen, et al.
Published: (2026)
Tabula: Efficiently Computing Nonlinear Activation Functions for Secure Neural Network Inference
by: Lam, Maximilian, et al.
Published: (2022)
by: Lam, Maximilian, et al.
Published: (2022)
Quant Fever, Reasoning Blackholes, Schrodinger's Compliance, and More: Probing GPT-OSS-20B
by: Lin, Shuyi, et al.
Published: (2025)
by: Lin, Shuyi, et al.
Published: (2025)
SUDP: Secret-Use Delegation Protocol for Agentic Systems
by: Yu, Xiaohang, et al.
Published: (2026)
by: Yu, Xiaohang, et al.
Published: (2026)
Efficient and Encrypted Inference using Binarized Neural Networks within In-Memory Computing Architectures
by: Rajendran, Gokulnath, et al.
Published: (2025)
by: Rajendran, Gokulnath, et al.
Published: (2025)
Memory-Efficient and Secure DNN Inference on TrustZone-enabled Consumer IoT Devices
by: Xie, Xueshuo, et al.
Published: (2024)
by: Xie, Xueshuo, et al.
Published: (2024)
Exposing Hidden Interfaces: LLM-Guided Type Inference for Reverse Engineering macOS Private Frameworks
by: Kharlamova, Arina, et al.
Published: (2026)
by: Kharlamova, Arina, et al.
Published: (2026)
MCP Security Bench (MSB): Benchmarking Attacks Against Model Context Protocol in LLM Agents
by: Zhang, Dongsen, et al.
Published: (2025)
by: Zhang, Dongsen, et al.
Published: (2025)
Nimbus: Secure and Efficient Two-Party Inference for Transformers
by: Li, Zhengyi, et al.
Published: (2024)
by: Li, Zhengyi, et al.
Published: (2024)
Membership Inference Attacks on Tokenizers of Large Language Models
by: Tong, Meng, et al.
Published: (2025)
by: Tong, Meng, et al.
Published: (2025)
OptMark: Robust Multi-bit Diffusion Watermarking via Inference Time Optimization
by: Xing, Jiazheng, et al.
Published: (2025)
by: Xing, Jiazheng, et al.
Published: (2025)
Towards Anonymous Neural Network Inference
by: Peiyuan, Liao
Published: (2025)
by: Peiyuan, Liao
Published: (2025)
PACZero: PAC-Private Fine-Tuning of Language Models via Sign Quantization
by: Ertan, Murat Bilgehan, et al.
Published: (2026)
by: Ertan, Murat Bilgehan, et al.
Published: (2026)
Similar Items
-
PrivCirNet: Efficient Private Inference via Block Circulant Transformation
by: Xu, Tianshi, et al.
Published: (2024) -
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference
by: Xu, Tianshi, et al.
Published: (2024) -
EQO: Exploring Ultra-Efficient Private Inference with Winograd-Based Protocol and Quantization Co-Optimization
by: Zeng, Wenxuan, et al.
Published: (2024) -
UFO: Unlocking Ultra-Efficient Quantized Private Inference with Protocol and Algorithm Co-Optimization
by: Zeng, Wenxuan, et al.
Published: (2026) -
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference
by: Lin, Chenqi, et al.
Published: (2024)