SecureRouter: Encrypted Routing for Efficient Secure Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yukuan, Zheng, Mengxin, Lou, Qian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Privacy-Preserving LLMs Routing
von: Wu, Xidong, et al.
Veröffentlicht: (2026)
von: Wu, Xidong, et al.
Veröffentlicht: (2026)
Encrypted Prompt: Securing LLM Applications Against Unauthorized Actions
von: Chan, Shih-Han
Veröffentlicht: (2025)
von: Chan, Shih-Han
Veröffentlicht: (2025)
Life-Cycle Routing Vulnerabilities of LLM Router
von: Lin, Qiqi, et al.
Veröffentlicht: (2025)
von: Lin, Qiqi, et al.
Veröffentlicht: (2025)
Nimbus: Secure and Efficient Two-Party Inference for Transformers
von: Li, Zhengyi, et al.
Veröffentlicht: (2024)
von: Li, Zhengyi, et al.
Veröffentlicht: (2024)
DESIGN: Encrypted GNN Inference via Server-Side Input Graph Pruning
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)
On the (In-)Security of the Shuffling Defense in the Transformer Secure Inference
von: Li, Zhengyi, et al.
Veröffentlicht: (2026)
von: Li, Zhengyi, et al.
Veröffentlicht: (2026)
PIR-RAG: A System for Private Information Retrieval in Retrieval-Augmented Generation
von: Wang, Baiqiang, et al.
Veröffentlicht: (2025)
von: Wang, Baiqiang, et al.
Veröffentlicht: (2025)
ENSI: Efficient Non-Interactive Secure Inference for Large Language Models
von: He, Zhiyu, et al.
Veröffentlicht: (2025)
von: He, Zhiyu, et al.
Veröffentlicht: (2025)
A Scalable Multi-GPU Framework for Encrypted Large-Model Inference
von: Jayashankar, Siddharth, et al.
Veröffentlicht: (2025)
von: Jayashankar, Siddharth, et al.
Veröffentlicht: (2025)
A First Look At Efficient And Secure On-Device LLM Inference Against KV Leakage
von: Yang, Huan, et al.
Veröffentlicht: (2024)
von: Yang, Huan, et al.
Veröffentlicht: (2024)
SecMoE: Communication-Efficient Secure MoE Inference via Select-Then-Compute
von: Shen, Bowen, et al.
Veröffentlicht: (2026)
von: Shen, Bowen, et al.
Veröffentlicht: (2026)
Encrypted Large Model Inference: The Equivariant Encryption Paradigm
von: Buban, James, et al.
Veröffentlicht: (2025)
von: Buban, James, et al.
Veröffentlicht: (2025)
Efficient and Encrypted Inference using Binarized Neural Networks within In-Memory Computing Architectures
von: Rajendran, Gokulnath, et al.
Veröffentlicht: (2025)
von: Rajendran, Gokulnath, et al.
Veröffentlicht: (2025)
Tabula: Efficiently Computing Nonlinear Activation Functions for Secure Neural Network Inference
von: Lam, Maximilian, et al.
Veröffentlicht: (2022)
von: Lam, Maximilian, et al.
Veröffentlicht: (2022)
PADER: Paillier-based Secure Decentralized Social Recommendation
von: Chen, Chaochao, et al.
Veröffentlicht: (2026)
von: Chen, Chaochao, et al.
Veröffentlicht: (2026)
Memory-Efficient and Secure DNN Inference on TrustZone-enabled Consumer IoT Devices
von: Xie, Xueshuo, et al.
Veröffentlicht: (2024)
von: Xie, Xueshuo, et al.
Veröffentlicht: (2024)
SentinelLMs: Encrypted Input Adaptation and Fine-tuning of Language Models for Private and Secure Inference
von: Mishra, Abhijit, et al.
Veröffentlicht: (2023)
von: Mishra, Abhijit, et al.
Veröffentlicht: (2023)
Towards Secure and Private AI: A Framework for Decentralized Inference
von: Zhang, Hongyang, et al.
Veröffentlicht: (2024)
von: Zhang, Hongyang, et al.
Veröffentlicht: (2024)
AttestLLM: Efficient Attestation Framework for Billion-scale On-device LLMs
von: Zhang, Ruisi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruisi, et al.
Veröffentlicht: (2025)
CryptoGen: Secure Transformer Generation with Encrypted KV-Cache Reuse
von: Zhang, Hedong, et al.
Veröffentlicht: (2026)
von: Zhang, Hedong, et al.
Veröffentlicht: (2026)
MPC-Minimized Secure LLM Inference
von: Rathee, Deevashwer, et al.
Veröffentlicht: (2024)
von: Rathee, Deevashwer, et al.
Veröffentlicht: (2024)
CryptoTrain: Fast Secure Training on Encrypted Dataset
von: Xue, Jiaqi, et al.
Veröffentlicht: (2024)
von: Xue, Jiaqi, et al.
Veröffentlicht: (2024)
SecPE: Secure Prompt Ensembling for Private and Robust Large Language Models
von: Zhang, Jiawen, et al.
Veröffentlicht: (2025)
von: Zhang, Jiawen, et al.
Veröffentlicht: (2025)
Information-Dense Reasoning for Efficient and Auditable Security Alert Triage
von: Zhao, Guangze, et al.
Veröffentlicht: (2025)
von: Zhao, Guangze, et al.
Veröffentlicht: (2025)
MAS-Shield: A Defense Framework for Secure and Efficient LLM MAS
von: Wang, Kaixiang, et al.
Veröffentlicht: (2025)
von: Wang, Kaixiang, et al.
Veröffentlicht: (2025)
VeriSplit: Secure and Practical Offloading of Machine Learning Inferences across IoT Devices
von: Zhang, Han, et al.
Veröffentlicht: (2024)
von: Zhang, Han, et al.
Veröffentlicht: (2024)
EP-HDC: Hyperdimensional Computing with Encrypted Parameters for High-Throughput Privacy-Preserving Inference
von: Park, Jaewoo, et al.
Veröffentlicht: (2025)
von: Park, Jaewoo, et al.
Veröffentlicht: (2025)
Privacy-Preserving AI for Encrypted Medical Imaging: A Framework for Secure Diagnosis and Learning
von: Siam, Abdullah Al, et al.
Veröffentlicht: (2025)
von: Siam, Abdullah Al, et al.
Veröffentlicht: (2025)
FedMPQ: Secure and Communication-Efficient Federated Learning with Multi-codebook Product Quantization
von: Yang, Xu, et al.
Veröffentlicht: (2024)
von: Yang, Xu, et al.
Veröffentlicht: (2024)
Provably Secure Agent Guardrail
von: Wu, Benlong, et al.
Veröffentlicht: (2026)
von: Wu, Benlong, et al.
Veröffentlicht: (2026)
Towards Secure Agent Skills: Architecture, Threat Taxonomy, and Security Analysis
von: Li, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Li, Zhiyuan, et al.
Veröffentlicht: (2026)
Secure and Efficient Watermarking for Latent Diffusion Models in Model Distribution Scenarios
von: Lei, Liangqi, et al.
Veröffentlicht: (2025)
von: Lei, Liangqi, et al.
Veröffentlicht: (2025)
A New Era in LLM Security: Exploring Security Concerns in Real-World LLM-based Systems
von: Wu, Fangzhou, et al.
Veröffentlicht: (2024)
von: Wu, Fangzhou, et al.
Veröffentlicht: (2024)
CodeBC: A More Secure Large Language Model for Smart Contract Code Generation in Blockchain
von: Wang, Lingxiang, et al.
Veröffentlicht: (2025)
von: Wang, Lingxiang, et al.
Veröffentlicht: (2025)
DictPFL: Efficient and Private Federated Learning on Encrypted Gradients
von: Xue, Jiaqi, et al.
Veröffentlicht: (2025)
von: Xue, Jiaqi, et al.
Veröffentlicht: (2025)
On the Security of Research Artifacts
von: Rani, Nanda, et al.
Veröffentlicht: (2026)
von: Rani, Nanda, et al.
Veröffentlicht: (2026)
Security of AI Agents
von: He, Yifeng, et al.
Veröffentlicht: (2024)
von: He, Yifeng, et al.
Veröffentlicht: (2024)
AI Security Map: Holistic Organization of AI Security Technologies and Impacts on Stakeholders
von: Kato, Hiroya, et al.
Veröffentlicht: (2025)
von: Kato, Hiroya, et al.
Veröffentlicht: (2025)
Efficient Skip Connections Realization for Secure Inference on Encrypted Data
von: Drucker, Nir, et al.
Veröffentlicht: (2023)
von: Drucker, Nir, et al.
Veröffentlicht: (2023)
LLM Agents Should Employ Security Principles
von: Zhang, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Kaiyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Privacy-Preserving LLMs Routing
von: Wu, Xidong, et al.
Veröffentlicht: (2026) -
Encrypted Prompt: Securing LLM Applications Against Unauthorized Actions
von: Chan, Shih-Han
Veröffentlicht: (2025) -
Life-Cycle Routing Vulnerabilities of LLM Router
von: Lin, Qiqi, et al.
Veröffentlicht: (2025) -
Nimbus: Secure and Efficient Two-Party Inference for Transformers
von: Li, Zhengyi, et al.
Veröffentlicht: (2024) -
DESIGN: Encrypted GNN Inference via Server-Side Input Graph Pruning
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2025)