Your Inference Request Will Become a Black Box: Confidential Inference for Cloud-based Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Chung-ju, Zhao, Huiqiang, He, Yuanpeng, Li, Lijian, Jiao, Wenpin, Jin, Zhi, Chen, Peixuan, Wang, Leye |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Confidential Prompting: Privacy-preserving LLM Inference on Cloud
von: Li, Caihua, et al.
Veröffentlicht: (2024)
von: Li, Caihua, et al.
Veröffentlicht: (2024)
Learning-based Privacy-Preserving Graph Publishing Against Sensitive Link Inference Attacks
von: Wu, Yucheng, et al.
Veröffentlicht: (2025)
von: Wu, Yucheng, et al.
Veröffentlicht: (2025)
Towards Confidential and Efficient LLM Inference with Dual Privacy Protection
von: Yu, Honglan, et al.
Veröffentlicht: (2025)
von: Yu, Honglan, et al.
Veröffentlicht: (2025)
Prompt Inference Attack on Distributed Large Language Model Inference Frameworks
von: Luo, Xinjian, et al.
Veröffentlicht: (2025)
von: Luo, Xinjian, et al.
Veröffentlicht: (2025)
E-MIA: Exam-Style Black-Box Membership Inference Attacks against RAG Systems
von: Guan, Zelin, et al.
Veröffentlicht: (2026)
von: Guan, Zelin, et al.
Veröffentlicht: (2026)
Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
CRISP: Confidentiality, Rollback, and Integrity Storage Protection for Confidential Cloud-Native Computing
von: Hartono, Ardhi Putra Pratama, et al.
Veröffentlicht: (2024)
von: Hartono, Ardhi Putra Pratama, et al.
Veröffentlicht: (2024)
Arca: A Lightweight Confidential Container Architecture for Cloud-Native Environments
von: Lu, Di, et al.
Veröffentlicht: (2026)
von: Lu, Di, et al.
Veröffentlicht: (2026)
Black-Box Membership Inference Attack for LVLMs via Prior Knowledge-Calibrated Memory Probing
von: Yin, Jinhua, et al.
Veröffentlicht: (2025)
von: Yin, Jinhua, et al.
Veröffentlicht: (2025)
UniTrans: A Unified Vertical Federated Knowledge Transfer Framework for Enhancing Cross-Hospital Collaboration
von: Huang, Chung-ju, et al.
Veröffentlicht: (2025)
von: Huang, Chung-ju, et al.
Veröffentlicht: (2025)
It's a Feature, Not a Bug: Secure and Auditable State Rollback for Confidential Cloud Applications
von: Burke, Quinn, et al.
Veröffentlicht: (2025)
von: Burke, Quinn, et al.
Veröffentlicht: (2025)
Enabling Performant and Secure EDA as a Service in Public Clouds Using Confidential Containers
von: Ye, Mengmei, et al.
Veröffentlicht: (2024)
von: Ye, Mengmei, et al.
Veröffentlicht: (2024)
Membership Inference Attacks Against Video Large Language Models
von: Song, Wei, et al.
Veröffentlicht: (2026)
von: Song, Wei, et al.
Veröffentlicht: (2026)
HasTEE+ : Confidential Cloud Computing and Analytics with Haskell
von: Sarkar, Abhiroop, et al.
Veröffentlicht: (2024)
von: Sarkar, Abhiroop, et al.
Veröffentlicht: (2024)
Towards Black-Box Membership Inference Attack for Diffusion Models
von: Li, Jingwei, et al.
Veröffentlicht: (2024)
von: Li, Jingwei, et al.
Veröffentlicht: (2024)
WebWeaver: Breaking Topology Confidentiality in LLM Multi-Agent Systems with Stealthy Context-Based Inference
von: Xiong, Zixun, et al.
Veröffentlicht: (2026)
von: Xiong, Zixun, et al.
Veröffentlicht: (2026)
The Challenge of Identifying the Origin of Black-Box Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
When Reasoning Leaks Membership: Membership Inference Attack on Black-box Large Reasoning Models
von: Hu, Ruihan, et al.
Veröffentlicht: (2026)
von: Hu, Ruihan, et al.
Veröffentlicht: (2026)
Variational Autoencoder-Based Black-Box Adversarial Attack on Collaborative DNN Inference
von: Yousefi, Shima, et al.
Veröffentlicht: (2025)
von: Yousefi, Shima, et al.
Veröffentlicht: (2025)
Prompt Inversion Attack against Collaborative Inference of Large Language Models
von: Qu, Wenjie, et al.
Veröffentlicht: (2025)
von: Qu, Wenjie, et al.
Veröffentlicht: (2025)
Confidential Wrapped Ethereum
von: Chystiakov, Artem, et al.
Veröffentlicht: (2025)
von: Chystiakov, Artem, et al.
Veröffentlicht: (2025)
Tempo: Confidentiality Preservation in Cloud-Based Neural Network Training
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
OnePath: Efficient and Privacy-Preserving Decision Tree Inference in the Cloud
von: Yuan, Shuai, et al.
Veröffentlicht: (2024)
von: Yuan, Shuai, et al.
Veröffentlicht: (2024)
Confidential LLM Inference: Performance and Cost Across CPU and GPU TEEs
von: Chrapek, Marcin, et al.
Veröffentlicht: (2025)
von: Chrapek, Marcin, et al.
Veröffentlicht: (2025)
Confidential Computing for Cloud Security: Exploring Hardware based Encryption Using Trusted Execution Environments
von: Agarwal, Dhruv Deepak, et al.
Veröffentlicht: (2025)
von: Agarwal, Dhruv Deepak, et al.
Veröffentlicht: (2025)
Thinking Inside The Box: Privacy Against Stronger Adversaries
von: Chung, Eldon
Veröffentlicht: (2024)
von: Chung, Eldon
Veröffentlicht: (2024)
InferDPT: Privacy-Preserving Inference for Closed-box Large Language Model
von: Tong, Meng, et al.
Veröffentlicht: (2023)
von: Tong, Meng, et al.
Veröffentlicht: (2023)
ENSI: Efficient Non-Interactive Secure Inference for Large Language Models
von: He, Zhiyu, et al.
Veröffentlicht: (2025)
von: He, Zhiyu, et al.
Veröffentlicht: (2025)
Fingerprinting Inference Systems of Large Language Models
von: Wimbauer, Anna, et al.
Veröffentlicht: (2026)
von: Wimbauer, Anna, et al.
Veröffentlicht: (2026)
Trustworthy and Confidential SBOM Exchange
von: Ishgair, Eman Abu, et al.
Veröffentlicht: (2025)
von: Ishgair, Eman Abu, et al.
Veröffentlicht: (2025)
Cachemir: Fully Homomorphic Encrypted Inference of Generative Large Language Model with KV Cache
von: Yu, Ye, et al.
Veröffentlicht: (2026)
von: Yu, Ye, et al.
Veröffentlicht: (2026)
Balancing Confidentiality and Transparency for Blockchain-based Process-Aware Information Systems
von: Marcelletti, Alessandro, et al.
Veröffentlicht: (2024)
von: Marcelletti, Alessandro, et al.
Veröffentlicht: (2024)
EvoDefense: Co-Evolving Black-Box Defense with Large Language Models
von: Li, Yu, et al.
Veröffentlicht: (2026)
von: Li, Yu, et al.
Veröffentlicht: (2026)
Distilled Large Language Model in Confidential Computing Environment for System-on-Chip Design
von: Ben, Dong, et al.
Veröffentlicht: (2025)
von: Ben, Dong, et al.
Veröffentlicht: (2025)
Inducing Overthink: Hierarchical Genetic Algorithm-based DoS Attack on Black-Box Large Language Reasoning Models
von: Wang, Shuqiang, et al.
Veröffentlicht: (2026)
von: Wang, Shuqiang, et al.
Veröffentlicht: (2026)
Towards Privacy-Preserving Split Learning: Destabilizing Adversarial Inference and Reconstruction Attacks in the Cloud
von: Higgins, Griffin, et al.
Veröffentlicht: (2025)
von: Higgins, Griffin, et al.
Veröffentlicht: (2025)
Automated Profile Inference with Language Model Agents
von: Du, Yuntao, et al.
Veröffentlicht: (2025)
von: Du, Yuntao, et al.
Veröffentlicht: (2025)
Membership Inference Attacks on Tokenizers of Large Language Models
von: Tong, Meng, et al.
Veröffentlicht: (2025)
von: Tong, Meng, et al.
Veröffentlicht: (2025)
Differentially Private and Communication Efficient Large Language Model Split Inference via Stochastic Quantization and Soft Prompt
von: Gu, Yujie, et al.
Veröffentlicht: (2026)
von: Gu, Yujie, et al.
Veröffentlicht: (2026)
PermLLM: Private Inference of Large Language Models within 3 Seconds under WAN
von: Zheng, Fei, et al.
Veröffentlicht: (2024)
von: Zheng, Fei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Confidential Prompting: Privacy-preserving LLM Inference on Cloud
von: Li, Caihua, et al.
Veröffentlicht: (2024) -
Learning-based Privacy-Preserving Graph Publishing Against Sensitive Link Inference Attacks
von: Wu, Yucheng, et al.
Veröffentlicht: (2025) -
Towards Confidential and Efficient LLM Inference with Dual Privacy Protection
von: Yu, Honglan, et al.
Veröffentlicht: (2025) -
Prompt Inference Attack on Distributed Large Language Model Inference Frameworks
von: Luo, Xinjian, et al.
Veröffentlicht: (2025) -
E-MIA: Exam-Style Black-Box Membership Inference Attacks against RAG Systems
von: Guan, Zelin, et al.
Veröffentlicht: (2026)