GPU Acceleration of TFHE-Based High-Precision Nonlinear Layers for Encrypted LLM Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Guoci, Pan, Xiurui, Li, Qiao, Mao, Bo, Gao, Congming, Huan, Chengying, Zhang, Mingzhe, Zhang, Jie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OFHE: An Electro-Optical Accelerator for Discretized TFHE
von: Zheng, Mengxin, et al.
Veröffentlicht: (2024)
von: Zheng, Mengxin, et al.
Veröffentlicht: (2024)
NTTSuite: Number Theoretic Transform Benchmarks for Accelerating Encrypted Computation
von: Ding, Juran, et al.
Veröffentlicht: (2024)
von: Ding, Juran, et al.
Veröffentlicht: (2024)
Formal Verification of Secure Encrypted Virtualization
von: Weerasena, Hansika, et al.
Veröffentlicht: (2026)
von: Weerasena, Hansika, et al.
Veröffentlicht: (2026)
GME: GPU-based Microarchitectural Extensions to Accelerate Homomorphic Encryption
von: Shivdikar, Kaustubh, et al.
Veröffentlicht: (2023)
von: Shivdikar, Kaustubh, et al.
Veröffentlicht: (2023)
Side-channel Inference of User Activities in AR/VR Using GPU Profiling
von: Son, Seonghun, et al.
Veröffentlicht: (2025)
von: Son, Seonghun, et al.
Veröffentlicht: (2025)
Trinity: A General Purpose FHE Accelerator
von: Deng, Xianglong, et al.
Veröffentlicht: (2024)
von: Deng, Xianglong, et al.
Veröffentlicht: (2024)
ALLMod: Exploring $\underline{\mathbf{A}}$rea-Efficiency of $\underline{\mathbf{L}}$UT-based $\underline{\mathbf{L}}$arge Number $\underline{\mathbf{Mod}}$ular Reduction via Hybrid Workloads
von: Liu, Fangxin, et al.
Veröffentlicht: (2025)
von: Liu, Fangxin, et al.
Veröffentlicht: (2025)
WebGPU-SPY: Finding Fingerprints in the Sandbox through GPU Cache Attacks
von: Ferguson, Ethan, et al.
Veröffentlicht: (2024)
von: Ferguson, Ethan, et al.
Veröffentlicht: (2024)
Vulnerabilities in Partial TEE-Shielded LLM Inference with Precomputed Noise
von: Saini, Abhishek, et al.
Veröffentlicht: (2026)
von: Saini, Abhishek, et al.
Veröffentlicht: (2026)
APINT: A Full-Stack Framework for Acceleration of Privacy-Preserving Inference of Transformers based on Garbled Circuits
von: Cho, Hyunjun, et al.
Veröffentlicht: (2025)
von: Cho, Hyunjun, et al.
Veröffentlicht: (2025)
FHECore: Rethinking GPU Microarchitecture for Fully Homomorphic Encryption
von: Daksha, Lohit, et al.
Veröffentlicht: (2026)
von: Daksha, Lohit, et al.
Veröffentlicht: (2026)
GPU in the Blind Spot: Overlooked Security Risks in Transportation
von: Puspa, Sefatun-Noor, et al.
Veröffentlicht: (2025)
von: Puspa, Sefatun-Noor, et al.
Veröffentlicht: (2025)
SoK: Fully Homomorphic Encryption Accelerators
von: Zhang, Junxue, et al.
Veröffentlicht: (2022)
von: Zhang, Junxue, et al.
Veröffentlicht: (2022)
Highly Efficient Parallel Row-Layered Min-Sum MDPC Decoder for McEliece Cryptosystem
von: Cai, Jiaxuan, et al.
Veröffentlicht: (2024)
von: Cai, Jiaxuan, et al.
Veröffentlicht: (2024)
Efficient Layered New Bit-Flipping QC-MDPC Decoder for BIKE Post-Quantum Cryptography
von: Cai, Jiaxuan, et al.
Veröffentlicht: (2024)
von: Cai, Jiaxuan, et al.
Veröffentlicht: (2024)
OpenGL GPU-Based Rowhammer Attack (Work in Progress)
von: Plin, Antoine, et al.
Veröffentlicht: (2025)
von: Plin, Antoine, et al.
Veröffentlicht: (2025)
Fastrack: Fast IO for Secure ML using GPU TEEs
von: Wang, Yongqin, et al.
Veröffentlicht: (2024)
von: Wang, Yongqin, et al.
Veröffentlicht: (2024)
Taiyi: A high-performance CKKS accelerator for Practical Fully Homomorphic Encryption
von: Fan, Shengyu, et al.
Veröffentlicht: (2024)
von: Fan, Shengyu, et al.
Veröffentlicht: (2024)
Confidential LLM Inference: Performance and Cost Across CPU and GPU TEEs
von: Chrapek, Marcin, et al.
Veröffentlicht: (2025)
von: Chrapek, Marcin, et al.
Veröffentlicht: (2025)
Confidential Computing on Heterogeneous CPU-GPU Systems: Survey and Future Directions
von: Wang, Qifan, et al.
Veröffentlicht: (2024)
von: Wang, Qifan, et al.
Veröffentlicht: (2024)
Enabling Deterministic User-Level Interrupts in Real-Time Processors via Hardware Extension
von: Yang, Hongbin, et al.
Veröffentlicht: (2026)
von: Yang, Hongbin, et al.
Veröffentlicht: (2026)
FLASH-FHE: A Heterogeneous Architecture for Fully Homomorphic Encryption Acceleration
von: Zhang, Junxue, et al.
Veröffentlicht: (2025)
von: Zhang, Junxue, et al.
Veröffentlicht: (2025)
EFFACT: A Highly Efficient Full-Stack FHE Acceleration Platform
von: Huang, Yi, et al.
Veröffentlicht: (2025)
von: Huang, Yi, et al.
Veröffentlicht: (2025)
Automated Physical Design Watermarking Leveraging Graph Neural Networks
von: Zhang, Ruisi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruisi, et al.
Veröffentlicht: (2024)
ICMarks: A Robust Watermarking Framework for Integrated Circuit Physical Design IP Protection
von: Zhang, Ruisi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruisi, et al.
Veröffentlicht: (2024)
Attacking AI Accelerators by Leveraging Arithmetic Properties of Addition
von: Heidary, Masoud, et al.
Veröffentlicht: (2026)
von: Heidary, Masoud, et al.
Veröffentlicht: (2026)
REDACTOR: eFPGA Redaction for DNN Accelerator Security
von: Baddour, Yazan, et al.
Veröffentlicht: (2025)
von: Baddour, Yazan, et al.
Veröffentlicht: (2025)
Presto: Hardware Acceleration of Ciphers for Hybrid Homomorphic Encryption
von: Jeon, Yeonsoo, et al.
Veröffentlicht: (2025)
von: Jeon, Yeonsoo, et al.
Veröffentlicht: (2025)
SZKP: A Scalable Accelerator Architecture for Zero-Knowledge Proofs
von: Daftardar, Alhad, et al.
Veröffentlicht: (2024)
von: Daftardar, Alhad, et al.
Veröffentlicht: (2024)
if-ZKP: Intel FPGA-Based Acceleration of Zero Knowledge Proofs
von: Butt, Shahzad Ahmad, et al.
Veröffentlicht: (2024)
von: Butt, Shahzad Ahmad, et al.
Veröffentlicht: (2024)
FAME: FPGA Acceleration of Secure Matrix Multiplication with Homomorphic Encryption
von: Xu, Zhihan, et al.
Veröffentlicht: (2025)
von: Xu, Zhihan, et al.
Veröffentlicht: (2025)
CIBPU: A Conflict-Invisible Secure Branch Prediction Unit
von: Zhou, Zhe, et al.
Veröffentlicht: (2025)
von: Zhou, Zhe, et al.
Veröffentlicht: (2025)
Theodosian: A Deep Dive into Memory-Hierarchy-Centric FHE Acceleration
von: Choi, Wonseok, et al.
Veröffentlicht: (2025)
von: Choi, Wonseok, et al.
Veröffentlicht: (2025)
Need for zkSpeed: Accelerating HyperPlonk for Zero-Knowledge Proofs
von: Daftardar, Alhad, et al.
Veröffentlicht: (2025)
von: Daftardar, Alhad, et al.
Veröffentlicht: (2025)
HF-NTT: Hazard-Free Dataflow Accelerator for Number Theoretic Transform
von: Meng, Xiangchen, et al.
Veröffentlicht: (2024)
von: Meng, Xiangchen, et al.
Veröffentlicht: (2024)
AMAZE: Accelerated MiMC Hardware Architecture for Zero-Knowledge Applications on the Edge
von: Ahmed, Anees, et al.
Veröffentlicht: (2024)
von: Ahmed, Anees, et al.
Veröffentlicht: (2024)
CiFHER: A Chiplet-Based FHE Accelerator with a Resizable Structure
von: Kim, Sangpyo, et al.
Veröffentlicht: (2023)
von: Kim, Sangpyo, et al.
Veröffentlicht: (2023)
Analysis of LLM Vulnerability to GPU Soft Errors: An Instruction-Level Fault Injection Study
von: Chai, Duo, et al.
Veröffentlicht: (2025)
von: Chai, Duo, et al.
Veröffentlicht: (2025)
DNA-HHE: Dual-mode Near-network Accelerator for Hybrid Homomorphic Encryption on the Edge
von: Zhao, Yifan, et al.
Veröffentlicht: (2025)
von: Zhao, Yifan, et al.
Veröffentlicht: (2025)
BOLT: Bandwidth-Optimized Lightning-Fast Oblivious Map powered by Secure HBM Accelerators
von: Guo, Yitong, et al.
Veröffentlicht: (2025)
von: Guo, Yitong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
OFHE: An Electro-Optical Accelerator for Discretized TFHE
von: Zheng, Mengxin, et al.
Veröffentlicht: (2024) -
NTTSuite: Number Theoretic Transform Benchmarks for Accelerating Encrypted Computation
von: Ding, Juran, et al.
Veröffentlicht: (2024) -
Formal Verification of Secure Encrypted Virtualization
von: Weerasena, Hansika, et al.
Veröffentlicht: (2026) -
GME: GPU-based Microarchitectural Extensions to Accelerate Homomorphic Encryption
von: Shivdikar, Kaustubh, et al.
Veröffentlicht: (2023) -
Side-channel Inference of User Activities in AR/VR Using GPU Profiling
von: Son, Seonghun, et al.
Veröffentlicht: (2025)