Power-Softmax: Towards Secure LLM Inference over Encrypted Data
Fuente:
arXiv
Salvato in:
| Autori principali: | Zimerman, Itamar, Adir, Allon, Aharoni, Ehud, Avitan, Matan, Baruch, Moran, Drucker, Nir, Lerner, Jenny, Masalha, Ramy, Meiri, Reut, Soceanu, Omri |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Efficient Decoding Methods for Language Models on Encrypted Data
di: Avitan, Matan, et al.
Pubblicazione: (2025)
di: Avitan, Matan, et al.
Pubblicazione: (2025)
Efficient Skip Connections Realization for Secure Inference on Encrypted Data
di: Drucker, Nir, et al.
Pubblicazione: (2023)
di: Drucker, Nir, et al.
Pubblicazione: (2023)
Generating One-Hot Maps under Encryption
di: Aharoni, Ehud, et al.
Pubblicazione: (2023)
di: Aharoni, Ehud, et al.
Pubblicazione: (2023)
The Hidden Attention of Mamba Models
di: Ali, Ameen, et al.
Pubblicazione: (2024)
di: Ali, Ameen, et al.
Pubblicazione: (2024)
Explaining Modern Gated-Linear RNNs via a Unified Implicit Attention Formulation
di: Zimerman, Itamar, et al.
Pubblicazione: (2024)
di: Zimerman, Itamar, et al.
Pubblicazione: (2024)
Overclocking LLM Reasoning: Monitoring and Controlling Thinking Path Lengths in LLMs
di: Eisenstadt, Roy, et al.
Pubblicazione: (2025)
di: Eisenstadt, Roy, et al.
Pubblicazione: (2025)
Efficient Pruning for Machine Learning Under Homomorphic Encryption
di: Aharoni, Ehud, et al.
Pubblicazione: (2022)
di: Aharoni, Ehud, et al.
Pubblicazione: (2022)
Revisiting LRP: Positional Attribution as the Missing Ingredient for Transformer Explainability
di: Bakish, Yarden, et al.
Pubblicazione: (2025)
di: Bakish, Yarden, et al.
Pubblicazione: (2025)
TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors
di: Atad, Ido Andrew, et al.
Pubblicazione: (2026)
di: Atad, Ido Andrew, et al.
Pubblicazione: (2026)
EncFormer: Secure and Efficient Transformer Inference over Encrypted Data
di: Zhu, Yufan, et al.
Pubblicazione: (2026)
di: Zhu, Yufan, et al.
Pubblicazione: (2026)
On the Expressivity of Selective State-Space Layers: A Multivariate Polynomial Approach
di: Cohen-Karlik, Edo, et al.
Pubblicazione: (2025)
di: Cohen-Karlik, Edo, et al.
Pubblicazione: (2025)
A QUESTÃO DA TERRA NO VALE DO PARAÍBA: HISTÓRIA DE UM ASSENTAMENTO DO MST
di: Adir de Almeida Mota
Pubblicazione: (2011)
di: Adir de Almeida Mota
Pubblicazione: (2011)
Mobile Phone Sensor-based Nigerian Driving Dataset to Detect Alcohol-influenced Behaviours
di: Thompson, Iniakpokeikiye Peter, et al.
Pubblicazione: (2025)
di: Thompson, Iniakpokeikiye Peter, et al.
Pubblicazione: (2025)
CBR -- Boosting Adaptive Classification By Retrieval of Encrypted Network Traffic with Out-of-distribution
di: Lukach, Amir, et al.
Pubblicazione: (2024)
di: Lukach, Amir, et al.
Pubblicazione: (2024)
Measuring the Availability and Response Times of Public Encrypted DNS Resolvers
di: Sharma, Ranya, et al.
Pubblicazione: (2022)
di: Sharma, Ranya, et al.
Pubblicazione: (2022)
Optimal Preprocessing for Answering On-Line Product Queries
di: Alon, Noga, et al.
Pubblicazione: (2024)
di: Alon, Noga, et al.
Pubblicazione: (2024)
An End-to-End System for Culturally-Attuned Driving Feedback using a Dual-Component NLG Engine
di: Thompson, Iniakpokeikiye Peter, et al.
Pubblicazione: (2025)
di: Thompson, Iniakpokeikiye Peter, et al.
Pubblicazione: (2025)
Bilinear Mamba-Koopman Neural MPC for Varying Dynamics
di: Pagi, Matan, et al.
Pubblicazione: (2026)
di: Pagi, Matan, et al.
Pubblicazione: (2026)
Video QoE Metrics from Encrypted Traffic: Application-agnostic Methodology
di: Berger, Tamir, et al.
Pubblicazione: (2025)
di: Berger, Tamir, et al.
Pubblicazione: (2025)
SMSI: System Model Security Inference: Automated Threat Modeling for Cyber-Physical Systems
di: Radaideh, RoÝah, et al.
Pubblicazione: (2026)
di: Radaideh, RoÝah, et al.
Pubblicazione: (2026)
Code Documentation and Analysis to Secure Software Development
di: Attie, Paul, et al.
Pubblicazione: (2024)
di: Attie, Paul, et al.
Pubblicazione: (2024)
On the Invariants of Softmax Attention
di: Lee, Wonsuk
Pubblicazione: (2026)
di: Lee, Wonsuk
Pubblicazione: (2026)
Determination of language families using deep learning
di: Lerner, Peter B.
Pubblicazione: (2024)
di: Lerner, Peter B.
Pubblicazione: (2024)
Hairpin Completion Distance Lower Bound
di: Boneh, Itai, et al.
Pubblicazione: (2024)
di: Boneh, Itai, et al.
Pubblicazione: (2024)
Dynamic Difficulty Adjustment With Brain Waves as a Tool for Optimizing Engagement
di: Cafri, Nir
Pubblicazione: (2025)
di: Cafri, Nir
Pubblicazione: (2025)
Towards Deep Encrypted Training: Low-Latency, Memory-Efficient, and High-Throughput Inference for Privacy-Preserving Neural Networks
di: Njungle, Nges Brian, et al.
Pubblicazione: (2026)
di: Njungle, Nges Brian, et al.
Pubblicazione: (2026)
The Containment Game in the plane: between the Firefighter Problem and Conway's Angel Problem
di: Feldheim, Ohad Noy, et al.
Pubblicazione: (2023)
di: Feldheim, Ohad Noy, et al.
Pubblicazione: (2023)
Photochemistry of 7-Alcoxy and Thioalcoxy-3,3-dimethoxibicyclo[2.2.2] oct-5-3n-2-one. Sequence 1,3-acyl shift-decarbonylation reaction
di: Arjona, Odón - Medel, Rocío - Plumet, Joaquín - Rojas, Jenny K
Pubblicazione: (2003)
di: Arjona, Odón - Medel, Rocío - Plumet, Joaquín - Rojas, Jenny K
Pubblicazione: (2003)
On the Complexity of Neural Computation in Superposition
di: Adler, Micah, et al.
Pubblicazione: (2024)
di: Adler, Micah, et al.
Pubblicazione: (2024)
ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree
di: Koc, Vincent, et al.
Pubblicazione: (2026)
di: Koc, Vincent, et al.
Pubblicazione: (2026)
ReBoot: Encrypted Training of Deep Neural Networks with CKKS Bootstrapping
di: Pirillo, Alberto, et al.
Pubblicazione: (2025)
di: Pirillo, Alberto, et al.
Pubblicazione: (2025)
Photochemistry of 7-Alcoxy and Thioalcoxy-3,3-dimethoxybicycol (2.2.2) oct-5-en-2-one. Sequence 1,3-acyl shift-decarbonylation reaction
di: Arjona Odón, Medel Rocío, Plumet Joaquín y K. Rojas, Jenny
Pubblicazione: (2003)
di: Arjona Odón, Medel Rocío, Plumet Joaquín y K. Rojas, Jenny
Pubblicazione: (2003)
Active Inference for an Intelligent Agent in Autonomous Reconnaissance Missions
di: Schubert, Johan, et al.
Pubblicazione: (2025)
di: Schubert, Johan, et al.
Pubblicazione: (2025)
Guardians of the Web: The Evolution and Future of Website Information Security
di: Islam, Md Saiful, et al.
Pubblicazione: (2025)
di: Islam, Md Saiful, et al.
Pubblicazione: (2025)
Partially Recentralization Softmax Loss for Vision-Language Models Robustness
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
Integrated Sensing and Communication for Low-Altitude Security
di: Ren, Ruixing
Pubblicazione: (2026)
di: Ren, Ruixing
Pubblicazione: (2026)
Targeted Nakamoto: A Bitcoin Protocol to Balance Network Security and Carbon Emissions
di: Aronoff, Daniel
Pubblicazione: (2024)
di: Aronoff, Daniel
Pubblicazione: (2024)
Softmax Linear Attention: Reclaiming Global Competition
di: Xu, Mingwei, et al.
Pubblicazione: (2026)
di: Xu, Mingwei, et al.
Pubblicazione: (2026)
TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back) using Taylor-Softmax
di: Nauen, Tobias Christian, et al.
Pubblicazione: (2024)
di: Nauen, Tobias Christian, et al.
Pubblicazione: (2024)
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
di: Mitchell, Rupert, et al.
Pubblicazione: (2025)
di: Mitchell, Rupert, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Efficient Decoding Methods for Language Models on Encrypted Data
di: Avitan, Matan, et al.
Pubblicazione: (2025) -
Efficient Skip Connections Realization for Secure Inference on Encrypted Data
di: Drucker, Nir, et al.
Pubblicazione: (2023) -
Generating One-Hot Maps under Encryption
di: Aharoni, Ehud, et al.
Pubblicazione: (2023) -
The Hidden Attention of Mamba Models
di: Ali, Ameen, et al.
Pubblicazione: (2024) -
Explaining Modern Gated-Linear RNNs via a Unified Implicit Attention Formulation
di: Zimerman, Itamar, et al.
Pubblicazione: (2024)