NeuroStrike: Neuron-Level Attacks on Aligned LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Lichao, Behrouzi, Sasha, Rostami, Mohamadreza, Thang, Maximilian, Picek, Stjepan, Sadeghi, Ahmad-Reza |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GateBreaker: Gate-Guided Attacks on Mixture-of-Expert LLMs
by: Wu, Lichao, et al.
Published: (2025)
by: Wu, Lichao, et al.
Published: (2025)
GoodVibe: Security-by-Vibe for LLM-Based Code Generation
by: Thang, Maximilian, et al.
Published: (2026)
by: Thang, Maximilian, et al.
Published: (2026)
NeST: Neuron Selective Tuning for LLM Safety
by: Behrouzi, Sasha, et al.
Published: (2026)
by: Behrouzi, Sasha, et al.
Published: (2026)
NeuroStrike: Neuron-Level Attacks on Aligned LLMs
by: Wu, Lichao
Published: (2025)
by: Wu, Lichao
Published: (2025)
Fuzzilicon: A Post-Silicon Microcode-Guided x86 CPU Fuzzer
by: Lenzen, Johannes, et al.
Published: (2025)
by: Lenzen, Johannes, et al.
Published: (2025)
GoldenFuzz: Generative Golden Reference Hardware Fuzzing
by: Wu, Lichao, et al.
Published: (2025)
by: Wu, Lichao, et al.
Published: (2025)
AegisSat: Securing AI-Enabled SoC FPGA Satellite Platforms
by: Li, Huimin, et al.
Published: (2026)
by: Li, Huimin, et al.
Published: (2026)
Large Language Lobotomy: Jailbreaking Mixture-of-Experts via Expert Silencing
by: Lintelo, Jona te, et al.
Published: (2026)
by: Lintelo, Jona te, et al.
Published: (2026)
Fuzzerfly Effect: Hardware Fuzzing for Memory Safety
by: Rostami, Mohamadreza, et al.
Published: (2024)
by: Rostami, Mohamadreza, et al.
Published: (2024)
BadPatches: Routing-aware Backdoor Attacks on Vision Mixture of Experts
by: Chan, Cedric, et al.
Published: (2025)
by: Chan, Cedric, et al.
Published: (2025)
MASCing: Configurable Mixture-of-Experts Behavior via Activation Steering Masks
by: Lintelo, Jona te, et al.
Published: (2026)
by: Lintelo, Jona te, et al.
Published: (2026)
CatBack: Universal Backdoor Attacks on Tabular Data via Categorical Encoding
by: Tajalli, Behrad, et al.
Published: (2025)
by: Tajalli, Behrad, et al.
Published: (2025)
EmoBack: Backdoor Attacks Against Speaker Identification Using Emotional Prosody
by: Schoof, Coen, et al.
Published: (2024)
by: Schoof, Coen, et al.
Published: (2024)
ReFuzz: Reusing Tests for Processor Fuzzing with Contextual Bandits
by: Chen, Chen, et al.
Published: (2025)
by: Chen, Chen, et al.
Published: (2025)
The SkipSponge Attack: Sponge Weight Poisoning of Deep Neural Networks
by: Lintelo, Jona te, et al.
Published: (2024)
by: Lintelo, Jona te, et al.
Published: (2024)
BAN: Detecting Backdoors Activated by Adversarial Neuron Noise
by: Xu, Xiaoyun, et al.
Published: (2024)
by: Xu, Xiaoyun, et al.
Published: (2024)
WhisperFuzz: White-Box Fuzzing for Detecting and Locating Timing Vulnerabilities in Processors
by: Borkar, Pallavi, et al.
Published: (2024)
by: Borkar, Pallavi, et al.
Published: (2024)
Context is the Key: Backdoor Attacks for In-Context Learning with Vision Transformers
by: Abad, Gorka, et al.
Published: (2024)
by: Abad, Gorka, et al.
Published: (2024)
NoMod: A Non-modular Attack on Module Learning With Errors
by: Bassotto, Cristian, et al.
Published: (2025)
by: Bassotto, Cristian, et al.
Published: (2025)
Let's Focus: Focused Backdoor Attack against Federated Transfer Learning
by: Arazzi, Marco, et al.
Published: (2024)
by: Arazzi, Marco, et al.
Published: (2024)
Flashy Backdoor: Real-world Environment Backdoor Attack on SNNs with DVS Cameras
by: Riaño, Roberto, et al.
Published: (2024)
by: Riaño, Roberto, et al.
Published: (2024)
Lost and Found in Speculation: Hybrid Speculative Vulnerability Detection
by: Rostami, Mohamadreza, et al.
Published: (2024)
by: Rostami, Mohamadreza, et al.
Published: (2024)
Time-Distributed Backdoor Attacks on Federated Spiking Learning
by: Abad, Gorka, et al.
Published: (2024)
by: Abad, Gorka, et al.
Published: (2024)
Beyond Random Inputs: A Novel ML-Based Hardware Fuzzing
by: Rostami, Mohamadreza, et al.
Published: (2024)
by: Rostami, Mohamadreza, et al.
Published: (2024)
A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models
by: Xu, Zihao, et al.
Published: (2024)
by: Xu, Zihao, et al.
Published: (2024)
Interpreting Emergent Features in Deep Learning-based Side-channel Analysis
by: Karayalçin, Sengim, et al.
Published: (2025)
by: Karayalçin, Sengim, et al.
Published: (2025)
$$\mathbf{L^2\cdot M = C^2}$$ Large Language Models are Covert Channels
by: Gaure, Simen, et al.
Published: (2024)
by: Gaure, Simen, et al.
Published: (2024)
Backdoor Attacks on Decentralised Post-Training
by: Ersoy, Oğuzhan, et al.
Published: (2026)
by: Ersoy, Oğuzhan, et al.
Published: (2026)
Phantom: Untargeted Poisoning Attacks on Semi-Supervised Learning (Full Version)
by: Knauer, Jonathan, et al.
Published: (2024)
by: Knauer, Jonathan, et al.
Published: (2024)
More is Better (Mostly): On the Backdoor Attacks in Federated Graph Neural Networks
by: Xu, Jing, et al.
Published: (2022)
by: Xu, Jing, et al.
Published: (2022)
Sneaky Spikes: Uncovering Stealthy Backdoor Attacks in Spiking Neural Networks with Neuromorphic Data
by: Abad, Gorka, et al.
Published: (2023)
by: Abad, Gorka, et al.
Published: (2023)
Towards Remote Attestation of Microarchitectural Attacks: The Case of Rowhammer
by: Herrmann, Martin, et al.
Published: (2026)
by: Herrmann, Martin, et al.
Published: (2026)
Backdoor Attacks on Transformers for Tabular Data: An Empirical Study
by: Pleiter, Bart, et al.
Published: (2023)
by: Pleiter, Bart, et al.
Published: (2023)
Towards Backdoor Stealthiness in Model Parameter Space
by: Xu, Xiaoyun, et al.
Published: (2025)
by: Xu, Xiaoyun, et al.
Published: (2025)
Membership Privacy Evaluation in Deep Spiking Neural Networks
by: Li, Jiaxin, et al.
Published: (2024)
by: Li, Jiaxin, et al.
Published: (2024)
Label Inference Attacks against Node-level Vertical Federated GNNs
by: Arazzi, Marco, et al.
Published: (2023)
by: Arazzi, Marco, et al.
Published: (2023)
You Snooze, You Lose: Automatic Safety Alignment Restoration through Neural Weight Translation
by: Arazzi, Marco, et al.
Published: (2026)
by: Arazzi, Marco, et al.
Published: (2026)
NeuroLip: An Event-driven Spatiotemporal Learning Framework for Cross-Scene Lip-Motion-based Visual Speaker Recognition
by: Yao, Junguang, et al.
Published: (2026)
by: Yao, Junguang, et al.
Published: (2026)
Removing the Trigger, Not the Backdoor: Alternative Triggers and Latent Backdoors
by: Abad, Gorka, et al.
Published: (2026)
by: Abad, Gorka, et al.
Published: (2026)
The Power of Bamboo: On the Post-Compromise Security for Searchable Symmetric Encryption
by: Chen, Tianyang, et al.
Published: (2024)
by: Chen, Tianyang, et al.
Published: (2024)
Similar Items
-
GateBreaker: Gate-Guided Attacks on Mixture-of-Expert LLMs
by: Wu, Lichao, et al.
Published: (2025) -
GoodVibe: Security-by-Vibe for LLM-Based Code Generation
by: Thang, Maximilian, et al.
Published: (2026) -
NeST: Neuron Selective Tuning for LLM Safety
by: Behrouzi, Sasha, et al.
Published: (2026) -
NeuroStrike: Neuron-Level Attacks on Aligned LLMs
by: Wu, Lichao
Published: (2025) -
Fuzzilicon: A Post-Silicon Microcode-Guided x86 CPU Fuzzer
by: Lenzen, Johannes, et al.
Published: (2025)