Misrouter: Exploiting Routing Mechanisms for Input-Only Attacks on Mixture-of-Experts LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fei, Zekun, Wang, Zihao, Liu, Weijie, He, Ruiqi, Geng, Jianing, Liu, Zheli, Wang, XiaoFeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Your Semantic-Independent Watermark is Fragile: A Semantic Perturbation Attack against EaaS Watermark
von: Fei, Zekun, et al.
Veröffentlicht: (2024)
von: Fei, Zekun, et al.
Veröffentlicht: (2024)
Hiding in Plain Sight: A Steganographic Approach to Stealthy LLM Jailbreaks
von: Geng, Jianing, et al.
Veröffentlicht: (2025)
von: Geng, Jianing, et al.
Veröffentlicht: (2025)
IndirectAD: Practical Data Poisoning Attacks against Recommender Systems for Item Promotion
von: Wang, Zihao, et al.
Veröffentlicht: (2025)
von: Wang, Zihao, et al.
Veröffentlicht: (2025)
LocalAlign: Enabling Generalizable Prompt Injection Defense via Generation of Near-Target Adversarial Examples for Alignment Training
von: Gong, Yuyang, et al.
Veröffentlicht: (2026)
von: Gong, Yuyang, et al.
Veröffentlicht: (2026)
Not All Entities are Created Equal: A Dynamic Anonymization Framework for Privacy-Preserving RAG
von: Zhu, Xinyuan, et al.
Veröffentlicht: (2026)
von: Zhu, Xinyuan, et al.
Veröffentlicht: (2026)
MoEcho: Exploiting Side-Channel Attacks to Compromise User Privacy in Mixture-of-Experts LLMs
von: Ding, Ruyi, et al.
Veröffentlicht: (2025)
von: Ding, Ruyi, et al.
Veröffentlicht: (2025)
Rethinking Side-Channel Analysis: Automated Discovery and Analysis of Side-Channel Leakage with LLM-Assisted Agents
von: Xu, Zhen, et al.
Veröffentlicht: (2026)
von: Xu, Zhen, et al.
Veröffentlicht: (2026)
Characterizing Trust Boundary Vulnerabilities in TEE Containers: An Empirical Study
von: Liu, Weijie, et al.
Veröffentlicht: (2025)
von: Liu, Weijie, et al.
Veröffentlicht: (2025)
BadPatches: Routing-aware Backdoor Attacks on Vision Mixture of Experts
von: Chan, Cedric, et al.
Veröffentlicht: (2025)
von: Chan, Cedric, et al.
Veröffentlicht: (2025)
GateBreaker: Gate-Guided Attacks on Mixture-of-Expert LLMs
von: Wu, Lichao, et al.
Veröffentlicht: (2025)
von: Wu, Lichao, et al.
Veröffentlicht: (2025)
Who Speaks for the Trigger? Dynamic Expert Routing in Backdoored Mixture-of-Experts Transformers
von: Zhao, Xin, et al.
Veröffentlicht: (2025)
von: Zhao, Xin, et al.
Veröffentlicht: (2025)
BadMoE: Backdooring Mixture-of-Experts LLMs via Optimizing Routing Triggers and Infecting Dormant Experts
von: Wang, Qingyue, et al.
Veröffentlicht: (2025)
von: Wang, Qingyue, et al.
Veröffentlicht: (2025)
GraphAttack: Exploiting Representational Blindspots in LLM Safety Mechanisms
von: He, Sinan, et al.
Veröffentlicht: (2025)
von: He, Sinan, et al.
Veröffentlicht: (2025)
Consiglieres in the Shadow: Understanding the Use of Uncensored Large Language Models in Cybercrimes
von: Lin, Zilong, et al.
Veröffentlicht: (2025)
von: Lin, Zilong, et al.
Veröffentlicht: (2025)
Prompt-in-Content Attacks: Exploiting Uploaded Inputs to Hijack LLM Behavior
von: Lian, Zhuotao, et al.
Veröffentlicht: (2025)
von: Lian, Zhuotao, et al.
Veröffentlicht: (2025)
Clues in Tweets: Twitter-Guided Discovery and Analysis of SMS Spam
von: Tang, Siyuan, et al.
Veröffentlicht: (2022)
von: Tang, Siyuan, et al.
Veröffentlicht: (2022)
An Automated Attack Investigation Approach Leveraging Threat-Knowledge-Augmented Large Language Models
von: Dai, Rujie, et al.
Veröffentlicht: (2025)
von: Dai, Rujie, et al.
Veröffentlicht: (2025)
CryptoMoE: Privacy-Preserving and Scalable Mixture of Experts Inference via Balanced Expert Routing
von: Zhou, Yifan, et al.
Veröffentlicht: (2025)
von: Zhou, Yifan, et al.
Veröffentlicht: (2025)
Sparse Models, Sparse Safety: Unsafe Routes in Mixture-of-Experts LLMs
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
QUIC-Exfil: Exploiting QUIC's Server Preferred Address Feature to Perform Data Exfiltration Attacks
von: Grübl, Thomas, et al.
Veröffentlicht: (2025)
von: Grübl, Thomas, et al.
Veröffentlicht: (2025)
RASA: Routing-Aware Safety Alignment for Mixture-of-Experts Models
von: Liang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Liang, Jiacheng, et al.
Veröffentlicht: (2026)
LLM-Enhanced Software Patch Localization
von: Yu, Jinhong, et al.
Veröffentlicht: (2024)
von: Yu, Jinhong, et al.
Veröffentlicht: (2024)
Malla: Demystifying Real-world Large Language Model Integrated Malicious Services
von: Lin, Zilong, et al.
Veröffentlicht: (2024)
von: Lin, Zilong, et al.
Veröffentlicht: (2024)
The Early Bird Catches the Leak: Unveiling Timing Side Channels in LLM Serving Systems
von: Song, Linke, et al.
Veröffentlicht: (2024)
von: Song, Linke, et al.
Veröffentlicht: (2024)
Breaking Euston: Recovering Private Inputs from Secure Inference by Exploiting Subspace Leakage
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2026)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2026)
Toward Efficient Inference Attacks: Shadow Model Sharing via Mixture-of-Experts
von: Bai, Li, et al.
Veröffentlicht: (2025)
von: Bai, Li, et al.
Veröffentlicht: (2025)
Stealthy Peers: Understanding Security Risks of WebRTC-Based Peer-Assisted Video Streaming
von: Tang, Siyuan, et al.
Veröffentlicht: (2022)
von: Tang, Siyuan, et al.
Veröffentlicht: (2022)
Understanding the Security Risks of Decentralized Exchanges by Uncovering Unfair Trades in the Wild
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
Disassembling Obfuscated Executables with LLM
von: Rong, Huanyao, et al.
Veröffentlicht: (2024)
von: Rong, Huanyao, et al.
Veröffentlicht: (2024)
A Simple and Efficient Jailbreak Method Exploiting LLMs' Helpfulness
von: Luo, Xuan, et al.
Veröffentlicht: (2025)
von: Luo, Xuan, et al.
Veröffentlicht: (2025)
Routing-Aware Explanations for Mixture of Experts Graph Models in Malware Detection
von: Shokouhinejad, Hossein, et al.
Veröffentlicht: (2026)
von: Shokouhinejad, Hossein, et al.
Veröffentlicht: (2026)
MAWSEO: Adversarial Wiki Search Poisoning for Illicit Online Promotion
von: Lin, Zilong, et al.
Veröffentlicht: (2023)
von: Lin, Zilong, et al.
Veröffentlicht: (2023)
AlignSentinel: Alignment-Aware Detection of Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2026)
von: Jia, Yuqi, et al.
Veröffentlicht: (2026)
Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
LLMs Can Unlearn Refusal with Only 1,000 Benign Samples
von: Guo, Yangyang, et al.
Veröffentlicht: (2026)
von: Guo, Yangyang, et al.
Veröffentlicht: (2026)
Leaky Cauldron on the Dark Land: Understanding Memory Side-Channel Hazards in SGX
von: Wang, Wenhao, et al.
Veröffentlicht: (2017)
von: Wang, Wenhao, et al.
Veröffentlicht: (2017)
DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box Retrieval-Augmented Generation
von: Gong, Yuyang, et al.
Veröffentlicht: (2026)
von: Gong, Yuyang, et al.
Veröffentlicht: (2026)
Exploiting Cross-Layer Vulnerabilities: Off-Path Attacks on the TCP/IP Protocol Suite
von: Feng, Xuewei, et al.
Veröffentlicht: (2024)
von: Feng, Xuewei, et al.
Veröffentlicht: (2024)
Picachv: Formally Verified Data Use Policy Enforcement for Secure Data Analytics
von: Chen, Haobin Hiroki, et al.
Veröffentlicht: (2025)
von: Chen, Haobin Hiroki, et al.
Veröffentlicht: (2025)
Traceback of Poisoning Attacks to Retrieval-Augmented Generation
von: Zhang, Baolei, et al.
Veröffentlicht: (2025)
von: Zhang, Baolei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Your Semantic-Independent Watermark is Fragile: A Semantic Perturbation Attack against EaaS Watermark
von: Fei, Zekun, et al.
Veröffentlicht: (2024) -
Hiding in Plain Sight: A Steganographic Approach to Stealthy LLM Jailbreaks
von: Geng, Jianing, et al.
Veröffentlicht: (2025) -
IndirectAD: Practical Data Poisoning Attacks against Recommender Systems for Item Promotion
von: Wang, Zihao, et al.
Veröffentlicht: (2025) -
LocalAlign: Enabling Generalizable Prompt Injection Defense via Generation of Near-Target Adversarial Examples for Alignment Training
von: Gong, Yuyang, et al.
Veröffentlicht: (2026) -
Not All Entities are Created Equal: A Dynamic Anonymization Framework for Privacy-Preserving RAG
von: Zhu, Xinyuan, et al.
Veröffentlicht: (2026)