Syntax- and Compilation-Preserving Evasion of LLM Vulnerability Detectors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Luze, Oprea, Alina, Wong, Eric |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem
von: Lin, Shuyi, et al.
Veröffentlicht: (2025)
von: Lin, Shuyi, et al.
Veröffentlicht: (2025)
StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026)
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026)
The Jailbreak Tax: How Useful are Your Jailbreak Outputs?
von: Nikolić, Kristina, et al.
Veröffentlicht: (2025)
von: Nikolić, Kristina, et al.
Veröffentlicht: (2025)
PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense
von: Cadet, Xavier, et al.
Veröffentlicht: (2025)
von: Cadet, Xavier, et al.
Veröffentlicht: (2025)
SAGA: A Security Architecture for Governing AI Agentic Systems
von: Syros, Georgios, et al.
Veröffentlicht: (2025)
von: Syros, Georgios, et al.
Veröffentlicht: (2025)
EGAN: Evolutional GAN for Ransomware Evasion
von: Commey, Daniel, et al.
Veröffentlicht: (2024)
von: Commey, Daniel, et al.
Veröffentlicht: (2024)
Your Compiler is Backdooring Your Model: Understanding and Exploiting Compilation Inconsistency Vulnerabilities in Deep Learning Compilers
von: Chen, Simin, et al.
Veröffentlicht: (2025)
von: Chen, Simin, et al.
Veröffentlicht: (2025)
Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
Efficient but Vulnerable: Benchmarking and Defending LLM Batch Prompting Attack
von: Yue, Murong, et al.
Veröffentlicht: (2025)
von: Yue, Murong, et al.
Veröffentlicht: (2025)
A Middle Path for On-Premises LLM Deployment: Preserving Privacy Without Sacrificing Model Confidentiality
von: Huang, Hanbo, et al.
Veröffentlicht: (2024)
von: Huang, Hanbo, et al.
Veröffentlicht: (2024)
Optimizing Privacy-Preserving Primitives to Support LLM-Scale Applications
von: Jandali, Yaman, et al.
Veröffentlicht: (2025)
von: Jandali, Yaman, et al.
Veröffentlicht: (2025)
Privacy-Preserving Dynamic Assortment Selection
von: Cho, Young Hyun, et al.
Veröffentlicht: (2024)
von: Cho, Young Hyun, et al.
Veröffentlicht: (2024)
A Novel Perturb-ability Score to Mitigate Evasion Adversarial Attacks on Flow-Based ML-NIDS
von: elShehaby, Mohamed, et al.
Veröffentlicht: (2024)
von: elShehaby, Mohamed, et al.
Veröffentlicht: (2024)
CLASP: Training-Free LLM-Assisted Source Code Watermarking via Semantic-Preserving Transformations
von: Xu, Rui, et al.
Veröffentlicht: (2025)
von: Xu, Rui, et al.
Veröffentlicht: (2025)
Probing Latent Subspaces in LLM for AI Security: Identifying and Manipulating Adversarial States
von: Chia, Xin Wei, et al.
Veröffentlicht: (2025)
von: Chia, Xin Wei, et al.
Veröffentlicht: (2025)
Statement-Level Vulnerability Detection: Learning Vulnerability Patterns Through Information Theory and Contrastive Learning
von: Nguyen, Van, et al.
Veröffentlicht: (2022)
von: Nguyen, Van, et al.
Veröffentlicht: (2022)
ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?
von: Wang, Zhun, et al.
Veröffentlicht: (2026)
von: Wang, Zhun, et al.
Veröffentlicht: (2026)
Benchmarking Misuse Mitigation Against Covert Adversaries
von: Brown, Davis, et al.
Veröffentlicht: (2025)
von: Brown, Davis, et al.
Veröffentlicht: (2025)
Adversarial Attacks on Transformers-Based Malware Detectors
von: Jakhotiya, Yash, et al.
Veröffentlicht: (2022)
von: Jakhotiya, Yash, et al.
Veröffentlicht: (2022)
Optimal Zero-Shot Detector for Multi-Armed Attacks
von: Granese, Federica, et al.
Veröffentlicht: (2024)
von: Granese, Federica, et al.
Veröffentlicht: (2024)
How to make Medical AI Systems safer? Simulating Vulnerabilities, and Threats in Multimodal Medical RAG System
von: Zuo, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zuo, Kaiwen, et al.
Veröffentlicht: (2025)
Jailbreaking and Mitigation of Vulnerabilities in Large Language Models
von: Peng, Benji, et al.
Veröffentlicht: (2024)
von: Peng, Benji, et al.
Veröffentlicht: (2024)
Finetuning Large Language Models for Vulnerability Detection
von: Shestov, Alexey, et al.
Veröffentlicht: (2024)
von: Shestov, Alexey, et al.
Veröffentlicht: (2024)
Enhancing Continual Learning for Software Vulnerability Prediction: Addressing Catastrophic Forgetting via Hybrid-Confidence-Aware Selective Replay for Temporal LLM Fine-Tuning
von: Dou, Xuhui, et al.
Veröffentlicht: (2026)
von: Dou, Xuhui, et al.
Veröffentlicht: (2026)
GOD model: Privacy Preserved AI School for Personal Assistant
von: PIN AI Team, et al.
Veröffentlicht: (2025)
von: PIN AI Team, et al.
Veröffentlicht: (2025)
AuthorMist: Evading AI Text Detectors with Reinforcement Learning
von: David, Isaac, et al.
Veröffentlicht: (2025)
von: David, Isaac, et al.
Veröffentlicht: (2025)
Rethinking the Vulnerability of Concept Erasure and a New Method
von: Richardson, Alex D., et al.
Veröffentlicht: (2025)
von: Richardson, Alex D., et al.
Veröffentlicht: (2025)
Enhancing Vulnerability Reports with Automated and Augmented Description Summarization
von: Althebeiti, Hattan, et al.
Veröffentlicht: (2025)
von: Althebeiti, Hattan, et al.
Veröffentlicht: (2025)
Exploiting Efficiency Vulnerabilities in Dynamic Deep Learning Systems
von: Rathnasuriya, Ravishka, et al.
Veröffentlicht: (2025)
von: Rathnasuriya, Ravishka, et al.
Veröffentlicht: (2025)
ARVO: Atlas of Reproducible Vulnerabilities for Open Source Software
von: Mei, Xiang, et al.
Veröffentlicht: (2024)
von: Mei, Xiang, et al.
Veröffentlicht: (2024)
Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems
von: Hackett, William, et al.
Veröffentlicht: (2025)
von: Hackett, William, et al.
Veröffentlicht: (2025)
TBDetector:Transformer-Based Detector for Advanced Persistent Threats with Provenance Graph
von: Wang, Nan, et al.
Veröffentlicht: (2023)
von: Wang, Nan, et al.
Veröffentlicht: (2023)
Adversarial Inception Backdoor Attacks against Reinforcement Learning
von: Rathbun, Ethan, et al.
Veröffentlicht: (2024)
von: Rathbun, Ethan, et al.
Veröffentlicht: (2024)
SleeperNets: Universal Backdoor Poisoning Attacks Against Reinforcement Learning Agents
von: Rathbun, Ethan, et al.
Veröffentlicht: (2024)
von: Rathbun, Ethan, et al.
Veröffentlicht: (2024)
Learnability and Privacy Vulnerability are Entangled in a Few Critical Weights
von: Fang, Xingli, et al.
Veröffentlicht: (2026)
von: Fang, Xingli, et al.
Veröffentlicht: (2026)
Exploiting Layer-Specific Vulnerabilities to Backdoor Attack in Federated Learning
von: Foroughi, Mohammad Hadi, et al.
Veröffentlicht: (2026)
von: Foroughi, Mohammad Hadi, et al.
Veröffentlicht: (2026)
Securing Large Language Models: Threats, Vulnerabilities and Responsible Practices
von: Abdali, Sara, et al.
Veröffentlicht: (2024)
von: Abdali, Sara, et al.
Veröffentlicht: (2024)
Large Language Models in Cybersecurity: Applications, Vulnerabilities, and Defense Techniques
von: Jaffal, Niveen O., et al.
Veröffentlicht: (2025)
von: Jaffal, Niveen O., et al.
Veröffentlicht: (2025)
Can Neural Decompilation Assist Vulnerability Prediction on Binary Code?
von: Cotroneo, D., et al.
Veröffentlicht: (2024)
von: Cotroneo, D., et al.
Veröffentlicht: (2024)
Weakest Link in the Chain: Security Vulnerabilities in Advanced Reasoning Models
von: Krishna, Arjun, et al.
Veröffentlicht: (2025)
von: Krishna, Arjun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem
von: Lin, Shuyi, et al.
Veröffentlicht: (2025) -
StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026) -
The Jailbreak Tax: How Useful are Your Jailbreak Outputs?
von: Nikolić, Kristina, et al.
Veröffentlicht: (2025) -
PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense
von: Cadet, Xavier, et al.
Veröffentlicht: (2025) -
SAGA: A Security Architecture for Governing AI Agentic Systems
von: Syros, Georgios, et al.
Veröffentlicht: (2025)