Developing Assurance Cases for Adversarial Robustness and Regulatory Compliance in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Momcilovic, Tomas Bueno, Balta, Dian, Buesser, Beat, Zizzo, Giulio, Purcell, Mark |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Assuring EU AI Act Compliance and Adversarial Robustness of LLMs
by: Momcilovic, Tomas Bueno, et al.
Published: (2024)
by: Momcilovic, Tomas Bueno, et al.
Published: (2024)
Knowledge-Augmented Reasoning for EUAIA Compliance and Adversarial Robustness of LLMs
by: Momcilovic, Tomas Bueno, et al.
Published: (2024)
by: Momcilovic, Tomas Bueno, et al.
Published: (2024)
Towards Assurance of LLM Adversarial Robustness using Ontology-Driven Argumentation
by: Momcilovic, Tomas Bueno, et al.
Published: (2024)
by: Momcilovic, Tomas Bueno, et al.
Published: (2024)
OntoGSN: An Ontology-Based Framework for Semantic Management and Extension of Assurance Cases
by: Momcilovic, Tomas Bueno, et al.
Published: (2025)
by: Momcilovic, Tomas Bueno, et al.
Published: (2025)
An Ontology-Based Approach to Security Risk Identification of Container Deployments in OT Contexts
by: Landeck, Yannick, et al.
Published: (2026)
by: Landeck, Yannick, et al.
Published: (2026)
Evaluating the Role of Security Assurance Cases in Agile Medical Device Development
by: Fransson, Max, et al.
Published: (2024)
by: Fransson, Max, et al.
Published: (2024)
LLMs + Security = Trouble
by: Livshits, Benjamin
Published: (2026)
by: Livshits, Benjamin
Published: (2026)
LLMs as verification oracles for Solidity
by: Bartoletti, Massimo, et al.
Published: (2025)
by: Bartoletti, Massimo, et al.
Published: (2025)
Harnessing the Power of LLMs in Source Code Vulnerability Detection
by: Mahyari, Andrew A
Published: (2024)
by: Mahyari, Andrew A
Published: (2024)
Evaluating LLMs for One-Shot Patching of Real and Artificial Vulnerabilities
by: Garg, Aayush, et al.
Published: (2025)
by: Garg, Aayush, et al.
Published: (2025)
SPDZCoder: Combining Expert Knowledge with LLMs for Generating Privacy-Computing Code
by: Dong, Xiaoning, et al.
Published: (2024)
by: Dong, Xiaoning, et al.
Published: (2024)
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
by: Yu, Jiongchi, et al.
Published: (2025)
by: Yu, Jiongchi, et al.
Published: (2025)
Beyond Classification: Evaluating LLMs for Fine-Grained Automatic Malware Behavior Auditing
by: Zheng, Xinran, et al.
Published: (2025)
by: Zheng, Xinran, et al.
Published: (2025)
TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs
by: Shen, Qingchao, et al.
Published: (2026)
by: Shen, Qingchao, et al.
Published: (2026)
Mitigating Sensitive Information Leakage in LLMs4Code through Machine Unlearning
by: Gu, Shanzhi, et al.
Published: (2025)
by: Gu, Shanzhi, et al.
Published: (2025)
Towards Secure Logging: Characterizing and Benchmarking Logging Code Security Issues with LLMs
by: Yuan, He Yang, et al.
Published: (2026)
by: Yuan, He Yang, et al.
Published: (2026)
CASTLE: Benchmarking Dataset for Static Code Analyzers and LLMs towards CWE Detection
by: Dubniczky, Richard A., et al.
Published: (2025)
by: Dubniczky, Richard A., et al.
Published: (2025)
Rethinking and Exploring String-Based Malware Family Classification in the Era of LLMs and RAG
by: Chen, Yufan, et al.
Published: (2025)
by: Chen, Yufan, et al.
Published: (2025)
MASKDROID: Robust Android Malware Detection with Masked Graph Representations
by: Zheng, Jingnan, et al.
Published: (2024)
by: Zheng, Jingnan, et al.
Published: (2024)
How Do Semantically Equivalent Code Transformations Impact Membership Inference on LLMs for Code?
by: Yang, Hua, et al.
Published: (2025)
by: Yang, Hua, et al.
Published: (2025)
Detecting Data Poisoning in Code Generation LLMs via Black-Box, Vulnerability-Oriented Scanning
by: Yan, Shenao, et al.
Published: (2026)
by: Yan, Shenao, et al.
Published: (2026)
LLM4Vuln: A Unified Evaluation Framework for Decoupling and Enhancing LLMs' Vulnerability Reasoning
by: Sun, Yuqiang, et al.
Published: (2024)
by: Sun, Yuqiang, et al.
Published: (2024)
CoDe-R: Refining Decompiler Output with LLMs via Rationale Guidance and Adaptive Inference
by: Zhang, Qiang, et al.
Published: (2026)
by: Zhang, Qiang, et al.
Published: (2026)
Scam2Prompt: A Scalable Framework for Auditing Malicious Scam Endpoints in Production LLMs
by: Chen, Zhiyang, et al.
Published: (2025)
by: Chen, Zhiyang, et al.
Published: (2025)
Beyond Trusting Trust: Multi-Model Validation for Robust Code Generation
by: McDanel, Bradley
Published: (2025)
by: McDanel, Bradley
Published: (2025)
A Blockchain-Enabled Approach to Cross-Border Compliance and Trust
by: Kulothungan, Vikram
Published: (2025)
by: Kulothungan, Vikram
Published: (2025)
Evaluating Implicit Regulatory Compliance in LLM Tool Invocation via Logic-Guided Synthesis
by: Song, Da, et al.
Published: (2026)
by: Song, Da, et al.
Published: (2026)
Toward Patch Robustness Certification and Detection for Deep Learning Systems Beyond Consistent Samples
by: Zhou, Qilin, et al.
Published: (2025)
by: Zhou, Qilin, et al.
Published: (2025)
`Do as I say not as I do': A Semi-Automated Approach for Jailbreak Prompt Attack against Multimodal LLMs
by: Chiu, Chun Wai, et al.
Published: (2025)
by: Chiu, Chun Wai, et al.
Published: (2025)
CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions
by: Shi, Jingwei, et al.
Published: (2026)
by: Shi, Jingwei, et al.
Published: (2026)
CrossCert: A Cross-Checking Detection Approach to Patch Robustness Certification for Deep Learning Models
by: Zhou, Qilin, et al.
Published: (2024)
by: Zhou, Qilin, et al.
Published: (2024)
Poisoned Identifiers Survive LLM Deobfuscation: A Case Study on Claude Opus 4.6
by: Lorenzo, Luis Guzmán
Published: (2026)
by: Lorenzo, Luis Guzmán
Published: (2026)
SOK: Exploring Hallucinations and Security Risks in AI-Assisted Software Development with Insights for LLM Deployment
by: Haque, Ariful, et al.
Published: (2025)
by: Haque, Ariful, et al.
Published: (2025)
Towards Understanding and Applying Security Assurance Cases for Automotive Systems
by: Mohamad, Mazen
Published: (2024)
by: Mohamad, Mazen
Published: (2024)
Leveraging Large Language Models for Cybersecurity Risk Assessment -- A Case from Forestry Cyber-Physical Systems
by: Gultekin, Fikret Mert, et al.
Published: (2025)
by: Gultekin, Fikret Mert, et al.
Published: (2025)
Understanding, Implementing, and Supporting Security Assurance Cases in Safety-Critical Domains
by: Mohamad, Mazen
Published: (2025)
by: Mohamad, Mazen
Published: (2025)
Securing the AI Frontier: Urgent Ethical and Regulatory Imperatives for AI-Driven Cybersecurity
by: Kulothungan, Vikram
Published: (2025)
by: Kulothungan, Vikram
Published: (2025)
LinuxArena: A Control Setting for AI Agents in Live Production Software Environments
by: Tracy, Tyler, et al.
Published: (2026)
by: Tracy, Tyler, et al.
Published: (2026)
Groot: Adversarial Testing for Generative Text-to-Image Models with Tree-based Semantic Transformation
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
Defending against Adversarial Malware Attacks on ML-based Android Malware Detection Systems
by: He, Ping, et al.
Published: (2025)
by: He, Ping, et al.
Published: (2025)
Similar Items
-
Towards Assuring EU AI Act Compliance and Adversarial Robustness of LLMs
by: Momcilovic, Tomas Bueno, et al.
Published: (2024) -
Knowledge-Augmented Reasoning for EUAIA Compliance and Adversarial Robustness of LLMs
by: Momcilovic, Tomas Bueno, et al.
Published: (2024) -
Towards Assurance of LLM Adversarial Robustness using Ontology-Driven Argumentation
by: Momcilovic, Tomas Bueno, et al.
Published: (2024) -
OntoGSN: An Ontology-Based Framework for Semantic Management and Extension of Assurance Cases
by: Momcilovic, Tomas Bueno, et al.
Published: (2025) -
An Ontology-Based Approach to Security Risk Identification of Container Deployments in OT Contexts
by: Landeck, Yannick, et al.
Published: (2026)