Toward Secure and Compliant AI: Organizational Standards and Protocols for NLP Model Lifecycle Management
Fuente:
arXiv
Saved in:
| Main Authors: | Arora, Sunil, Hastings, John |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
by: Gu, Yongtong, et al.
Published: (2026)
by: Gu, Yongtong, et al.
Published: (2026)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
by: Chen, Renmiao, et al.
Published: (2025)
by: Chen, Renmiao, et al.
Published: (2025)
Securing Agentic AI Systems -- A Multilayer Security Framework
by: Arora, Sunil, et al.
Published: (2025)
by: Arora, Sunil, et al.
Published: (2025)
Benchmarking Large Language Models for IoC Recovery under Adversarial Code Obfuscation and Encryption
by: Morales, Jaime, et al.
Published: (2026)
by: Morales, Jaime, et al.
Published: (2026)
CEKER: A Generalizable LLM Framework for Literature Analysis with a Case Study in Unikernel Security
by: Wollman, Alex, et al.
Published: (2024)
by: Wollman, Alex, et al.
Published: (2024)
Towards Modeling Cybersecurity Behavior of Humans in Organizations
by: Kürtz, Klaas Ole
Published: (2026)
by: Kürtz, Klaas Ole
Published: (2026)
Countermind: A Multi-Layered Security Architecture for Large Language Models
by: Schwarz, Dominik
Published: (2025)
by: Schwarz, Dominik
Published: (2025)
The Need for Standardized Evidence Sampling in CMMC Assessments: A Survey-Based Analysis of Assessor Practices
by: Therrien, Logan, et al.
Published: (2026)
by: Therrien, Logan, et al.
Published: (2026)
Whisper Leak: a side-channel attack on Large Language Models
by: McDonald, Geoff, et al.
Published: (2025)
by: McDonald, Geoff, et al.
Published: (2025)
A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts
by: Young, Richard J., et al.
Published: (2026)
by: Young, Richard J., et al.
Published: (2026)
The Automation Advantage in AI Red Teaming
by: Mulla, Rob, et al.
Published: (2025)
by: Mulla, Rob, et al.
Published: (2025)
AIRTBench: Measuring Autonomous AI Red Teaming Capabilities in Language Models
by: Dawson, Ads, et al.
Published: (2025)
by: Dawson, Ads, et al.
Published: (2025)
Towards Agentic Investigation of Security Alerts
by: Eilertsen, Even, et al.
Published: (2026)
by: Eilertsen, Even, et al.
Published: (2026)
Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models
by: Syed, Mohammed Sameer, et al.
Published: (2026)
by: Syed, Mohammed Sameer, et al.
Published: (2026)
Amplifying Training Data Exposure through Fine-Tuning with Pseudo-Labeled Memberships
by: Oh, Myung Gyo, et al.
Published: (2024)
by: Oh, Myung Gyo, et al.
Published: (2024)
Organizational Adaptation to Generative AI in Cybersecurity
by: Nott, Christopher
Published: (2025)
by: Nott, Christopher
Published: (2025)
Send to which account? Evaluation of an LLM-based Scambaiting System
by: Siadati, Hossein, et al.
Published: (2025)
by: Siadati, Hossein, et al.
Published: (2025)
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps
by: Chona, Alankrit, et al.
Published: (2026)
by: Chona, Alankrit, et al.
Published: (2026)
Measuring Harmfulness of Computer-Using Agents
by: Tian, Aaron Xuxiang, et al.
Published: (2025)
by: Tian, Aaron Xuxiang, et al.
Published: (2025)
Sola-Visibility-ISPM: Benchmarking Agentic AI for Identity Security Posture Management Visibility
by: Engelberg, Gal, et al.
Published: (2026)
by: Engelberg, Gal, et al.
Published: (2026)
Towards Zero Trust Architecture: A Pilot Study on Information Systems Security Readiness amongst Small and Medium Enterprises
by: Deng, Yu, et al.
Published: (2026)
by: Deng, Yu, et al.
Published: (2026)
A Survey-Based Quantitative Analysis of Stress Factors and Their Impacts Among Cybersecurity Professionals
by: Arora, Sunil, et al.
Published: (2024)
by: Arora, Sunil, et al.
Published: (2024)
Semantic Superiority vs. Forensic Efficiency: A Comparative Analysis of Deep Learning and Psycholinguistics for Business Email Compromise Detection
by: Adjei, Yaw Osei, et al.
Published: (2025)
by: Adjei, Yaw Osei, et al.
Published: (2025)
AegisShield: Democratizing Cyber Threat Modeling with Generative AI
by: Grofsky, Matthew
Published: (2025)
by: Grofsky, Matthew
Published: (2025)
Multilingual AI-Driven Password Strength Estimation with Similarity-Based Detection
by: Palaniappan, Nikitha M., et al.
Published: (2026)
by: Palaniappan, Nikitha M., et al.
Published: (2026)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
by: Othman, Refat
Published: (2026)
by: Othman, Refat
Published: (2026)
Safeguarding Virtual Healthcare: A Novel Attacker-Centric Model for Data Security and Privacy
by: Herath, Suvineetha, et al.
Published: (2024)
by: Herath, Suvineetha, et al.
Published: (2024)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
by: Usman, Rana Muhammad
Published: (2026)
by: Usman, Rana Muhammad
Published: (2026)
Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers
by: Wang, Haochuan Kevin, et al.
Published: (2026)
by: Wang, Haochuan Kevin, et al.
Published: (2026)
Code as a Weapon: A Consensus-Labeled Prompt Bank for Measuring Coding-Model Compliance with Malicious-Code Requests
by: Young, Richard J., et al.
Published: (2026)
by: Young, Richard J., et al.
Published: (2026)
SBASH: a Framework for Designing and Evaluating RAG vs. Prompt-Tuned LLM Honeypots
by: Adebimpe, Adetayo, et al.
Published: (2025)
by: Adebimpe, Adetayo, et al.
Published: (2025)
RouteScan: A Non-Intrusive Approach to Auditing MoE LLMs Safety via Expert Routing Telemetry
by: Lv, Bo, et al.
Published: (2026)
by: Lv, Bo, et al.
Published: (2026)
LLM Detectors Still Fall Short of Real World: Case of LLM-Generated Short News-Like Posts
by: Gameiro, Henrique Da Silva, et al.
Published: (2024)
by: Gameiro, Henrique Da Silva, et al.
Published: (2024)
Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk
by: Wu, Shuai, et al.
Published: (2026)
by: Wu, Shuai, et al.
Published: (2026)
Evaluating the Reliability of Digital Forensic Evidence Discovered by Large Language Model: A Case Study
by: Khatiwala, Jeel Piyushkumar, et al.
Published: (2026)
by: Khatiwala, Jeel Piyushkumar, et al.
Published: (2026)
AgentSentry: Mitigating Indirect Prompt Injection in LLM Agents via Temporal Causal Diagnostics and Context Purification
by: Zhang, Tian, et al.
Published: (2026)
by: Zhang, Tian, et al.
Published: (2026)
CritBench: A Framework for Evaluating Cybersecurity Capabilities of Large Language Models in IEC 61850 Digital Substation Environments
by: Keppler, Gustav, et al.
Published: (2026)
by: Keppler, Gustav, et al.
Published: (2026)
Who Shares What? An Empirical Analysis of Security Conference Content Across Academia and Industry
by: Walter, Lukas, et al.
Published: (2024)
by: Walter, Lukas, et al.
Published: (2024)
Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection
by: Lelle, Travis
Published: (2026)
by: Lelle, Travis
Published: (2026)
Weak Enforcement and Low Compliance in PCI DSS: A Comparative Security Study
by: Park, Soonwon, et al.
Published: (2025)
by: Park, Soonwon, et al.
Published: (2025)
Similar Items
-
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
by: Gu, Yongtong, et al.
Published: (2026) -
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
by: Chen, Renmiao, et al.
Published: (2025) -
Securing Agentic AI Systems -- A Multilayer Security Framework
by: Arora, Sunil, et al.
Published: (2025) -
Benchmarking Large Language Models for IoC Recovery under Adversarial Code Obfuscation and Encryption
by: Morales, Jaime, et al.
Published: (2026) -
CEKER: A Generalizable LLM Framework for Literature Analysis with a Case Study in Unikernel Security
by: Wollman, Alex, et al.
Published: (2024)