AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Zhiqiang, Sun, Huan, Shroff, Ness |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Securing the Future of GenAI: Policy and Technology
by: Christodorescu, Mihai, et al.
Published: (2024)
by: Christodorescu, Mihai, et al.
Published: (2024)
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
by: Tong, Haibo, et al.
Published: (2026)
by: Tong, Haibo, et al.
Published: (2026)
The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI
by: Hall, Peter, et al.
Published: (2025)
by: Hall, Peter, et al.
Published: (2025)
Coordinated Flaw Disclosure for AI: Beyond Security Vulnerabilities
by: Cattell, Sven, et al.
Published: (2024)
by: Cattell, Sven, et al.
Published: (2024)
Governable AI: Provable Safety Under Extreme Threat Models
by: Wang, Donglin, et al.
Published: (2025)
by: Wang, Donglin, et al.
Published: (2025)
Clear, Compelling Arguments: Rethinking the Foundations of Frontier AI Safety Cases
by: Feakins, Shaun, et al.
Published: (2026)
by: Feakins, Shaun, et al.
Published: (2026)
Secure and Trustworthy Artificial Intelligence-Extended Reality (AI-XR) for Metaverses
by: Qayyum, Adnan, et al.
Published: (2022)
by: Qayyum, Adnan, et al.
Published: (2022)
Security, privacy, and agentic AI in a regulatory view: From definitions and distinctions to provisions and reflections
by: Zhang, Shiliang, et al.
Published: (2026)
by: Zhang, Shiliang, et al.
Published: (2026)
Securing the AI Frontier: Urgent Ethical and Regulatory Imperatives for AI-Driven Cybersecurity
by: Kulothungan, Vikram
Published: (2025)
by: Kulothungan, Vikram
Published: (2025)
AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
by: Barrett, Anthony M., et al.
Published: (2025)
by: Barrett, Anthony M., et al.
Published: (2025)
Black-Box Access is Insufficient for Rigorous AI Audits
by: Casper, Stephen, et al.
Published: (2024)
by: Casper, Stephen, et al.
Published: (2024)
Private, Verifiable, and Auditable AI Systems
by: South, Tobin
Published: (2025)
by: South, Tobin
Published: (2025)
Frontier AI's Impact on the Cybersecurity Landscape
by: Potter, Yujin, et al.
Published: (2025)
by: Potter, Yujin, et al.
Published: (2025)
Red Teaming AI Red Teaming
by: Majumdar, Subhabrata, et al.
Published: (2025)
by: Majumdar, Subhabrata, et al.
Published: (2025)
AI Propaganda factories with language models
by: Olejnik, Lukasz
Published: (2025)
by: Olejnik, Lukasz
Published: (2025)
Accelerating AI Development with Cyber Arenas
by: Cashman, William, et al.
Published: (2025)
by: Cashman, William, et al.
Published: (2025)
Comparing AI Agents to Cybersecurity Professionals in Real-World Penetration Testing
by: Lin, Justin W., et al.
Published: (2025)
by: Lin, Justin W., et al.
Published: (2025)
AI-Driven Cyber Threat Intelligence Automation
by: Shah, Shrit, et al.
Published: (2024)
by: Shah, Shrit, et al.
Published: (2024)
Generative AI Security: Challenges and Countermeasures
by: Zhu, Banghua, et al.
Published: (2024)
by: Zhu, Banghua, et al.
Published: (2024)
SecGenAI: Enhancing Security of Cloud-based Generative AI Applications within Australian Critical Technologies of National Interest
by: Haryanto, Christoforus Yoga, et al.
Published: (2024)
by: Haryanto, Christoforus Yoga, et al.
Published: (2024)
Interplay of ISMS and AIMS in context of the EU AI Act
by: Pötsch, Jordan
Published: (2024)
by: Pötsch, Jordan
Published: (2024)
The End Of Universal Lifelong Identifiers: Identity Systems For The AI Era
by: Palakodety, Shriphani
Published: (2025)
by: Palakodety, Shriphani
Published: (2025)
Enabling External Scrutiny of AI Systems with Privacy-Enhancing Technologies
by: Beers, Kendrea, et al.
Published: (2025)
by: Beers, Kendrea, et al.
Published: (2025)
RedTeamLLM: an Agentic AI framework for offensive security
by: Challita, Brian, et al.
Published: (2025)
by: Challita, Brian, et al.
Published: (2025)
ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety
by: Lee, Michael S., et al.
Published: (2026)
by: Lee, Michael S., et al.
Published: (2026)
Naming is framing: How cybersecurity's language problems are repeating in AI governance
by: Potter, Lianne
Published: (2025)
by: Potter, Lianne
Published: (2025)
The Impact of AI on the Cyber Offense-Defense Balance and the Character of Cyber Conflict
by: Lohn, Andrew J.
Published: (2025)
by: Lohn, Andrew J.
Published: (2025)
Is Your AI Truly Yours? Leveraging Blockchain for Copyrights, Provenance, and Lineage
by: Wang, Qin, et al.
Published: (2024)
by: Wang, Qin, et al.
Published: (2024)
Love, Lies, and Language Models: Investigating AI's Role in Romance-Baiting Scams
by: Gressel, Gilad, et al.
Published: (2025)
by: Gressel, Gilad, et al.
Published: (2025)
Can AI Models be Jailbroken to Phish Elderly Victims? An End-to-End Evaluation
by: Heiding, Fred, et al.
Published: (2025)
by: Heiding, Fred, et al.
Published: (2025)
Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
by: Glukhov, David, et al.
Published: (2024)
by: Glukhov, David, et al.
Published: (2024)
To Defend Against Cyber Attacks, We Must Teach AI Agents to Hack
by: Zhuo, Terry Yue, et al.
Published: (2026)
by: Zhuo, Terry Yue, et al.
Published: (2026)
Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
by: Bucknall, Ben, et al.
Published: (2025)
by: Bucknall, Ben, et al.
Published: (2025)
Cryptographic Runtime Governance for Autonomous AI Systems: The Aegis Architecture for Verifiable Policy Enforcement
by: Mazzocchetti, Adam Massimo
Published: (2026)
by: Mazzocchetti, Adam Massimo
Published: (2026)
What Security and Privacy Transparency Users Need from Consumer-Facing Generative AI
by: Cao, Jiaxun, et al.
Published: (2026)
by: Cao, Jiaxun, et al.
Published: (2026)
Trust and Dependability in Blockchain & AI Based MedIoT Applications: Research Challenges and Future Directions
by: Solaiman, Ellis, et al.
Published: (2025)
by: Solaiman, Ellis, et al.
Published: (2025)
Global Challenge for Safe and Secure LLMs Track 1
by: Jia, Xiaojun, et al.
Published: (2024)
by: Jia, Xiaojun, et al.
Published: (2024)
Cyber Shadows: Neutralizing Security Threats with AI and Targeted Policy Measures
by: Schmitt, Marc, et al.
Published: (2025)
by: Schmitt, Marc, et al.
Published: (2025)
A Whole New World: Creating a Parallel-Poisoned Web Only AI-Agents Can See
by: Zychlinski, Shaked
Published: (2025)
by: Zychlinski, Shaked
Published: (2025)
Authenticity Debt and the Synthetic Content Threat Landscape: A Layered Framework for Trust, Provenance, and IP Governance in the Generative AI Era
by: Sengupta, Shubhashis, et al.
Published: (2026)
by: Sengupta, Shubhashis, et al.
Published: (2026)
Similar Items
-
Securing the Future of GenAI: Policy and Technology
by: Christodorescu, Mihai, et al.
Published: (2024) -
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
by: Tong, Haibo, et al.
Published: (2026) -
The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI
by: Hall, Peter, et al.
Published: (2025) -
Coordinated Flaw Disclosure for AI: Beyond Security Vulnerabilities
by: Cattell, Sven, et al.
Published: (2024) -
Governable AI: Provable Safety Under Extreme Threat Models
by: Wang, Donglin, et al.
Published: (2025)