AI threats to national security can be countered through an incident regime
Fuente:
arXiv
Saved in:
| Main Author: | Ortega, Alejandro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Combatting deepfakes: Policies to address national security threats and rights violations
by: Miotti, Andrea, et al.
Published: (2024)
by: Miotti, Andrea, et al.
Published: (2024)
Deep opacity and AI: A threat to XAI and to privacy protection mechanisms
by: Müller, Vincent C.
Published: (2025)
by: Müller, Vincent C.
Published: (2025)
AI incidents and 'networked trouble': The case for a research agenda
by: Shane, Tommy Shaffer
Published: (2024)
by: Shane, Tommy Shaffer
Published: (2024)
Sleeper Social Bots: a new generation of AI disinformation bots are already a political threat
by: Doshi, Jaiv, et al.
Published: (2024)
by: Doshi, Jaiv, et al.
Published: (2024)
Designing escalation criteria for international AI incident response: criteria, triggers, and thresholds
by: Gomez, Francesca, et al.
Published: (2026)
by: Gomez, Francesca, et al.
Published: (2026)
Scheming in the wild: detecting real-world AI scheming incidents with open-source intelligence
by: Shane, Tommy Shaffer, et al.
Published: (2026)
by: Shane, Tommy Shaffer, et al.
Published: (2026)
What AI evaluations for preventing catastrophic risks can and cannot do
by: Barnett, Peter, et al.
Published: (2024)
by: Barnett, Peter, et al.
Published: (2024)
Governing dual-use technologies: Case studies of international security agreements and lessons for AI governance
by: Wasil, Akash R., et al.
Published: (2024)
by: Wasil, Akash R., et al.
Published: (2024)
AI for All: Identifying AI incidents Related to Diversity and Inclusion
by: Shams, Rifat Ara, et al.
Published: (2024)
by: Shams, Rifat Ara, et al.
Published: (2024)
The threat of analytic flexibility in using large language models to simulate human data
by: Cummins, Jamie
Published: (2025)
by: Cummins, Jamie
Published: (2025)
Deceptive AI systems that give explanations are more convincing than honest AI systems and can amplify belief in misinformation
by: Danry, Valdemar, et al.
Published: (2024)
by: Danry, Valdemar, et al.
Published: (2024)
Where can AI be used? Insights from a deep ontology of work activities
by: Cai, Alice, et al.
Published: (2026)
by: Cai, Alice, et al.
Published: (2026)
A new framework for prognostics in decentralized industries: Enhancing fairness, security, and transparency through Blockchain and Federated Learning
by: Pham, T. Q. D., et al.
Published: (2025)
by: Pham, T. Q. D., et al.
Published: (2025)
Standardised schema and taxonomy for AI incident databases in critical digital infrastructure
by: Agarwal, Avinash, et al.
Published: (2025)
by: Agarwal, Avinash, et al.
Published: (2025)
RedTeamLLM: an Agentic AI framework for offensive security
by: Challita, Brian, et al.
Published: (2025)
by: Challita, Brian, et al.
Published: (2025)
Incorporating AI incident reporting into telecommunications law and policy: Insights from India
by: Agarwal, Avinash, et al.
Published: (2025)
by: Agarwal, Avinash, et al.
Published: (2025)
AI Emergency Preparedness: Examining the federal government's ability to detect and respond to AI-related national security threats
by: Wasil, Akash, et al.
Published: (2024)
by: Wasil, Akash, et al.
Published: (2024)
Using AI Alignment Theory to understand the potential pitfalls of regulatory frameworks
by: Tlaie, Alejandro
Published: (2024)
by: Tlaie, Alejandro
Published: (2024)
The Loss of Control Playbook: Degrees, Dynamics, and Preparedness
by: Stix, Charlotte, et al.
Published: (2025)
by: Stix, Charlotte, et al.
Published: (2025)
Why can't Epidemiology be automated (yet)?
by: Bann, David, et al.
Published: (2025)
by: Bann, David, et al.
Published: (2025)
AI Sandbagging: Language Models can Strategically Underperform on Evaluations
by: van der Weij, Teun, et al.
Published: (2024)
by: van der Weij, Teun, et al.
Published: (2024)
Deconstructing Student Perceptions of Generative AI (GenAI) through an Expectancy Value Theory (EVT)-based Instrument
by: Chan, Cecilia Ka Yuk, et al.
Published: (2023)
by: Chan, Cecilia Ka Yuk, et al.
Published: (2023)
Distributed agency in second language learning and teaching through generative AI
by: Godwin-Jones, Robert
Published: (2024)
by: Godwin-Jones, Robert
Published: (2024)
Tracing GenAI Literacy: Uncovering Student-AI Interaction Patterns in Academic Writing through Epistemic Network Analysis
by: Chen, Angxuan, et al.
Published: (2026)
by: Chen, Angxuan, et al.
Published: (2026)
General-purpose AI models can generate actionable knowledge on agroecological crop protection
by: Wyckhuys, Kris A. G.
Published: (2025)
by: Wyckhuys, Kris A. G.
Published: (2025)
AI tutoring can safely and effectively support students: An exploratory RCT in UK classrooms
by: LearnLM Team, et al.
Published: (2025)
by: LearnLM Team, et al.
Published: (2025)
Mitigating loss of control in advanced AI systems through instrumental goal trajectories
by: Fourie, Willem
Published: (2026)
by: Fourie, Willem
Published: (2026)
Achieving Responsible AI through ESG: Insights and Recommendations from Industry Engagement
by: Perera, Harsha, et al.
Published: (2024)
by: Perera, Harsha, et al.
Published: (2024)
Quota-based debiasing can decrease representation of already underrepresented groups
by: Smirnov, Ivan, et al.
Published: (2020)
by: Smirnov, Ivan, et al.
Published: (2020)
Commercial Persuasion in AI-Mediated Conversations
by: Salvi, Francesco, et al.
Published: (2026)
by: Salvi, Francesco, et al.
Published: (2026)
Exploring Societal Concerns and Perceptions of AI: A Thematic Analysis through the Lens of Problem-Seeking
by: Kayembe, Naomi Omeonga wa
Published: (2025)
by: Kayembe, Naomi Omeonga wa
Published: (2025)
Between Fear and Desire, the Monster Artificial Intelligence (AI): Analysis through the Lenses of Monster Theory
by: Tlili, Ahmed
Published: (2025)
by: Tlili, Ahmed
Published: (2025)
"My Kind of Woman": Analysing Gender Stereotypes in AI through The Averageness Theory and EU Law
by: Doh, Miriam, et al.
Published: (2024)
by: Doh, Miriam, et al.
Published: (2024)
Foundation models may exhibit staged progression in novel CBRN threat disclosure
by: Esvelt, Kevin M
Published: (2025)
by: Esvelt, Kevin M
Published: (2025)
Imitating AI agents increase diversity in homogeneous information environments but can reduce it in heterogeneous ones
by: Johansen, Emil Bakkensen, et al.
Published: (2025)
by: Johansen, Emil Bakkensen, et al.
Published: (2025)
What can large language models do for sustainable food?
by: Thomas, Anna T., et al.
Published: (2025)
by: Thomas, Anna T., et al.
Published: (2025)
How malicious AI swarms can threaten democracy: The fusion of agentic AI and LLMs marks a new frontier in information warfare
by: Schroeder, Daniel Thilo, et al.
Published: (2025)
by: Schroeder, Daniel Thilo, et al.
Published: (2025)
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
by: Nghiem, Huy, et al.
Published: (2025)
by: Nghiem, Huy, et al.
Published: (2025)
Increasing intelligence in AI agents can worsen collective outcomes
by: Johnson, Neil F.
Published: (2026)
by: Johnson, Neil F.
Published: (2026)
Antisocial Analagous Behavior, Alignment and Human Impact of Google AI Systems: Evaluating through the lens of modified Antisocial Behavior Criteria by Human Interaction, Independent LLM Analysis, and AI Self-Reflection
by: Ogilvie, Alan D.
Published: (2024)
by: Ogilvie, Alan D.
Published: (2024)
Similar Items
-
Combatting deepfakes: Policies to address national security threats and rights violations
by: Miotti, Andrea, et al.
Published: (2024) -
Deep opacity and AI: A threat to XAI and to privacy protection mechanisms
by: Müller, Vincent C.
Published: (2025) -
AI incidents and 'networked trouble': The case for a research agenda
by: Shane, Tommy Shaffer
Published: (2024) -
Sleeper Social Bots: a new generation of AI disinformation bots are already a political threat
by: Doshi, Jaiv, et al.
Published: (2024) -
Designing escalation criteria for international AI incident response: criteria, triggers, and thresholds
by: Gomez, Francesca, et al.
Published: (2026)