FORTRESS: Frontier Risk Evaluation for National Security and Public Safety
Fuente:
arXiv
Saved in:
| Main Authors: | Knight, Christina Q., Deshpande, Kaustubh, Sirdeshmukh, Ved, Mankikar, Meher, Team, Scale Red, Team, SEAL Research, Michael, Julian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety
by: Lee, Michael S., et al.
Published: (2026)
by: Lee, Michael S., et al.
Published: (2026)
Implicit Intelligence -- Evaluating Agents on What Users Don't Say
by: Sirdeshmukh, Ved, et al.
Published: (2026)
by: Sirdeshmukh, Ved, et al.
Published: (2026)
Reliable Weak-to-Strong Monitoring of LLM Agents
by: Kale, Neil, et al.
Published: (2025)
by: Kale, Neil, et al.
Published: (2025)
MultiChallenge: A Realistic Multi-Turn Conversation Evaluation Benchmark Challenging to Frontier LLMs
by: Sirdeshmukh, Ved, et al.
Published: (2025)
by: Sirdeshmukh, Ved, et al.
Published: (2025)
Search-Time Data Contamination
by: Han, Ziwen, et al.
Published: (2025)
by: Han, Ziwen, et al.
Published: (2025)
Assurance of Frontier AI Built for National Security
by: Pistillo, Matteo, et al.
Published: (2025)
by: Pistillo, Matteo, et al.
Published: (2025)
Physical Grounding of Neural-Plasma Algorithms via Lead-Free KNN Piezoelectric Motile Heterojunctions for National Security Resilience and Critical Materials Independence
by: Venerable, Denise, et al.
Published: (2026)
by: Venerable, Denise, et al.
Published: (2026)
Evaluating AI Providers' Frontier Safety Frameworks
by: Stelling, Lily, et al.
Published: (2025)
by: Stelling, Lily, et al.
Published: (2025)
Enabling Frontier Lab Collaboration to Mitigate AI Safety Risks
by: Felstead, Nicholas
Published: (2025)
by: Felstead, Nicholas
Published: (2025)
Evaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
by: Krishna, Satyapriya, et al.
Published: (2025)
by: Krishna, Satyapriya, et al.
Published: (2025)
Safety Co-Option and Compromised National Security: The Self-Fulfilling Prophecy of Weakened AI Risk Thresholds
by: Khlaaf, Heidy, et al.
Published: (2025)
by: Khlaaf, Heidy, et al.
Published: (2025)
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
by: Tong, Haibo, et al.
Published: (2026)
by: Tong, Haibo, et al.
Published: (2026)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
by: Brundage, Miles, et al.
Published: (2026)
by: Brundage, Miles, et al.
Published: (2026)
Towards Frontier Safety Policies Plus
by: Pistillo, Matteo
Published: (2025)
by: Pistillo, Matteo
Published: (2025)
AI-Powered Upskilling at Larsen & Toubro
by: CoachBoTs Research Team
Published: (2025)
by: CoachBoTs Research Team
Published: (2025)
The Role of AI Safety Institutes in Contributing to International Standards for Frontier AI Safety
by: Fort, Kristina
Published: (2024)
by: Fort, Kristina
Published: (2024)
Safety Cases: A Scalable Approach to Frontier AI Safety
by: Hilton, Benjamin, et al.
Published: (2025)
by: Hilton, Benjamin, et al.
Published: (2025)
Emerging Practices in Frontier AI Safety Frameworks
by: Buhl, Marie Davidsen, et al.
Published: (2025)
by: Buhl, Marie Davidsen, et al.
Published: (2025)
Jailbreaking to Jailbreak
by: Kritz, Jeremy, et al.
Published: (2025)
by: Kritz, Jeremy, et al.
Published: (2025)
Astrophysics Research Organizations in the 21st Century: Database and Comparative Dashboards
by: Kurtz, Michael J., et al.
Published: (2026)
by: Kurtz, Michael J., et al.
Published: (2026)
Aim High, Stay Private: Differentially Private Synthetic Data Enables Public Release of Behavioral Health Information with High Utility
by: Ghasemizade, Mohsen, et al.
Published: (2025)
by: Ghasemizade, Mohsen, et al.
Published: (2025)
Safety and Security Analysis of Large Language Models: Benchmarking Risk Profile and Harm Potential
by: Akiri, Charankumar, et al.
Published: (2025)
by: Akiri, Charankumar, et al.
Published: (2025)
Global Cybercrime Damages: A Baseline for Frontier AI Risk Assessment
by: Lukošiūtė, Kamilė, et al.
Published: (2026)
by: Lukošiūtė, Kamilė, et al.
Published: (2026)
Sabotage Evaluations for Frontier Models
by: Benton, Joe, et al.
Published: (2024)
by: Benton, Joe, et al.
Published: (2024)
An Empirical Analysis on the Use and Reporting of National Security Letters
by: Bellon, Alex, et al.
Published: (2024)
by: Bellon, Alex, et al.
Published: (2024)
Generative AI in Saudi Arabia: A National Survey of Adoption, Risks, and Public Perceptions
by: AlDakheel, Abdulaziz, et al.
Published: (2026)
by: AlDakheel, Abdulaziz, et al.
Published: (2026)
Clear, Compelling Arguments: Rethinking the Foundations of Frontier AI Safety Cases
by: Feakins, Shaun, et al.
Published: (2026)
by: Feakins, Shaun, et al.
Published: (2026)
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
by: Chen, Zhongren, et al.
Published: (2026)
by: Chen, Zhongren, et al.
Published: (2026)
Expanding External Access To Frontier AI Models For Dangerous Capability Evaluations
by: Charnock, Jacob, et al.
Published: (2026)
by: Charnock, Jacob, et al.
Published: (2026)
From Frontier to Shadow AI: A Simmering Threat to Assurance and Security in Critical Infrastructure
by: Chhetri, Mohan Baruwal, et al.
Published: (2026)
by: Chhetri, Mohan Baruwal, et al.
Published: (2026)
Astra: AI Safety, Trust, & Risk Assessment
by: Aggarwal, Pranav, et al.
Published: (2026)
by: Aggarwal, Pranav, et al.
Published: (2026)
Open Problems in Frontier AI Risk Management
by: Ziosi, Marta, et al.
Published: (2026)
by: Ziosi, Marta, et al.
Published: (2026)
Evaluating the Clinical Safety of LLMs in Response to High-Risk Mental Health Disclosures
by: Shah, Siddharth, et al.
Published: (2025)
by: Shah, Siddharth, et al.
Published: (2025)
DOI Identifiers ...Now available for every published article
by: Hwalgi Team
Published: (2025)
by: Hwalgi Team
Published: (2025)
AmesPAHdbIDLSuite
by: The PAHdb Team
Published: (2020)
by: The PAHdb Team
Published: (2020)
github.com/lilab-bcb/cumulus/Cellranger_atac_aggr
by: Cumulus Team
Published: (2025)
by: Cumulus Team
Published: (2025)
Qwen3.5-Omni Technical Report
by: Qwen Team
Published: (2026)
by: Qwen Team
Published: (2026)
Chameleon: Mixed-Modal Early-Fusion Foundation Models
by: Chameleon Team
Published: (2024)
by: Chameleon Team
Published: (2024)
Psychosocial Intervention enters a new phase
by: The Editorial Team
Published: (2011)
by: The Editorial Team
Published: (2011)
Chemical and mineral composition of rocks and sediments from the Amirante Arc
by: FEGI Team
Published: (1997)
by: FEGI Team
Published: (1997)
Similar Items
-
ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety
by: Lee, Michael S., et al.
Published: (2026) -
Implicit Intelligence -- Evaluating Agents on What Users Don't Say
by: Sirdeshmukh, Ved, et al.
Published: (2026) -
Reliable Weak-to-Strong Monitoring of LLM Agents
by: Kale, Neil, et al.
Published: (2025) -
MultiChallenge: A Realistic Multi-Turn Conversation Evaluation Benchmark Challenging to Frontier LLMs
by: Sirdeshmukh, Ved, et al.
Published: (2025) -
Search-Time Data Contamination
by: Han, Ziwen, et al.
Published: (2025)