Report on NSF Workshop on Science of Safe AI
Fuente:
arXiv
Saved in:
| Main Authors: | Alur, Rajeev, Durrett, Greg, Kress-Gazit, Hadas, Păsăreanu, Corina, Vidal, René |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fighting AI with AI: Leveraging Foundation Models for Assuring AI-Enabled Safety-Critical Systems
by: Mavridou, Anastasia, et al.
Published: (2025)
by: Mavridou, Anastasia, et al.
Published: (2025)
Report from Workshop on Dialogue alongside Artificial Intelligence
by: McKenna, Thomas J, et al.
Published: (2025)
by: McKenna, Thomas J, et al.
Published: (2025)
Legal Zero-Days: A Novel Risk Vector for Advanced AI Systems
by: Sadler, Greg, et al.
Published: (2025)
by: Sadler, Greg, et al.
Published: (2025)
Molecular Facts: Desiderata for Decontextualization in LLM Fact Verification
by: Gunjal, Anisha, et al.
Published: (2024)
by: Gunjal, Anisha, et al.
Published: (2024)
Contemporary AI foundation models increase biological weapons risk
by: Brent, Roger, et al.
Published: (2025)
by: Brent, Roger, et al.
Published: (2025)
Towards Safe Multilingual Frontier AI
by: Kanepajs, Artūrs, et al.
Published: (2024)
by: Kanepajs, Artūrs, et al.
Published: (2024)
Advancing Trustworthy AI in Healthcare Through Meta-Research: Results of an Interdisciplinary Design-Thinking Workshop
by: Bürger, Valerie, et al.
Published: (2026)
by: Bürger, Valerie, et al.
Published: (2026)
GeoAI in Social Science
by: Li, Wenwen
Published: (2023)
by: Li, Wenwen
Published: (2023)
The Contribution of XAI for the Safe Development and Certification of AI: An Expert-Based Analysis
by: Fresz, Benjamin, et al.
Published: (2024)
by: Fresz, Benjamin, et al.
Published: (2024)
SynthesizRR: Generating Diverse Datasets with Retrieval Augmentation
by: Divekar, Abhishek, et al.
Published: (2024)
by: Divekar, Abhishek, et al.
Published: (2024)
AI Literacy, Safety Awareness, and STEM Career Aspirations of Australian Secondary Students: Evaluating the Impact of Workshop Interventions
by: Bergh, Christian, et al.
Published: (2026)
by: Bergh, Christian, et al.
Published: (2026)
AI-washing: The Asymmetric Effects of Its Two Types on Consumer Moral Judgments
by: Nyilasy, Greg, et al.
Published: (2025)
by: Nyilasy, Greg, et al.
Published: (2025)
Report on the NSF Workshop on Sustainable Computing for Sustainability (NSF WSCS 2024)
by: Guérin, Roch, et al.
Published: (2024)
by: Guérin, Roch, et al.
Published: (2024)
Report for NSF Workshop on AI for Electronic Design Automation
by: Chen, Deming, et al.
Published: (2026)
by: Chen, Deming, et al.
Published: (2026)
Toward Safe and Responsible AI Agents: A Three-Pillar Model for Transparency, Accountability, and Trustworthiness
by: Cheng, Edward C., et al.
Published: (2026)
by: Cheng, Edward C., et al.
Published: (2026)
Can an AI Agent Safely Run a Government? Existence of Probably Approximately Aligned Policies
by: Berdoz, Frédéric, et al.
Published: (2024)
by: Berdoz, Frédéric, et al.
Published: (2024)
Assessing AI vs Human-Authored Spear Phishing SMS Attacks: An Empirical Study
by: Francia, Jerson, et al.
Published: (2024)
by: Francia, Jerson, et al.
Published: (2024)
International Scientific Report on the Safety of Advanced AI (Interim Report)
by: Bengio, Yoshua, et al.
Published: (2024)
by: Bengio, Yoshua, et al.
Published: (2024)
Designing Safe and Accountable GenAI as a Learning Companion with Women Banned from Formal Education
by: Behmanush, Hamayoon, et al.
Published: (2026)
by: Behmanush, Hamayoon, et al.
Published: (2026)
Calibrate-Then-Act: Cost-Aware Exploration in LLM Agents
by: Ding, Wenxuan, et al.
Published: (2026)
by: Ding, Wenxuan, et al.
Published: (2026)
MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents
by: Tang, Liyan, et al.
Published: (2024)
by: Tang, Liyan, et al.
Published: (2024)
Responsible Reporting for Frontier AI Development
by: Kolt, Noam, et al.
Published: (2024)
by: Kolt, Noam, et al.
Published: (2024)
Toward an Engineering of Science: Rebalancing Generation and Verification in the Age of AI
by: Ma, Jiaqi W.
Published: (2026)
by: Ma, Jiaqi W.
Published: (2026)
The Landscape of AI in Science Education: What is Changing and How to Respond
by: Zhai, Xiaoming, et al.
Published: (2026)
by: Zhai, Xiaoming, et al.
Published: (2026)
A Moonshot for AI Oracles in the Sciences
by: Kaiser, Bryan, et al.
Published: (2024)
by: Kaiser, Bryan, et al.
Published: (2024)
Agentic AI Software Engineers: Programming with Trust
by: Roychoudhury, Abhik, et al.
Published: (2025)
by: Roychoudhury, Abhik, et al.
Published: (2025)
Safe and Certifiable AI Systems: Concepts, Challenges, and Lessons Learned
by: Schweighofer, Kajetan, et al.
Published: (2025)
by: Schweighofer, Kajetan, et al.
Published: (2025)
Ensuring Computer Science Learning in the AI Era: Open Generative AI Policies and Assignment-Driven Written Quizzes
by: Chung, Chan-Jin
Published: (2026)
by: Chung, Chan-Jin
Published: (2026)
STREAM (ChemBio): A Standard for Transparently Reporting Evaluations in AI Model Reports
by: McCaslin, Tegan, et al.
Published: (2025)
by: McCaslin, Tegan, et al.
Published: (2025)
Breaking Up with Normatively Monolithic Agency with GRACE: A Reason-Based Neuro-Symbolic Architecture for Safe and Ethical AI Alignment
by: Jahn, Felix, et al.
Published: (2026)
by: Jahn, Felix, et al.
Published: (2026)
Responsible AI Governance: A Response to UN Interim Report on Governing AI for Humanity
by: Kiden, Sarah, et al.
Published: (2024)
by: Kiden, Sarah, et al.
Published: (2024)
From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law
by: Mavi, John, et al.
Published: (2025)
by: Mavi, John, et al.
Published: (2025)
Risk Reporting for Developers' Internal AI Model Use
by: Delaney, Oscar, et al.
Published: (2026)
by: Delaney, Oscar, et al.
Published: (2026)
AI Mentors for Student Projects: Spotting Early Issues in Computer Science Proposals
by: Aher, Gati, et al.
Published: (2025)
by: Aher, Gati, et al.
Published: (2025)
Charting the Future of AI-supported Science Education: A Human-Centered Vision
by: Zhai, Xiaoming, et al.
Published: (2026)
by: Zhai, Xiaoming, et al.
Published: (2026)
Towards a Healthy AI Tradition: Lessons from Biology and Biomedical Science
by: Kasif, Simon
Published: (2024)
by: Kasif, Simon
Published: (2024)
Envisioning National Resources for Artificial Intelligence Research: NSF Workshop Report
by: Jha, Shantenu, et al.
Published: (2024)
by: Jha, Shantenu, et al.
Published: (2024)
Computational Thinking with Computer Vision: Developing AI Competency in an Introductory Computer Science Course
by: Chowdhury, Tahiya
Published: (2025)
by: Chowdhury, Tahiya
Published: (2025)
Google, AI Literacy, and the Learning Sciences: Multiple Modes of Research, Industry, and Practice Partnerships
by: Lee, Victor R., et al.
Published: (2026)
by: Lee, Victor R., et al.
Published: (2026)
AI Safety is Stuck in Technical Terms -- A System Safety Response to the International AI Safety Report
by: Dobbe, Roel
Published: (2025)
by: Dobbe, Roel
Published: (2025)
Similar Items
-
Fighting AI with AI: Leveraging Foundation Models for Assuring AI-Enabled Safety-Critical Systems
by: Mavridou, Anastasia, et al.
Published: (2025) -
Report from Workshop on Dialogue alongside Artificial Intelligence
by: McKenna, Thomas J, et al.
Published: (2025) -
Legal Zero-Days: A Novel Risk Vector for Advanced AI Systems
by: Sadler, Greg, et al.
Published: (2025) -
Molecular Facts: Desiderata for Decontextualization in LLM Fact Verification
by: Gunjal, Anisha, et al.
Published: (2024) -
Contemporary AI foundation models increase biological weapons risk
by: Brent, Roger, et al.
Published: (2025)