Assessing a Safety Case: Bottom-up Guidance for Claims and Evidence Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Schnelle, Scott, Favaro, Francesca, Fraade-Blanar, Laura, Wichner, David, Broce, Holland, Miranda, Justin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Determining Absence of Unreasonable Risk: Approval Guidelines for an Automated Driving System Deployment
by: Favaro, Francesca, et al.
Published: (2025)
by: Favaro, Francesca, et al.
Published: (2025)
Safety Drift After Fine-Tuning: Evidence from High-Stakes Domains
by: Khan, Emaan Bilal, et al.
Published: (2026)
by: Khan, Emaan Bilal, et al.
Published: (2026)
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
by: Vishwarupe, Varad, et al.
Published: (2026)
by: Vishwarupe, Varad, et al.
Published: (2026)
Measuring What Matters: A Framework for Evaluating Safety Risks in Real-World LLM Applications
by: Goh, Jia Yi, et al.
Published: (2025)
by: Goh, Jia Yi, et al.
Published: (2025)
Being good (at driving): Characterizing behavioral expectations on automated and human driven vehicles
by: Fraade-Blanar, Laura, et al.
Published: (2025)
by: Fraade-Blanar, Laura, et al.
Published: (2025)
Counting the Trees in the Forest: Evaluating Prompt Segmentation for Classifying Code Comprehension Level
by: Smith IV, David H., et al.
Published: (2025)
by: Smith IV, David H., et al.
Published: (2025)
What Should Frontier AI Developers Disclose About Internal Deployments?
by: Charnock, Jacob, et al.
Published: (2026)
by: Charnock, Jacob, et al.
Published: (2026)
How frontier AI companies could implement an internal audit function
by: Gomez, Francesca, et al.
Published: (2025)
by: Gomez, Francesca, et al.
Published: (2025)
Assessing Simulation Knowledge and Proficiency Among Undergraduate Computing Students in Brazil: Insights and Results from a Survey Research
by: Rodrigues, Fernando Brito, et al.
Published: (2025)
by: Rodrigues, Fernando Brito, et al.
Published: (2025)
LLM Contribution Summarization in Software Projects
by: Ferrao, Rafael Corsi, et al.
Published: (2025)
by: Ferrao, Rafael Corsi, et al.
Published: (2025)
Towards More Empathic Programming Environments: An Experimental Empathic AI-Enhanced IDE
by: Go, Justin Rainier, et al.
Published: (2026)
by: Go, Justin Rainier, et al.
Published: (2026)
The Evaluation of Open Source Software Innovativeness
by: Benkeltoum, Nordine
Published: (2025)
by: Benkeltoum, Nordine
Published: (2025)
Generative AI Adoption in an Energy Company: Exploring Challenges and Use Cases
by: Sami, Malik Abdul, et al.
Published: (2026)
by: Sami, Malik Abdul, et al.
Published: (2026)
Challenges for Inclusion in Software Engineering: The Case of the Emerging Papua New Guinean Society
by: Kula, Raula Gaikovina, et al.
Published: (2019)
by: Kula, Raula Gaikovina, et al.
Published: (2019)
The Case for Contextual Copyleft: Licensing Open Source Training Data and Generative AI
by: Shanklin, Grant, et al.
Published: (2025)
by: Shanklin, Grant, et al.
Published: (2025)
Evaluation of Systems Programming Exercises through Tailored Static Analysis
by: Natella, Roberto
Published: (2024)
by: Natella, Roberto
Published: (2024)
Open, Small, Rigmarole -- Evaluating Llama 3.2 3B's Feedback for Programming Exercises
by: Azaiz, Imen, et al.
Published: (2025)
by: Azaiz, Imen, et al.
Published: (2025)
An Integrated Usability Framework for Evaluating Open Government Data Portals: Comparative Analysis of EU and GCC Countries
by: Molodtsov, Fillip, et al.
Published: (2024)
by: Molodtsov, Fillip, et al.
Published: (2024)
Discovering Ideologies of the Open Source Software Movement
by: Yue, Yang, et al.
Published: (2025)
by: Yue, Yang, et al.
Published: (2025)
Evaluating the Effectiveness of OpenAI's Parental Control System
by: Ersoz, Kerem, et al.
Published: (2026)
by: Ersoz, Kerem, et al.
Published: (2026)
PRISM: A Design Framework for Open-Source Foundation Model Safety
by: Neumann, Terrence, et al.
Published: (2024)
by: Neumann, Terrence, et al.
Published: (2024)
Ideology in Open Source Development
by: Yue, Yang, et al.
Published: (2021)
by: Yue, Yang, et al.
Published: (2021)
An AI System Evaluation Framework for Advancing AI Safety: Terminology, Taxonomy, Lifecycle Mapping
by: Xia, Boming, et al.
Published: (2024)
by: Xia, Boming, et al.
Published: (2024)
Safety Analysis in the Era of Large Language Models: A Case Study of STPA using ChatGPT
by: Qi, Yi, et al.
Published: (2023)
by: Qi, Yi, et al.
Published: (2023)
Ten simple rules for training scientists to make better software
by: Gallagher, Kit, et al.
Published: (2024)
by: Gallagher, Kit, et al.
Published: (2024)
CS1-LLM: Integrating LLMs into CS1 Instruction
by: Vadaparty, Annapurna, et al.
Published: (2024)
by: Vadaparty, Annapurna, et al.
Published: (2024)
Robustness tests for biomedical foundation models should tailor to specifications
by: Xian, R. Patrick, et al.
Published: (2025)
by: Xian, R. Patrick, et al.
Published: (2025)
To Err is AI : A Case Study Informing LLM Flaw Reporting Practices
by: McGregor, Sean, et al.
Published: (2024)
by: McGregor, Sean, et al.
Published: (2024)
Why Companies "Democratise" Artificial Intelligence: The Case of Open Source Software Donations
by: Osborne, Cailean
Published: (2024)
by: Osborne, Cailean
Published: (2024)
Developing Compelling Safety Cases
by: Hawkins, Richard
Published: (2025)
by: Hawkins, Richard
Published: (2025)
Fuzzy Intelligent System for Student Software Project Evaluation
by: Ogorodova, Anna, et al.
Published: (2024)
by: Ogorodova, Anna, et al.
Published: (2024)
GenAI Integration into Engineering Education: A Case Study of an Introductory Undergraduate Engineering Course
by: Kozan, Kadir, et al.
Published: (2026)
by: Kozan, Kadir, et al.
Published: (2026)
LLM Use, Cheating, and Academic Integrity in Software Engineering Education
by: Santos, Ronnie de Souza, et al.
Published: (2026)
by: Santos, Ronnie de Souza, et al.
Published: (2026)
FastFixer: An Efficient and Effective Approach for Repairing Programming Assignments
by: Liu, Fang, et al.
Published: (2024)
by: Liu, Fang, et al.
Published: (2024)
A Scenario Analysis of Ethical Issues in Dark Patterns and Their Research
by: Ruohonen, Jukka, et al.
Published: (2025)
by: Ruohonen, Jukka, et al.
Published: (2025)
Towards a Knowledge Base of Common Sustainability Weaknesses in Green Software Development
by: Pathania, Priyavanshi, et al.
Published: (2025)
by: Pathania, Priyavanshi, et al.
Published: (2025)
From Generation to Adaptation: Comparing AI-Assisted Strategies in High School Programming Education
by: Hu, Tong, et al.
Published: (2025)
by: Hu, Tong, et al.
Published: (2025)
Overcoming Obstacles: Challenges of Gender Inequality in Undergraduate ICT Programs
by: Souza, Angelica Pereira, et al.
Published: (2025)
by: Souza, Angelica Pereira, et al.
Published: (2025)
The State of Computational Science in Fission and Fusion Energy
by: Coto, Andrea Morales, et al.
Published: (2025)
by: Coto, Andrea Morales, et al.
Published: (2025)
Understanding: reframing automation and assurance
by: Bloomfield, Robin
Published: (2026)
by: Bloomfield, Robin
Published: (2026)
Similar Items
-
Determining Absence of Unreasonable Risk: Approval Guidelines for an Automated Driving System Deployment
by: Favaro, Francesca, et al.
Published: (2025) -
Safety Drift After Fine-Tuning: Evidence from High-Stakes Domains
by: Khan, Emaan Bilal, et al.
Published: (2026) -
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
by: Vishwarupe, Varad, et al.
Published: (2026) -
Measuring What Matters: A Framework for Evaluating Safety Risks in Real-World LLM Applications
by: Goh, Jia Yi, et al.
Published: (2025) -
Being good (at driving): Characterizing behavioral expectations on automated and human driven vehicles
by: Fraade-Blanar, Laura, et al.
Published: (2025)