PRISM: A Design Framework for Open-Source Foundation Model Safety
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Neumann, Terrence, Jones, Bryan |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Characterising Open Source Co-opetition in Company-hosted Open Source Software Projects: The Cases of PyTorch, TensorFlow, and Transformers
par: Osborne, Cailean, et autres
Publié: (2024)
par: Osborne, Cailean, et autres
Publié: (2024)
Why Companies "Democratise" Artificial Intelligence: The Case of Open Source Software Donations
par: Osborne, Cailean
Publié: (2024)
par: Osborne, Cailean
Publié: (2024)
Narrowing the Gap: Supervised Fine-Tuning of Open-Source LLMs as a Viable Alternative to Proprietary Models for Pedagogical Tools
par: Solano, Lorenzo Lee, et autres
Publié: (2025)
par: Solano, Lorenzo Lee, et autres
Publié: (2025)
The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence
par: White, Matt, et autres
Publié: (2024)
par: White, Matt, et autres
Publié: (2024)
Benchmarking Open-Source Safety Guard Models: A Comprehensive Evaluation
par: Harsh, Reetu Raj, et autres
Publié: (2026)
par: Harsh, Reetu Raj, et autres
Publié: (2026)
Safety Analysis in the Era of Large Language Models: A Case Study of STPA using ChatGPT
par: Qi, Yi, et autres
Publié: (2023)
par: Qi, Yi, et autres
Publié: (2023)
An AI System Evaluation Framework for Advancing AI Safety: Terminology, Taxonomy, Lifecycle Mapping
par: Xia, Boming, et autres
Publié: (2024)
par: Xia, Boming, et autres
Publié: (2024)
Design of a Quality Management System based on the EU Artificial Intelligence Act
par: Mustroph, Henryk, et autres
Publié: (2024)
par: Mustroph, Henryk, et autres
Publié: (2024)
Bridging the Skills Gap: A Course Model for Modern Generative AI Education
par: Bardach, Anya, et autres
Publié: (2025)
par: Bardach, Anya, et autres
Publié: (2025)
The Systems Engineering Approach in Times of Large Language Models
par: Cabrera, Christian, et autres
Publié: (2024)
par: Cabrera, Christian, et autres
Publié: (2024)
Computational Foundations for Strategic Coopetition: Formalizing Trust and Reputation Dynamics
par: Pant, Vik, et autres
Publié: (2025)
par: Pant, Vik, et autres
Publié: (2025)
Computational Foundations for Strategic Coopetition: Formalizing Collective Action and Loyalty
par: Pant, Vik, et autres
Publié: (2026)
par: Pant, Vik, et autres
Publié: (2026)
Artificial Intelligence in Open Source Software Engineering: A Foundation for Sustainability
par: Karim, S M Rakib UI, et autres
Publié: (2026)
par: Karim, S M Rakib UI, et autres
Publié: (2026)
$\texttt{Droid}$: A Resource Suite for AI-Generated Code Detection
par: Orel, Daniil, et autres
Publié: (2025)
par: Orel, Daniil, et autres
Publié: (2025)
Fairness Concerns in App Reviews: A Study on AI-based Mobile Apps
par: Nasab, Ali Rezaei, et autres
Publié: (2024)
par: Nasab, Ali Rezaei, et autres
Publié: (2024)
A governance horizon for ethical-use constraints in open-weight AI models
par: Xu, Weiwei, et autres
Publié: (2026)
par: Xu, Weiwei, et autres
Publié: (2026)
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
A Survey for What Developers Require in AI-powered Tools that Aid in Component Selection in CBSD
par: Ansari, Mahdi Jaberzadeh, et autres
Publié: (2025)
par: Ansari, Mahdi Jaberzadeh, et autres
Publié: (2025)
Enhancing Debugging Skills with AI-Powered Assistance: A Real-Time Tool for Debugging Support
par: Artser, Elizaveta, et autres
Publié: (2026)
par: Artser, Elizaveta, et autres
Publié: (2026)
GenAI Integration into Engineering Education: A Case Study of an Introductory Undergraduate Engineering Course
par: Kozan, Kadir, et autres
Publié: (2026)
par: Kozan, Kadir, et autres
Publié: (2026)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
par: Jewitt, James, et autres
Publié: (2026)
par: Jewitt, James, et autres
Publié: (2026)
Predicting Likely-Vulnerable Code Changes: Machine Learning-based Vulnerability Protections for Android Open Source Project
par: Yim, Keun Soo
Publié: (2024)
par: Yim, Keun Soo
Publié: (2024)
Granite Code Models: A Family of Open Foundation Models for Code Intelligence
par: Mishra, Mayank, et autres
Publié: (2024)
par: Mishra, Mayank, et autres
Publié: (2024)
A Conceptual Framework for Ethical Evaluation of Machine Learning Systems
par: Gupta, Neha R., et autres
Publié: (2024)
par: Gupta, Neha R., et autres
Publié: (2024)
Fuzzy Intelligent System for Student Software Project Evaluation
par: Ogorodova, Anna, et autres
Publié: (2024)
par: Ogorodova, Anna, et autres
Publié: (2024)
Impact of AI-tooling on the Engineering Workspace
par: Chretien, Lena, et autres
Publié: (2024)
par: Chretien, Lena, et autres
Publié: (2024)
Using AI-Based Coding Assistants in Practice: State of Affairs, Perceptions, and Ways Forward
par: Sergeyuk, Agnia, et autres
Publié: (2024)
par: Sergeyuk, Agnia, et autres
Publié: (2024)
Rapid Mobile App Development for Generative AI Agents on MIT App Inventor
par: Gao, Jaida, et autres
Publié: (2024)
par: Gao, Jaida, et autres
Publié: (2024)
Benefits and Risks of Using ChatGPT4 as a Teaching Assistant for Computer Science Students
par: Aragonés-Soria, Yaiza, et autres
Publié: (2024)
par: Aragonés-Soria, Yaiza, et autres
Publié: (2024)
Navigating Fairness: Practitioners' Understanding, Challenges, and Strategies in AI/ML Development
par: Pant, Aastha, et autres
Publié: (2024)
par: Pant, Aastha, et autres
Publié: (2024)
Logic Error Localization in Student Programming Assignments Using Pseudocode and Graph Neural Networks
par: Xu, Zhenyu, et autres
Publié: (2024)
par: Xu, Zhenyu, et autres
Publié: (2024)
Balancing Innovation and Ethics in AI-Driven Software Development
par: Baqar, Mohammad
Publié: (2024)
par: Baqar, Mohammad
Publié: (2024)
Scaling CS1 Support with Compiler-Integrated Conversational AI
par: Renzella, Jake, et autres
Publié: (2024)
par: Renzella, Jake, et autres
Publié: (2024)
Enhancing Educational Efficiency: Generative AI Chatbots and DevOps in Education 4.0
par: Mekić, Edis, et autres
Publié: (2024)
par: Mekić, Edis, et autres
Publié: (2024)
Code-Driven Law NO, Normware SI!
par: Sileno, Giovanni
Publié: (2024)
par: Sileno, Giovanni
Publié: (2024)
An Approach to Detect Abnormal Submissions for CodeWorkout Dataset
par: Hicks, Alex, et autres
Publié: (2024)
par: Hicks, Alex, et autres
Publié: (2024)
Can LLMs Identify Gaps and Misconceptions in Students' Code Explanations?
par: Oli, Priti, et autres
Publié: (2024)
par: Oli, Priti, et autres
Publié: (2024)
An AIC-based approach for articulating unpredictable problems in open complex environments
par: AL-Shareefy, Haider, et autres
Publié: (2024)
par: AL-Shareefy, Haider, et autres
Publié: (2024)
Trustworthy AI in practice: an analysis of practitioners' needs and challenges
par: Baldassarre, Maria Teresa, et autres
Publié: (2024)
par: Baldassarre, Maria Teresa, et autres
Publié: (2024)
AI Act for the Working Programmer
par: Hermanns, Holger, et autres
Publié: (2024)
par: Hermanns, Holger, et autres
Publié: (2024)
Documents similaires
-
Characterising Open Source Co-opetition in Company-hosted Open Source Software Projects: The Cases of PyTorch, TensorFlow, and Transformers
par: Osborne, Cailean, et autres
Publié: (2024) -
Why Companies "Democratise" Artificial Intelligence: The Case of Open Source Software Donations
par: Osborne, Cailean
Publié: (2024) -
Narrowing the Gap: Supervised Fine-Tuning of Open-Source LLMs as a Viable Alternative to Proprietary Models for Pedagogical Tools
par: Solano, Lorenzo Lee, et autres
Publié: (2025) -
The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence
par: White, Matt, et autres
Publié: (2024) -
Benchmarking Open-Source Safety Guard Models: A Comprehensive Evaluation
par: Harsh, Reetu Raj, et autres
Publié: (2026)