NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vishwarupe, Varad, Shadbolt, Nigel, Jirotka, Marina, Flechais, Ivan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
The Collaboration Gap in Human-AI Work
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
From Rights to Rites: Expectations Management in Smart-Home AI
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
An AI System Evaluation Framework for Advancing AI Safety: Terminology, Taxonomy, Lifecycle Mapping
von: Xia, Boming, et al.
Veröffentlicht: (2024)
von: Xia, Boming, et al.
Veröffentlicht: (2024)
The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence
von: White, Matt, et al.
Veröffentlicht: (2024)
von: White, Matt, et al.
Veröffentlicht: (2024)
Standardization Trends on Safety and Trustworthiness Technology for Advanced AI
von: Jeon, Jonghong
Veröffentlicht: (2024)
von: Jeon, Jonghong
Veröffentlicht: (2024)
A Survey for What Developers Require in AI-powered Tools that Aid in Component Selection in CBSD
von: Ansari, Mahdi Jaberzadeh, et al.
Veröffentlicht: (2025)
von: Ansari, Mahdi Jaberzadeh, et al.
Veröffentlicht: (2025)
Securing the AI Frontier: Urgent Ethical and Regulatory Imperatives for AI-Driven Cybersecurity
von: Kulothungan, Vikram
Veröffentlicht: (2025)
von: Kulothungan, Vikram
Veröffentlicht: (2025)
Rethinking Technological Readiness in the Era of AI Uncertainty
von: Browne, S. Tucker, et al.
Veröffentlicht: (2025)
von: Browne, S. Tucker, et al.
Veröffentlicht: (2025)
Green AI in Action: Strategic Model Selection for Ensembles in Production
von: Nijkamp, Nienke, et al.
Veröffentlicht: (2024)
von: Nijkamp, Nienke, et al.
Veröffentlicht: (2024)
Privacy and Copyright Protection in Generative AI: A Lifecycle Perspective
von: Zhang, Dawen, et al.
Veröffentlicht: (2023)
von: Zhang, Dawen, et al.
Veröffentlicht: (2023)
Legal Aspects for Software Developers Interested in Generative AI Applications
von: Herbold, Steffen, et al.
Veröffentlicht: (2024)
von: Herbold, Steffen, et al.
Veröffentlicht: (2024)
Leveraging Generative AI for Enhancing Automated Assessment in Programming Education Contests
von: Dascalescu, Stefan, et al.
Veröffentlicht: (2025)
von: Dascalescu, Stefan, et al.
Veröffentlicht: (2025)
NeurIPS should lead scientific consensus on AI policy
von: Bommasani, Rishi
Veröffentlicht: (2025)
von: Bommasani, Rishi
Veröffentlicht: (2025)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
von: Jewitt, James, et al.
Veröffentlicht: (2026)
von: Jewitt, James, et al.
Veröffentlicht: (2026)
DC-Check: A Data-Centric AI checklist to guide the development of reliable machine learning systems
von: Seedat, Nabeel, et al.
Veröffentlicht: (2022)
von: Seedat, Nabeel, et al.
Veröffentlicht: (2022)
Optimizing Travel Itineraries with AI Algorithms in a Microservices Architecture: Balancing Cost, Time, Preferences, and Sustainability
von: Barua, Biman, et al.
Veröffentlicht: (2024)
von: Barua, Biman, et al.
Veröffentlicht: (2024)
What Should Frontier AI Developers Disclose About Internal Deployments?
von: Charnock, Jacob, et al.
Veröffentlicht: (2026)
von: Charnock, Jacob, et al.
Veröffentlicht: (2026)
AI-augmented Cybersecurity Requirements Generation using LLMs | Reproducible Research Package
von: Yelmo, Juan Carlos, et al.
Veröffentlicht: (2025)
von: Yelmo, Juan Carlos, et al.
Veröffentlicht: (2025)
PRISM: A Design Framework for Open-Source Foundation Model Safety
von: Neumann, Terrence, et al.
Veröffentlicht: (2024)
von: Neumann, Terrence, et al.
Veröffentlicht: (2024)
MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench Competition
von: Cofala, Tim, et al.
Veröffentlicht: (2025)
von: Cofala, Tim, et al.
Veröffentlicht: (2025)
Towards Equitable Agile Research and Development of AI and Robotics
von: Hundt, Andrew, et al.
Veröffentlicht: (2024)
von: Hundt, Andrew, et al.
Veröffentlicht: (2024)
ReqBrain: Task-Specific Instruction Tuning of LLMs for AI-Assisted Requirements Generation
von: Habib, Mohammad Kasra, et al.
Veröffentlicht: (2025)
von: Habib, Mohammad Kasra, et al.
Veröffentlicht: (2025)
Natural Language Requirements Testability Measurement Based on Requirement Smells
von: Zakeri-Nasrabadi, Morteza, et al.
Veröffentlicht: (2024)
von: Zakeri-Nasrabadi, Morteza, et al.
Veröffentlicht: (2024)
Foundational Analysis of Safety Engineering Requirements (SAFER)
von: Chemo, Noga, et al.
Veröffentlicht: (2026)
von: Chemo, Noga, et al.
Veröffentlicht: (2026)
On the Replicability and Reproducibility of Deep Learning in Software Engineering
von: Liu, Chao, et al.
Veröffentlicht: (2020)
von: Liu, Chao, et al.
Veröffentlicht: (2020)
AI Act for the Working Programmer
von: Hermanns, Holger, et al.
Veröffentlicht: (2024)
von: Hermanns, Holger, et al.
Veröffentlicht: (2024)
Automated Knowledge Component Generation for Interpretable Knowledge Tracing in Coding Problems
von: Duan, Zhangqi, et al.
Veröffentlicht: (2025)
von: Duan, Zhangqi, et al.
Veröffentlicht: (2025)
Impact of AI-tooling on the Engineering Workspace
von: Chretien, Lena, et al.
Veröffentlicht: (2024)
von: Chretien, Lena, et al.
Veröffentlicht: (2024)
Balancing Innovation and Ethics in AI-Driven Software Development
von: Baqar, Mohammad
Veröffentlicht: (2024)
von: Baqar, Mohammad
Veröffentlicht: (2024)
LeafTutor: An AI Agent for Programming Assignment Tutoring
von: Bochard, Madison, et al.
Veröffentlicht: (2025)
von: Bochard, Madison, et al.
Veröffentlicht: (2025)
Trustworthy AI in practice: an analysis of practitioners' needs and challenges
von: Baldassarre, Maria Teresa, et al.
Veröffentlicht: (2024)
von: Baldassarre, Maria Teresa, et al.
Veröffentlicht: (2024)
Machine Learning with Requirements: a Manifesto
von: Giunchiglia, Eleonora, et al.
Veröffentlicht: (2023)
von: Giunchiglia, Eleonora, et al.
Veröffentlicht: (2023)
Scaling CS1 Support with Compiler-Integrated Conversational AI
von: Renzella, Jake, et al.
Veröffentlicht: (2024)
von: Renzella, Jake, et al.
Veröffentlicht: (2024)
Can Requirements Engineering Support Explainable Artificial Intelligence? Towards a User-Centric Approach for Explainability Requirements
von: Umm-e-Habiba, et al.
Veröffentlicht: (2022)
von: Umm-e-Habiba, et al.
Veröffentlicht: (2022)
$\texttt{Droid}$: A Resource Suite for AI-Generated Code Detection
von: Orel, Daniil, et al.
Veröffentlicht: (2025)
von: Orel, Daniil, et al.
Veröffentlicht: (2025)
Navigating Fairness: Practitioners' Understanding, Challenges, and Strategies in AI/ML Development
von: Pant, Aastha, et al.
Veröffentlicht: (2024)
von: Pant, Aastha, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026) -
The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026) -
The Collaboration Gap in Human-AI Work
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026) -
From Rights to Rites: Expectations Management in Smart-Home AI
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026) -
From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)