Preliminary suggestions for rigorous GPAI model evaluations
Fuente:
arXiv
Saved in:
| Main Authors: | Paskov, Patricia, Byun, Michael J., Wei, Kevin, Webster, Toby |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GPAI Evaluations Standards Taskforce: Towards Effective AI Governance
by: Paskov, Patricia, et al.
Published: (2024)
by: Paskov, Patricia, et al.
Published: (2024)
Recommendations and Reporting Checklist for Rigorous & Transparent Human Baselines in Model Evaluations
by: Wei, Kevin L., et al.
Published: (2025)
by: Wei, Kevin L., et al.
Published: (2025)
Quality Assessment of Public Summary of Training Content for GPAI models required by AI Act Article 53(1)(d)
by: Blankvoort, Dick A. H., et al.
Published: (2026)
by: Blankvoort, Dick A. H., et al.
Published: (2026)
Mapping Industry Practices to the EU AI Act's GPAI Code of Practice Safety and Security Measures
by: Stelling, Lily, et al.
Published: (2025)
by: Stelling, Lily, et al.
Published: (2025)
AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
by: Barrett, Anthony M., et al.
Published: (2025)
by: Barrett, Anthony M., et al.
Published: (2025)
The Reasoning Under Uncertainty Trap: A Structural AI Risk
by: Pilditch, Toby D.
Published: (2024)
by: Pilditch, Toby D.
Published: (2024)
Red Lines and Grey Zones in the Fog of War: Benchmarking Legal Risk, Moral Harm, and Regional Bias in Large Language Model Military Decision-Making
by: Drinkall, Toby
Published: (2025)
by: Drinkall, Toby
Published: (2025)
Mapping AI Risk Mitigations: Evidence Scan and Preliminary AI Risk Mitigation Taxonomy
by: Saeri, Alexander K., et al.
Published: (2025)
by: Saeri, Alexander K., et al.
Published: (2025)
Emergent evaluation hubs in a decentralizing large language model ecosystem
by: Cebrian, Manuel, et al.
Published: (2025)
by: Cebrian, Manuel, et al.
Published: (2025)
The Sentience Readiness Index: A Preliminary Framework for Measuring National Preparedness for the Possibility of Artificial Sentience
by: Rost, Tony
Published: (2026)
by: Rost, Tony
Published: (2026)
Standing on FURM ground -- A framework for evaluating Fair, Useful, and Reliable AI Models in healthcare systems
by: Callahan, Alison, et al.
Published: (2024)
by: Callahan, Alison, et al.
Published: (2024)
The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems
by: Staufer, Leon, et al.
Published: (2026)
by: Staufer, Leon, et al.
Published: (2026)
Trust, Experience, and Innovation: Key Factors Shaping American Attitudes About AI
by: Palm, Risa, et al.
Published: (2025)
by: Palm, Risa, et al.
Published: (2025)
Fairness Is Not Enough: Auditing Competence and Intersectional Bias in AI-powered Resume Screening
by: Webster, Kevin T
Published: (2025)
by: Webster, Kevin T
Published: (2025)
Gender Bias in LLMs: Preliminary Evidence from Shared Parenting Scenario in Czech Family Law
by: Harasta, Jakub, et al.
Published: (2026)
by: Harasta, Jakub, et al.
Published: (2026)
Declare and Justify: Explicit assumptions in AI evaluations are necessary for effective regulation
by: Barnett, Peter, et al.
Published: (2024)
by: Barnett, Peter, et al.
Published: (2024)
What AI evaluations for preventing catastrophic risks can and cannot do
by: Barnett, Peter, et al.
Published: (2024)
by: Barnett, Peter, et al.
Published: (2024)
Justifications for Democratizing AI Alignment and Their Prospects
by: Steingrüber, André, et al.
Published: (2025)
by: Steingrüber, André, et al.
Published: (2025)
LLM hallucinations in the wild: Large-scale evidence from non-existent citations
by: Zhao, Zhenyue, et al.
Published: (2026)
by: Zhao, Zhenyue, et al.
Published: (2026)
Scientific production in the era of Large Language Models
by: Kusumegi, Keigo, et al.
Published: (2026)
by: Kusumegi, Keigo, et al.
Published: (2026)
Artificial intelligence and machine learning applications for cultured meat
by: Todhunter, Michael E., et al.
Published: (2024)
by: Todhunter, Michael E., et al.
Published: (2024)
Reducing research bureaucracy in UK higher education: Can generative AI assist with the internal evaluation of quality?
by: Fletcher, Gordon, et al.
Published: (2025)
by: Fletcher, Gordon, et al.
Published: (2025)
Implications for Governance in Public Perceptions of Societal-scale AI Risks
by: Gruetzemacher, Ross, et al.
Published: (2024)
by: Gruetzemacher, Ross, et al.
Published: (2024)
Foundation models may exhibit staged progression in novel CBRN threat disclosure
by: Esvelt, Kevin M
Published: (2025)
by: Esvelt, Kevin M
Published: (2025)
Securing External Deeper-than-black-box GPAI Evaluations
by: Tlaie, Alejandro, et al.
Published: (2025)
by: Tlaie, Alejandro, et al.
Published: (2025)
Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
by: Bucknall, Ben, et al.
Published: (2025)
by: Bucknall, Ben, et al.
Published: (2025)
First, do NOHARM: towards clinically safe large language models
by: Wu, David, et al.
Published: (2025)
by: Wu, David, et al.
Published: (2025)
Preliminary Study of the Impact of AI-Based Interventions on Health and Behavioral Outcomes in Maternal Health Programs
by: Dasgupta, Arpan, et al.
Published: (2024)
by: Dasgupta, Arpan, et al.
Published: (2024)
Large language models accurately predict public perceptions of support for climate action worldwide
by: Powdthavee, Nattavudh, et al.
Published: (2026)
by: Powdthavee, Nattavudh, et al.
Published: (2026)
(A)I Am Not a Lawyer, But...: Engaging Legal Experts towards Responsible LLM Policies for Legal Advice
by: Cheong, Inyoung, et al.
Published: (2024)
by: Cheong, Inyoung, et al.
Published: (2024)
Do AI Companies Make Good on Voluntary Commitments to the White House?
by: Wang, Jennifer, et al.
Published: (2025)
by: Wang, Jennifer, et al.
Published: (2025)
Generative AI in K-12 Classrooms: A Midyear Implementation Report
by: Esbenshade, Lief, et al.
Published: (2026)
by: Esbenshade, Lief, et al.
Published: (2026)
Evaluating Large Language Models for Fair and Reliable Organ Allocation
by: Kim, Brian Hyeongseok, et al.
Published: (2025)
by: Kim, Brian Hyeongseok, et al.
Published: (2025)
New Tools are Needed for Tracking Adherence to AI Model Behavioral Use Clauses
by: McDuff, Daniel, et al.
Published: (2025)
by: McDuff, Daniel, et al.
Published: (2025)
An evaluation of LLMs for political bias in Western media: Israel-Hamas and Ukraine-Russia wars
by: Chandra, Rohitash, et al.
Published: (2026)
by: Chandra, Rohitash, et al.
Published: (2026)
Breaking Up with Normatively Monolithic Agency with GRACE: A Reason-Based Neuro-Symbolic Architecture for Safe and Ethical AI Alignment
by: Jahn, Felix, et al.
Published: (2026)
by: Jahn, Felix, et al.
Published: (2026)
A Mechanism-Based Approach to Mitigating Harms from Persuasive Generative AI
by: El-Sayed, Seliem, et al.
Published: (2024)
by: El-Sayed, Seliem, et al.
Published: (2024)
On the Regulatory Potential of User Interfaces for AI Agent Governance
by: Feng, K. J. Kevin, et al.
Published: (2025)
by: Feng, K. J. Kevin, et al.
Published: (2025)
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
by: Baum, Kevin
Published: (2025)
by: Baum, Kevin
Published: (2025)
Decomposed evaluations of geographic disparities in text-to-image models
by: Sureddy, Abhishek, et al.
Published: (2024)
by: Sureddy, Abhishek, et al.
Published: (2024)
Similar Items
-
GPAI Evaluations Standards Taskforce: Towards Effective AI Governance
by: Paskov, Patricia, et al.
Published: (2024) -
Recommendations and Reporting Checklist for Rigorous & Transparent Human Baselines in Model Evaluations
by: Wei, Kevin L., et al.
Published: (2025) -
Quality Assessment of Public Summary of Training Content for GPAI models required by AI Act Article 53(1)(d)
by: Blankvoort, Dick A. H., et al.
Published: (2026) -
Mapping Industry Practices to the EU AI Act's GPAI Code of Practice Safety and Security Measures
by: Stelling, Lily, et al.
Published: (2025) -
AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
by: Barrett, Anthony M., et al.
Published: (2025)