Toward Third-Party Assurance of AI Systems: Design Requirements, Prototype, and Early Testing
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Rachel M., Kuehnert, Blaine, Lai, Alice, Holstein, Kenneth, Heidari, Hoda, Ghani, Rayid |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The "Who", "What", and "How" of Responsible AI Governance: A Systematic Review and Meta-Analysis of (Actor, Stage)-Specific Tools
by: Kuehnert, Blaine, et al.
Published: (2025)
by: Kuehnert, Blaine, et al.
Published: (2025)
Disclosure or Marketing? Analyzing the Efficacy of Vendor Self-reports for Vetting Public-sector AI
by: Kuehnert, Blaine, et al.
Published: (2026)
by: Kuehnert, Blaine, et al.
Published: (2026)
Towards Automated Scoping of AI for Social Good Projects
by: Emmerson, Jacob, et al.
Published: (2025)
by: Emmerson, Jacob, et al.
Published: (2025)
The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder Early-stage Deliberations Around Public Sector AI Proposals
by: Kawakami, Anna, et al.
Published: (2024)
by: Kawakami, Anna, et al.
Published: (2024)
Toward Valid Measurement Of (Un)fairness For Generative AI: A Proposal For Systematization Through The Lens Of Fair Equality of Chances
by: Truong, Kimberly Le, et al.
Published: (2025)
by: Truong, Kimberly Le, et al.
Published: (2025)
Preventing Eviction-Caused Homelessness through ML-Informed Distribution of Rental Assistance
by: Vajiac, Catalina, et al.
Published: (2024)
by: Vajiac, Catalina, et al.
Published: (2024)
Studying Up Public Sector AI: How Networks of Power Relations Shape Agency Decisions Around AI Design and Use
by: Kawakami, Anna, et al.
Published: (2024)
by: Kawakami, Anna, et al.
Published: (2024)
Enabling the AI Revolution in Healthcare
by: Singh, Mona, et al.
Published: (2025)
by: Singh, Mona, et al.
Published: (2025)
Requirements for Quality Assurance of AI Models for Early Detection of Lung Cancer
by: Hahn, Horst K., et al.
Published: (2025)
by: Hahn, Horst K., et al.
Published: (2025)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
by: Brundage, Miles, et al.
Published: (2026)
by: Brundage, Miles, et al.
Published: (2026)
Designing Algorithmic Delegates: The Role of Indistinguishability in Human-AI Handoff
by: Greenwood, Sophie, et al.
Published: (2025)
by: Greenwood, Sophie, et al.
Published: (2025)
E-LENS: User Requirements-Oriented AI Ethics Assurance
by: Zhou, Jianlong, et al.
Published: (2025)
by: Zhou, Jianlong, et al.
Published: (2025)
The Backfiring Effect of Weak AI Safety Regulation
by: Laufer, Benjamin, et al.
Published: (2025)
by: Laufer, Benjamin, et al.
Published: (2025)
Modeling the Economic Impacts of AI Openness Regulation
by: Qiu, Tori, et al.
Published: (2025)
by: Qiu, Tori, et al.
Published: (2025)
Assessing Computer Science Student Attitudes Towards AI Ethics and Policy
by: Weichert, James, et al.
Published: (2025)
by: Weichert, James, et al.
Published: (2025)
AI Failure Loops in Devalued Work: The Confluence of Overconfidence in AI and Underconfidence in Worker Expertise
by: Kawakami, Anna, et al.
Published: (2025)
by: Kawakami, Anna, et al.
Published: (2025)
A Closer Look at the Existing Risks of Generative AI: Mapping the Who, What, and How of Real-World Incidents
by: Li, Megan, et al.
Published: (2025)
by: Li, Megan, et al.
Published: (2025)
Twitch Third-Party Developers' Support Seeking and Provision Practices on Discord
by: Cai, Jie, et al.
Published: (2026)
by: Cai, Jie, et al.
Published: (2026)
Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations
by: Reuel, Anka, et al.
Published: (2025)
by: Reuel, Anka, et al.
Published: (2025)
Red-Teaming for Generative AI: Silver Bullet or Security Theater?
by: Feffer, Michael, et al.
Published: (2024)
by: Feffer, Michael, et al.
Published: (2024)
Funding AI for Good: A Call for Meaningful Engagement
by: Lin, Hongjin, et al.
Published: (2025)
by: Lin, Hongjin, et al.
Published: (2025)
Assurance of Frontier AI Built for National Security
by: Pistillo, Matteo, et al.
Published: (2025)
by: Pistillo, Matteo, et al.
Published: (2025)
Attacks on Third-Party APIs of Large Language Models
by: Zhao, Wanru, et al.
Published: (2024)
by: Zhao, Wanru, et al.
Published: (2024)
Third-Party Developers and Tool Development For Community Management on Live Streaming Platform Twitch
by: Cai, Jie, et al.
Published: (2024)
by: Cai, Jie, et al.
Published: (2024)
Towards Comprehensive Legislative Requirements for Cyber Physical Systems Testing in the European Union
by: Nguyen, Guillaume, et al.
Published: (2024)
by: Nguyen, Guillaume, et al.
Published: (2024)
Trustworthy AI Posture (TAIP): A Framework for Continuous AI Assurance of Agentic Systems at Horizontal and Vertical scale
by: Lupo, Guy, et al.
Published: (2026)
by: Lupo, Guy, et al.
Published: (2026)
A Framework for Assurance Audits of Algorithmic Systems
by: Lam, Khoa, et al.
Published: (2024)
by: Lam, Khoa, et al.
Published: (2024)
Aequitas Flow: Streamlining Fair ML Experimentation
by: Jesus, Sérgio, et al.
Published: (2024)
by: Jesus, Sérgio, et al.
Published: (2024)
Educating a Responsible AI Workforce: Piloting a Curricular Module on AI Policy in a Graduate Machine Learning Course
by: Weichert, James, et al.
Published: (2025)
by: Weichert, James, et al.
Published: (2025)
Evaluating AI-Generated Images of Cultural Artifacts with Community-Informed Rubrics
by: Johnson, Nari, et al.
Published: (2026)
by: Johnson, Nari, et al.
Published: (2026)
Legacy Procurement Practices Shape How U.S. Cities Govern AI: Understanding Government Employees' Practices, Challenges, and Needs
by: Johnson, Nari, et al.
Published: (2024)
by: Johnson, Nari, et al.
Published: (2024)
Can AI Model the Complexities of Human Moral Decision-Making? A Qualitative Study of Kidney Allocation Decisions
by: Keswani, Vijay, et al.
Published: (2025)
by: Keswani, Vijay, et al.
Published: (2025)
The Democratic Ontology Deficit: How AI Systems Fail to Represent What Democracy Requires
by: Ceresa, Robert M., et al.
Published: (2026)
by: Ceresa, Robert M., et al.
Published: (2026)
Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback
by: Keswani, Vijay, et al.
Published: (2025)
by: Keswani, Vijay, et al.
Published: (2025)
Economic Competition, EU Regulation, and Executive Orders: A Framework for Discussing AI Policy Implications in CS Courses
by: Weichert, James, et al.
Published: (2025)
by: Weichert, James, et al.
Published: (2025)
Predictive Performance Comparison of Decision Policies Under Confounding
by: Guerdan, Luke, et al.
Published: (2024)
by: Guerdan, Luke, et al.
Published: (2024)
AI Safety Assurance for Automated Vehicles: A Survey on Research, Standardization, Regulation
by: Ullrich, Lars, et al.
Published: (2025)
by: Ullrich, Lars, et al.
Published: (2025)
Taking a Bite Out of the Forbidden Fruit: Characterizing Third-Party Iranian iOS App Stores
by: Khanlari, Amirhossein, et al.
Published: (2026)
by: Khanlari, Amirhossein, et al.
Published: (2026)
Prototyping Multimodal GenAI Real-Time Agents with Counterfactual Replays and Hybrid Wizard-of-Oz
by: Gmeiner, Frederic, et al.
Published: (2025)
by: Gmeiner, Frederic, et al.
Published: (2025)
The Third-Party Access Effect: An Overlooked Challenge in Secondary Use of Educational Real-World Data
by: Ito, Hibiki, et al.
Published: (2026)
by: Ito, Hibiki, et al.
Published: (2026)
Similar Items
-
The "Who", "What", and "How" of Responsible AI Governance: A Systematic Review and Meta-Analysis of (Actor, Stage)-Specific Tools
by: Kuehnert, Blaine, et al.
Published: (2025) -
Disclosure or Marketing? Analyzing the Efficacy of Vendor Self-reports for Vetting Public-sector AI
by: Kuehnert, Blaine, et al.
Published: (2026) -
Towards Automated Scoping of AI for Social Good Projects
by: Emmerson, Jacob, et al.
Published: (2025) -
The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder Early-stage Deliberations Around Public Sector AI Proposals
by: Kawakami, Anna, et al.
Published: (2024) -
Toward Valid Measurement Of (Un)fairness For Generative AI: A Proposal For Systematization Through The Lens Of Fair Equality of Chances
by: Truong, Kimberly Le, et al.
Published: (2025)