Enabling External Scrutiny of AI Systems with Privacy-Enhancing Technologies
Fuente:
arXiv
Saved in:
| Main Authors: | Beers, Kendrea, Toner, Helen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OpenAI's Approach to External Red Teaming for AI Models and Systems
by: Ahmad, Lama, et al.
Published: (2025)
by: Ahmad, Lama, et al.
Published: (2025)
Securing the Future of GenAI: Policy and Technology
by: Christodorescu, Mihai, et al.
Published: (2024)
by: Christodorescu, Mihai, et al.
Published: (2024)
User Privacy and Large Language Models: An Analysis of Frontier Developers' Privacy Policies
by: King, Jennifer, et al.
Published: (2025)
by: King, Jennifer, et al.
Published: (2025)
Private, Verifiable, and Auditable AI Systems
by: South, Tobin
Published: (2025)
by: South, Tobin
Published: (2025)
Synthetic Data and Health Privacy
by: Abgrall, Gwénolé, et al.
Published: (2025)
by: Abgrall, Gwénolé, et al.
Published: (2025)
The End Of Universal Lifelong Identifiers: Identity Systems For The AI Era
by: Palakodety, Shriphani
Published: (2025)
by: Palakodety, Shriphani
Published: (2025)
Optimal Allocation of Privacy Budget on Hierarchical Data Release
by: Ko, Joonhyuk, et al.
Published: (2025)
by: Ko, Joonhyuk, et al.
Published: (2025)
SecGenAI: Enhancing Security of Cloud-based Generative AI Applications within Australian Critical Technologies of National Interest
by: Haryanto, Christoforus Yoga, et al.
Published: (2024)
by: Haryanto, Christoforus Yoga, et al.
Published: (2024)
A Longitudinal Measurement of Privacy Policy Evolution for Large Language Models
by: Tao, Zhen, et al.
Published: (2025)
by: Tao, Zhen, et al.
Published: (2025)
Fingerprinting and Tracing Shadows: The Development and Impact of Browser Fingerprinting on Digital Privacy
by: Lawall, Alexander
Published: (2024)
by: Lawall, Alexander
Published: (2024)
Synthetic Data, Similarity-based Privacy Metrics, and Regulatory (Non-)Compliance
by: Ganev, Georgi
Published: (2024)
by: Ganev, Georgi
Published: (2024)
Cryptographic Runtime Governance for Autonomous AI Systems: The Aegis Architecture for Verifiable Policy Enforcement
by: Mazzocchetti, Adam Massimo
Published: (2026)
by: Mazzocchetti, Adam Massimo
Published: (2026)
Negotiating Privacy with Smart Voice Assistants: Risk-Benefit and Control-Acceptance Tensions
by: Campbell, Molly, et al.
Published: (2026)
by: Campbell, Molly, et al.
Published: (2026)
Covert Surveillance in Smart Devices: A SCOUR Framework Analysis of Youth Privacy Implications
by: Shouli, Austin, et al.
Published: (2025)
by: Shouli, Austin, et al.
Published: (2025)
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation
by: Brundage, Miles, et al.
Published: (2018)
by: Brundage, Miles, et al.
Published: (2018)
Privacy at a Price: Exploring its Dual Impact on AI Fairness
by: Yang, Mengmeng, et al.
Published: (2024)
by: Yang, Mengmeng, et al.
Published: (2024)
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
by: Lin, Zhiqiang, et al.
Published: (2025)
by: Lin, Zhiqiang, et al.
Published: (2025)
Gender-Based Heterogeneity in Youth Privacy-Protective Behavior for Smart Voice Assistants: Evidence from Multigroup PLS-SEM
by: Campbell, Molly, et al.
Published: (2026)
by: Campbell, Molly, et al.
Published: (2026)
AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
by: Barrett, Anthony M., et al.
Published: (2025)
by: Barrett, Anthony M., et al.
Published: (2025)
Frontier AI's Impact on the Cybersecurity Landscape
by: Potter, Yujin, et al.
Published: (2025)
by: Potter, Yujin, et al.
Published: (2025)
Red Teaming AI Red Teaming
by: Majumdar, Subhabrata, et al.
Published: (2025)
by: Majumdar, Subhabrata, et al.
Published: (2025)
AI Propaganda factories with language models
by: Olejnik, Lukasz
Published: (2025)
by: Olejnik, Lukasz
Published: (2025)
Accelerating AI Development with Cyber Arenas
by: Cashman, William, et al.
Published: (2025)
by: Cashman, William, et al.
Published: (2025)
AI-Driven Cyber Threat Intelligence Automation
by: Shah, Shrit, et al.
Published: (2024)
by: Shah, Shrit, et al.
Published: (2024)
The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI
by: Hall, Peter, et al.
Published: (2025)
by: Hall, Peter, et al.
Published: (2025)
Black-Box Access is Insufficient for Rigorous AI Audits
by: Casper, Stephen, et al.
Published: (2024)
by: Casper, Stephen, et al.
Published: (2024)
Coordinated Flaw Disclosure for AI: Beyond Security Vulnerabilities
by: Cattell, Sven, et al.
Published: (2024)
by: Cattell, Sven, et al.
Published: (2024)
Interplay of ISMS and AIMS in context of the EU AI Act
by: Pötsch, Jordan
Published: (2024)
by: Pötsch, Jordan
Published: (2024)
What Security and Privacy Transparency Users Need from Consumer-Facing Generative AI
by: Cao, Jiaxun, et al.
Published: (2026)
by: Cao, Jiaxun, et al.
Published: (2026)
RedTeamLLM: an Agentic AI framework for offensive security
by: Challita, Brian, et al.
Published: (2025)
by: Challita, Brian, et al.
Published: (2025)
Governable AI: Provable Safety Under Extreme Threat Models
by: Wang, Donglin, et al.
Published: (2025)
by: Wang, Donglin, et al.
Published: (2025)
Comparing AI Agents to Cybersecurity Professionals in Real-World Penetration Testing
by: Lin, Justin W., et al.
Published: (2025)
by: Lin, Justin W., et al.
Published: (2025)
Naming is framing: How cybersecurity's language problems are repeating in AI governance
by: Potter, Lianne
Published: (2025)
by: Potter, Lianne
Published: (2025)
The Impact of AI on the Cyber Offense-Defense Balance and the Character of Cyber Conflict
by: Lohn, Andrew J.
Published: (2025)
by: Lohn, Andrew J.
Published: (2025)
Secure and Trustworthy Artificial Intelligence-Extended Reality (AI-XR) for Metaverses
by: Qayyum, Adnan, et al.
Published: (2022)
by: Qayyum, Adnan, et al.
Published: (2022)
Is Your AI Truly Yours? Leveraging Blockchain for Copyrights, Provenance, and Lineage
by: Wang, Qin, et al.
Published: (2024)
by: Wang, Qin, et al.
Published: (2024)
Clear, Compelling Arguments: Rethinking the Foundations of Frontier AI Safety Cases
by: Feakins, Shaun, et al.
Published: (2026)
by: Feakins, Shaun, et al.
Published: (2026)
Children's Voice Privacy: First Steps And Emerging Challenges
by: Kulkarni, Ajinkya, et al.
Published: (2025)
by: Kulkarni, Ajinkya, et al.
Published: (2025)
Love, Lies, and Language Models: Investigating AI's Role in Romance-Baiting Scams
by: Gressel, Gilad, et al.
Published: (2025)
by: Gressel, Gilad, et al.
Published: (2025)
Can AI Models be Jailbroken to Phish Elderly Victims? An End-to-End Evaluation
by: Heiding, Fred, et al.
Published: (2025)
by: Heiding, Fred, et al.
Published: (2025)
Similar Items
-
OpenAI's Approach to External Red Teaming for AI Models and Systems
by: Ahmad, Lama, et al.
Published: (2025) -
Securing the Future of GenAI: Policy and Technology
by: Christodorescu, Mihai, et al.
Published: (2024) -
User Privacy and Large Language Models: An Analysis of Frontier Developers' Privacy Policies
by: King, Jennifer, et al.
Published: (2025) -
Private, Verifiable, and Auditable AI Systems
by: South, Tobin
Published: (2025) -
Synthetic Data and Health Privacy
by: Abgrall, Gwénolé, et al.
Published: (2025)