Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Bucknall, Ben, Trager, Robert F., Osborne, Michael A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Security, privacy, and agentic AI in a regulatory view: From definitions and distinctions to provisions and reflections
di: Zhang, Shiliang, et al.
Pubblicazione: (2026)
di: Zhang, Shiliang, et al.
Pubblicazione: (2026)
"I'm not for sale" -- Perceptions and limited awareness of privacy risks by digital natives about location data
di: Boutet, Antoine, et al.
Pubblicazione: (2025)
di: Boutet, Antoine, et al.
Pubblicazione: (2025)
Black-Box Access is Insufficient for Rigorous AI Audits
di: Casper, Stephen, et al.
Pubblicazione: (2024)
di: Casper, Stephen, et al.
Pubblicazione: (2024)
Accelerating AI Development with Cyber Arenas
di: Cashman, William, et al.
Pubblicazione: (2025)
di: Cashman, William, et al.
Pubblicazione: (2025)
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
di: Lin, Zhiqiang, et al.
Pubblicazione: (2025)
di: Lin, Zhiqiang, et al.
Pubblicazione: (2025)
AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
di: Barrett, Anthony M., et al.
Pubblicazione: (2025)
di: Barrett, Anthony M., et al.
Pubblicazione: (2025)
Private, Verifiable, and Auditable AI Systems
di: South, Tobin
Pubblicazione: (2025)
di: South, Tobin
Pubblicazione: (2025)
Frontier AI's Impact on the Cybersecurity Landscape
di: Potter, Yujin, et al.
Pubblicazione: (2025)
di: Potter, Yujin, et al.
Pubblicazione: (2025)
Red Teaming AI Red Teaming
di: Majumdar, Subhabrata, et al.
Pubblicazione: (2025)
di: Majumdar, Subhabrata, et al.
Pubblicazione: (2025)
AI Propaganda factories with language models
di: Olejnik, Lukasz
Pubblicazione: (2025)
di: Olejnik, Lukasz
Pubblicazione: (2025)
AI-Driven Cyber Threat Intelligence Automation
di: Shah, Shrit, et al.
Pubblicazione: (2024)
di: Shah, Shrit, et al.
Pubblicazione: (2024)
Securing the Future of GenAI: Policy and Technology
di: Christodorescu, Mihai, et al.
Pubblicazione: (2024)
di: Christodorescu, Mihai, et al.
Pubblicazione: (2024)
The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI
di: Hall, Peter, et al.
Pubblicazione: (2025)
di: Hall, Peter, et al.
Pubblicazione: (2025)
Coordinated Flaw Disclosure for AI: Beyond Security Vulnerabilities
di: Cattell, Sven, et al.
Pubblicazione: (2024)
di: Cattell, Sven, et al.
Pubblicazione: (2024)
Interplay of ISMS and AIMS in context of the EU AI Act
di: Pötsch, Jordan
Pubblicazione: (2024)
di: Pötsch, Jordan
Pubblicazione: (2024)
The End Of Universal Lifelong Identifiers: Identity Systems For The AI Era
di: Palakodety, Shriphani
Pubblicazione: (2025)
di: Palakodety, Shriphani
Pubblicazione: (2025)
Enabling External Scrutiny of AI Systems with Privacy-Enhancing Technologies
di: Beers, Kendrea, et al.
Pubblicazione: (2025)
di: Beers, Kendrea, et al.
Pubblicazione: (2025)
RedTeamLLM: an Agentic AI framework for offensive security
di: Challita, Brian, et al.
Pubblicazione: (2025)
di: Challita, Brian, et al.
Pubblicazione: (2025)
Governable AI: Provable Safety Under Extreme Threat Models
di: Wang, Donglin, et al.
Pubblicazione: (2025)
di: Wang, Donglin, et al.
Pubblicazione: (2025)
Comparing AI Agents to Cybersecurity Professionals in Real-World Penetration Testing
di: Lin, Justin W., et al.
Pubblicazione: (2025)
di: Lin, Justin W., et al.
Pubblicazione: (2025)
Naming is framing: How cybersecurity's language problems are repeating in AI governance
di: Potter, Lianne
Pubblicazione: (2025)
di: Potter, Lianne
Pubblicazione: (2025)
The Impact of AI on the Cyber Offense-Defense Balance and the Character of Cyber Conflict
di: Lohn, Andrew J.
Pubblicazione: (2025)
di: Lohn, Andrew J.
Pubblicazione: (2025)
Secure and Trustworthy Artificial Intelligence-Extended Reality (AI-XR) for Metaverses
di: Qayyum, Adnan, et al.
Pubblicazione: (2022)
di: Qayyum, Adnan, et al.
Pubblicazione: (2022)
Is Your AI Truly Yours? Leveraging Blockchain for Copyrights, Provenance, and Lineage
di: Wang, Qin, et al.
Pubblicazione: (2024)
di: Wang, Qin, et al.
Pubblicazione: (2024)
Clear, Compelling Arguments: Rethinking the Foundations of Frontier AI Safety Cases
di: Feakins, Shaun, et al.
Pubblicazione: (2026)
di: Feakins, Shaun, et al.
Pubblicazione: (2026)
Love, Lies, and Language Models: Investigating AI's Role in Romance-Baiting Scams
di: Gressel, Gilad, et al.
Pubblicazione: (2025)
di: Gressel, Gilad, et al.
Pubblicazione: (2025)
Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
di: Glukhov, David, et al.
Pubblicazione: (2024)
di: Glukhov, David, et al.
Pubblicazione: (2024)
Can AI Models be Jailbroken to Phish Elderly Victims? An End-to-End Evaluation
di: Heiding, Fred, et al.
Pubblicazione: (2025)
di: Heiding, Fred, et al.
Pubblicazione: (2025)
To Defend Against Cyber Attacks, We Must Teach AI Agents to Hack
di: Zhuo, Terry Yue, et al.
Pubblicazione: (2026)
di: Zhuo, Terry Yue, et al.
Pubblicazione: (2026)
Cryptographic Runtime Governance for Autonomous AI Systems: The Aegis Architecture for Verifiable Policy Enforcement
di: Mazzocchetti, Adam Massimo
Pubblicazione: (2026)
di: Mazzocchetti, Adam Massimo
Pubblicazione: (2026)
Trust and Dependability in Blockchain & AI Based MedIoT Applications: Research Challenges and Future Directions
di: Solaiman, Ellis, et al.
Pubblicazione: (2025)
di: Solaiman, Ellis, et al.
Pubblicazione: (2025)
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
di: Tong, Haibo, et al.
Pubblicazione: (2026)
di: Tong, Haibo, et al.
Pubblicazione: (2026)
Artificial Intelligence in Cybersecurity: Building Resilient Cyber Diplomacy Frameworks
di: Stoltz, Michael
Pubblicazione: (2024)
di: Stoltz, Michael
Pubblicazione: (2024)
A Whole New World: Creating a Parallel-Poisoned Web Only AI-Agents Can See
di: Zychlinski, Shaked
Pubblicazione: (2025)
di: Zychlinski, Shaked
Pubblicazione: (2025)
Authenticity Debt and the Synthetic Content Threat Landscape: A Layered Framework for Trust, Provenance, and IP Governance in the Generative AI Era
di: Sengupta, Shubhashis, et al.
Pubblicazione: (2026)
di: Sengupta, Shubhashis, et al.
Pubblicazione: (2026)
Preserving Decision Sovereignty in Military AI: A Trade-Secret-Safe Architectural Framework for Model Replaceability, Human Authority, and State Control
di: Wei, Peng, et al.
Pubblicazione: (2026)
di: Wei, Peng, et al.
Pubblicazione: (2026)
On Technique Identification and Threat-Actor Attribution using LLMs and Embedding Models
di: Guru, Kyla, et al.
Pubblicazione: (2025)
di: Guru, Kyla, et al.
Pubblicazione: (2025)
How Do Data Owners Say No? A Case Study of Data Consent Mechanisms in Web-Scraped Vision-Language AI Training Datasets
di: Lee, Chung Peng, et al.
Pubblicazione: (2025)
di: Lee, Chung Peng, et al.
Pubblicazione: (2025)
OpenAI's Approach to External Red Teaming for AI Models and Systems
di: Ahmad, Lama, et al.
Pubblicazione: (2025)
di: Ahmad, Lama, et al.
Pubblicazione: (2025)
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation
di: Brundage, Miles, et al.
Pubblicazione: (2018)
di: Brundage, Miles, et al.
Pubblicazione: (2018)
Documenti analoghi
-
Security, privacy, and agentic AI in a regulatory view: From definitions and distinctions to provisions and reflections
di: Zhang, Shiliang, et al.
Pubblicazione: (2026) -
"I'm not for sale" -- Perceptions and limited awareness of privacy risks by digital natives about location data
di: Boutet, Antoine, et al.
Pubblicazione: (2025) -
Black-Box Access is Insufficient for Rigorous AI Audits
di: Casper, Stephen, et al.
Pubblicazione: (2024) -
Accelerating AI Development with Cyber Arenas
di: Cashman, William, et al.
Pubblicazione: (2025) -
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
di: Lin, Zhiqiang, et al.
Pubblicazione: (2025)