Dual-Use AI Face Swap Apps Are Mostly Unsafe: A Systematic Safety Audit
Fuente:
arXiv
Salvato in:
| Autori principali: | Daffalla, Alaa, Chao, Sarah, Zeng, Eric |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Auditing Agent Harness Safety
di: Liu, Chengzhi, et al.
Pubblicazione: (2026)
di: Liu, Chengzhi, et al.
Pubblicazione: (2026)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
di: Brundage, Miles, et al.
Pubblicazione: (2026)
di: Brundage, Miles, et al.
Pubblicazione: (2026)
AuditMAI: Towards An Infrastructure for Continuous AI Auditing
di: Waltersdorfer, Laura, et al.
Pubblicazione: (2024)
di: Waltersdorfer, Laura, et al.
Pubblicazione: (2024)
From Transparency to Accountability and Back: A Discussion of Access and Evidence in AI Auditing
di: Cen, Sarah H., et al.
Pubblicazione: (2024)
di: Cen, Sarah H., et al.
Pubblicazione: (2024)
Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
di: Glukhov, David, et al.
Pubblicazione: (2024)
di: Glukhov, David, et al.
Pubblicazione: (2024)
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset
di: Chehbouni, Khaoula, et al.
Pubblicazione: (2024)
di: Chehbouni, Khaoula, et al.
Pubblicazione: (2024)
Audit Cards: Contextualizing AI Evaluations
di: Staufer, Leon, et al.
Pubblicazione: (2025)
di: Staufer, Leon, et al.
Pubblicazione: (2025)
A Tale of Two Identities: An Ethical Audit of Human and AI-Crafted Personas
di: Venkit, Pranav Narayanan, et al.
Pubblicazione: (2025)
di: Venkit, Pranav Narayanan, et al.
Pubblicazione: (2025)
Investigating Youth AI Auditing
di: Solyst, Jaemarie, et al.
Pubblicazione: (2025)
di: Solyst, Jaemarie, et al.
Pubblicazione: (2025)
Dual Use Concerns of Generative AI and Large Language Models
di: Grinbaum, Alexei, et al.
Pubblicazione: (2023)
di: Grinbaum, Alexei, et al.
Pubblicazione: (2023)
A Blueprint for Auditing Generative AI
di: Mokander, Jakob, et al.
Pubblicazione: (2024)
di: Mokander, Jakob, et al.
Pubblicazione: (2024)
Can AI be Auditable?
di: Verma, Himanshu, et al.
Pubblicazione: (2025)
di: Verma, Himanshu, et al.
Pubblicazione: (2025)
Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit Tooling
di: Ojewale, Victor, et al.
Pubblicazione: (2024)
di: Ojewale, Victor, et al.
Pubblicazione: (2024)
AuditWen:An Open-Source Large Language Model for Audit
di: Huang, Jiajia, et al.
Pubblicazione: (2024)
di: Huang, Jiajia, et al.
Pubblicazione: (2024)
The AI-Fraud Diamond: A Novel Lens for Auditing Algorithmic Deception
di: Zweers, Benjamin, et al.
Pubblicazione: (2025)
di: Zweers, Benjamin, et al.
Pubblicazione: (2025)
Face Recognition: to Deploy or not to Deploy? A Framework for Assessing the Proportional Use of Face Recognition Systems in Real-World Scenarios
di: Negri, Pablo, et al.
Pubblicazione: (2024)
di: Negri, Pablo, et al.
Pubblicazione: (2024)
Towards Understanding Unsafe Video Generation
di: Pang, Yan, et al.
Pubblicazione: (2024)
di: Pang, Yan, et al.
Pubblicazione: (2024)
Safety Co-Option and Compromised National Security: The Self-Fulfilling Prophecy of Weakened AI Risk Thresholds
di: Khlaaf, Heidy, et al.
Pubblicazione: (2025)
di: Khlaaf, Heidy, et al.
Pubblicazione: (2025)
Food Recognition and Nutritional Apps
di: Rahman, Lubnaa Abdur, et al.
Pubblicazione: (2023)
di: Rahman, Lubnaa Abdur, et al.
Pubblicazione: (2023)
Fairness Concerns in App Reviews: A Study on AI-based Mobile Apps
di: Nasab, Ali Rezaei, et al.
Pubblicazione: (2024)
di: Nasab, Ali Rezaei, et al.
Pubblicazione: (2024)
AI Safety for Everyone
di: Gyevnar, Balint, et al.
Pubblicazione: (2025)
di: Gyevnar, Balint, et al.
Pubblicazione: (2025)
Towards an Automated Framework to Audit Youth Safety on TikTok
di: Xue, Linda, et al.
Pubblicazione: (2025)
di: Xue, Linda, et al.
Pubblicazione: (2025)
Global AI Bias Audit for Technical Governance
di: Hung, Jason
Pubblicazione: (2026)
di: Hung, Jason
Pubblicazione: (2026)
Auditing of AI: Legal, Ethical and Technical Approaches
di: Mokander, Jakob
Pubblicazione: (2024)
di: Mokander, Jakob
Pubblicazione: (2024)
The Role of AI Safety Institutes in Contributing to International Standards for Frontier AI Safety
di: Fort, Kristina
Pubblicazione: (2024)
di: Fort, Kristina
Pubblicazione: (2024)
Safety First: Psychological Safety as the Key to AI Transformation
di: Reich, Aaron, et al.
Pubblicazione: (2026)
di: Reich, Aaron, et al.
Pubblicazione: (2026)
How Should AI Safety Benchmarks Benchmark Safety?
di: Yu, Cheng, et al.
Pubblicazione: (2026)
di: Yu, Cheng, et al.
Pubblicazione: (2026)
Unsafe2Safe: Controllable Image Anonymization for Downstream Utility
di: Dinh, Mih, et al.
Pubblicazione: (2026)
di: Dinh, Mih, et al.
Pubblicazione: (2026)
AI Safety, Alignment, and Ethics (AI SAE)
di: Waldner, Dylan
Pubblicazione: (2025)
di: Waldner, Dylan
Pubblicazione: (2025)
Safety cases for frontier AI
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2024)
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2024)
Safety Features for a Centralised AGI Project
di: Hastings-Woodhouse, Sarah
Pubblicazione: (2025)
di: Hastings-Woodhouse, Sarah
Pubblicazione: (2025)
A Grading Rubric for AI Safety Frameworks
di: Alaga, Jide, et al.
Pubblicazione: (2024)
di: Alaga, Jide, et al.
Pubblicazione: (2024)
Green AI: Which Programming Language Consumes the Most?
di: Marini, Niccolò, et al.
Pubblicazione: (2024)
di: Marini, Niccolò, et al.
Pubblicazione: (2024)
Learning AI Auditing: A Case Study of Teenagers Auditing a Generative AI Model
di: Morales-Navarro, Luis, et al.
Pubblicazione: (2025)
di: Morales-Navarro, Luis, et al.
Pubblicazione: (2025)
Unveiling User Perceptions in the Generative AI Era: A Sentiment-Driven Evaluation of AI Educational Apps' Role in Digital Transformation of e-Teaching
di: Mazaherian, Adeleh, et al.
Pubblicazione: (2025)
di: Mazaherian, Adeleh, et al.
Pubblicazione: (2025)
Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI
di: O'Brien, Joe, et al.
Pubblicazione: (2024)
di: O'Brien, Joe, et al.
Pubblicazione: (2024)
Reporting Non-Consensual Intimate Media: An Audit Study of Deepfakes
di: Qiwei, Li, et al.
Pubblicazione: (2024)
di: Qiwei, Li, et al.
Pubblicazione: (2024)
(Mis-)Informed Consent: Predatory Apps and the Exploitation of Populations with Limited Literacy
di: Pervez, Muhammad Muneeb, et al.
Pubblicazione: (2026)
di: Pervez, Muhammad Muneeb, et al.
Pubblicazione: (2026)
Weapons of Online Harassment: Menacing and Profiling Users via Social Apps
di: Cheerla, Sanjana, et al.
Pubblicazione: (2025)
di: Cheerla, Sanjana, et al.
Pubblicazione: (2025)
Factors Influencing the Usage of Mobile Banking Apps Among Malaysian Consumers
di: Jalani, Siti Nurdianah binti Mohamad, et al.
Pubblicazione: (2024)
di: Jalani, Siti Nurdianah binti Mohamad, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Auditing Agent Harness Safety
di: Liu, Chengzhi, et al.
Pubblicazione: (2026) -
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
di: Brundage, Miles, et al.
Pubblicazione: (2026) -
AuditMAI: Towards An Infrastructure for Continuous AI Auditing
di: Waltersdorfer, Laura, et al.
Pubblicazione: (2024) -
From Transparency to Accountability and Back: A Discussion of Access and Evidence in AI Auditing
di: Cen, Sarah H., et al.
Pubblicazione: (2024) -
Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
di: Glukhov, David, et al.
Pubblicazione: (2024)