Geographic Blind Spots in AI Control Monitors: A Cross-National Audit of Claude Opus 4.6
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Hung, Jason |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Global AI Bias Audit for Technical Governance
von: Hung, Jason
Veröffentlicht: (2026)
von: Hung, Jason
Veröffentlicht: (2026)
AI's Blind Spots: Geographic Knowledge and Diversity Deficit in Generated Urban Scenario
von: Beneduce, Ciro, et al.
Veröffentlicht: (2025)
von: Beneduce, Ciro, et al.
Veröffentlicht: (2025)
All Eyes on the Ranker: Participatory Auditing to Surface Blind Spots in Ranked Search Results
von: Rezk, Anna Marie, et al.
Veröffentlicht: (2026)
von: Rezk, Anna Marie, et al.
Veröffentlicht: (2026)
The DSA's Blind Spot: Algorithmic Audit of Advertising and Minor Profiling on TikTok
von: Solarova, Sara, et al.
Veröffentlicht: (2026)
von: Solarova, Sara, et al.
Veröffentlicht: (2026)
From Defense to Advocacy: Empowering Users to Leverage the Blind Spot of AI Inference
von: Wei, Yumou, et al.
Veröffentlicht: (2026)
von: Wei, Yumou, et al.
Veröffentlicht: (2026)
Investigating Youth AI Auditing
von: Solyst, Jaemarie, et al.
Veröffentlicht: (2025)
von: Solyst, Jaemarie, et al.
Veröffentlicht: (2025)
Trajectories and Comparative Analysis of Global Countries Dominating AI Publications, 2000-2025
von: Hung, Jason
Veröffentlicht: (2025)
von: Hung, Jason
Veröffentlicht: (2025)
The Algorithmic Blind Spot: Bias, Moral Status, and the Future of Robot Rights
von: Karthikeyan, Rahulrajan, et al.
Veröffentlicht: (2026)
von: Karthikeyan, Rahulrajan, et al.
Veröffentlicht: (2026)
Belief in the Machine: Investigating Epistemological Blind Spots of Language Models
von: Suzgun, Mirac, et al.
Veröffentlicht: (2024)
von: Suzgun, Mirac, et al.
Veröffentlicht: (2024)
AuditMAI: Towards An Infrastructure for Continuous AI Auditing
von: Waltersdorfer, Laura, et al.
Veröffentlicht: (2024)
von: Waltersdorfer, Laura, et al.
Veröffentlicht: (2024)
AI Governance and Accountability: An Analysis of Anthropic's Claude
von: Priyanshu, Aman, et al.
Veröffentlicht: (2024)
von: Priyanshu, Aman, et al.
Veröffentlicht: (2024)
ZPD-SCA: Unveiling the Blind Spots of LLMs in Assessing Students' Cognitive Abilities
von: Dong, Wenhan, et al.
Veröffentlicht: (2025)
von: Dong, Wenhan, et al.
Veröffentlicht: (2025)
Audit Cards: Contextualizing AI Evaluations
von: Staufer, Leon, et al.
Veröffentlicht: (2025)
von: Staufer, Leon, et al.
Veröffentlicht: (2025)
Poisoned Identifiers Survive LLM Deobfuscation: A Case Study on Claude Opus 4.6
von: Lorenzo, Luis Guzmán
Veröffentlicht: (2026)
von: Lorenzo, Luis Guzmán
Veröffentlicht: (2026)
Mind the Blind Spot
von: Kelly Singleton, et al.
Veröffentlicht: (2026)
von: Kelly Singleton, et al.
Veröffentlicht: (2026)
A Blueprint for Auditing Generative AI
von: Mokander, Jakob, et al.
Veröffentlicht: (2024)
von: Mokander, Jakob, et al.
Veröffentlicht: (2024)
Can AI be Auditable?
von: Verma, Himanshu, et al.
Veröffentlicht: (2025)
von: Verma, Himanshu, et al.
Veröffentlicht: (2025)
Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit Tooling
von: Ojewale, Victor, et al.
Veröffentlicht: (2024)
von: Ojewale, Victor, et al.
Veröffentlicht: (2024)
The Biosecurity Blind Spot: Systematic Dual-use Detection in Open Science Infrastructure
von: Sharma, Vasudha, et al.
Veröffentlicht: (2026)
von: Sharma, Vasudha, et al.
Veröffentlicht: (2026)
Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems
von: Alamdari, Parand A., et al.
Veröffentlicht: (2026)
von: Alamdari, Parand A., et al.
Veröffentlicht: (2026)
The AI-Fraud Diamond: A Novel Lens for Auditing Algorithmic Deception
von: Zweers, Benjamin, et al.
Veröffentlicht: (2025)
von: Zweers, Benjamin, et al.
Veröffentlicht: (2025)
Putnam 2025 Problems in Rocq using Opus 4.6 and Rocq-MCP
von: Baudart, Guillaume, et al.
Veröffentlicht: (2026)
von: Baudart, Guillaume, et al.
Veröffentlicht: (2026)
Correctness Comparison of ChatGPT-4, Gemini, Claude-3, and Copilot for Spatial Tasks
von: Hochmair, Hartwig H., et al.
Veröffentlicht: (2024)
von: Hochmair, Hartwig H., et al.
Veröffentlicht: (2024)
Catching The Correct Answer Trap: Characterising AI Tutor Blind Spots When Analysing Student Reasoning
von: Imran, Moiz, et al.
Veröffentlicht: (2026)
von: Imran, Moiz, et al.
Veröffentlicht: (2026)
Bias in Decision-Making for AI's Ethical Dilemmas: A Comparative Study of ChatGPT and Claude
von: Xu, Wentao, et al.
Veröffentlicht: (2025)
von: Xu, Wentao, et al.
Veröffentlicht: (2025)
Auditing of AI: Legal, Ethical and Technical Approaches
von: Mokander, Jakob
Veröffentlicht: (2024)
von: Mokander, Jakob
Veröffentlicht: (2024)
Authority Signals in Claude AI Health Citations: A Descriptive Analysis Using the Authority Signals Framework
von: Jacques, Erin T., et al.
Veröffentlicht: (2026)
von: Jacques, Erin T., et al.
Veröffentlicht: (2026)
When AI Speaks, Whose Values Does It Express? A Cross-Cultural Audit of Individualism-Collectivism Bias in Large Language Models
von: Venkata, Pruthvinath Jeripity
Veröffentlicht: (2026)
von: Venkata, Pruthvinath Jeripity
Veröffentlicht: (2026)
Dual-Use AI Face Swap Apps Are Mostly Unsafe: A Systematic Safety Audit
von: Daffalla, Alaa, et al.
Veröffentlicht: (2026)
von: Daffalla, Alaa, et al.
Veröffentlicht: (2026)
Learning AI Auditing: A Case Study of Teenagers Auditing a Generative AI Model
von: Morales-Navarro, Luis, et al.
Veröffentlicht: (2025)
von: Morales-Navarro, Luis, et al.
Veröffentlicht: (2025)
Towards Environmentally Equitable AI via Geographical Load Balancing
von: Li, Pengfei, et al.
Veröffentlicht: (2023)
von: Li, Pengfei, et al.
Veröffentlicht: (2023)
Misaligned Roles, Misplaced Images: Structural Input Perturbations Expose Multimodal Alignment Blind Spots
von: Shayegani, Erfan, et al.
Veröffentlicht: (2025)
von: Shayegani, Erfan, et al.
Veröffentlicht: (2025)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
von: Brundage, Miles, et al.
Veröffentlicht: (2026)
von: Brundage, Miles, et al.
Veröffentlicht: (2026)
Private, Verifiable, and Auditable AI Systems
von: South, Tobin
Veröffentlicht: (2025)
von: South, Tobin
Veröffentlicht: (2025)
Impact Matters! An Audit Method to Evaluate AI Projects and their Impact for Sustainability and Public Interest
von: Züger, Theresa, et al.
Veröffentlicht: (2026)
von: Züger, Theresa, et al.
Veröffentlicht: (2026)
Frontier Lag: A Bibliometric Audit of Capability Misrepresentation in Academic AI Evaluation
von: Gringras, David, et al.
Veröffentlicht: (2026)
von: Gringras, David, et al.
Veröffentlicht: (2026)
A Tale of Two Identities: An Ethical Audit of Human and AI-Crafted Personas
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2025)
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2025)
AI Governance in the GCC States: A Comparative Analysis of National AI Strategies
von: Albous, Mohammad Rashed, et al.
Veröffentlicht: (2025)
von: Albous, Mohammad Rashed, et al.
Veröffentlicht: (2025)
From Transparency to Accountability and Back: A Discussion of Access and Evidence in AI Auditing
von: Cen, Sarah H., et al.
Veröffentlicht: (2024)
von: Cen, Sarah H., et al.
Veröffentlicht: (2024)
Does Claude's Constitution Have a Culture?
von: Pourdavood, Parham
Veröffentlicht: (2026)
von: Pourdavood, Parham
Veröffentlicht: (2026)
Ähnliche Einträge
-
Global AI Bias Audit for Technical Governance
von: Hung, Jason
Veröffentlicht: (2026) -
AI's Blind Spots: Geographic Knowledge and Diversity Deficit in Generated Urban Scenario
von: Beneduce, Ciro, et al.
Veröffentlicht: (2025) -
All Eyes on the Ranker: Participatory Auditing to Surface Blind Spots in Ranked Search Results
von: Rezk, Anna Marie, et al.
Veröffentlicht: (2026) -
The DSA's Blind Spot: Algorithmic Audit of Advertising and Minor Profiling on TikTok
von: Solarova, Sara, et al.
Veröffentlicht: (2026) -
From Defense to Advocacy: Empowering Users to Leverage the Blind Spot of AI Inference
von: Wei, Yumou, et al.
Veröffentlicht: (2026)