Measuring the Authority Stack of AI Systems: Empirical Analysis of 366,120 Forced-Choice Responses Across 8 AI Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Lee, Seulki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Comparing Fairness of Generative Mobility Models
von: Wang, Daniel, et al.
Veröffentlicht: (2024)
von: Wang, Daniel, et al.
Veröffentlicht: (2024)
Operational AI Deployment Assurance: Governance-State Orchestration Under Threshold-Sensitive Deployment Conditions -- A Governance Framework for High-Stakes AI Systems
von: Alsayed, Khalid Adnan
Veröffentlicht: (2026)
von: Alsayed, Khalid Adnan
Veröffentlicht: (2026)
AI Integrity: A New Paradigm for Verifiable AI Governance
von: Lee, Seulki
Veröffentlicht: (2026)
von: Lee, Seulki
Veröffentlicht: (2026)
Toward Individual Fairness Without Centralized Data: Selective Counterfactual Consistency for Vertical Federated Learning
von: Wasif, Dawood, et al.
Veröffentlicht: (2026)
von: Wasif, Dawood, et al.
Veröffentlicht: (2026)
PRISM Risk Signal Framework: Hierarchy-Based Red Lines for AI Behavioral Risk
von: Lee, Seulki
Veröffentlicht: (2026)
von: Lee, Seulki
Veröffentlicht: (2026)
Futurity as Infrastructure: A Techno-Philosophical Interpretation of the AI Lifecycle
von: Cote, Mark, et al.
Veröffentlicht: (2025)
von: Cote, Mark, et al.
Veröffentlicht: (2025)
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
von: Beltoft, Stine, et al.
Veröffentlicht: (2025)
von: Beltoft, Stine, et al.
Veröffentlicht: (2025)
Fairness Interventions: A Study in AI Explainability
von: Souverain, Thomas, et al.
Veröffentlicht: (2024)
von: Souverain, Thomas, et al.
Veröffentlicht: (2024)
Membership Inference Attacks against Large Audio Language Models
von: Dong, Jia-Kai, et al.
Veröffentlicht: (2026)
von: Dong, Jia-Kai, et al.
Veröffentlicht: (2026)
Fairness Is Not Enough: Auditing Competence and Intersectional Bias in AI-powered Resume Screening
von: Webster, Kevin T
Veröffentlicht: (2025)
von: Webster, Kevin T
Veröffentlicht: (2025)
Inference Scaling Reshapes AI Governance
von: Ord, Toby
Veröffentlicht: (2025)
von: Ord, Toby
Veröffentlicht: (2025)
Bias Mitigation for AI-Feedback Loops in Recommender Systems: A Systematic Literature Review and Taxonomy
von: Stoecker, Theodor, et al.
Veröffentlicht: (2025)
von: Stoecker, Theodor, et al.
Veröffentlicht: (2025)
Scalable and Verifiable Federated Learning for Cross-Institution Financial Fraud Detection
von: Panth, Prajwal, et al.
Veröffentlicht: (2026)
von: Panth, Prajwal, et al.
Veröffentlicht: (2026)
LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability
von: Lucas, Tom, et al.
Veröffentlicht: (2026)
von: Lucas, Tom, et al.
Veröffentlicht: (2026)
The AI Fiction Paradox
von: Elkins, Katherine
Veröffentlicht: (2026)
von: Elkins, Katherine
Veröffentlicht: (2026)
Selecting for Less Discriminatory Algorithms: A Relational Search Framework for Navigating Fairness-Accuracy Trade-offs in Practice
von: Samad, Hana, et al.
Veröffentlicht: (2025)
von: Samad, Hana, et al.
Veröffentlicht: (2025)
Learning the Signature of Memorization in Autoregressive Language Models
von: Ilić, David, et al.
Veröffentlicht: (2026)
von: Ilić, David, et al.
Veröffentlicht: (2026)
Human Values in a Single Sentence: Moral Presence, Hierarchies, and Transformer Ensembles on the Schwartz Continuum
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
REMIND: Input Loss Landscapes Reveal Residual Memorization in Post-Unlearning LLMs
von: Cohen, Liran, et al.
Veröffentlicht: (2025)
von: Cohen, Liran, et al.
Veröffentlicht: (2025)
Privacy as Commodity: MFG-RegretNet for Large-Scale Privacy Trading in Federated Learning
von: Sun, Kangkang, et al.
Veröffentlicht: (2026)
von: Sun, Kangkang, et al.
Veröffentlicht: (2026)
HybridVFL: Disentangled Feature Learning for Edge-Enabled Vertical Federated Multimodal Classification
von: Anoosha, Mostafa, et al.
Veröffentlicht: (2025)
von: Anoosha, Mostafa, et al.
Veröffentlicht: (2025)
Tatemae: Detecting Alignment Faking via Tool Selection in LLMs
von: Leonesi, Matteo, et al.
Veröffentlicht: (2026)
von: Leonesi, Matteo, et al.
Veröffentlicht: (2026)
From Safety Risk to Design Principle: Peer-Preservation in Multi-Agent LLM Systems and Its Implications for Orchestrated Democratic Discourse Analysis
von: Dietrich, Juergen
Veröffentlicht: (2026)
von: Dietrich, Juergen
Veröffentlicht: (2026)
Uncovering Bugs in Formal Explainers: A Case Study with PyXAI
von: Huang, Xuanxiang, et al.
Veröffentlicht: (2025)
von: Huang, Xuanxiang, et al.
Veröffentlicht: (2025)
How Does Environmental Information Disclosure Affect Corporate Environmental Performance? Evidence from Chinese A-Share Listed Companies
von: Lin, Zehao
Veröffentlicht: (2025)
von: Lin, Zehao
Veröffentlicht: (2025)
Learning from Discriminatory Training Data
von: Grabowicz, Przemyslaw A., et al.
Veröffentlicht: (2019)
von: Grabowicz, Przemyslaw A., et al.
Veröffentlicht: (2019)
How Worrying Are Privacy Attacks Against Machine Learning?
von: Domingo-Ferrer, Josep
Veröffentlicht: (2025)
von: Domingo-Ferrer, Josep
Veröffentlicht: (2025)
Closing the SNAP Gap: Identifying Under-Enrollment in High-Poverty ZIP Codes
von: Ray, Auyona
Veröffentlicht: (2025)
von: Ray, Auyona
Veröffentlicht: (2025)
Will Humanity Be Rendered Obsolete by AI?
von: Louadi, Mohamed El, et al.
Veröffentlicht: (2025)
von: Louadi, Mohamed El, et al.
Veröffentlicht: (2025)
More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
Who Governs the Machine? A Machine Identity Governance Taxonomy (MIGT) for AI Systems Operating Across Enterprise and Geopolitical Boundaries
von: Kurtz, Andrew, et al.
Veröffentlicht: (2026)
von: Kurtz, Andrew, et al.
Veröffentlicht: (2026)
Privately Fine-Tuned LLMs Preserve Temporal Dynamics in Tabular Data
von: Rosenblatt, Lucas, et al.
Veröffentlicht: (2026)
von: Rosenblatt, Lucas, et al.
Veröffentlicht: (2026)
The Epistemic Suite: A Post-Foundational Diagnostic Methodology for Assessing AI Knowledge Claims
von: Kelly, Matthew
Veröffentlicht: (2025)
von: Kelly, Matthew
Veröffentlicht: (2025)
SafetyDrift: Predicting When AI Agents Cross the Line Before They Actually Do
von: Dhodapkar, Aditya, et al.
Veröffentlicht: (2026)
von: Dhodapkar, Aditya, et al.
Veröffentlicht: (2026)
Why we need an AI-resilient society
von: Bartz-Beielstein, Thomas
Veröffentlicht: (2019)
von: Bartz-Beielstein, Thomas
Veröffentlicht: (2019)
BioRefusalAudit: Auditing Biosecurity Refusal Depth Using General and Domain-Fine-Tuned Sparse Autoencoders
von: DeLeeuw, Caleb
Veröffentlicht: (2026)
von: DeLeeuw, Caleb
Veröffentlicht: (2026)
Do Schwartz Higher-Order Values Help Sentence-Level Human Value Detection? A Study of Hierarchical Gating and Calibration
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
Digital Forgetting in Large Language Models: A Survey of Unlearning Methods
von: Blanco-Justicia, Alberto, et al.
Veröffentlicht: (2024)
von: Blanco-Justicia, Alberto, et al.
Veröffentlicht: (2024)
The Necessity of AI Audit Standards Boards
von: Manheim, David, et al.
Veröffentlicht: (2024)
von: Manheim, David, et al.
Veröffentlicht: (2024)
Engineering Trustworthy AI: A Developer Guide for Empirical Risk Minimization
von: Pfau, Diana, et al.
Veröffentlicht: (2024)
von: Pfau, Diana, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Comparing Fairness of Generative Mobility Models
von: Wang, Daniel, et al.
Veröffentlicht: (2024) -
Operational AI Deployment Assurance: Governance-State Orchestration Under Threshold-Sensitive Deployment Conditions -- A Governance Framework for High-Stakes AI Systems
von: Alsayed, Khalid Adnan
Veröffentlicht: (2026) -
AI Integrity: A New Paradigm for Verifiable AI Governance
von: Lee, Seulki
Veröffentlicht: (2026) -
Toward Individual Fairness Without Centralized Data: Selective Counterfactual Consistency for Vertical Federated Learning
von: Wasif, Dawood, et al.
Veröffentlicht: (2026) -
PRISM Risk Signal Framework: Hierarchy-Based Red Lines for AI Behavioral Risk
von: Lee, Seulki
Veröffentlicht: (2026)