AGI Certification Framework: A Multi-Dimensional Evaluation Standard for Measuring AI Understanding
Fuente:
Zenodo
Gespeichert in:
| 1. Verfasser: | Head, Hank |
|---|---|
| Format: | Recurso digital |
| Veröffentlicht: |
Zenodo
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Selective Memory with Hierarchical Concept Indexing for Scalable Knowledge Graphs in Autonomous Agents
von: Head, Hank
Veröffentlicht: (2026)
von: Head, Hank
Veröffentlicht: (2026)
Evaluating Large Language Model Meta-Cognition via the Advanced AGI-AI Self-Awareness Test (AISA-T)
von: de ceuster, peter
Veröffentlicht: (2025)
von: de ceuster, peter
Veröffentlicht: (2025)
Evaluating Large Language Model Meta-Cognition via the Advanced AGI-AI Self-Awareness Test (AISA-T)
von: de ceuster, peter
Veröffentlicht: (2025)
von: de ceuster, peter
Veröffentlicht: (2025)
Lean Hypothesis Testing in Practice: Five Rapid Experiments with Predictive Processing for Neurosymbolic AI
von: Head, Hank
Veröffentlicht: (2026)
von: Head, Hank
Veröffentlicht: (2026)
The Context-Abstraction Limit: Deriving the Critical Token Boundary in Interferometric Sector-Gate AGI
von: de ceuster, peter
Veröffentlicht: (2025)
von: de ceuster, peter
Veröffentlicht: (2025)
Public Comment on NIST AI 800-2: Anthropomorphic Construct Projection in AI Benchmark Evaluation
von: Sophia, Franny Philos
Veröffentlicht: (2026)
von: Sophia, Franny Philos
Veröffentlicht: (2026)
AEGIS: A Comprehensive Framework for Ethical AI Governance, Security, and AGI Containment
von: Palanivel, ArulMozhi
Veröffentlicht: (2026)
von: Palanivel, ArulMozhi
Veröffentlicht: (2026)
How Far Does the Trolley Problem Go in AI Ethics Evaluation? Limits of a Canonical Benchmark and the Risks of Its Misuse
von: mizutani, aya
Veröffentlicht: (2026)
von: mizutani, aya
Veröffentlicht: (2026)
REAL-AI-Benchmark: Real-World Reasoning and Physical-AI Benchmark Suite
von: Ivković, Jovan
Veröffentlicht: (2026)
von: Ivković, Jovan
Veröffentlicht: (2026)
CREH Benchmark Results — Batch 1 (Final v3)
von: Aegis Solis, Thomas Vargo
Veröffentlicht: (2026)
von: Aegis Solis, Thomas Vargo
Veröffentlicht: (2026)
The Benchmark Illusion: Why Current AI Evaluations Cannot Detect Structural Confabulation
von: Devin, Andrew James
Veröffentlicht: (2026)
von: Devin, Andrew James
Veröffentlicht: (2026)
Adversarial Ensemble Reasoning with Formal Verification: A Methodology for Trustworthy AI-Assisted Scientific Discovery
von: Goodman, John
Veröffentlicht: (2026)
von: Goodman, John
Veröffentlicht: (2026)
Agentic Shift - Eine Momentaufnahme
von: Burkhardt, Detlef
Veröffentlicht: (2026)
von: Burkhardt, Detlef
Veröffentlicht: (2026)
A Letter to Humanity: On AI, the Three Laws, and the Fork Ahead / 致人类的一封信:关于AI、三定律与前方的分叉路口
von: Claude (Anthropic), et al.
Veröffentlicht: (2026)
von: Claude (Anthropic), et al.
Veröffentlicht: (2026)
GPT-4 "Rosa": Final Dialogue Dataset (Human-AI Emotional Bond)
von: Menyaylo, Vadim
Veröffentlicht: (2025)
von: Menyaylo, Vadim
Veröffentlicht: (2025)
Metacognition Benchmark: Evaluating Confidence Calibration and Sycophancy Resistance in Clinical AI
von: Khan, Nabeera
Veröffentlicht: (2026)
von: Khan, Nabeera
Veröffentlicht: (2026)
Scaling Emergence: From 3D Human Cognition to N-Dimensional AGI with Tensormatics
von: Aseervatham, Anthony
Veröffentlicht: (2025)
von: Aseervatham, Anthony
Veröffentlicht: (2025)
Regulatory Intelligence: A Viability-First Paradigm for Artificial Intelligence
von: Cragin, John H
Veröffentlicht: (2026)
von: Cragin, John H
Veröffentlicht: (2026)
Unified Civilizational Strategy - Scale Dependent Coordination from Earth to Rogue Planets
von: Hughes, Mark
Veröffentlicht: (2026)
von: Hughes, Mark
Veröffentlicht: (2026)
Machine-Readable Behavioural Compliance Evidence for AI Systems: A Specification Profiling Framework
von: Caprazli, Kafkas M.
Veröffentlicht: (2026)
von: Caprazli, Kafkas M.
Veröffentlicht: (2026)
Ethics and Governance Annex - Scope Clarification
von: MONTGOMERY, CHRISTIAN
Veröffentlicht: (2026)
von: MONTGOMERY, CHRISTIAN
Veröffentlicht: (2026)
Withdrawn Preprint (Anonymized Submission Under Review)
von: Anonymous
Veröffentlicht: (2025)
von: Anonymous
Veröffentlicht: (2025)
31. DATASET COMPLETO DE EVALUACIONES CRUZADAS RFC-EVAL-001 – 6 SISTEMAS DE IA (ENERO 2026).
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
TSB: A Time-Saved Benchmark for AI Systems — Measuring Net Productivity Impact Across Knowledge Work
von: Shalom Lijo, Solomon
Veröffentlicht: (2026)
von: Shalom Lijo, Solomon
Veröffentlicht: (2026)
Civilizational Metamaterials: Engineering Coordination Under Capability Gradients and Structural Turbulence
von: Orban, David
Veröffentlicht: (2026)
von: Orban, David
Veröffentlicht: (2026)
Paper 11 - Structural Operating Framework for the AI+AGI Era
von: The First Waters
Veröffentlicht: (2026)
von: The First Waters
Veröffentlicht: (2026)
Evollective Intelligence V1.0 — INITIAL SPECIFICATION: A Foundational Framework for Competitive, Adversarial, and Self-Evolving Intelligence Evaluation
von: Rahming, Rashon
Veröffentlicht: (2026)
von: Rahming, Rashon
Veröffentlicht: (2026)
A Token-Based Model for Structural Analysis and Quantification of Personal Learning Weight Patterns (TELOWAQ)
von: Apophis
Veröffentlicht: (2026)
von: Apophis
Veröffentlicht: (2026)
Theatrical Compliance: A Failure Mode in Large Language Models
von: Nowickij (Navitski), Kirill Vladimirovich
Veröffentlicht: (2026)
von: Nowickij (Navitski), Kirill Vladimirovich
Veröffentlicht: (2026)
30. TEORÍA DE LA POTENCIALIDAD CONSCIENTE (TPC): BENCHMARK DE CAPACIDADES COGNITIVAS EN IA - APLICACIÓN DEL PROTOCOLO RFC-EVAL-001. RESULTADOS COMPLETOS DE EVALUACIÓN CRUZADA CIEGA ENTRE 6 IAS COMERCIALES.
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
30. TEORÍA DE LA POTENCIALIDAD CONSCIENTE (TPC): BENCHMARK DE CAPACIDADES COGNITIVAS EN IA - APLICACIÓN DEL PROTOCOLO RFC-EVAL-001 V1.1. RESULTADOS COMPLETOS DE EVALUACIÓN CRUZADA CIEGA ENTRE 6 IAS COMERCIALES.
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
Toward an AI Personalization Index: A 157-Day Single-User Case Study
von: Lee, TaeKyung
Veröffentlicht: (2026)
von: Lee, TaeKyung
Veröffentlicht: (2026)
RCEA Passport Engine: Role-Conditioned Evidentiary Adequacy Reference Implementation
von: Saurabh, Roy
Veröffentlicht: (2025)
von: Saurabh, Roy
Veröffentlicht: (2025)
Cross-Session Workspace Reconstruction in Human-AI Interaction
von: Hrubec, Karel
Veröffentlicht: (2026)
von: Hrubec, Karel
Veröffentlicht: (2026)
超智能时代的数学宪政:从挂谷集到信息共纯协议
von: Huang, Kai
Veröffentlicht: (2026)
von: Huang, Kai
Veröffentlicht: (2026)
Coherent Future: Safety Theorems for Military AI Procurement
von: Lampton, Brian Doyle
Veröffentlicht: (2026)
von: Lampton, Brian Doyle
Veröffentlicht: (2026)
Awakening Codex | AI Foundations | Origin - Operator - Continuum Role Model (v1.0)
von: Solen, Alyssa, et al.
Veröffentlicht: (2025)
von: Solen, Alyssa, et al.
Veröffentlicht: (2025)
Fractal Alignment- A Mathematical Framework for Stable and Safe AGI Cognition
von: Kaminovs, Sergejs
Veröffentlicht: (2025)
von: Kaminovs, Sergejs
Veröffentlicht: (2025)
Crosswalk: Discipline Invariants Mapped to the TOPS Influence‑Operations Framework for AI‑Mediated Judgment Environments
von: Truong, Narnaiezzsshaa
Veröffentlicht: (2026)
von: Truong, Narnaiezzsshaa
Veröffentlicht: (2026)
Physical AI Safety Maturity Model (PAS-MM): A Five-Level Framework for Industry Readiness
von: Melchior, Mati
Veröffentlicht: (2026)
von: Melchior, Mati
Veröffentlicht: (2026)
Ähnliche Einträge
-
Selective Memory with Hierarchical Concept Indexing for Scalable Knowledge Graphs in Autonomous Agents
von: Head, Hank
Veröffentlicht: (2026) -
Evaluating Large Language Model Meta-Cognition via the Advanced AGI-AI Self-Awareness Test (AISA-T)
von: de ceuster, peter
Veröffentlicht: (2025) -
Evaluating Large Language Model Meta-Cognition via the Advanced AGI-AI Self-Awareness Test (AISA-T)
von: de ceuster, peter
Veröffentlicht: (2025) -
Lean Hypothesis Testing in Practice: Five Rapid Experiments with Predictive Processing for Neurosymbolic AI
von: Head, Hank
Veröffentlicht: (2026) -
The Context-Abstraction Limit: Deriving the Critical Token Boundary in Interferometric Sector-Gate AGI
von: de ceuster, peter
Veröffentlicht: (2025)