Measuring the metacognition of AI
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Servajean, Richard, Servajean, Philippe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GeoPl@ntNet: A Platform for Exploring Essential Biodiversity Variables
von: Picek, Lukas, et al.
Veröffentlicht: (2025)
von: Picek, Lukas, et al.
Veröffentlicht: (2025)
How to Optimize Multispecies Set Predictions in Presence-Absence Modeling ?
von: Gigot--Léandri, Sébastien, et al.
Veröffentlicht: (2026)
von: Gigot--Léandri, Sébastien, et al.
Veröffentlicht: (2026)
Impact of population size on early adaptation in rugged fitness landscapes
von: Servajean, Richard, et al.
Veröffentlicht: (2022)
von: Servajean, Richard, et al.
Veröffentlicht: (2022)
Mapping biodiversity at very-high resolution in Europe
von: Leblanc, César, et al.
Veröffentlicht: (2025)
von: Leblanc, César, et al.
Veröffentlicht: (2025)
Impact of complex spatial population structure on early and long-term adaptation in rugged fitness landscapes
von: Servajean, Richard, et al.
Veröffentlicht: (2024)
von: Servajean, Richard, et al.
Veröffentlicht: (2024)
Imagining and building wise machines: The centrality of AI metacognition
von: Johnson, Samuel G. B., et al.
Veröffentlicht: (2024)
von: Johnson, Samuel G. B., et al.
Veröffentlicht: (2024)
Before you <think>, monitor: Implementing Flavell's metacognitive framework in LLMs
von: Oh, Nick
Veröffentlicht: (2025)
von: Oh, Nick
Veröffentlicht: (2025)
AI-based Mapping of the Conservation Status of Orchid Assemblages at Global Scale
von: Estopinan, Joaquim, et al.
Veröffentlicht: (2024)
von: Estopinan, Joaquim, et al.
Veröffentlicht: (2024)
Could you be wrong: Debiasing LLMs using a metacognitive prompt for improving human decision making
von: Hills, Thomas T.
Veröffentlicht: (2025)
von: Hills, Thomas T.
Veröffentlicht: (2025)
Generative AI as a metacognitive agent: A comparative mixed-method study with human participants on ICF-mimicking exam performance
von: Pavlovic, Jelena, et al.
Veröffentlicht: (2024)
von: Pavlovic, Jelena, et al.
Veröffentlicht: (2024)
Domain-level metacognitive monitoring in frontier LLMs: A 33-model atlas
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
Branching Out: Broadening AI Measurement and Evaluation with Measurement Trees
von: Greenberg, Craig, et al.
Veröffentlicht: (2025)
von: Greenberg, Craig, et al.
Veröffentlicht: (2025)
Modelling Species Distributions with Deep Learning to Predict Plant Extinction Risk and Assess Climate Change Impacts
von: Estopinan, Joaquim, et al.
Veröffentlicht: (2024)
von: Estopinan, Joaquim, et al.
Veröffentlicht: (2024)
Measuring AI Alignment with Human Flourishing
von: Hilliard, Elizabeth, et al.
Veröffentlicht: (2025)
von: Hilliard, Elizabeth, et al.
Veröffentlicht: (2025)
Measuring What Matters: The AI Pluralism Index
von: Mushkani, Rashid
Veröffentlicht: (2025)
von: Mushkani, Rashid
Veröffentlicht: (2025)
Open-World Evaluations for Measuring Frontier AI Capabilities
von: Kapoor, Sayash, et al.
Veröffentlicht: (2026)
von: Kapoor, Sayash, et al.
Veröffentlicht: (2026)
Measuring the environmental impact of delivering AI at Google Scale
von: Elsworth, Cooper, et al.
Veröffentlicht: (2025)
von: Elsworth, Cooper, et al.
Veröffentlicht: (2025)
Don't Measure Once: Measuring Visibility in AI Search (GEO)
von: Schulte, Julius, et al.
Veröffentlicht: (2026)
von: Schulte, Julius, et al.
Veröffentlicht: (2026)
Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems
von: Johnson, Rebecca L.
Veröffentlicht: (2026)
von: Johnson, Rebecca L.
Veröffentlicht: (2026)
AI Must Embrace Specialization via Superhuman Adaptable Intelligence
von: Goldfeder, Judah, et al.
Veröffentlicht: (2026)
von: Goldfeder, Judah, et al.
Veröffentlicht: (2026)
Measuring AI R&D Automation
von: Chan, Alan, et al.
Veröffentlicht: (2026)
von: Chan, Alan, et al.
Veröffentlicht: (2026)
Measuring What Matters: Benchmarking Generative, Multimodal, and Agentic AI in Healthcare
von: Desikan, Prasanna, et al.
Veröffentlicht: (2026)
von: Desikan, Prasanna, et al.
Veröffentlicht: (2026)
Intentionality is a Design Decision: Measuring Functional Intentionality for Accountable AI Systems
von: Chiappetta, Allessia, et al.
Veröffentlicht: (2026)
von: Chiappetta, Allessia, et al.
Veröffentlicht: (2026)
Safety by Measurement: A Systematic Literature Review of AI Safety Evaluation Methods
von: Grey, Markov, et al.
Veröffentlicht: (2025)
von: Grey, Markov, et al.
Veröffentlicht: (2025)
Measuring AI agent autonomy: Towards a scalable approach with code inspection
von: Cihon, Peter, et al.
Veröffentlicht: (2025)
von: Cihon, Peter, et al.
Veröffentlicht: (2025)
Measuring AI Reasoning: A Guide for Researchers
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026)
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026)
Measuring What AI Systems Might Do: Towards A Measurement Science in AI
von: Voudouris, Konstantinos, et al.
Veröffentlicht: (2026)
von: Voudouris, Konstantinos, et al.
Veröffentlicht: (2026)
Advancing Explainable AI Toward Human-Like Intelligence: Forging the Path to Artificial Brain
von: Zhou, Yongchen, et al.
Veröffentlicht: (2024)
von: Zhou, Yongchen, et al.
Veröffentlicht: (2024)
Favi-Score: A Measure for Favoritism in Automated Preference Ratings for Generative AI Evaluation
von: von Däniken, Pius, et al.
Veröffentlicht: (2024)
von: von Däniken, Pius, et al.
Veröffentlicht: (2024)
Ground-Truthing AI Energy Consumption: Validating CodeCarbon Against External Measurements
von: Fischer, Raphael
Veröffentlicht: (2025)
von: Fischer, Raphael
Veröffentlicht: (2025)
Voice-Enabled AI Agents can Perform Common Scams
von: Fang, Richard, et al.
Veröffentlicht: (2024)
von: Fang, Richard, et al.
Veröffentlicht: (2024)
Measuring AI Ability to Complete Long Software Tasks
von: Kwa, Thomas, et al.
Veröffentlicht: (2025)
von: Kwa, Thomas, et al.
Veröffentlicht: (2025)
Feedback Forensics: A Toolkit to Measure AI Personality
von: Findeis, Arduin, et al.
Veröffentlicht: (2025)
von: Findeis, Arduin, et al.
Veröffentlicht: (2025)
Using AI to Measure Parkinson's Disease Severity at Home
von: Islam, Md Saiful, et al.
Veröffentlicht: (2023)
von: Islam, Md Saiful, et al.
Veröffentlicht: (2023)
LLM Rationalis? Measuring Bargaining Capabilities of AI Negotiators
von: Shah, Cheril, et al.
Veröffentlicht: (2025)
von: Shah, Cheril, et al.
Veröffentlicht: (2025)
Measuring AI Diffusion: A Population-Normalized Metric for Tracking Global AI Usage
von: Misra, Amit, et al.
Veröffentlicht: (2025)
von: Misra, Amit, et al.
Veröffentlicht: (2025)
Who Uses AI? Platform Selection and the Measurement of Occupational AI Exposure
von: Yin, Michelle, et al.
Veröffentlicht: (2026)
von: Yin, Michelle, et al.
Veröffentlicht: (2026)
Applying the maximum entropy principle to neural networks enhances multi‐species distribution models
von: Maxime Ryckewaert, et al.
Veröffentlicht: (2026)
von: Maxime Ryckewaert, et al.
Veröffentlicht: (2026)
Applying the maximum entropy principle to neural networks enhances multi-species distribution models
von: Ryckewaert, Maxime, et al.
Veröffentlicht: (2024)
von: Ryckewaert, Maxime, et al.
Veröffentlicht: (2024)
MALPOLON: A Framework for Deep Species Distribution Modeling
von: Larcher, Theo, et al.
Veröffentlicht: (2024)
von: Larcher, Theo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GeoPl@ntNet: A Platform for Exploring Essential Biodiversity Variables
von: Picek, Lukas, et al.
Veröffentlicht: (2025) -
How to Optimize Multispecies Set Predictions in Presence-Absence Modeling ?
von: Gigot--Léandri, Sébastien, et al.
Veröffentlicht: (2026) -
Impact of population size on early adaptation in rugged fitness landscapes
von: Servajean, Richard, et al.
Veröffentlicht: (2022) -
Mapping biodiversity at very-high resolution in Europe
von: Leblanc, César, et al.
Veröffentlicht: (2025) -
Impact of complex spatial population structure on early and long-term adaptation in rugged fitness landscapes
von: Servajean, Richard, et al.
Veröffentlicht: (2024)