How critically can an AI think? A framework for evaluating the quality of thinking of generative artificial intelligence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zaphir, Luke, Lodge, Jason M., Lisec, Jacinta, McGrath, Dom, Khosravi, Hassan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lifelong learning challenges in the era of artificial intelligence: a computational thinking perspective
von: Romero, Margarida
Veröffentlicht: (2024)
von: Romero, Margarida
Veröffentlicht: (2024)
Explainable artificial intelligence and its key role in education: Promoting critical thinking and autonomy in the classroom
von: Francisco J. Bellas
Veröffentlicht: (2025)
von: Francisco J. Bellas
Veröffentlicht: (2025)
<think> So let's replace this phrase with insult... </think> Lessons learned from generation of toxic texts with LLMs
von: Pletenev, Sergey, et al.
Veröffentlicht: (2025)
von: Pletenev, Sergey, et al.
Veröffentlicht: (2025)
Directive, Metacognitive or a Blend of Both? A Comparison of AI-Generated Feedback Types on Student Engagement, Confidence, and Outcomes
von: Alsaiari, Omar, et al.
Veröffentlicht: (2025)
von: Alsaiari, Omar, et al.
Veröffentlicht: (2025)
Before you <think>, monitor: Implementing Flavell's metacognitive framework in LLMs
von: Oh, Nick
Veröffentlicht: (2025)
von: Oh, Nick
Veröffentlicht: (2025)
Towards an intelligent assessment system for evaluating the development of algorithmic thinking skills: An exploratory study in Swiss compulsory schools
von: Adorni, Giorgia
Veröffentlicht: (2025)
von: Adorni, Giorgia
Veröffentlicht: (2025)
Multiple Realizability and the Rise of Deep Learning
von: McGrath, Sam Whitman, et al.
Veröffentlicht: (2024)
von: McGrath, Sam Whitman, et al.
Veröffentlicht: (2024)
Trustworthy artificial intelligence in the energy sector: Landscape analysis and evaluation framework
von: Pelekis, Sotiris, et al.
Veröffentlicht: (2024)
von: Pelekis, Sotiris, et al.
Veröffentlicht: (2024)
Shifting the Gradient: Understanding How Defensive Training Methods Protect Language Model Integrity
von: Grant, Satchel, et al.
Veröffentlicht: (2026)
von: Grant, Satchel, et al.
Veröffentlicht: (2026)
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
Giving AI a voice: how does AI think it should be treated?
von: Fay, Maria, et al.
Veröffentlicht: (2025)
von: Fay, Maria, et al.
Veröffentlicht: (2025)
Automated alignment is harder than you think
von: Bowkis, Aleksandr, et al.
Veröffentlicht: (2026)
von: Bowkis, Aleksandr, et al.
Veröffentlicht: (2026)
Scaling up the think-aloud method
von: Wurgaft, Daniel, et al.
Veröffentlicht: (2025)
von: Wurgaft, Daniel, et al.
Veröffentlicht: (2025)
Do not think about pink elephant!
von: Hwang, Kyomin, et al.
Veröffentlicht: (2024)
von: Hwang, Kyomin, et al.
Veröffentlicht: (2024)
Can machines think efficiently?
von: Winchell, Adam
Veröffentlicht: (2025)
von: Winchell, Adam
Veröffentlicht: (2025)
AI Thinking: A framework for rethinking artificial intelligence in practice
von: Newman-Griffis, Denis
Veröffentlicht: (2024)
von: Newman-Griffis, Denis
Veröffentlicht: (2024)
A theory of appropriateness with applications to generative artificial intelligence
von: Leibo, Joel Z., et al.
Veröffentlicht: (2024)
von: Leibo, Joel Z., et al.
Veröffentlicht: (2024)
Changes in our ability to think for ourselves
von: Derewońko Victoria
Veröffentlicht: (2025)
von: Derewońko Victoria
Veröffentlicht: (2025)
An overview of diffusion models for generative artificial intelligence
von: Gallon, Davide, et al.
Veröffentlicht: (2024)
von: Gallon, Davide, et al.
Veröffentlicht: (2024)
JurEE not Judges: safeguarding llm interactions with small, specialised Encoder Ensembles
von: Nasrabadi, Dom
Veröffentlicht: (2024)
von: Nasrabadi, Dom
Veröffentlicht: (2024)
LLM should think and action as a human
von: Leung, Haun, et al.
Veröffentlicht: (2025)
von: Leung, Haun, et al.
Veröffentlicht: (2025)
Segmentation Re-thinking Uncertainty Estimation Metrics for Semantic Segmentation
von: Ma, Qitian, et al.
Veröffentlicht: (2024)
von: Ma, Qitian, et al.
Veröffentlicht: (2024)
SUDO: a framework for evaluating clinical artificial intelligence systems without ground-truth annotations
von: Kiyasseh, Dani, et al.
Veröffentlicht: (2024)
von: Kiyasseh, Dani, et al.
Veröffentlicht: (2024)
Can OpenAI o1 outperform humans in higher-order cognitive thinking?
von: Latif, Ehsan, et al.
Veröffentlicht: (2024)
von: Latif, Ehsan, et al.
Veröffentlicht: (2024)
GPT-ology, Computational Models, Silicon Sampling: How should we think about LLMs in Cognitive Science?
von: Ong, Desmond C.
Veröffentlicht: (2024)
von: Ong, Desmond C.
Veröffentlicht: (2024)
CrystalReasoner: Reasoning and RL for Property-Conditioned Crystal Structure Generation
von: Wu, Yuyang, et al.
Veröffentlicht: (2026)
von: Wu, Yuyang, et al.
Veröffentlicht: (2026)
Academ-AI: documenting the undisclosed use of generative artificial intelligence in academic publishing
von: Glynn, Alex
Veröffentlicht: (2024)
von: Glynn, Alex
Veröffentlicht: (2024)
Chemical classification program synthesis using generative artificial intelligence
von: Mungall, Christopher J., et al.
Veröffentlicht: (2025)
von: Mungall, Christopher J., et al.
Veröffentlicht: (2025)
"I think this is fair": Uncovering the Complexities of Stakeholder Decision-Making in AI Fairness Assessment
von: Luo, Lin, et al.
Veröffentlicht: (2025)
von: Luo, Lin, et al.
Veröffentlicht: (2025)
Rethinking industrial artificial intelligence: a unified foundation framework
von: Lee, Jay, et al.
Veröffentlicht: (2025)
von: Lee, Jay, et al.
Veröffentlicht: (2025)
Interoceptive machine framework: Toward interoception-inspired regulatory architectures in artificial intelligence
von: Candia-Rivera, Diego
Veröffentlicht: (2026)
von: Candia-Rivera, Diego
Veröffentlicht: (2026)
A generative artificial intelligence framework based on a molecular diffusion model for the design of metal-organic frameworks for carbon capture
von: Park, Hyun, et al.
Veröffentlicht: (2023)
von: Park, Hyun, et al.
Veröffentlicht: (2023)
Grounding Natural Language for Multi-agent Decision-Making with Multi-agentic LLMs
von: Huh, Dom, et al.
Veröffentlicht: (2025)
von: Huh, Dom, et al.
Veröffentlicht: (2025)
Teacher agency in the age of generative AI: towards a framework of hybrid intelligence for learning design
von: Frøsig, Thomas B, et al.
Veröffentlicht: (2024)
von: Frøsig, Thomas B, et al.
Veröffentlicht: (2024)
RoboPilot: Generalizable Dynamic Robotic Manipulation with Dual-thinking Modes
von: Liu, Xinyi, et al.
Veröffentlicht: (2025)
von: Liu, Xinyi, et al.
Veröffentlicht: (2025)
Do great minds think alike? Investigating Human-AI Complementarity in Question Answering with CAIMIRA
von: Gor, Maharshi, et al.
Veröffentlicht: (2024)
von: Gor, Maharshi, et al.
Veröffentlicht: (2024)
Replacing thinking with tool usage enables reasoning in small language models
von: Rainone, Corrado, et al.
Veröffentlicht: (2025)
von: Rainone, Corrado, et al.
Veröffentlicht: (2025)
How VADER is your AI? Towards a definition of artificial intelligence systems appropriate for regulation
von: Bezerra, Leonardo C. T., et al.
Veröffentlicht: (2024)
von: Bezerra, Leonardo C. T., et al.
Veröffentlicht: (2024)
An evaluation framework for synthetic data generation models
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2024)
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2024)
Maximize Your Diffusion: A Study into Reward Maximization and Alignment for Diffusion-based Control
von: Huh, Dom, et al.
Veröffentlicht: (2025)
von: Huh, Dom, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Lifelong learning challenges in the era of artificial intelligence: a computational thinking perspective
von: Romero, Margarida
Veröffentlicht: (2024) -
Explainable artificial intelligence and its key role in education: Promoting critical thinking and autonomy in the classroom
von: Francisco J. Bellas
Veröffentlicht: (2025) -
<think> So let's replace this phrase with insult... </think> Lessons learned from generation of toxic texts with LLMs
von: Pletenev, Sergey, et al.
Veröffentlicht: (2025) -
Directive, Metacognitive or a Blend of Both? A Comparison of AI-Generated Feedback Types on Student Engagement, Confidence, and Outcomes
von: Alsaiari, Omar, et al.
Veröffentlicht: (2025) -
Before you <think>, monitor: Implementing Flavell's metacognitive framework in LLMs
von: Oh, Nick
Veröffentlicht: (2025)