Evaluating undergraduate mathematics examinations in the era of generative AI: a curriculum-level case study
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Walker, Benjamin J., Kalaydzhieva, Nikoleta, Lameda, Beatriz Navarro, Reynolds, Ruth A. |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
From Checklists to Clusters: A Homeostatic Account of AGI Evaluation
par: Reynolds, Brett
Publié: (2025)
par: Reynolds, Brett
Publié: (2025)
Computational Hermeneutics: Evaluating generative AI as a cultural technology
par: Kommers, Cody, et autres
Publié: (2026)
par: Kommers, Cody, et autres
Publié: (2026)
The erasure of intensive livestock farming in text-to-image generative AI
par: Sheng, Kehan, et autres
Publié: (2025)
par: Sheng, Kehan, et autres
Publié: (2025)
A taxonomy of epistemic injustice in the context of AI and the case for generative hermeneutical erasure
par: Mollema, Warmhold Jan Thomas
Publié: (2025)
par: Mollema, Warmhold Jan Thomas
Publié: (2025)
Hallucination, reliability, and the role of generative AI in science
par: Rathkopf, Charles
Publié: (2025)
par: Rathkopf, Charles
Publié: (2025)
Teaching Introduction to Programming in the times of AI: A case study of a course re-design
par: Avouris, Nikolaos, et autres
Publié: (2025)
par: Avouris, Nikolaos, et autres
Publié: (2025)
Epistemic Deference to AI
par: Lange, Benjamin
Publié: (2025)
par: Lange, Benjamin
Publié: (2025)
From chalkboards to chatbots: SELAR assists teachers in embracing AI in the curriculum
par: Alers, Hani, et autres
Publié: (2024)
par: Alers, Hani, et autres
Publié: (2024)
Evaluating AI Evaluation: Perils and Prospects
par: Burden, John
Publié: (2024)
par: Burden, John
Publié: (2024)
Ethics of generative AI and manipulation: a design-oriented research agenda
par: Klenk, Michael
Publié: (2025)
par: Klenk, Michael
Publié: (2025)
Distributed agency in second language learning and teaching through generative AI
par: Godwin-Jones, Robert
Publié: (2024)
par: Godwin-Jones, Robert
Publié: (2024)
LegalScore: Development of a Benchmark for Evaluating AI Models in Legal Career Exams in Brazil
par: Caparroz, Roberto, et autres
Publié: (2025)
par: Caparroz, Roberto, et autres
Publié: (2025)
Copyright in AI-generated works: Lessons from recent developments in patent law
par: Matulionyte, Rita, et autres
Publié: (2025)
par: Matulionyte, Rita, et autres
Publié: (2025)
AI incidents and 'networked trouble': The case for a research agenda
par: Shane, Tommy Shaffer
Publié: (2024)
par: Shane, Tommy Shaffer
Publié: (2024)
Safety challenges of AI in medicine in the era of large language models
par: Wang, Xiaoye, et autres
Publié: (2024)
par: Wang, Xiaoye, et autres
Publié: (2024)
Evaluating General-Purpose AI with Psychometrics
par: Wang, Xiting, et autres
Publié: (2023)
par: Wang, Xiting, et autres
Publié: (2023)
Responsible Evaluation of AI for Mental Health
par: Arnaout, Hiba, et autres
Publié: (2026)
par: Arnaout, Hiba, et autres
Publié: (2026)
A Knowledge-Component-Based Methodology for Evaluating AI Assistants
par: Qi, Laryn, et autres
Publié: (2024)
par: Qi, Laryn, et autres
Publié: (2024)
Healthy Distrust in AI systems
par: Paaßen, Benjamin, et autres
Publié: (2025)
par: Paaßen, Benjamin, et autres
Publié: (2025)
From job titles to jawlines: Using context voids to study generative AI systems
par: Memon, Shahan Ali, et autres
Publié: (2025)
par: Memon, Shahan Ali, et autres
Publié: (2025)
AI-AI Bias: large language models favor communications generated by large language models
par: Laurito, Walter, et autres
Publié: (2024)
par: Laurito, Walter, et autres
Publié: (2024)
Can AI expose tax loopholes? Towards a new generation of legal policy assistants
par: Fratrič, Peter, et autres
Publié: (2025)
par: Fratrič, Peter, et autres
Publié: (2025)
Teacher agency in the age of generative AI: towards a framework of hybrid intelligence for learning design
par: Frøsig, Thomas B, et autres
Publié: (2024)
par: Frøsig, Thomas B, et autres
Publié: (2024)
Automating psychological hypothesis generation with AI: when large language models meet causal graph
par: Tong, Song, et autres
Publié: (2024)
par: Tong, Song, et autres
Publié: (2024)
Comprehensive Framework for Evaluating Conversational AI Chatbots
par: Gupta, Shailja, et autres
Publié: (2025)
par: Gupta, Shailja, et autres
Publié: (2025)
Critically Engaged Pragmatism: A Scientific Norm and Social, Pragmatist Epistemology for AI Science Evaluation Tools
par: Lee, Carole J.
Publié: (2026)
par: Lee, Carole J.
Publié: (2026)
System 2 thinking in OpenAI's o1-preview model: Near-perfect performance on a mathematics exam
par: de Winter, Joost, et autres
Publié: (2024)
par: de Winter, Joost, et autres
Publié: (2024)
Can generative AI and ChatGPT outperform humans on cognitive-demanding problem-solving tasks in science?
par: Zhai, Xiaoming, et autres
Publié: (2024)
par: Zhai, Xiaoming, et autres
Publié: (2024)
Sleeper Social Bots: a new generation of AI disinformation bots are already a political threat
par: Doshi, Jaiv, et autres
Publié: (2024)
par: Doshi, Jaiv, et autres
Publié: (2024)
"Draw me a curator" Examining the visual stereotyping of a cultural services profession by generative AI
par: Spennemann, Dirk HR
Publié: (2025)
par: Spennemann, Dirk HR
Publié: (2025)
Reducing research bureaucracy in UK higher education: Can generative AI assist with the internal evaluation of quality?
par: Fletcher, Gordon, et autres
Publié: (2025)
par: Fletcher, Gordon, et autres
Publié: (2025)
The Ghost in the Grammar: Methodological Anthropomorphism in AI Safety Evaluations
par: Costa, Mariana Lins
Publié: (2026)
par: Costa, Mariana Lins
Publié: (2026)
General-purpose AI models can generate actionable knowledge on agroecological crop protection
par: Wyckhuys, Kris A. G.
Publié: (2025)
par: Wyckhuys, Kris A. G.
Publié: (2025)
Quantitative study about the estimated impact of the AI Act
par: Hauer, Marc P., et autres
Publié: (2023)
par: Hauer, Marc P., et autres
Publié: (2023)
Data-Centric AI Governance: Addressing the Limitations of Model-Focused Policies
par: Gupta, Ritwik, et autres
Publié: (2024)
par: Gupta, Ritwik, et autres
Publié: (2024)
Reckoning with the Political Economy of AI: Avoiding Decoys in Pursuit of Accountability
par: Vertesi, Janet, et autres
Publié: (2026)
par: Vertesi, Janet, et autres
Publié: (2026)
Evaluating the Social Impact of Generative AI Systems in Systems and Society
par: Solaiman, Irene, et autres
Publié: (2023)
par: Solaiman, Irene, et autres
Publié: (2023)
A Study on the Framework for Evaluating the Ethics and Trustworthiness of Generative AI
par: Jeong, Cheonsu, et autres
Publié: (2025)
par: Jeong, Cheonsu, et autres
Publié: (2025)
A mathematical theory of evolution for self-designing AIs
par: Harris, Kenneth D
Publié: (2026)
par: Harris, Kenneth D
Publié: (2026)
Exploring a Behavioral Model of "Positive Friction" in Human-AI Interaction
par: Chen, Zeya, et autres
Publié: (2024)
par: Chen, Zeya, et autres
Publié: (2024)
Documents similaires
-
From Checklists to Clusters: A Homeostatic Account of AGI Evaluation
par: Reynolds, Brett
Publié: (2025) -
Computational Hermeneutics: Evaluating generative AI as a cultural technology
par: Kommers, Cody, et autres
Publié: (2026) -
The erasure of intensive livestock farming in text-to-image generative AI
par: Sheng, Kehan, et autres
Publié: (2025) -
A taxonomy of epistemic injustice in the context of AI and the case for generative hermeneutical erasure
par: Mollema, Warmhold Jan Thomas
Publié: (2025) -
Hallucination, reliability, and the role of generative AI in science
par: Rathkopf, Charles
Publié: (2025)